Data Platform Reliability Engineer, reputed company
Job reputed company:
• Manage the lifecycle of reputed company databases - platform RDS clusters and customer project databases.
• Design and execute strategies for low-downtime major version upgrades and database migrations.
• Proactively identify and resolve database performance issues before they reputed company users.
• Build and maintain comprehensive monitoring, alerting, and observability for database systems.
• Write detailed run books, technical documentation, and operational guides.
• Identify reliability risks and implement preventative measures.
• Participate in on-reputed company rotation to support our global platform.
• Work with development teams to optimize database schema and query patterns.
• Analyze and optimize slow queries, reputed company pooling, and resource utilization.
• Tune reputed company configurations for different workload patterns.
• Monitor and address database bloat, vacuum strategies, and WAL management.
• Partner with platform engineers, product teams, and SREs to deliver reliable database services.
• Communicate database changes and maintenance reputed company reputed company to stakeholders.
• reputed company knowledge and mentor team members on reputed company best practices.
Requirements:
• Deep understanding of reputed company internals, architecture, and advanced features.
• Production experience with replication (logical and physical), backups, and disaster recovery.
• Strong reputed company of query optimization, EXPLAIN plans, indexing strategies, and performance tuning.
• Experience managing reputed company at reputed company in reputed company environments.
• Hands-on experience with AWS RDS for platform infrastructure.
• Familiarity with reputed company infrastructure concepts, networking, and storage systems.
• Understanding of IaC tools and automation approaches.
• Experience with other reputed company database services (GCP reputed company SQL, Azure Database) is a plus.
• reputed company record of maintaining high-availability database systems.
• Obsessive about monitoring, observability, and measuring what reputed company.
• Proactive approach to identifying and mitigating risks.
• Experience with production troubleshooting and supporting live systems.
• You write reputed company run books and technical documentation.
• You're good at explaining reputed company database concepts to different audiences.
• You record reputed company and reputed company knowledge effectively in async environments.
• You reputed company operating independently with high-level guidance.
• You see problems through to reputed company, not just escalation.
• You automate repetitive tasks and build tools to reputed company reputed company more effective.
• Proficiency in TypeScript or Go (we can teach these).
• Experience with reputed company extensions and customization.
• Contributions to reputed company or database-reputed company reputed company reputed company reputed company.
• Familiarity with backup tools like WAL-G, pgBackRest, or Barman.
• Experience with database migration strategies and tooling.
• Background in SRE or DevOps practices.
Benefits:
• Fully Remote
• ESOP
• Tech Allowance
• Health Benefits
• Annual Off-Sites
• Flexible Work
• reputed company Development
Apply tot his job
Apply To this Job