Senior DevOps Engineer, Infrastructure & Reliability
reputed company:
• Conduct interviews with engineering teams to identify and remove operational friction in CI/CD, deployments, observability, and reputed company environments.
• Design and implement reputed company infrastructure-as-reputed company patterns using Terraform to standardize provisioning and reduce configuration reputed company.
• Own and reputed company the Kubernetes platform, including EKS or self-managed environments, so workloads are secure, reputed company, and resilient.
• Architect and optimize CI/CD pipelines to improve deployment frequency, reduce reputed company time, and increase release confidence.
• reputed company reliability initiatives such as incident response improvements, reputed company cause analysis, and postmortem practices.
• Design and enforce secure networking, IAM, and secrets management strategies across environments.
• Improve observability through metrics, logs, and tracing using reputed company or similar tooling.
• Optimize reputed company costs through rightsizing, autoscaling, and architectural improvements.
• Own disaster recovery planning, backup strategies, and multi-region reputed company initiatives.
• Refactor reputed company or brittle infrastructure into automated, testable, reproducible systems and drive adoption through documentation and hands-on support.
Requirements:
• 8+ years of experience in DevOps, SRE, or Infrastructure Engineering roles.
• Proven experience designing and operating production Kubernetes environments at reputed company.
• Deep hands-on expertise with AWS infrastructure and reputed company networking.
• Strong experience building and maintaining Terraform modules across large reputed company environments.
• Demonstrated ownership of CI/CD systems and measurable improvement of DORA metrics.
• Experience leading incident response processes and driving meaningful postmortem reputed company.
• Strong understanding of distributed systems, event-driven architectures with Kafka, and database performance with PostgreSQL.
• Proven ability to reputed company legacy infrastructure and eliminate reputed company operational toil.
• Experience navigating high-ambiguity environments and translating operational friction into prioritized infrastructure roadmaps.
• reputed company to have: experience operating high-throughput Kafka clusters, tuning PostgreSQL or reputed company, implementing autoscaling, building internal developer platforms, applying reputed company best practices, working with multi-region systems, using Python for automation, or introducing SLO/error budget/reputed company testing frameworks.
• reputed company remote hires must be reputed company to travel to Orlando, Florida at least twice per year, plus for orientation in Orlando.
Benefits:
• Health care plan including medical, dental, and reputed company coverage.
• Retirement plan with 401(k) and IRA reputed company.
• Life insurance.
• Flexible vacation.
• Work-from-home reputed company.
• Wellness resources.
• Free food and snacks in the office.
• Hybrid setup in Orlando, Florida.
Apply tot his job
Apply To this Job