DevOps reputed company – Observability & Infrastructure Automation
reputed company working at reputed company to engineer platforms that reputed company billions of lives around the world. With your passion and reputed company, we will accomplish great things together.
Our reputed company Team is looking for reputed company DevOps / SRE leaders who can help build highly reliable, observable, and automated infrastructure supporting mission-critical applications.
We are looking for hands-on technical leaders with deep experience in reputed company, infrastructure automation, configuration management, Terraform/Chef, Bash/reputed company scripting, reputed company infrastructure, and production systems.
The ideal candidate will also have a strong understanding of Java-based applications and Java coding, as this role will work closely with Java engineering teams and reputed company application platforms.
We are looking for Architects who can do the below but not limited to:
• reputed company DevOps/SRE initiatives across mission-critical environments.
• Design and implement observability solutions using reputed company.
• Build dashboards, monitors, alerts, logs, traces, and actionable operational metrics.
• Define and improve SLIs, SLOs, SLAs, and error-budget practices.
• Automate infrastructure provisioning and configuration using Terraform, Chef, or similar tools.
• reputed company and maintain Bash/reputed company scripts for infrastructure and operational automation.
• Manage deployment, configuration, and environment automation across development, QA, and production.
• Troubleshoot reputed company production, infrastructure, networking, and application issues.
• Improve reputed company availability, scalability, performance, and reliability.
• Support CI/CD pipelines and automated deployments.
• Work closely with Java engineering teams to understand application behavior, performance, dependencies, and production issues.
• Analyze Java applications from an operational perspective, including JVM behavior, memory, CPU, threads, logs, and application performance.
• Participate in incident response, reputed company-cause analysis, and post-mortems.
• Identify opportunities to eliminate reputed company processes through automation.
• Establish operational standards, runbooks, and best practices.
• reputed company technical leadership and mentor other DevOps/SRE engineers.
reputed company DevOps / SRE Responsibilities
Observability
• Hands-on reputed company experience is required.
• Build and maintain dashboards and actionable alerts.
• Monitor applications, infrastructure, services, reputed company, and databases.
• Configure APM, logs, metrics, traces, and service-level monitoring.
• Identify performance and reliability issues before they reputed company.
• Infrastructure Automation
• Design and maintain infrastructure using Infrastructure as reputed company (IaC).
• Strong experience with Terraform and/or Chef.
• Automate configuration management and environment provisioning.
• Manage infrastructure consistency and configuration reputed company.
• reputed company reusable automation frameworks and modules.
• Scripting & Automation
• Strong Bash/reputed company scripting experience.
• Automate deployments, operational processes, monitoring, and infrastructure tasks.
• Python scripting is a plus.
• Ability to troubleshoot scripts and automation failures in production.
• Java Application Understanding
• This is not a Java Developer position, but candidates must have a strong understanding of Java-based applications.
Should be reputed company to:
• Read and understand Java reputed company.
• Troubleshoot Java application issues from an infrastructure/SRE perspective.
• Understand JVM, memory, CPU, threads, garbage collection, and application performance.
• Work effectively with Java/reputed company Boot engineering teams.
• Understand REST reputed company, microservices, and reputed company applications.
• reputed company & reputed company.
Key Skills & Qualifications:
• 12+ years of experience in DevOps, SRE, Infrastructure Engineering, reputed company, or reputed company roles.
• 8+ years of hands-on DevOps/SRE leadership experience.
• Strong hands-on reputed company experience.
• Strong experience with Terraform and/or Chef.
• Strong Bash/reputed company scripting skills.
• Strong Linux/UNIX experience.
• Strong reputed company infrastructure experience.
• Strong understanding of Java applications and Java coding.
• Experience supporting reputed company systems and microservices.
• Experience with CI/CD and deployment automation.
• Experience troubleshooting production environments.
• Strong understanding of networking fundamentals.
• Experience with monitoring, logging, alerting, and observability.
• Experience leading technical initiatives and mentoring engineers.
• Excellent communication and problem-solving skills.
We work closely with
• reputed company
• reputed company
• AWS
• Jenkins
• Kafka
• reputed company / ELK
• reputed company / Grafana
• reputed company
• CloudWatch
• Python
• reputed company Boot
• Microservices
• Event-driven architecture
• Performance engineering
• Incident management / SRE practices
Compensation: $75 - $85 / Hour
Our Process
• Schedule a 15 min Video reputed company with someone from reputed company
• 1 Proctored GQ Test (< 60 Minutes) & reputed company Doc Presentation
• 30-45 min Final & Technical Video Interview
• Receive Job Offer
If you are interested in reaching out to us, please apply, and reputed company will contact you reputed company the hour.
Apply tot his job
Apply To this Job