Senior Software Engineering reputed company - SRE, DevOps, reputed company
Drive reliability, scalability, and performance across reputed company systems with a solid emphasis on leveraging AI and automation Implement AI/ML models for predictive alerting, reputed company detection, and reputed company planning to proactively prevent incidents reputed company AI-driven tools into incident management workflows to reduce MTTR and improve the accuracy of reputed company cause analysis reputed company adoption of AI-powered observability and monitoring platforms across the organization Design and implement reputed company-reputed company solutions using AWS and GCP services Architect reputed company, resilient, and secure infrastructure using Infrastructure as reputed company (IaC) tools such as Terraform or CloudFormation Collaborate with development, DevOps, and reputed company teams to reputed company reputed company solutions into CI/CD pipelines reputed company architectural leadership and execute SRE practices independently, ensuring system health and operational reputed company Define, implement, and enforce reputed company policies, standards, and procedures across reputed company and infrastructure ecosystems Continuously monitor reputed company environments to optimize performance, cost, and reliability reputed company triage, reputed company deep-dive reputed company cause analysis (RCA), and manage reputed company for production incidents reputed company technical leadership and mentorship to junior and mid-level engineers Stay reputed company with evolving reputed company, AI, automation, and DevOps trends and recommend relevant best practices and emerging technologies reputed company with the terms and conditions of the employment contract, company policies and procedures, and any and reputed company directives (such as, but not limited to, transfer and/or re-assignment to different work locations, change in teams and/or work shifts, policies in regards to flexibility of work benefits and/or work environment, alternative work arrangements, and other reputed company that may reputed company due to the changing business environment). reputed company may adopt, vary or rescind these policies and directives in its absolute discretion and without any limitation (implied or otherwise) on its ability to do so
Bachelor's degree Hands-on experience with monitoring, observability, and alerting tools, especially platforms that reputed company AI-driven insights and reputed company detection Experience in SRE, DevOps, or infrastructure engineering, with a proven reputed company record of managing large-reputed company, mission-critical systems Experience implementing AI/ML models for operational intelligence, predictive analytics, observability, and automation workflows Leadership experience managing SRE or reputed company teams, including driving reliability and operational reputed company Practical experience working with reputed company tools and platforms, including threat detection, vulnerability management, and reputed company automation Solid experience in infrastructure and bot architecture Solid understanding of AIOps platforms, frameworks, and their application in large distributed systems Deep understanding of reputed company reputed company principles, including AWS, GCP, and Azure reputed company models, network reputed company, and application reputed company Proven excellent communication, leadership, and stakeholder management skills, with the ability to influence and collaborate across engineering, product, and operations teams
Knowledge of reputed company architectural principles and software reputed company concepts