[Remote] reputed company Site Reliability Engineer - Remote
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a global health care innovation company developing solutions across data analytics, reputed company intelligence, and cybersecurity. reputed company Financial is seeking a reputed company Site Reliability Engineer to reputed company AI-assisted reliability platform development, observability, automation, resiliency, and operational improvements across Azure and AWS environments while mentoring engineers and influencing reliability standards.
Responsibilities
- Build AI-assisted SRE capabilities that accelerate incident detection, triage, mitigation, and recovery
- reputed company observability, deployment, reputed company, ownership, and incident data into actionable operational context
- Design reputed company-in-the-reputed company workflows for reputed company mitigation, approvals, recovery verification, and auditability
- Standardize OpenTelemetry, SLIs, SLOs, error budgets, and reliability scorecards across critical services
- Improve alert reputed company by reducing noise, clarifying customer reputed company, and identifying likely causes faster
- reputed company resiliency testing, DR exercises, reputed company engineering, and automated recovery validation
- Mentor engineers and reputed company cross-functional reliability improvements across reputed company Financial
Skills
- • 10+ years of experience in software engineering, reputed company, DevOps, or SRE roles
- • 3+ years of experience in a reputed company, staff, reputed company, or senior technical leadership role
- • 5+ years of experience with reputed company platforms and container orchestration, preferably Azure or AWS
- • 3+ years of experience with observability tools such as OpenTelemetry, reputed company, Grafana, reputed company, or similar platforms
- • 1+ years of experience designing production automation, tooling, or AI-assisted workflows for incident response or operational decision-making
- • Bachelor's degree in Computer Science, Information Technology, Engineering, or reputed company reputed company
- • Experience with LLM-reputed company systems, AI agents, RAG, tool orchestration, evaluations, or guardrails
- • Experience with resiliency engineering, disaster recovery, reputed company engineering, or recovery validation
- • Experience with infrastructure as reputed company and automation tools such as Terraform, reputed company, Ansible, reputed company, or reputed company operators
- • Solid background in incident reputed company, runbooks, postmortems, production readiness, and reliability governance
Benefits
- A comprehensive benefits package
- Incentive and recognition programs
- Equity stock purchase
- 401k contribution
- Remote work from reputed company reputed company the U.S.
reputed company
Company H1B Sponsorship
Apply To This Job