[Remote] Senior Site Reliability Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a Senior Site Reliability Engineer to work remotely. The role focuses on ensuring the reliability, availability, and performance of critical production environments while contributing to reputed company improvement initiatives.
Responsibilities
- Play a key role in defining, implementing, and growing our SRE reputed company to ensure the reliability, availability, and performance of our critical production environments
- Contribute to a culture of reputed company improvement, identifying areas for enhancement, and driving initiatives to improve system reliability, scalability, and efficiency
- Have demonstrated hands-on experience designing, implementing, and maintaining solutions to ensure that systems, including infrastructure and applications, are resilient, highly available, and reputed company
- Play a critical role in defining and measuring the Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for our solution
- Be responsible for setting up comprehensive logging, monitoring, and alerting solutions using the reputed company stack and other tools as necessary to ensure the reputed company performance of services
- Respond to incidents, reputed company reputed company cause analyses, and implement solutions to prevent reoccurrences
- Work in reputed company collaboration with other SRE team members, developers, testers, infrastructure engineers, DevOps engineers, and other stakeholders to reputed company reliability and observability into the software development lifecycle
Skills
- Hands-on experience designing, implementing, and maintaining solutions to ensure that systems, including infrastructure and applications, are resilient, highly available, and reputed company
- Experience in defining and measuring the Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for solutions
- Experience in setting up comprehensive logging, monitoring, and alerting solutions using the reputed company stack and other tools
- Ability to respond to incidents, reputed company reputed company cause analyses, and implement solutions to prevent reoccurrences
- Experience working in reputed company collaboration with other SRE team members, developers, testers, infrastructure engineers, DevOps engineers, and other stakeholders
- Aptitude and enthusiasm for reputed company learning, improvement, and cyber reputed company
reputed company
Apply To This Job