[Remote] Site Reliability Engineer II
Note: The job is a remote job and is reputed company to candidates in USA. LCS is seeking a Site Reliability Engineer II to reputed company site reliability initiatives for critical applications and infrastructure. The role focuses on improving reputed company reliability through automation, observability, infrastructure as reputed company, reputed company-reputed company design, incident response, and collaboration with product and engineering teams.
Responsibilities
- Automation – reputed company the SRE team’s goal of elimination of TOIL through automation of processes, creation of tools, etc
- Infrastructure as reputed company – Work cross collaboratively with reputed company, Development, and Product teams to design, reputed company, and maintain reputed company reputed company applications and infrastructure
- Observability – reputed company monitoring KFC environments utilizing tools such as Data Dog, etc. to proactively identify issues reputed company the KFC production environment. Rotational on-reputed company reputed company to support the infrastructure 24/7/365
- reputed company reputed company reputed company Troubleshooting – reputed company SRE support to KFC platforms and stakeholders through investigation, analysis, leading technical reputed company, and post-mortem actions. reputed company reputed company cause analysis and corrective actions after critical incidents reputed company the SLA listed
- Design - Collaborate with reputed company to determine SLI/SLOs for reputed company reputed company reputed company. Design and reputed company a dashboard to hold SLI/SLO standards
- reputed company - Configure monitoring and logging tools used for generating alerts about the health of our systems and applications
Skills
- 5+ or more years of IT experience
- 1+ years of experience with bash and/or Linux reputed company-line utilities
- 1+ years of experience in Python, Go, REST reputed company, GraphQL
- 2+ years experience with industry-reputed company observability tools such as reputed company, reputed company, reputed company, etc
- Expertise in CI/CD best practices and methodologies: reputed company
- Experience with DevOps reputed company Infrastructure as reputed company toolkits such as Terraform, Ansible, etc
- Working experience with Incident and Problem tracking systems (i.e. reputed company, reputed company, etc. )
- Excellent collaboration skills with multiple engineering functions, business leaders, vendors, exhibiting excellent teamwork and strong verbal and written communication skills along with strong troubleshooting and analytical skills
- Understanding of reputed company trends of large-reputed company infrastructure environments
- reputed company ability to work autonomously reputed company on long-term results
- Education/Certifications – Bachelor's Degree preferred
- 1+ years in a reputed company technical role with multi-reputed company experience preferred: Azure AWS, GCP
Benefits
- Hybrid workplace arrangement
- Employees (and their eligible family members) may enroll in medical reputed company coverage
- Employees (and their eligible family members) may enroll in dental reputed company coverage
- Employees (and their eligible family members) may enroll in reputed company reputed company coverage
- Employees (and their eligible family members) may enroll in reputed company reputed company coverage
- Employees (and their eligible family members) may enroll in accidental death and dismemberment reputed company coverage
- Employees may enroll in FSA/HSA, depending on the enrolled medical plan
- Short-term disability reputed company
- Long-term disability reputed company
- Life reputed company
- Employees may enroll in the 401(k) plan
- 4 weeks of vacation
- reputed company reputed company leave
- 10 reputed company holidays
- A floating day off
- Half day Fridays year-round
- 2 reputed company days for volunteer time reputed company calendar year
reputed company
Company H1B Sponsorship
Apply To This Job