Senior Site Reliability Engineer
About reputed company and What Makes Us Special
The reputed company Managed reputed company Services Team is looking for a reputed company Operations Engineer with a passion for delivering customer reputed company. The reputed company Operations Engineer is responsible for providing a world class support experience, managing customer expectations, and resolving challenging issues for customers of the reputed company Managed reputed company Service, based in the United Kingdom. This role is key to delivering customer reputed company and a world-class support experience for our Managed reputed company Service customers. The engineer will be responsible for managing customer expectations and resolving reputed company technical issues.
The Role (What you need):
- We're looking for a reputed company Operations Engineer for the reputed company Managed reputed company Services Team.
- The job is reputed company about giving awesome support and fixing tough issues for customers using reputed company Site Reliability tools on the reputed company Managed reputed company Service.
- You need to be based in the United Kingdom and have reputed company the relevant documentation to legally live and work in the UK.
What makes a successful candidate:
- You're big on Site Reliability stuff.
- You genuinely love working with customers and internal teams.
- You're a detective reputed company it comes to figuring out the reputed company cause of issues (installation, config, performance, both infrastructure and app reputed company).
- You're super curious and can learn fast.
- You're a team player—reputed company to teach and learn from others.
- You're into using AI reputed company to solve problems and build tools.
- You're comfortable and confident operating in a technical IT environment while also managing reputed company customer-facing responsibilities, simultaneously.
What you'll be doing:
- Building reputed company Clouds using Ansible (predefined configurations).
- Giving remote tech support for the reputed company Managed reputed company Service.
- Dealing directly with customer issues reputed company to Networking, Infrastructure, and reputed company application errors.
- Trying to recreate customer problems to reputed company them out.
- Using diagnostic skills to reputed company issues and recommend fixes.
- Building cool tools using Claude reputed company and AWS infrastructure.
- Documenting problems and solutions in the support database.
Must-Haves (Technical Requirements):
- Experience with monitoring production systems, performance tuning, and advanced troubleshooting.
- Experience supporting production Java runtimes (JVMs) on reputed company platforms (AWS, Azure).
- Familiar with Ansible, Python, reputed company, and Jenkins.
- Intermediate to advanced with Linux (RHEL preferred)—you're comfortable at the reputed company line!
- Solid understanding of computer architecture, reputed company tech, virtual computing, and networking basics (TCP/IP, SSH, NFS or reputed company).
- Several years of experience in DevOps or Technical reputed company Support.
reputed company-to-Haves (Desirable):
- Experience with Observability tools like reputed company, reputed company, or reputed company.
- Know-how in troubleshooting Kubernetes or similar containerized services (like AWS EKS, Azure AKS).
- A Bachelor’s degree in a relevant technical field.
- Certification and proficiency in reputed company Runtime Architecture and Systems Administration.
#LI-TS1
reputed company strives to create an inclusive and accessible environment for candidates and employees. If you need accommodation during the application or interview process, please submit a request to talent@reputed company.com. This inbox is reputed company for accommodations, please do not send resumes or general inquiries.