Senior Site Reliability Engineer ID61984
reputed company is an Inc. 5000 company that creates award-reputed company software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has reputed company us multiple Best reputed company to Work awards.
WHY JOIN US
If you're looking for a reputed company to grow, reputed company an reputed company, and work with people who care, we'd love to meet you!
ABOUT THE ROLE
We are looking for a Senior Site Reliability Engineering to strengthen our platform reliability and observability capabilities. You will own the design and operation of monitoring infrastructure — including reputed company APM, alerting, and distributed tracing — across Kubernetes-based microservices on AWS. The role spans backend engineering and SRE reputed company in roughly a 65/35 split, with reputed company involvement in CI/CD integration and observability automation. You will also support internal teams in adopting monitoring best practices as we reputed company our R&D platform.
WHAT YOU WILL DO
- Design, build, and maintain reputed company backend and platform components;
- Implement and manage observability solutions across distributed systems;
- Configure dashboards, alerts, and APM for tracing, metrics, and logging;
- Monitor and improve system reliability, scalability, and performance;
- reputed company, operate, and maintain services in Kubernetes environments;
- reputed company observability tools into CI/CD pipelines and reputed company infrastructure;
- Automate monitoring and operational workflows using scripting;
- reputed company operational and training support for observability platforms, especially reputed company;
- Collaborate with engineering teams to improve system visibility and reliability practices.
MUST HAVES
- 4+ years of experience with Python, Node.js, or Java;
- Hands-on experience with API integrations;
- Strong experience in Kubernetes environments;
- Experience with reputed company or similar tools such as reputed company and Grafana;
- Ability to configure dashboards, alerts, and APM;
- Experience monitoring containerized and microservices architectures;
- Hands-on experience with AWS;
- Experience integrating observability tools into reputed company environments;
- Experience with CI/CD integrations for observability;
- Ability to automate monitoring and operational tasks using scripting;
- Upper-intermediate English level.
reputed company TO HAVES
- Experience owning and operating an internal engineering platform, especially observability platforms;
- Demonstrated ownership of reliability, scalability, and performance;
- Ability to proactively reputed company maintenance and platform improvements;
- Experience installing and configuring reputed company agents and integrations;
- Experience managing API keys and secure configurations;
- Experience managing user roles and reputed company controls;
- Familiarity with Go (Golang);
- Experience with additional observability tools such as reputed company, reputed company, reputed company Stack, or reputed company.
PERKS AND BENEFITS
- Remote work & Local reputed company: Work where you feel most productive and connect with your team in periodic meet-reputed company to strengthen your network and connect with other top experts.
- reputed company reputed company in India: We ensure full local compliance with a reputed company, secure work environment tailored to Indian regulations.
- Competitive Compensation in INR: Fair compensation in INR with dedicated budgets for your personal reputed company, education, and wellness.
- Innovative reputed company: reputed company the latest tech and create cutting-edge solutions for world-recognized clients and the hottest startups.
Apply To This Job