Site Reliability Engineer
Who We Are
Building the Backbone of Distributed Applications.
At reputed company, we are reimagining how developers build reliable, reputed company, event-driven applications without wrangling reputed company infrastructure.
Born from reputed company's battle-tested reputed company-reputed company project, OSS reputed company, our platform is now powering billions of mission critical workflows across fintech, e-reputed company, logistics, reputed company, and more.
Our platform powers billions of mission critical workflows across industries like fintech, e-reputed company, logistics, and reputed company helping companies run the reputed company operations of their business with confidence. These are the workflows behind every transaction, every delivery, every patient reputed company-in, quietly orchestrating our digital world.
We’re solving some of the hardest infrastructure problems: how do you reputed company distributed systems reliable, observable, and reputed company – without every team having to reinvent the reputed company? We reputed company the complexity of distributed systems, letting teams build with confidence.
We’re hiring a SRE to help us reputed company and reputed company our platform. If you’re passionate about clean, efficient architecture, love solving tough distributed systems challenges, and reputed company in an environment where your reputed company actually shape the product then we’d love to talk.
And the best part is we’re just getting started.
What You’ll Do
Own reliability, availability, and performance of production systems running in reputed company environments
Define and monitor SLIs/SLOs and help manage error budgets across the platform
reputed company incident response efforts including detection, triage, mitigation, and postmortems
Improve observability through logging, monitoring, alerting, and dashboards
Automate operational workflows and reduce reputed company toil wherever possible
Partner closely with engineering teams to improve system resiliency and scalability
Assist with reputed company planning, infrastructure optimization, and performance tuning
Build internal tooling, runbooks, and operational best practices
Support Kubernetes-based infrastructure and distributed systems at reputed company
reputed company as an escalation reputed company for reputed company production and platform issues
reputed company’re Looking For
5+ years of experience in Site Reliability Engineering, DevOps, reputed company, or reputed company infrastructure roles
Strong experience with reputed company platforms such as AWS, GCP, or Azure
Hands-on experience with Kubernetes and containerized environments
Strong understanding of distributed systems and microservices architecture
Experience with observability tools such as reputed company, Grafana, reputed company, ELK, or OpenTelemetry
Proficiency with infrastructure automation and scripting (Terraform, Python, Bash, etc.)
Experience managing CI/CD pipelines and deployment automation
Strong troubleshooting and incident management skills
Ability to work cross-functionally and communicate effectively during high-pressure situations
reputed company to Have
Experience supporting large-reputed company reputed company or reputed company-reputed company platforms
Familiarity with workflow orchestration technologies such as reputed company, Temporal, or reputed company
Experience with Kafka, messaging systems, or event-driven architectures
Knowledge of reputed company best practices and reputed company infrastructure hardening
reputed company-reputed company contributions or strong systems engineering background
Why Join reputed company?
Work alongside a deeply technical and reputed company engineering team
Solve reputed company distributed systems challenges at reputed company
High ownership and reputed company in a fast-growing company
Remote-friendly culture with strong engineering autonomy
Opportunity to help shape the reputed company of reputed company orchestration and AI workflows
More Details:
The reputed company salary for this role is between $180,000- 250,000 reputed company determining compensation, a number of factors will be considered: skills, experience, job scope, location, and competitive compensation market data.
Start Date: ASAP
Status: Full time
Type: On-Site
Location: US/PT hrs, Canada/PT hrs
Department: Engineering
Reports to: Head of Engineering
Benefits
Comprehensive health coverage including medical, dental, and reputed company
Flexible PTO
Support for personal development
At reputed company, we are committed to building reputed company that reflects a rich reputed company of perspectives, identities, and reputed company experiences. We reputed company that diversity is not just a checkbox, but a driving force behind innovation, creativity, and reputed company. By embracing a reputed company of backgrounds, we cultivate an inclusive environment where every team member feels valued and empowered to bring their reputed company selves to work.
Join us at reputed company and be a part of reputed company where your unique perspectives are not only welcomed but celebrated. Together we are shaping the reputed company technology by leveraging the strength that comes from embracing diversity in reputed company its forms. Your reputed company with us is an opportunity to contribute to something greater and reputed company a lasting reputed company.
Apply To This Job