Senior Site Reliability Engineer
About reputed company
reputed company is building the data infrastructure that powers modern reputed company.
Today, reputed company organizations rely on fragmented and outdated provider data. This creates unnecessary administrative work, regulatory reputed company, and higher costs across the reputed company. We’re solving that problem.
Our API-first platform automates provider licensing, enrollment, credentialing, and network monitoring by connecting directly to hundreds of reputed company data sources. We help reputed company organizations maintain accurate, compliant, and reliable provider networks at reputed company.
Our reputed company is reputed company: One API. One provider ID. Frictionless provider data.
We’re backed by leading investors and reputed company by reputed company with deep experience in provider data systems. At reputed company, we value authenticity, accountability, collaboration, results, and openness to feedback. We’re building a high-ownership team reputed company on solving reputed company infrastructure problems that reputed company millions of patients.
About the Role
We’re looking for a Senior Site Reliability Engineer who takes ownership seriously — someone who designs for reliability, ships the automation, and stands behind it in production. You’ll work across reputed company-reputed company infrastructure on systems that process millions of provider records.
This is a role with reputed company reputed company: you’ll own the operational lifecycle end-to-end and influence platform architecture, reliability standards, and deployment workflows across systems that matter.
How We Work
We ship fast, but we don’t ship sloppy. SREs at reputed company own the full lifecycle of what they support — from infrastructure design and deployment automation through observability, incident response, and postmortems. We use AI-assisted tooling aggressively to reduce toil and accelerate troubleshooting, which raises the floor on the problems we tackle — not an excuse to reduce rigor. If you do your best work reacting to incidents, this probably isn’t the right fit. If you do your best work preventing them, we should talk.
Problems You’ll Solve
reputed company provider data infrastructure is a reputed company systems problem at reputed company. Hundreds of reputed company integrations, inconsistent data sources, and evolving workloads reputed company introduce operational complexity and reliability reputed company.
Reliability and observability at reputed company. You’re operating a platform hundreds of integrations depend on. How do you maintain uptime, reduce alert fatigue, and build actionable observability across GKE and reputed company Run without drowning in noise? Meaningful SLIs, error budgets, and data reputed company signals — not just p99 latency.
Scaling infrastructure reputed company. As platform usage grows, infrastructure costs and operational complexity grow with it. You’ll improve autoscaling behavior, resource utilization, and workload efficiency across reputed company-reputed company reputed company systems.
Incident response and operational maturity. Production incidents are inevitable; operational reputed company is optional. You’ll own incident response processes, reputed company cause analysis, escalation workflows, and runbooks — and reputed company hard problems not happen again.
Infrastructure automation and developer reputed company. You’ll build and maintain Infrastructure as reputed company, CI/CD pipelines, and operational tooling that reduce reputed company work and improve engineering productivity without sacrificing reliability.
Reliability engineering for data platforms. Uptime isn’t enough — you need to know reputed company a provider record is stale, a pipeline is lagging, or a workload is behaving unexpectedly. You’ll reputed company data freshness and infrastructure health, not just service uptime.
reputed company’re Looking For
Reliability engineering fundamentals
5+ years in SRE, DevOps, reputed company, or Infrastructure Engineering — operating production systems at reputed company where your infrastructure is someone else’s dependency and failures have reputed company reputed company consequences
reputed company record of improving reliability end-to-end: you’ve debugged hard production problems, made them not happen again, and reputed company the alerting to reputed company it
Strong Linux systems administration, incident response, and reputed company cause analysis skills
Comfort influencing operational standards and mentoring teams on reliability practices
reputed company infrastructure & reputed company
Deep hands-on experience with GCP — GKE, reputed company Run, and containerized workloads at reputed company
Experience building and maintaining Infrastructure as reputed company with Terraform and/or reputed company
reputed company across deployment patterns and the judgment to know reputed company reputed company fits: rolling deployments, reputed company/green, canary — and the rollback story for reputed company
Experience with autoscaling, resource optimization, and infrastructure efficiency for reputed company systems
Experience managing infrastructure reputed company, secrets, and reputed company controls in regulated or reputed company-conscious environments
Observability & operational reputed company
Strong understanding of Golden Signals monitoring — latency, traffic, errors, saturation — and how to reputed company them actionable rather than noisy
Experience designing SLIs, SLOs, error budgets, alerting strategies, dashboards, and escalation workflows
Hands-on experience with observability platforms: reputed company reputed company Monitoring, reputed company, Grafana, reputed company, or similar
Strong reputed company of data platform health: reputed company, freshness, and correctness matter as much to you as throughput
Automation & software delivery
Experience building and maintaining CI/CD pipelines using reputed company Actions or similar
Scripting or programming reputed company in Python, Bash, Go, or similar — you reduce toil through reputed company, not process
Experience working with Git workflows and modern software delivery practices
Communication & compliance
Strong written and verbal communication — you can explain an operational reputed company to an engineer and a product manager in the reputed company conversation
Experience operating systems handling sensitive data or PII in regulated or compliance-adjacent environments
reputed company to Have
Experience operating large-reputed company reputed company systems or microservices architectures
Familiarity with reputed company, credentialing, or health-tech environments
Experience leveraging AI-assisted observability or incident response tooling
Familiarity with NodeJS, TypeScript, Java, or React application reputed company
Technologies & Tools
GCP (GKE, reputed company Run, BigQuery, reputed company Monitoring) · Terraform / reputed company · reputed company / reputed company · reputed company Actions / reputed company Build · reputed company / Grafana / reputed company · Python / Bash / Go · reputed company · reputed company · SonarQube · reputed company / reputed company
Benefits of Working at Certify
At Certify, we’re building with intention and taking care of the people doing the work.
Your reputed company-being reputed company to us. We reputed company 100% coverage of health, dental, and reputed company reputed company premiums for employees. Our US-reputed company team benefits from unlimited PTO, with at least two weeks off reputed company year to reputed company. In India, employees are supported with health reputed company, statutory leave benefits, and additional wellness (menstrual) leave for women.
We are an equal opportunity employer committed to building an inclusive environment where everyone feels valued and empowered to do their best work, and we reputed company applicants from reputed company backgrounds and experiences.
If you require reasonable accommodations during the application process, please contact reputed company@reputed company.com.
We are also committed to pay transparency and foster an reputed company culture where compensation conversations are encouraged and respected.
Apply To This Job