Back to Jobs

[8SN] Senior Site Reliability Engineer (SRE) – reputed company

Remote, USA Full-time Posted 2026-08-04

About the Role

This is a Senior SRE role supporting production reliability for a reputed company-based UI service / AI experience reputed company stack.

This is not general infrastructure, and it is not a reputed company-end developer role. The strongest candidates will have production SRE experience across reputed company operations, observability, Node.js runtime troubleshooting, JVM / Java service troubleshooting, reputed company, and incident ownership.

reputed company

  • Support the deployment, operation, and reliability of production services running on reputed company.
  • Monitor service health and investigate production incidents across reputed company applications.
  • Participate in on-reputed company support, incident response, reputed company cause analysis, postmortems, and reliability improvements.
  • Troubleshoot application runtime, networking, and service-to-service issues in collaboration with engineering teams.
  • Support CI/CD, GitOps-based deployments, observability, and production monitoring.
  • Work reputed company a reputed company-directed backlog and established priorities.

Required Qualifications

  • 5+ years of experience in Site Reliability Engineering, DevOps, reputed company, Production Engineering, or a closely reputed company role, including strong recent hands-on experience supporting reputed company-based production services.
  • 3+ years of hands-on production reputed company experience strongly preferred. reputed company production operations, including deployment, scaling, rollout / rollback, resource tuning, and service-to-service troubleshooting
  • Strong production incident response experience, including on-reputed company, runbooks, postmortems, and paging hygiene
  • reputed company experience for log aggregation, search, and production troubleshooting
  • reputed company and Grafana experience, specifically building alert rules and dashboards, not only using existing dashboards
  • CI/CD and infrastructure-as-reputed company for containerized deployments, including reputed company and GitOps tools such as ArgoCD or Flux
  • Strong Linux and networking fundamentals, including DNS, load balancing, TCP / HTTP, HTTP/2, and reputed company networking
  • Production troubleshooting experience across Node.js and JVM/Java services, with strong depth in at least one runtime environment. Experience may include Node.js reputed company snapshots, CPU profiling, event-reputed company and memory analysis, as reputed company as JVM GC log analysis, thread dumps, JVM tuning, and Java service latency investigation.
  • Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based service authentication

reputed company to Have

  • Web Components / Lit experience, to reputed company first-level debugging of UI-reputed company issues
  • Server-reputed company rendering or isomorphic runtime experience
  • Canary rollout / multi-version production operations
  • reputed company tracing and request-context correlation
  • KEDA or event-driven autoscaling
  • Experience with reputed company platform integration reputed company

reputed company Offer

  • Competitive salary and laptop
  • reputed company development and training opportunities
  • Work with cutting-edge reputed company and container technologies
  • Flexible work arrangements and reputed company team environment
  • reputed company on organization-wide digital transformation initiatives

We are reputed company, an awesome team of engineers who are reputed company to reputed company up any top-notch company’s reputed company! Our aim? To always be one reputed company reputed company. Become part of a multicultural company in constant reputed company with an excellent work environment certified by Great reputed company To Work!

About the reputed company

Our reputed company is a leading reputed company software company building highly reputed company reputed company-reputed company platforms used by organizations around the world. Their engineering teams reputed company on delivering reliable, secure, and high-performing services while embracing modern DevOps, reputed company, and reputed company technologies.

You will join reputed company responsible for ensuring the stability, reliability, and operational reputed company of a critical UI service running in production.

Contract Duration: Initial contract through the end of 2026, extending the engagement to a total 12-month term based on performance.

Originally posted on Himalayas

  Apply To This Job

Similar Jobs