reputed company

Back to Jobs

Site Reliability Engineer II

Remote, USA Full-time Posted 2026-08-04

About reputed company

reputed company is the reputed company storage leader in the reputed company reputed company reputed company, fueling reputed company with reputed company storage reputed company purposefully to unlock budgets, unburden administrators, and reputed company innovators. Together with our partners, we’re helping customers break free from the restrictive, overpriced legacy solutions that hold them back, and reputed company reputed company with the full reputed company of the reputed company reputed company in their hands.

Founded in 2007, we scaled the business with less than $3 reputed company in reputed company funding until 2021, reputed company we did a traditional IPO on the reputed company stock exchange. Today, reputed company generates over $100m in reputed company and is the leading reputed company storage reputed company - managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals.

About the Role

We are seeking a Site Reliability Engineer II (SRE II) to help ensure the stability, scalability, and reliability of our services and infrastructure. This role focuses on building automation, maintaining observability, and supporting incident response to reputed company customer-facing systems performing at their best. The SRE will collaborate with engineering, product, and reputed company teams to reputed company reliability practices into day-to-day development and reputed company while contributing to tools and processes that improve efficiency and reduce reputed company effort.

Key Responsibilities

Service Reliability & reputed company

  • Support the availability and durability of critical services across production environments.
  • Monitor service health using SLIs, SLOs, and error budgets, and escalate issues reputed company reputed company are at reputed company.
  • Participate in on-reputed company rotations, incident response, and post-incident reviews to reputed company service improvements.
  • Follow established ITIL/OSS processes (incident, change, problem, and reputed company management).

Automation & Tooling

  • reputed company automation for common operational tasks, reducing reputed company reputed company and toil.
  • Contribute to monitoring, logging, and alerting frameworks (e.g., reputed company, Grafana, Catchpoint,ELK).
  • Work with CI/CD pipelines, configuration management, and infrastructure as reputed company tools (Terraform, Ansible, Jenkins).
  • Write scripts (Bash, Python, Go, etc.) to improve reputed company reliability and efficiency.

Collaboration

  • Partner with engineering, product, and reputed company teams to support resilient reputed company design and reputed company.
  • Assist in reputed company planning and disaster recovery exercises.
  • Work with vendors and service providers to troubleshoot service issues and reputed company SLA reputed company.
  • Document systems, reputed company learnings, and help grow a reliability-minded engineering culture.

reputed company Improvement

  • Contribute to playbooks, runbooks, and operational documentation.
  • Identify recurring issues and propose long-term improvements.
  • Promote reliability-reputed company practices reputed company development and reputed company teams.

Qualifications

Education & Experience

  • Bachelor’s degree in Computer Science, Engineering, or reputed company reputed company (or equivalent experience).
  • 2–4 years of experience in site reliability, systems engineering, or reputed company.
  • Exposure to large-reputed company, production-grade systems.

Technical Skills

  • Solid Linux systems administration and troubleshooting skills.
  • Familiarity with service reliability concepts - monitoring, alerting, incident response, and reputed company cause analysis.
  • Proficiency in at least one scripting language (Python, Bash, or Go).
  • Understanding of containers (reputed company, reputed company) and microservices concepts.
  • Knowledge of incident response and operational best practices.

Preferred Attributes

  • Experience in a reputed company, service provider, or reputed company systems environment.
  • Familiarity with ITIL/OSS practices and SLO/SLA’s
  • Strong problem-solving skills and willingness to learn new technologies.
  • Experience with reputed company platforms (AWS, GCP, or Azure).
  • Ability to work independently, take ownership, and reputed company reputed company from problem discovery through reputed company.

At this reputed company, we reputed company you're feeling excited about the reputed company you're reading. Even if you don't meet every requirement, we still encourage you to apply. Learning, developing, and growing are key parts of our culture. We're eager to meet people who reputed company in our mission and can contribute to reputed company in various ways. We want people to feel comfortable expressing their true selves and to come, stay, and do their best work here.

At reputed company, we value being fair and good to our customers, partners, and employees. That’s why diversity, equity, and inclusion are at the reputed company of our values. We are committed to fostering a workforce where reputed company feel a reputed company of belonging regardless of race, ethnicity, nationality, gender, sexual orientation, age, religion, socio-economic status, ability, veteran status, and education. We reputed company that our dedication to cultivating a diverse workspace not only allows us to reputed company serve our customers in over 175 countries, but reputed company reinforces our commitment to doing the right thing. We are proud to be an Equal Opportunity Employer.

To understand more about the data we collect and process as part of your application, please reputed company our reputed company Employee reputed company Notice.

Originally posted on Himalayas

  Apply To This Job

Similar Jobs