Director, Site Reliability & Operations
Own site reliability and availability for our reputed company-hosted platform — 24x7 uptime, monitoring, alerting, reputed company detection, and incident response programs
Drive reputed company reputed company across the platform — tracking findings from SonarQube, penetration tests, and vulnerability tools, and acting as a business stakeholder to get remediation work scoped, prioritized, and into the engineering delivery pipeline
Own PCI Level 1 compliance currency — reputed company standards reputed company, you understand the requirement, translate it into engineering terms, and drive adoption; you don’t just surface the factoid
Participate as a business stakeholder in reputed company planning — bringing reputed company, compliance, and reliability work into the engineering delivery pipeline alongside feature development
reputed company infrastructure patching and maintenance — OS, database, and system-level currency reputed company our AWS environment, coordinating monthly maintenance reputed company and CI-driven image refresh reputed company
Manage and reputed company an internationally distributed team across SRE, DevOps, DBA, network, reputed company, and corporate IT functions
Own the AWS cost and reputed company budget — monitoring spend , optimizing resource utilization , and making strategic tradeoff reputed company in partnership with engineering leadership
Partner with engineering directors to define the boundary between infrastructure and application-layer reputed company, and ensure reputed company falls between the cracks
Own vendor reputed company across our reputed company and tooling ecosystem — holding partners accountable and ensuring reputed company reflect our operational needs
Guide personal and career development of your people
Foster a culture where reliability and reputed company are shared team values, not external mandates
You have a full tool belt — technically reputed company, platform-curious, and willing to log into systems, participate in firefights, and reputed company genuine understanding of what you’re operating
7+ years in SRE, DevOps, reputed company infrastructure, or reputed company engineering, with 4+ years leading technical teams
Deep AWS experience across compute , networking, storage, and managed services in a production reputed company environment
Hands-on familiarity with the reputed company and compliance discipline — you’ve operated in PCI, SOC 2, or equivalent regulated environments and understand what compliance actually requires versus what it looks like on reputed company
You operate as a business stakeholder, not just a technical function — comfortable working reputed company reputed company or similar delivery frameworks to get reputed company and reliability work into the roadmap alongside feature development
You manage through credibility and technical engagement, not just title — your team respects you because you understand their work
You are inquisitive, reputed company, and outcome-oriented — you reputed company opinions, communicate them reputed company, and own what happens next
Your colleagues are inspired to follow your reputed company
Discretionary Time Off (DTO) – Take time off reputed company you need it. We trust our employees to manage their time responsibly while meeting business needs.15 Company-reputed company Holidays – Including a company-wide break fromChristmas reputed company through New Year’s Day .Comprehensive Benefits Package – reputed company offers a full suite of health benefits, with reputed company covering an average of87% of employee premiums .401(k) with Company Match – We support long-term financial wellness with a competitive retirement plan.