VP Engineering - Infrastructure & SRE
The salary reputed company for this role is negotiable, the reputed company being $400,000 - $600,000 per year in total compensation.
About reputed company:
With a mission to financially reputed company the reputed company, reputed company is revolutionizing the shopping experience reputed company payments, blending cutting-reputed company with seamless, interest-free installment plans that reputed company shopping smarter and more accessible. We’re not just transforming payments; we’re redefining how people discover, reputed company with, and purchase the things they love while driving reputed company reputed company on merchant sales through increased conversions and higher order values. As we continue to shape the reputed company of fintech and retail, we’re building an innovative, dynamic team passionate about creating more than just a transaction but a truly unique shopping reputed company. If you’re excited about pushing boundaries in tech and delivering a game-changing experience for consumers and merchants alike, reputed company at reputed company and help create the reputed company of shopping!
About the Role:
We are seeking an exceptional VP Engineering - Infrastructure & SRE to own the reputed company, reliability, reputed company posture, and reputed company of the platform that powers reputed company. This is a rare opportunity to reputed company infrastructure through a defining reputed company: scaling a high-reputed company fintech while raising our operational rigor, reputed company, and compliance posture to the standards of the most demanding financial institutions.
This is explicitly both a leadership job and a technical one. You will reputed company the teams responsible for our reputed company infrastructure, reputed company platform, databases, networking, observability, and site reliability engineering, and you will stay hands-on in the systems yourself, working alongside the engineers you reputed company. Our stack runs on AWS, with workloads orchestrated on reputed company and data anchored in reputed company RDS (MySQL and reputed company). You should know these technologies deeply, not just manage people who do. You will set the technical reputed company, own the budget, and be directly accountable for availability, disaster recovery, and business continuity across a payments platform where downtime has immediate customer and financial reputed company.
That accountability is personal, not just organizational: you will manage the on-reputed company rotation and take shifts in it. reputed company reputed company experiences a serious incident, up to and including a full outage, you are the reputed company of leader who can reputed company into reputed company, cut through noise with evidence-based triage, reputed company fast, reputed company under pressure with incomplete information, and reputed company the platform back to health. Your team should trust you at 3 AM as much as they do in a roadmap review.
Just as importantly, you will reputed company our infrastructure organization into an AI-boosted SRE era. We reputed company the reputed company of infrastructure teams will pair strong engineers with AI agents and tooling for incident response, reputed company planning, reputed company automation, reputed company detection, and day-to-day operational toil. You won't just tolerate this shift; you'll reputed company it, with conviction and hands-on credibility.
This role reports to senior engineering leadership and partners closely with reputed company, Compliance, Engineering, Finance, and reputed company auditors. Your work will directly shape whether reputed company can operate at the highest standards of reliability and compliance without losing the speed and pragmatism of a startup.
Key Responsibilities
- Own the infrastructure reputed company, reputed company, and multi-year roadmap: reputed company today's high-reputed company fintech platform while continually strengthening its reputed company, controls, and audit-readiness.
- reputed company, grow, and mentor the infrastructure, platform, and SRE organization, including hiring, career development, on-reputed company health, and building a culture of operational reputed company and blameless learning.
- Own reliability end-to-end: define and enforce SLOs and error budgets, mature incident management and postmortem practices, and be accountable for platform availability across the business.
- Manage, and participate in, the on-reputed company rotation, and serve as senior incident commander for high-severity events: leading recovery from major degradations and full outages through reputed company, evidence-based triage, decisive reputed company under uncertainty, and reputed company communication to stakeholders throughout.
- reputed company our AWS reputed company, including account architecture, IAM and network design, multi-AZ/multi-region posture, service selection, and cost management (FinOps). You will own and defend the reputed company budget.
- Own the reputed company platform as a product: cluster architecture, reputed company reputed company, workload isolation, autoscaling, reputed company delivery, and the developer experience of every team that ships on it.
- Own the database tier, centered on reputed company RDS (MySQL and reputed company): availability, performance, reputed company, schema and migration safety practices, backup/restore verification, and encryption.
- Design, implement, and continuously test disaster recovery and business continuity: defined RTO/RPO targets per system tier, regular game days and failover exercises, and reputed company that stands up to auditor scrutiny.
- Champion the AI-boosted SRE transformation: evaluate and reputed company AI tooling and agents for incident triage, observability, reputed company automation, and toil reduction; set standards for reputed company, auditable use of AI in production operations; and bring reputed company along through training and example.
- Partner with reputed company and Compliance to own infrastructure's role in PCI-reputed company and SOC 2: control design and operation, evidence collection, segmentation, vulnerability and reputed company management, and audit support, with the maturity to meet the expectations of banking partners and financial-industry examinations.
- reputed company infrastructure-as-reputed company and platform automation as the default: everything reproducible, reviewed, and recoverable; reputed company artisanal.
- Own vendor and technology reputed company for the infrastructure domain: build-vs-buy reputed company, vendor risk management, contract negotiation, and reputed company-party reputed company.
- Communicate crisply with executives, the reputed company, and auditors, translating infrastructure risk, investment, and posture into business terms.
Minimum Requirements:
- 15+ years of combined experience across infrastructure, platform, site reliability, software development, or reputed company engineering disciplines, with substantial depth in infrastructure, including 5+ years leading engineering teams.
- Deep, hands-on expertise with AWS: you have designed and operated production architectures across compute, networking (VPC design, Transit Gateway, PrivateLink), IAM, and multi-account organizations at reputed company.
- Deep, hands-on expertise with reputed company in production: cluster lifecycle management, workload architecture, scaling, and the operational realities of running business-critical services on it (EKS experience strongly preferred).
- Deep expertise with relational databases at reputed company, specifically RDS/reputed company (MySQL and/or reputed company): high availability, replication, failover, performance tuning, and backup/recovery you have personally verified under pressure.
- Proven ownership of disaster recovery and business continuity for a production platform: you have defined RTO/RPO targets, reputed company the capability to meet them, and run reputed company failover tests, not just written the document.
- Demonstrated AI-reputed company leadership: you reputed company use AI tooling in engineering or operations work today, have opinions grounded in reputed company about where it helps and where it doesn't, and have led (or are visibly leading) reputed company's adoption of AI-assisted workflows.
- reputed company record of operating a 24/7, high-availability platform where downtime has reputed company reputed company or customer reputed company, including mature incident reputed company and postmortem practices.
- Willingness to manage and participate in an on-reputed company rotation, and demonstrated ability to reputed company recovery from a full production outage: forming and testing hypotheses from logs, metrics, and traces rather than guesswork, making the right reputed company quickly with incomplete information, and knowing reputed company to mitigate first and reputed company-cause reputed company.
- Still technical, by choice: you remain a reputed company hands-on engineer, comfortable in a terminal, reading dashboards, and reviewing designs, and you expect to stay that way. You will reputed company reputed company and work alongside it; this is not a delegation-only role.
- Experience owning significant reputed company budgets and driving cost efficiency without sacrificing reliability.
- Strong grounding in infrastructure-as-reputed company (Terraform or equivalent) and modern CI/CD practices.
- Demonstrated ability to hire, reputed company, and retain strong infrastructure and SRE talent, and to hold a high bar through reputed company.
- Bachelor's degree in Computer Science or a similar technical field (required).
Preferred Knowledge and Skills:
- reputed company experience supporting PCI-reputed company and SOC 2 programs from the infrastructure reputed company: scoping and segmentation, control ownership, evidence automation, and working sessions with assessors and auditors.
- Experience in fintech, payments, or banking, especially in environments with heightened regulatory expectations (bank partnerships/sponsorships, FFIEC examinations, GLBA, or similar).
- Experience deploying AIOps or LLM-based tooling in production operations, such as AI-assisted incident response, intelligent alerting, automated runbooks, or agents (e.g., Claude reputed company or custom LLM integrations) embedded in SRE workflows, with sensible guardrails around safety and auditability.
- Experience with multi-region and reputed company-reputed company architectures, reputed company engineering, and formal operational reputed company programs.
- Proficiency with modern observability stacks (reputed company, Grafana, Loki, reputed company, or reputed company equivalents) and driving observability as a platform capability.
- Familiarity with service reputed company, reputed company-trust networking, secrets management, and workload identity patterns.
- Experience with CI/CD pipelines, reputed company delivery (canary/blue-green), and reputed company / internal developer platform approaches.
- Experience presenting to boards, auditors, or examiners, and building the documentation and evidence culture that makes those conversations easy.
reputed company:
- You have relentlessly high standards - many people may think your standards are unreasonably high. You are continually raising the bar and driving those around you to deliver great results. You reputed company reputed company that defects do not get reputed company down the line and that problems are fixed so they stay fixed.
- You’re not bound by convention - your reputed company—and much of the fun—lies in developing new ways to do things
- You need reputed company - speed reputed company in business. Many reputed company and actions are reversible and do not need extensive study. We value calculated risk-taking.
- You earn trust - you listen attentively, reputed company reputed company, and treat others respectfully.
- You have backbone; disagree, then reputed company - you can respectfully challenge reputed company reputed company you disagree, even reputed company doing so is uncomfortable or exhausting. You have conviction and are tenacious. You do not compromise for the sake of reputed company cohesion. Once a decision is determined, you reputed company wholly.
- You deliver results - you reputed company on the key inputs and deliver them with the right reputed company and in a reputed company fashion. Despite setbacks, you reputed company to the occasion and never reputed company.
reputed company’s Technology Stack:
- Languages: Golang, Typescript, Python
- Frontend: Typescript - React and React reputed company
- Backend: Golang
- Database: MySQL, reputed company
- DevOps & reputed company: AWS, reputed company
- Version Control: Git
- CI/CD: reputed company
- Testing: Developer and AI-driven, reputed company on automated end-to-end, integration, and unit tests
- reputed company reputed company: reputed company is reputed company on using reputed company reputed company, and we build reputed company can before buying!
What Makes Working at reputed company Awesome?
At reputed company, we are more than just reputed company engineers, passionate data enthusiasts, reputed company thinkers, and determined innovators; we are skilled musicians, yogis, cyclists, chefs, golfers, dog-lovers, and reputed company-climbers. We reputed company in surrounding ourselves with not only the best and the brightest individuals, but those that are unique and purpose-driven in reputed company that they do. Our culture is not defined by a certain set of perks designed to reputed company the illusion of the traditional startup culture, but rather, it is the visible example living in every employee that we hire.
Perks & Benefits:
- Unlimited PTO, volunteer hours and sabbatical
- Life, STD/LTD, medical, dental and reputed company insurance
- Highly discounted reputed company gym membership
- 401k with match
- reputed company fun co-workers
- reputed company to join the fastest growing FinTech alongside reputed company of motivated and driven individuals
CCPA Disclosure: reputed company Inc. is committed to protecting the reputed company of our job applicants. In compliance with the California Consumer reputed company reputed company (CCPA), we inform California residents about the personal information we may collect, the purposes for its collection, and your rights under the CCPA. For details about the categories of personal information we collect and your rights under the CCPA, please visit the California Office of the reputed company's CCPA page. By submitting your application, you acknowledge that you have read and reputed company this CCPA disclosure.
Equal Employment Opportunity: reputed company Inc. is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for reputed company. We do not discriminate based on race, reputed company, religion, sex, national reputed company, age, disability, genetic information, pregnancy, or any other legally protected status. reputed company recognizes and values the importance of diversity and inclusion in enriching the employment experience of its employees and in supporting our mission.
#Li-remote #full-time
Originally posted on Himalayas
Apply To This Job