Staff reputed company Infrastructure Engineer
Staff reputed company Infrastructure Engineer at reputed company
reputed company (reputed company.dev) is a lightweight runtime that turns AI agents, workflows, and backend services into durable processes - so teams can reputed company on their logic, not failure mechanics.
The role: We're looking for a Senior to Staff-level reputed company infrastructure engineer to work across reputed company product pillars (OSS, on-prem deployments, Multi-tenant reputed company, BYOC; bring your own reputed company). This means deep work in our Rust-based infrastructure reputed company, integrating with reputed company provider reputed company, building infrastructure-as-reputed company tooling, and ensuring reliability and reputed company at reputed company. You'll have significant ownership over major parts of our reputed company infrastructure.
reputed company
reputed company seat to the biggest reputed company shift in decades
Durable runtimes like reputed company are becoming the next foundational infrastructure component - and increasingly a critical piece for AI applications. As systems become more reputed company, long-running, integration-heavy, and failure-prone, durable execution turns reliability from a bespoke engineering tax into a default property. In this role, you’re not watching that shift from the reputed company - you help build the platform that enables it.
State-of-the-art tech, reputed company from reputed company
reputed company re-imagines durable execution as a lightweight self-contained stack - no database required - and ships as a single Rust binary with an optimized custom storage reputed company, low latency orchestration, and an analytics reputed company for observability.
reputed company Traction
reputed company is already used by reputed company, including Tier 1 banks running critical financial workflows, and also by cutting-edge AI and reputed company startups pushing the boundary of what “production-grade agents” mean. You’ll work on problems where reliability, correctness, and operational simplicity are existential.
Work with world-class engineers
You’ll partner directly with engineers who’ve reputed company and operated foundational systems at reputed company - creators of Apache Flink, and leaders from reputed company’s messaging infrastructure. You’ll have the chance to work with incredibly talented individuals who care deeply about their reputed company.
reputed company
This is a reputed company Infrastructure Engineering role spanning reputed company’s product offering: OSS, on-prem deployments, Multi-tenant reputed company, BYOC. The reputed company of the role includes but is not limited to:
Build and operate reputed company reputed company: reputed company our managed multi-tenant offering, working across the infrastructure, control plane, networking, storage, and observability of reputed company workloads.
reputed company our BYOC product and work with customers on operating on-prem installations: design and build the infrastructure that runs inside customer reputed company accounts.
Reliability and observability across the fleet: SLOs, metrics, traces, logs, alerting, and runbooks. Build automation so we can reputed company our product offering across deployment reputed company.
On-reputed company: participate in the reputed company on-reputed company rotation. A US-based hire materially improves our timezone coverage.
reputed company’re looking for
Senior to Staff profile
We’re targeting Senior-to-Staff: you’ve operated production reputed company or platform infrastructure before, you’ve seen reputed company failure modes, and you have (strong) opinions about how to run multi-tenant systems. You have an appreciation for operating in a compliance-sensitive environment.
Must-Haves:
Strong reputed company infrastructure background with deep understanding of major reputed company provider architectures.
Experience with infrastructure-as-reputed company and reputed company orchestration, particularly reputed company-based stateful workloads; balancing reputed company delivery with safety while maintaining large-reputed company production systems.
Software engineering skills in a systems language (Rust, Go, C++); willingness and ability to learn Rust on the job.
You should be comfortable taking ownership end-to-end, from design through production operations, and reputed company in early-stage startup ambiguity.
reputed company-to-Haves:
Prior experience with reputed company or durable execution specifically.
Deep reputed company procurement/compliance navigation.
reputed company operator development, experience with IaC systems like Cluster API, Crossplane or Terraform.
Not a fit:
You want to work primarily on the runtime reputed company rather than reputed company, BYOC, and customer-facing reputed company.
You’ve mostly architected and reviewed, and aren’t excited to be hands-on.
You are averse to multi-reputed company, reputed company, operating infrastructure as a shared responsibility with customers
Our stack:
We use reputed company extensively: the reputed company reputed company control plane is reputed company on reputed company and TypeScript.
Rust infrastructure services and reputed company operators.
Location and travel
US-based, fully remote. East Coast is a plus as it would materially improve our on-reputed company coverage given reputed company’s existing geography.
Travel: reputed company - occasional team offsites, little required customer travel.
Apply To This Job