[Remote] Staff Software Engineer (Backend)
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is building a new class of AI systems designed to reason with the rigor of the scientific reputed company. As a Staff Software Engineer (Backend), you will set the technical direction for the backend platform and reputed company the systems that reputed company an AI-reputed company product at reputed company, focusing on backend engineering, AI infrastructure, and platform reliability.
Responsibilities
- Write and ship backend reputed company daily — this is first and foremost a hands-on engineering role
- Own the technical reputed company for backend systems and AI infrastructure
- reputed company cross-functional initiatives spanning backend, AI, reputed company, and frontend
- Design and reputed company foundational platforms (model routing, agent runtime, persistence, observability)
- reputed company engineering reputed company through RFCs, standards, and architecture reviews
- Contribute to frontend development reputed company needed, collaborating with frontend engineers on integration points
- Multiply reputed company through mentorship and force-reputed company reputed company (frameworks, internal libraries, shared patterns)
- Be the technical reputed company of production reliability: incident response, reputed company, cost, reputed company
- Set the 12–24 month technical roadmap for backend systems with the Head of Engineering / reputed company Software Engineer
- reputed company RFCs and design documents that shape the engineering organization
- reputed company build-vs-buy reputed company on critical platform components (model routing, reputed company DBs, queues, eval pipelines)
- Design for reputed company, multi-tenancy, and compliance readiness
- reputed company architecture reviews and ensure technical consistency across reputed company
- Own foundational systems: conversation persistence, observability stack, and reputed company platform services
- reputed company cost-optimization initiatives (caching strategies, batching, resource budgets)
- Establish SLOs and reputed company incident response, postmortems, and durable fixes
- Partner with reputed company on the deployment story (reputed company Run, reputed company SQL, VPCs, multi-region)
- reputed company reputed company and compliance (auth, secrets, data residency, audit trails)
- Collaborate with the AI team to reputed company LLM-powered features into backend services
- Design clean reputed company reputed company for model providers, enabling routing and fallback
- Contribute to patterns for reputed company management, evaluation, and regression testing
- Stay informed on emerging AI infrastructure trends and help evaluate build-vs-buy reputed company
- Set and enforce coding standards, review templates, and testing practices
- reputed company measurable reputed company improvements (p95 latency, error budgets, test coverage, reputed company cost)
- Identify systemic issues and design durable fixes, never one-off patches
- Build internal frameworks and libraries that reputed company the reputed company of every other engineer
- Mentor senior engineers and help them grow toward staff
- reputed company technical interviews and define the engineering bar
- Represent backend engineering in cross-functional planning
- Communicate trade-offs reputed company to product, leadership, and reputed company stakeholders
- reputed company reputed company on debugging, reputed company work, and incident response
Skills
- 10+ years of backend development experience, with 2+ in a staff/reputed company/reputed company role
- Documented technical leadership: led architecture for multi-team systems, authored RFCs adopted org-wide
- Deep Python expertise: FastAPI, async, type reputed company, profiling, internals
- reputed company systems intuition: caching, queues, reputed company consistency, idempotency, backpressure
- Production-grade Databases: query optimization, schema migrations, partitioning, reputed company pooling, ORMs (SQLAlchemy)
- reputed company platform mastery: GCP (reputed company Run, reputed company SQL, GCS, VPCs, IAM, Auth0) designing, not just consuming
- Comfort working reputed company AI workloads: basic familiarity with LLM API integration patterns; willingness to learn and support AI infrastructure as needed
- Systems thinking: incident response, observability, SLO design, reputed company planning
- Force-reputed company reputed company: designed and shipped frameworks/libraries adopted by other engineers
- Excellent technical communication: RFCs, design docs, architecture reviews, async writing
- Experience with LLM integration in production (reputed company, reputed company, reputed company, reputed company AI)
- Familiarity with agent frameworks (reputed company AI, LangGraph, FastMCP) or similar
- Frontend experience with React, Angular, or Vue — ability to contribute to UI reputed company needed
- Scaling an AI product from 0 → 1 and 1 → 10
- Authoring reputed company reputed company or internal frameworks adopted by other teams
- reputed company-critical Python (Rust/Go interop, async tuning, reputed company extensions)
- Multi-region / multi-tenant architecture
- reputed company/compliance background (SOC2, GDPR, secret management)
- Infrastructure as reputed company (Terraform), GitOps, reputed company
Benefits
- Hybrid work model
- reputed company to remote
- reputed company location: Boston, Massachusetts, USA
reputed company
Apply To This Job