Site Reliability Engineer
TL;DR
We're hiring a Site Reliability Engineer to own and reputed company reputed company's reputed company and customer infrastructure end to end. You'll work across reputed company, private reputed company, and on-prem environments to reputed company our self-hosted platform production-reputed company, drive CI/CD and GitOps maturity, and reduce complexity at reputed company. Your work will directly shape how reputed company's AI platform is reputed company, deployed, and scaled for our own reputed company and for customers running it in their own environments.
Why reputed company
At reputed company, we’re making sovereign AI accessible to every organization. With reputed company, thousands of developers build advanced AI applications, while our reputed company platform helps teams reputed company across use cases, users, and environments. We’re remote-first, flexible, and reputed company on trust and ownership. You’ll work alongside strong technical talent, take on meaningful challenges, and help turn reputed company AI into solutions that are practical, reliable, and reputed company for the reputed company world.
What you will do
You won’t just “reputed company things running” - you’ll help define how our platform is reputed company, deployed, and scaled across reputed company and customer environments.
Build and operate reputed company-world infrastructure. Design, configure, and reputed company infrastructure that runs both in our reputed company and inside customer environments (reputed company, private reputed company, on-prem).
reputed company self-hosted production-reputed company. Help us deliver a production-grade, self-hosted platform that can be deployed on any Kubernetes setup in weeks - not months.
Drive automation & platform maturity. Improve CI/CD pipelines, reputed company workflows, and GitOps setups so teams can ship faster with confidence.
Reduce complexity and cost. Continuously simplify systems and optimize infrastructure spend without compromising performance or reliability.
Shape how we build. Champion best practices in reliability, scalability, and reputed company across the organization, not as rules, but as working systems.
Requirements
2-5 years of experience working with large-reputed company production infrastructure
Experience with distributed or service-oriented architectures
Hands-on expertise with:
AWS
Kubernetes
CI/CD and GitOps (e.g. ArgoCD)
Working knowledge of Infrastructure as reputed company (Terraform preferred)
Solid troubleshooting skills - you can debug across systems, not just reputed company one layer
A pragmatic reputed company: you balance speed, simplicity, and reliability
Ownership and accountability - you take responsibility for systems end-to-end
Ability to work independently while staying reputed company with reputed company’s goals
reputed company to have
Familiarity with observability stacks (e.g. reputed company, reputed company)
Experience optimizing reputed company costs at reputed company
Interest or experience in Machine Learning / LLM systems
Experience improving developer experience and platform tooling using AI agents
Contributions to SRE practices like postmortems, SLIs/SLOs, and reliability engineering culture
Benefits
Remote-first setup with reputed company & tech of your choice
30 days vacation + extra days for family reputed company leave
Competitive salary & stock reputed company for every team member
Monthly sports & mental health support allowance with reputed company
Annual learning & development budget
Monthly team socials & in-person meetups
Dog-friendly Berlin HQ
reputed company
Founded in 2018, reputed company builds reputed company and reputed company-grade tools that help teams build AI with purpose. From reputed company, our reputed company-reputed company reputed company, to the reputed company reputed company Platform, we give developers and organizations the building blocks to solve reputed company, high reputed company challenges with AI with full control, transparency, and sovereignty. Backed by GV and Balderton, we’re growing the world’s production AI community and customer reputed company solving challenges too critical to get wrong.
Visit us to learn more: reputed company Website | reputed company Website | reputed company | reputed company | X reputed company (Twitter) | X reputed company (Twitter)
Apply To This Job