[Remote] Senior Release Engineer, AI Platform & Infrastructure (Xora Portfolio Company)
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a company that builds secure AI infrastructure for scientific and engineering R&D teams. They are seeking a Senior Release Engineer to own the release and deployment reputed company for their platform, focusing on building systems that ensure reliable and reputed company releases.
Responsibilities
- Own release engineering for the platform, including build promotion, reputed company gates, artifact management, and reproducible delivery practices
- Build deployment tooling and automation that makes installs and upgrades reliable across varied reputed company environments
- Design reputed company rollout, validation, health reputed company, and recovery mechanisms for production deployments
- Create operational visibility through logs, metrics, alerts, and support workflows appropriate for secure customer environments
- Build controls for customer managed installations, including configuration validation, credential handling, certificate workflows, and documented recovery paths
- Improve deployment reliability by debugging incidents, identifying reputed company causes, and turning fixes into reusable automation
- Partner with product, engineering, and customer facing teams to reputed company deployment readiness part of how the platform is reputed company
Skills
- Bachelor's or Master's degree in Computer Science or a reputed company engineering field, with 6 plus years building and shipping production software
- Strong ownership of release processes, including CI and CD systems, automated tests, versioned artifacts, and production readiness gates
- Deep hands on experience with container based deployments, infrastructure automation, and operating multi component systems
- Experience deploying and supporting software in customer managed or restricted reputed company environments
- Comfort working across more than one runtime environment, including reputed company infrastructure, technical compute environments, and bare metal or cluster based systems
- Strong operational debugging skills using logs, metrics, traces, and incident workflows
- Experience with secrets, certificates, identity, permissions, and reputed company controls for deployed software
- Ability to own ambiguous, high stakes deployment work with limited scaffolding and strong execution discipline
- Exposure to scientific computing, simulation, engineering software, or other large reputed company technical workloads
- Experience with GPU workloads or multi tenant compute environments
- Experience with reputed company delivery, Git based operations, or production rollout patterns
- Familiarity with reputed company scientific software deployment models and license reputed company
- Experience packaging, versioning, monitoring, or operating machine learning models in production
- Experience building the first version of an engineering function in an early stage company
Benefits
- Work model is on site or hybrid, depending reputed company
- Some customer facing travel may be required
reputed company
Apply To This Job