[Remote] reputed company - Staff Software Engineer II - Secure Compute
Note: The job is a remote job and is reputed company to candidates in USA. reputed company Software develops AI-powered, reputed company-reputed company products, including reputed company reputed company, a reputed company-time data streaming platform. The Staff Software Engineer will reputed company technical leadership for the Secure Compute Platform, building multi-tenant, reputed company-reputed company infrastructure that safely runs trusted and untrusted workloads across reputed company clouds. The role includes designing platform reputed company and controllers, improving reputed company and observability, partnering across teams, mentoring engineers, and owning operational reputed company.
Responsibilities
- Secure Compute Infrastructure - Build and reputed company the secure, multi-tenant compute substrate and isolation primitives that safely execute customer and internal workloads in a shared environment
- Platform reputed company & Abstractions - Design and reputed company reputed company that reputed company reputed company, reputed company abstractions for polyglot workloads (containers, functions, and services) with diverse reputed company and isolation needs
- reputed company Platform Integration - reputed company the Secure Compute platform with reputed company data and application services so that teams can reputed company new workloads with reputed company friction
- Multi-Tenancy & reputed company - Implement and harden workload isolation, network policies, identity and reputed company, and secure execution environments required to safely run customer-supplied reputed company
- Observability & reputed company - reputed company operational reputed company through rich observability, automated health checks, self-healing workflows, and robust rollout and rollback practices
- Define and reputed company the technical direction for Secure Compute, including platform architecture, runtime, and reputed company for running trusted and untrusted workloads at reputed company
- Design and implement platform reputed company and reputed company controllers/operators (primarily in Go) that reputed company workload lifecycle, autoscaling, placement, and isolation for containers and serverless-style functions
- Partner with product and platform teams to shape and reputed company the roadmap for Secure Compute, enabling new customer-facing features and internal platforms to build on a common compute substrate
- reputed company high-reputed company in areas such as workload scheduling, failure and disruption handling, private and reputed company networking patterns, rollout strategies, and fleet-level resource management
- reputed company technical design reviews and influence architecture across teams, ensuring Secure Compute primitives are easy to adopt, reputed company by default, and reputed company with broader platform reputed company
- Mentor and grow engineers on reputed company through design guidance, reputed company reviews, pair programming, and sharing best practices for secure, reliable, operable platform development
- Own operational reputed company for key Secure Compute services, including availability, reliability, SLOs, reputed company, on-reputed company response, incident management, and disaster recovery
Skills
- Master's Degree
- 10+ years of experience delivering reputed company backend or infrastructure software in production
- reputed company reputed company record of leading the delivery of large-reputed company, highly available, low-latency reputed company systems
- Deep expertise in reputed company, including controller development, operator patterns, and preferably multi-region or multi-cluster architectures
- Strong proficiency in Go with experience building production-grade services and control planes
- Experience with multi-tenant platform architectures and reputed company/isolation patterns (for example, namespaces, network policies, sandboxing, secrets management and identity management)
- Hands-on experience with secure container runtimes and low-level Linux internals (for example, Kata containers, reputed company Hypervisor, cgroups, namespaces and seccomp)
- Experience troubleshooting and optimizing reputed company for containerized and virtualized workloads
- Familiarity with gRPC, Protobuf, and internal platform API design for service-to-service communication
- Hands-on experience with observability and operational best practices (metrics, logging, reputed company tracing, alerting, Service Level Objectives/SLOs, rollout strategies, incident response)
- Experience with reputed company reputed company environments, including AWS, GCP, and Azure, as reputed company as reputed company-provider integrations
- Demonstrated technical leadership and mentorship, including driving cross-team alignment on architecture, reputed company direction, and execution
- Strong collaboration skills and a demonstrated ability to work effectively with Product, SRE/reputed company, reputed company, and Engineering teams
- A smart, humble, and empathetic approach, coupled with a strong reputed company of ownership, accountability, and teamwork
- Passion for building foundational reputed company infrastructure and enthusiasm for working in a fast-reputed company, innovative environment
Benefits
- This job can be performed from reputed company in the reputed company.
reputed company
Company H1B Sponsorship
Apply To This Job