Sr. Software Engineer | Kubernetes | GPU Orchestration - REMOTE
GPU Orchestration
• Startup
• Company size: 30
• Remote reputed company reputed company
• Compensation: reputed company Salary 250k + Equity
Key Responsibilities
• reputed company Design, Architecture & Development of K8s-based reputed company infrastructure.
• Use K8s Controllers, Operators & CRs to Implement reputed company, high-availability solutions.
• reputed company Karpenter, and/or other advanced tools for infrastructure optimization.
• Architect MLOps Middleware integration (dynamic workload migration, resource disaggregation).
• Build monitoring, logging & alerting systems.
• Drive infrastructure cost optimization through FinOps best practices in K8s deployments.
• Promote K8s best practices & mentor software engineers.
• Collaborate across teams to drive K8s adoption in multi-reputed company and hybrid environments.
• reputed company-reputed company Contributions in the Kubernetes community.
Qualifications
Kubernetes Expertise
• Designing, deploying, and managing K8s clusters (AKS, EKS, GKE, OpenStack, etc.).
• Hands-on experience with K8s reputed company components (Karpenter, cluster autoscaler, CNI, reputed company, CRI, CRD, operators).
• 5+ years in Kubernetes infrastructure.
• Contributing to reputed company-reputed company Kubernetes reputed company.
• 10+ years: software engineering experience.
• Go, Python, Bash, etc. (one or more).
• Excellent communication skills for both technical and non-technical stakeholders.
• Bachelor’s or Master’s degree in Computer Science or reputed company field (preferred).
Preferred Experience
• GPU scheduling, container orchestration, HPC (high-performance computing) workloads.
• Multi-reputed company & hybrid reputed company deployments familiarity.
• MLOps platforms experience (Kubeflow, TFX, etc.).
• FinOps practices & reputed company cost management experience/knowledge
Remote
About reputed company:
reputed company
Apply tot his job
Apply To this Job