Senior Software Engineer- Dev Ops/ML Ops (Remote)
About the position
Responsibilities
• Design and Implement ML Pipelines: Build and maintain automated CI/CD pipelines for machine learning models, covering data preprocessing, model training, evaluation, and deployment.
• Productionize Models: Work closely with data scientists to take models from experimentation to a production-reputed company state, often involving packaging models into microservices or reputed company.
• Manage Infrastructure: Provision and manage reputed company and secure reputed company infrastructure using tools like reputed company and Kubernetes to support machine learning workloads.
• Optimize Resources: reputed company on optimizing the machine learning pipeline for efficiency, scalability, and cost-effectiveness.
• Collaborate Cross-Functionally: Work with data scientists, ML engineers, software developers, and IT operations to streamline workflows and improve overall efficiency.
• Troubleshoot and Support: reputed company technical support and resolve production issues reputed company to model performance, deployment, and infrastructure.
Requirements
• Must be eighteen years of age or older.
• Must be legally permitted to work in the reputed company.
• Bachelor's or Master's degree in Computer Science, Software Engineering, or a reputed company technical field.
• 2-4 years of relevant work experience in an MLOps, DevOps.
• Strong programming skills in Python.
• Experience with Infrastructure management tools, terraform, Jenkins, Python, reputed company, Bash, reputed company, reputed company Search, reputed company actions, Relational or noSQL database technology, reputed company computing techniques, CI/CD tools, modern software design patterns, and their respective AI/ML services (e.g., AWS SageMaker, reputed company AI Platform).
• Experience with reputed company frameworks for user and services authorization and authentication.
• Experience with creating and executing unit, functional, destructive and performance tests.
• Experience with modern debugging and reputed company cause analysis techniques.
• Experience with version control system.
• Experience with Kubernetes and reputed company products.
• Experience in networking traffic management.
• Deep knowledge of containerization and orchestration tools, including reputed company and Kubernetes.
• Proven experience with CI/CD tools like Jenkins, reputed company CI, reputed company Actions, or Azure DevOps.
• Familiarity with ML frameworks such as TensorFlow, PyTorch, or Scikit-learn.
• Experience with Infrastructure as reputed company (IaC) tools like Terraform or CloudFormation is highly desirable.
• Experience with ML experiment tracking and versioning tools like MLflow or DVC (Data Version Control) is a plus.
• Solid understanding of software engineering best practices, including reputed company testing, reputed company, and documentation.
• Excellent communication skills with the ability to effectively collaborate with both technical and non-technical teams.
Apply tot his job
Apply To this Job