29 Aug
|
Mobileprogramming
|
Navi Mumbai
29 Aug
Mobileprogramming
Navi Mumbai
MLOps Engineer
Location: Navi Mumbai, India
Job Type: Full-time
Experience: 3+ Years
Department: Machine Learning Engineering / DevOps
About the Role
We are seeking an experienced MLOps Engineer to help build and operate reliable machine learning platforms and production ML workflows. This role sits at the intersection of software engineering, machine learning, cloud infrastructure, and DevOps.
The successful candidate will help data science teams move models from experimentation into production while establishing repeatable processes for deployment, monitoring, versioning, scalability, and operational support.
Key Responsibilities
- Build and maintain automated workflows for machine learning model development and deployment.
- Collaborate with data scientists and ML engineers to productionize trained models.
- Design CI/CD pipelines for machine learning applications, models, and supporting services.
- Containerize ML workloads using Docker and manage deployments through Kubernetes.
- Implement experiment tracking, model versioning, and lifecycle management using platforms such as MLflow or Kubeflow.
- Automate infrastructure and application provisioning using infrastructure-as-code practices.
- Develop reliable deployment processes across development, testing, staging, and production environments.
- Work with AWS, Azure, or Google Cloud to deploy and operate machine learning workloads.
- Build and maintain data and model pipelines using technologies such as Airflow.
- Configure monitoring and observability for machine learning services and infrastructure.
- Track model performance, system health, resource utilization,
and operational metrics.
- Investigate failed pipelines, deployment issues, infrastructure problems, and production incidents.
- Implement appropriate logging, alerting, rollback, and recovery mechanisms.
- Maintain secure and reproducible development and deployment environments.
- Work with engineering and data teams to improve the scalability and reliability of ML platforms.
- Contribute to platform documentation, operational procedures, and engineering standards.
Required Skills & Experience
- 3+ years of professional experience in MLOps, ML engineering, DevOps, or a closely related field.
- Strong Python programming and scripting skills.
- Good understanding of machine learning development and deployment workflows.
- Hands-on experience with Docker and Kubernetes.
- Experience designing and maintaining CI/CD pipelines.
- Practical experience with Jenkins, Git, or comparable DevOps tooling.
- Familiarity with MLflow, Kubeflow, or similar machine learning lifecycle platforms.
- Experience working with at least one major cloud platform: AWS, Azure, or GCP.
- Good Linux administration and troubleshooting skills.
- Understanding of infrastructure automation and configuration management.
- Experience monitoring production systems and investigating operational issues.
Preferred Skills
- Experience with Terraform or another infrastructure-as-code solution.
- Knowledge of Airflow and workflow orchestration.
- Experience supporting TensorFlow, PyTorch, or similar ML frameworks.
- Familiarity with model serving and inference architectures.
- Understanding of Kubernetes operators and cloud-native patterns.
- Knowledge of model monitoring, data drift, and model performance tracking.
- Experience with security practices for cloud and ML environments.
- Exposure to scalable distributed systems.
Key Responsibilities in Production The role will involve ensuring that machine learning solutions remain reliable after deployment. This includes:
- Improving deployment repeatability.
- Reducing manual intervention in ML releases.
- Establishing effective monitoring and alerting.
- Supporting model rollback and version management.
- Identifying infrastructure and model-related performance issues.
- Helping teams achieve reliable and efficient production ML operations.
Key Qualities
- Robust automation mindset.
- Comfortable working across software, infrastructure, and machine learning environments.
- Excellent troubleshooting and incident-analysis skills.
- Attention to reliability, security, and scalability.
- Ability to collaborate effectively with data scientists, developers, and infrastructure teams.
- Willingness to continuously improve engineering processes.
Pay: ₹850,000.00 - ₹1,100,000.00 per year
Work Location: In person
📌 MLOps Engineer (Navi Mumbai)
🏢 Mobileprogramming
📍 Navi Mumbai