Rubiscape Private Limited Logo

Rubiscape Private Limited

MLOps Engineer

Posted 9 Hours Ago
Be an Early Applicant
In-Office
Pune, Maharashtra, IND
Mid level
In-Office
Pune, Maharashtra, IND
Mid level
Design and operate production MLOps infrastructure for large-scale model training, deployment, monitoring, and rollback across SaaS, cloud, on-premises, and air-gapped environments. Build ML pipelines, model registries, containerized serving platforms, observability systems, and governance controls. Collaborate with engineering, security, compliance, and customer teams while managing production model reliability, performance SLOs, and incident response.
The summary above was generated by AI
About the Role

Rubiscape’s RubiStudio studio promises enterprises a path from experiment to production in under 90 days — and the MLOps Engineer is the person who makes that promise real. You will design the CI/CD pipelines, model registries, deployment orchestration, and monitoring infrastructure that keep hundreds of ML models running reliably across SaaS, BYOC, on-premises, and air-gap deployments. You will work closely with ML Engineers, Platform Engineers, and enterprise customer success teams to eliminate the gap between model training and business value.

 

Key Responsibilities

·         Build and maintain end-to-end ML pipelines using MLflow, Kubeflow, or Airflow that handle training, validation, packaging, and deployment of models at scale.

·         Design the model registry architecture within RubiStudio: versioning strategies, stage transitions (staging → canary → production), approval gates, and rollback mechanisms.

·         Implement automated model monitoring for data drift, concept drift, and prediction quality degradation, surfacing alerts into RubiSight operational dashboards.

·         Manage containerised model serving infrastructure (Docker + Kubernetes) across multi-cloud and on-premises deployment topologies aligned with Rubiscape’s deployment flexibility.

·         Define and enforce MLOps best practices: reproducible experiments, environment parity, feature store integration, and audit-ready lineage for regulated-sector customers.

·         Collaborate with security and compliance teams to ensure model artefacts, training data references, and inference logs meet enterprise data governance standards.

·         Instrument inference endpoints with latency, throughput, and error-rate SLOs; own on-call response for production model degradation incidents.

Nice to Have

·         Experience operating ML infrastructure in air-gap or on-premises environments for government or defence customers.

·         Knowledge of feature stores (Feast, Tecton, or a custom implementation) and their integration into training and online inference paths.

·         Exposure to GPU cluster management and optimising inference throughput for large model serving.

·         Certification in AWS Machine Learning Specialty, Google Professional ML Engineer, or equivalent.

 

 

 

About Rubiscape

Rubiscape is India’s leading Decision Intelligence Platform, unifying data engineering, BI, machine learning, and agentic AI in a single governed platform. Built in Pune and trusted by Fortune 500 enterprises across BFSI, manufacturing, healthcare, and government. 8 international innovation patents. 10 Industry-Academia Labs & COEs. From BI to AI — One Platform. Every Decision.



RequirementsRequirements

·         3+ years in MLOps, ML infrastructure, or ML platform engineering roles with demonstrable production deployments.

·         Proficiency with MLflow (or similar experiment tracking + registry tools) and workflow orchestration frameworks such as Airflow, Kubeflow Pipelines, or Prefect.

·         Strong container and Kubernetes skills: writing Helm charts, managing model-serving deployments, horizontal pod autoscaling for inference workloads.

·         Experience with at least one model-serving framework: TorchServe, Triton Inference Server, BentoML, or Seldon Core.

·         Working knowledge of Python and shell scripting sufficient to own pipeline code, not just configure GUI tools.

·         Familiarity with observability tooling (Prometheus, Grafana, OpenTelemetry) applied to ML workloads.



Similar Jobs

9 Hours Ago
In-Office
Senior level
Senior level
HR Tech • Information Technology • Professional Services
Build and manage AWS cloud-native and serverless applications, including Lambda, API Gateway, DynamoDB, SQS, ECS/EKS, and related services. Implement CI/CD pipelines and infrastructure as code using AWS CDK, Terraform, and CloudFormation. Support AI/ML workloads, model deployment, monitoring, and lifecycle management. Develop microservices and event-driven architectures, improve scalability and reliability, implement observability, troubleshoot infrastructure and ML pipelines, and follow cloud security and governance practices.
Top Skills: Ai/MlAmazon CloudwatchAmazon CognitoAmazon EcsAmazon EksAmazon S3Api GatewayAWSAws CdkAws CloudformationAws GlueAws LambdaAws Step FunctionsCi/CdDistributed SystemsDynamoDBEvent-Driven ArchitectureGoInfrastructure As CodeJavaMicroservicesMlopsObservabilityPythonSqsTerraformTest-Driven DevelopmentTypescript
10 Hours Ago
In-Office or Remote
2 Locations
Mid level
Mid level
Cloud • Fintech • Software
Build and operate scalable MLOps infrastructure for machine learning development, training, inference, and monitoring. Responsibilities include Kubernetes orchestration, AWS resource management, GPU optimization, prediction and inference platforms, ML workflows, CI/CD pipelines, observability, security compliance, and LLM serving clusters. The role collaborates with ML engineers and supports reliable, high-performance AI services and production operations.
Top Skills: AWSCi/CdGpu OptimizationInfrastructure As CodeKubeflowKubernetesLlm Inference ServersMlflowOpentelemetryPythonRay ServeTerraformTerragruntVllm
13 Days Ago
In-Office
Pune, Mahārāshtra, IND
Senior level
Senior level
Artificial Intelligence • Consulting
Senior Databricks MLOps Engineer responsible for automating the end-to-end ML lifecycle: CI/CD, packaging, deployment, scheduling, monitoring, and reproducible experiments. Build frameworks and pipelines, integrate embedding/LLM components (RAG, vector DBs, LangChain), optimize ML workloads, and collaborate with data scientists and platform teams to productionize models for Claims Payment Integrity.
Top Skills: AzureAzure DevopsAzure OpenaiDatabricksEmbedding ModelsGitGithub ActionsJenkinsLangchainMlflowOpenai ApiPythonRagScalaSparkTerraformVector Databases

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account