Optum Logo

Optum

Senior Data Engineering Lead

Posted 4 Days Ago
Be an Early Applicant
In-Office
Pune, Mahārāshtra, IND
Senior level
In-Office
Pune, Mahārāshtra, IND
Senior level
Lead the design and implementation of resilient batch, micro-batch, and event-driven healthcare data pipelines on AWS and Databricks. Build ingestion frameworks, Spark/PySpark transformations, Delta Lake and Iceberg tables, healthcare data models, quality and lineage frameworks, and privacy-aware PHI processing. Deliver CI/CD, infrastructure automation, testing, observability, monitoring, and incident recovery. Optimize governed data consumption while ensuring compliance with healthcare and regulated-data requirements.
The summary above was generated by AI
Requisition Number: 2384847
Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together.
Primary Responsibilities:
  • Build reliable batch, micro-batch, and event-driven pipelines on AWS and Databricks
  • Develop reusable ingestion frameworks for REST APIs, FHIR Bulk Export, HL7 interfaces, databases, SFTP/file exchange, JSON/NDJSON, CSV, XML, PDFs, and clinical text
  • Implement scalable Spark/PySpark and SQL transformations, including schema inference/evolution, checkpointing, idempotency, retries, backfills, and replay
  • Design and maintain Delta Lake and Apache Iceberg tables, including physical design, partitioning/clustering, compaction, file sizing, incremental reads/writes, performance tuning, and cost optimization
  • Build curated healthcare data models and transformations using FHIR, HL7, OMOP CDM, and clinical terminology mappings
  • Implement data-quality frameworks: schema validation, referential-integrity checks, business-rule testing, anomaly detection, reconciliation, completeness checks, and data-quality observability
  • Build metadata and lineage capture across source systems, pipeline runs, code versions, transformation rules, mappings, and published data products
  • Implement privacy-aware data processing for PHI, including access controls, masking, tokenization/pseudonymization, de-identification, and auditable handling patterns
  • Deliver CI/CD pipelines, automated unit/integration/data tests, Terraform or CloudFormation, containerized services, monitoring, alerting, runbooks, and incident-recovery procedures
  • Integrate and optimize governed data consumption through Databricks SQL, Athena, Snowflake, Trino, or equivalent engines
  • Comply with all applicable Company policies, procedures, and business directives, changes including those relating to work location, team assignments, work schedules, and flexible work arrangements

Eligibility
To apply to an internal job, employees must meet the following criteria:
  • The candidate should have completed 12 months in the current role
  • The candidate should not be on any active CAP/ PIP
  • The performance review of the candidate must be ME & Above in the last common review

Required Qualifications:
  • Graduate degree or equivalent experience
  • 5+ years of production data-engineering experience
  • Hands-on AWS experience with S3, IAM, VPC, ECS/EKS or Lambda, Step Functions, EventBridge, CloudWatch, Secrets Manager, and KMS
  • Experience with Git, pull requests, CI/CD, Docker, Terraform/CloudFormation, automated testing, observability, and incident response
  • Experience processing semi-structured data and building resilient ingestion pipelines with quality controls, error handling, and replay capability
  • Production Databricks experience with Auto Loader, Delta Lake, Workflows, Unity Catalog, notebooks/jobs, SQL Warehouses, and cluster/job optimization
  • Advanced Python, SQL, and Apache Spark/PySpark; strong understanding of distributed processing and performance tuning
  • Practical Apache Iceberg knowledge, including tables, catalogs, snapshots, schema/partition evolution, compaction, and interoperability with query engines
  • Healthcare data knowledge: FHIR R4 and/or HL7 v2, OMOP CDM, clinical terminologies, PHI, HIPAA-aligned engineering controls, and de-identification concepts
  • Demonstrated responsible use of AI coding tools and the ability to critically review, test, and productionize generated code

Required Qualifications:
  • Healthcare or regulated-data experience

At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.

Optum Pune, Maharashtra, IND Office

Pune, India, India

Similar Jobs at Optum

Yesterday
In-Office
Pune, Mahārāshtra, IND
Senior level
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads the design, development, modernization, and support of scalable enterprise applications using Java, Spring Boot, microservices, Kafka, and cloud platforms. Drives architecture decisions, cloud migration, CI/CD, DevOps automation, performance optimization, security, monitoring, and production support. Collaborates with stakeholders, mentors engineers, conducts code reviews, and promotes engineering best practices while contributing to Agile planning and project execution.
Top Skills: Apache KafkaApi GatewaysArm TemplatesAzure DevopsBicepBigQueryCloud Deployment ManagerDockerDynatraceGithub ActionsGitlab CiGoogle Cloud PlatformGrafanaJ2EeJavaJenkinsKubernetesMicroservicesAzureMongoDBOraclePostgresPrometheusRest ApisService MeshSplunkSpring BootSpring MvcSpring SecuritySQL ServerTerraform
19 Days Ago
In-Office
Pune, Mahārāshtra, IND
Senior level
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Manage statutory HR compliance across India, including labor-law audits, inspections, internal reviews, vendor oversight, legislative monitoring, documentation, and risk mitigation. Interpret employment laws during workforce changes, coordinate with Legal and other stakeholders, manage compliance documentation, and support new legal-entity setup. The role requires extensive experience with Contract Labour, ESI, Provident Fund, and Shops and Establishments regulations, as well as engagement with labor authorities and compliance vendors.
Top Skills: Crm ToolsExcelPeoplesoftPowerPointSAP
29 Days Ago
In-Office
Pune, Mahārāshtra, IND
Senior level
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Leads design, development, testing, deployment, and maintenance of cloud-native enterprise applications and APIs. Builds Java and Python backend services, React-based front ends, microservices, CI/CD pipelines, and Azure infrastructure. Develops agentic AI and RAG solutions, supports AI lifecycle operations, and ensures secure, scalable, high-quality delivery. Provides technical leadership, mentorship, architecture guidance, production troubleshooting, and continuous improvement across Agile teams.
Top Skills: Agentic AiAksAngularArtifactoryAzureDockerGenerative AiGithub ActionsJavaJfrogKubernetesLlm OrchestrationMicroservicesNext.JsPrompt EngineeringPythonRagReactRestful ApisSpring BootTerraformVector Search

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account