Develops, tests, validates, and supports large-scale batch and real-time data platforms using Spark, Kafka, and NiFi. Builds resilient data pipelines, validation frameworks, distributed storage solutions, deployment automation, and operational tooling. Optimizes performance, scalability, and reliability; supports production troubleshooting and incident resolution; contributes to architecture, CI/CD, documentation, code reviews, and engineering standards; and mentors junior engineers.
Our Purpose
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
Title and Summary
Senior Data Engineer - Spark / Kafka / NiFi
Job Posting Title: Senior Data Engineer - Spark / Kafka / NiF
Who is Mastercard?
Mastercard is a global technology company in the payments industry. Our mission is to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our culture is guided by the Mastercard Way-own it, simplify it, sense of urgency, thoughtful risk taking, unlock potential, and be inclusive.Job Summary
We are seeking a highly skilled Senior Data Engineer with deep expertise in Testing, validating, Developing and supporting large-scale batch and real-time data platforms built on Apache Spark, Apache Kafka, and Apache NiFi.
The ideal candidate will have a strong background in distributed systems, streaming architectures, event-driven platforms, and cloud-native data processing.
The successful candidate will collaborate with cross-functional teams to design innovative solutions, enhance platform capabilities, and ensure operational excellence in production environments.
Key Responsibilities
Develop, test, validate and maintain, scalable batch and real-time data processing applications using Apache Spark, Kafka, and NiFi.
Build validation suite for high-performance, fault-tolerant, and resilient distributed systems capable of handling large data volumes.
Develop reusable frameworks, libraries, and shared platform components to accelerate engineering productivity.
Translate business and technical requirements into scalable software solutions.
Contribute to architecture discussions, technical designs, and engineering standards.
Streaming & Data Platform Engineering
Create and implement validation suite for Spark batch and streaming applications to process high-volume datasets efficiently.
Validate and support Apache NiFi data ingestion, transformation, and routing workflows.
Build reliable data pipelines supporting real-time and near-real-time processing requirements.
Implement solutions for data replay, recovery, checkpoint management, and failure handling.
Performance & Scalability
Analyze system bottlenecks and optimize application performance, throughput, and resource utilization.
Improve scalability, reliability, and availability of distributed applications.
Develop and test solutions leveraging modern storage technologies including Apache Ozone, Ceph, and cloud-native storage platforms.
Build deployment automation and operational tooling to improve platform reliability.
Monitor production environments and proactively address operational concerns.
Participate in troubleshooting, root-cause analysis, and incident resolution activities.
Engineering Excellence
Participate in code reviews and promote engineering best practices.
Maintain high-quality documentation for systems, APIs, and platform components.
Collaborate closely with Product, Architecture, Platform, and DevOps teams.
Contribute to CI/CD processes and continuous improvement initiatives.
Mentor junior engineers and foster a culture of technical excellence and innovation.
Required Qualifications
Bachelor's degree in Computer Science, Engineering, or related technical field.
6+ years of hands-on software deployment, validation, testing, development experience on large-scale data-intensive applications.
Strong expertise in Apache Spark (Batch and Structured Streaming).
Strong experience with Apache Kafka and event-driven architectures.
Experience developing, testing and supporting Apache NiFi data flows.
Proficiency in Scala, pyspark or Python.
Strong SQL and data modeling skills.
Experience working with distributed storage systems such as Apache Ozone, Ceph, HDFS, or cloud-based object stores.
Hands-on experience with Linux, Git, shell scripting, and CI/CD pipelines.
Strong analytical, problem-solving, and communication skills.
Experience operating production-grade distributed systems in cloud or hybrid-cloud environments.
Preferred Qualifications
Experience building observability solutions using monitoring and logging platforms.
Knowledge of data governance, metadata management, and data platform best practices.
Experience with containerization and orchestration technologies (Docker, Kubernetes).
Knowledge of performance engineering, resilience testing, and production readiness assessments.
Experience building engineering automation frameworks and reusable validation platforms.
Familiarity with observability tools, monitoring systems, and operational analytics.
Experience in highly regulated, transaction-processing, or large-scale enterprise environments.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we're helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.
Title and Summary
Senior Data Engineer - Spark / Kafka / NiFi
Job Posting Title: Senior Data Engineer - Spark / Kafka / NiF
Who is Mastercard?
Mastercard is a global technology company in the payments industry. Our mission is to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential. Our culture is guided by the Mastercard Way-own it, simplify it, sense of urgency, thoughtful risk taking, unlock potential, and be inclusive.Job Summary
We are seeking a highly skilled Senior Data Engineer with deep expertise in Testing, validating, Developing and supporting large-scale batch and real-time data platforms built on Apache Spark, Apache Kafka, and Apache NiFi.
The ideal candidate will have a strong background in distributed systems, streaming architectures, event-driven platforms, and cloud-native data processing.
The successful candidate will collaborate with cross-functional teams to design innovative solutions, enhance platform capabilities, and ensure operational excellence in production environments.
Key Responsibilities
Develop, test, validate and maintain, scalable batch and real-time data processing applications using Apache Spark, Kafka, and NiFi.
Build validation suite for high-performance, fault-tolerant, and resilient distributed systems capable of handling large data volumes.
Develop reusable frameworks, libraries, and shared platform components to accelerate engineering productivity.
Translate business and technical requirements into scalable software solutions.
Contribute to architecture discussions, technical designs, and engineering standards.
Streaming & Data Platform Engineering
Create and implement validation suite for Spark batch and streaming applications to process high-volume datasets efficiently.
Validate and support Apache NiFi data ingestion, transformation, and routing workflows.
Build reliable data pipelines supporting real-time and near-real-time processing requirements.
Implement solutions for data replay, recovery, checkpoint management, and failure handling.
Performance & Scalability
Analyze system bottlenecks and optimize application performance, throughput, and resource utilization.
Improve scalability, reliability, and availability of distributed applications.
Develop and test solutions leveraging modern storage technologies including Apache Ozone, Ceph, and cloud-native storage platforms.
Build deployment automation and operational tooling to improve platform reliability.
Monitor production environments and proactively address operational concerns.
Participate in troubleshooting, root-cause analysis, and incident resolution activities.
Engineering Excellence
Participate in code reviews and promote engineering best practices.
Maintain high-quality documentation for systems, APIs, and platform components.
Collaborate closely with Product, Architecture, Platform, and DevOps teams.
Contribute to CI/CD processes and continuous improvement initiatives.
Mentor junior engineers and foster a culture of technical excellence and innovation.
Required Qualifications
Bachelor's degree in Computer Science, Engineering, or related technical field.
6+ years of hands-on software deployment, validation, testing, development experience on large-scale data-intensive applications.
Strong expertise in Apache Spark (Batch and Structured Streaming).
Strong experience with Apache Kafka and event-driven architectures.
Experience developing, testing and supporting Apache NiFi data flows.
Proficiency in Scala, pyspark or Python.
Strong SQL and data modeling skills.
Experience working with distributed storage systems such as Apache Ozone, Ceph, HDFS, or cloud-based object stores.
Hands-on experience with Linux, Git, shell scripting, and CI/CD pipelines.
Strong analytical, problem-solving, and communication skills.
Experience operating production-grade distributed systems in cloud or hybrid-cloud environments.
Preferred Qualifications
Experience building observability solutions using monitoring and logging platforms.
Knowledge of data governance, metadata management, and data platform best practices.
Experience with containerization and orchestration technologies (Docker, Kubernetes).
Knowledge of performance engineering, resilience testing, and production readiness assessments.
Experience building engineering automation frameworks and reusable validation platforms.
Familiarity with observability tools, monitoring systems, and operational analytics.
Experience in highly regulated, transaction-processing, or large-scale enterprise environments.
Corporate Security Responsibility
All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:
- Abide by Mastercard's security policies and practices;
- Ensure the confidentiality and integrity of the information being accessed;
- Report any suspected information security violation or breach, and
- Complete all periodic mandatory security trainings in accordance with Mastercard's guidelines.
Mastercard Pune, Mahārāshtra, IND Office



Poona Club Road, Pune, Maharashtra, India, 411001
Similar Jobs at Mastercard
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Develop and maintain scalable cloud-ready applications and microservices using Java, Angular, Spring Boot, REST APIs, and related technologies. Design reusable frameworks, improve software performance and reliability, troubleshoot root causes, conduct code reviews, create proofs of concept, and collaborate in Agile teams. The role also involves introducing technologies, leading engineering improvements, maintaining code quality and automation, and supporting architecture and integration initiatives.
Top Skills:
AgileAngularCloud TechnologiesGitJavaJavaScriptJenkinsMicroservicesPivotal Cloud FoundryRest ApiSpring BootSwagger
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Lead backend engineer responsible for designing, building, and owning Java-based microservices and full-stack applications on PCF/Kubernetes. Deliver high-quality, secure, scalable code, participate in code reviews, work in Agile teams, and follow Mastercard engineering and 12-factor principles. Collaborate with engineers, testers, TPMs, and PMs while leveraging PostgreSQL/Oracle storage, automated testing, and CI/CD practices.
Top Skills:
AngularAutomated TestingCi/CdCSSGitHTML5JavaJavaScriptKubernetesMicroservicesOraclePcfPivotal Cloud FoundryPostgres
Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Design, develop, test, deploy, and support secure, scalable payment-network software. Build Java and Spring Boot microservices, REST APIs, Kafka integrations, and responsive Angular interfaces. Create automated tests, troubleshoot production issues, contribute to CI/CD, observability, resiliency, and automation, and collaborate with globally distributed engineering, product, architecture, and operations teams.
Top Skills:
AngularAutomated TestingCi/CdClaude CliCloud-Native PlatformsContainersCSSEvent-Driven ArchitectureFigmaGitGithub CopilotHTMLJavaKafkaMicro Front End ArchitectureMicroservicesNoSQLRest ApisSpring BootSQLTypescript
What you need to know about the Pune Tech Scene
Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.




