Design, develop, and maintain data solutions and ETL pipelines using AWS services, PySpark, SQL, and Glue. Migrate and transform high-volume data from on-premises systems to AWS, validate data quality, troubleshoot workflows, optimize performance and costs, and support incremental loads. Collaborate with solution designers, DevOps engineers, and cross-functional teams to test, deploy, document, and improve scalable data pipelines.
Project Role : Data Engineer
Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems.
Must have skills : AWS Glue, PySpark
Good to have skills : NA
Minimum 3 year(s) of experience is required
Educational Qualification : 15 years full time education
Summary:
As a Data Engineer, you will design, develop, and maintain data solutions that facilitate data generation, collection, and processing. Your typical day will involve creating data pipelines, ensuring data quality, and implementing ETL processes to migrate and deploy data across various systems. You will collaborate with cross-functional teams to understand data requirements and provide innovative solutions to enhance data accessibility and usability.
We are looking for skilled Data Engineer to join data migration project. In this developer-focused role, you will build and optimize ETL pipelines to migrate and transform data from legacy on-prem systems to AWS, leveraging native AWS transformation technologies
Key Responsibilities
Implement end-to-end ETL pipelines using AWS native services like Step Function, EventBridge, Glue (PySpark, SQL scripts), and Lambda for data extraction, transformation, and loading.
Use pre-created utility & for seamless migration, handling high-volume datasets with error handling and retry mechanisms.
Collaborate with the Solution Designer to test pipelines for performance, and deploy with help of devOps engineer.
Monitor and troubleshoot pipelines using CloudWatch, optimize for cost and process scalability, and document code for team handover.
Support data quality validation and incremental loads to maintain data integrity during the hydration process.
Engage with multiple teams and contribute on key decisions. Provide solutions to problems for their immediate team and across multiple teams.
Professional & Technical Skills:
- Must To Have Skills: Proficiency in AWS Glue, Pyspark, SQL, ETL, Unix, Iceberg, Astronomer, DW concepts
- Strong understanding of data pipeline architecture and design principles.
- Experience with data warehousing solutions and ETL processes.
- Familiarity with cloud computing services and data storage solutions.
- Ability to troubleshoot and optimize data workflows for performance.
Additional Information:
- The candidate should have minimum 3 years of experience in AWS Glue and Pyspark.
- This position is based at our Pune office.
- A 15 years full time education is required.
15 years full time education
Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems.
Must have skills : AWS Glue, PySpark
Good to have skills : NA
Minimum 3 year(s) of experience is required
Educational Qualification : 15 years full time education
Summary:
As a Data Engineer, you will design, develop, and maintain data solutions that facilitate data generation, collection, and processing. Your typical day will involve creating data pipelines, ensuring data quality, and implementing ETL processes to migrate and deploy data across various systems. You will collaborate with cross-functional teams to understand data requirements and provide innovative solutions to enhance data accessibility and usability.
We are looking for skilled Data Engineer to join data migration project. In this developer-focused role, you will build and optimize ETL pipelines to migrate and transform data from legacy on-prem systems to AWS, leveraging native AWS transformation technologies
Key Responsibilities
Implement end-to-end ETL pipelines using AWS native services like Step Function, EventBridge, Glue (PySpark, SQL scripts), and Lambda for data extraction, transformation, and loading.
Use pre-created utility & for seamless migration, handling high-volume datasets with error handling and retry mechanisms.
Collaborate with the Solution Designer to test pipelines for performance, and deploy with help of devOps engineer.
Monitor and troubleshoot pipelines using CloudWatch, optimize for cost and process scalability, and document code for team handover.
Support data quality validation and incremental loads to maintain data integrity during the hydration process.
Engage with multiple teams and contribute on key decisions. Provide solutions to problems for their immediate team and across multiple teams.
Professional & Technical Skills:
- Must To Have Skills: Proficiency in AWS Glue, Pyspark, SQL, ETL, Unix, Iceberg, Astronomer, DW concepts
- Strong understanding of data pipeline architecture and design principles.
- Experience with data warehousing solutions and ETL processes.
- Familiarity with cloud computing services and data storage solutions.
- Ability to troubleshoot and optimize data workflows for performance.
Additional Information:
- The candidate should have minimum 3 years of experience in AWS Glue and Pyspark.
- This position is based at our Pune office.
- A 15 years full time education is required.
15 years full time education
About Accenture
Accenture is a leading global professional services company that helps the world’s leading businesses, governments and other organizations build their digital core, optimize their operations, accelerate revenue growth and enhance citizen services—creating tangible value at speed and scale. We are a talent- and innovation-led company with approximately 791,000 people serving clients in more than 120 countries. Technology is at the core of change today, and we are one of the world’s leaders in helping drive that change, with strong ecosystem relationships. We combine our strength in technology and leadership in cloud, data and AI with unmatched industry experience, functional expertise and global delivery capability. Our broad range of services, solutions and assets across Strategy & Consulting, Technology, Operations, Industry X and Song, together with our culture of shared success and commitment to creating 360° value, enable us to help our clients reinvent and build trusted, lasting relationships. We measure our success by the 360° value we create for our clients, each other, our shareholders, partners and communities.Visit us at www.accenture.com
Equal Employment Opportunity Statement
We believe that no one should be discriminated against because of their differences. All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, military veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by applicable law. Our rich diversity makes us more innovative, more competitive, and more creative, which helps us better serve our clients and our communities.
Accenture Mahārāshtra, IND Office
India
Accenture Pune, Mahārāshtra, IND Office
Building B-1, Magarpatta City (SEZ, Mundhwa Rd, Magarpatta, Hadapsar, Pune, Maharashtra, India, 411013
Similar Jobs
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Analyze business and user requirements across financial services value streams, define technical requirements, develop cloud solution options, and manage business and IT stakeholders. The role requires understanding existing framework capabilities, complex data models, cross-functional solution design, clear documentation, team coordination across regions, and familiarity with the SDLC from design through testing and implementation.
Top Skills:
Google Cloud Platform
Software • Hospitality
Build, test, and maintain scalable ETL/ELT pipelines and data models using Python, PySpark, Databricks, and AWS. Lead data platform features from architecture through deployment, implement CI/CD and monitoring, optimize workloads and cloud costs, and ensure data quality, security, and governance. Collaborate with product managers, data scientists, and engineers while taking technical ownership of major platform components.
Top Skills:
AirflowAWSDatabricksDatabricks Asset BundlesDbtDelta LakeDelta Live TablesPysparkPythonSnowflakeSpark Structured StreamingSQLTerraformUnity Catalog
Automotive
Leads the design, development, deployment, and maintenance of enterprise data platforms and scalable data pipelines. Builds ETL/ELT, streaming, data quality, governance, storage, and analytics solutions using Azure, Databricks, Spark, Python, Scala, and SQL. Optimizes distributed workloads and cloud infrastructure, implements CI/CD and DevOps practices, ensures compliance and security, collaborates with technical and business stakeholders, and mentors data engineering team members.
Top Skills:
Apache IcebergSparkAzure Blob StorageAzure Data Lake StorageAzure DatabricksAzure Sql Data WarehouseAzure Synapse AnalyticsDelta LakeGitHbaseHiveInfrastructure As CodeJenkinsKafkaMapreduceOrcParquetPythonQlik ReplicateScalaSnowflakeSQL
What you need to know about the Pune Tech Scene
Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.



