Aligned Automation Logo

Aligned Automation

Pyspark developer

Posted 27 Days Ago
Be an Early Applicant
In-Office
Pune, Maharashtra, IND
Senior level
In-Office
Pune, Maharashtra, IND
Senior level
Design, develop, and optimize large-scale PySpark ETL/ELT pipelines on Databricks. Implement Delta Lake lakehouse architectures, tune Spark jobs, write complex SQL, monitor production, enable CI/CD, mentor junior engineers, and collaborate with stakeholders on scalable data solutions.
The summary above was generated by AI

About the Job

About Aligned Automation

At Aligned Automation, we live by our "Better Together" philosophy to build a better world. As a strategic service provider to Fortune 500 companies, we help digitize enterprise operations and drive impactful business strategies. Our purpose goes beyond projects—we strive to deliver meaningful, sustainable change that shapes a more optimistic and equitable future.

Our culture is deeply rooted in our 4Cs—Care, Courage, Curiosity, and Collaboration—ensuring that each employee is empowered to grow, innovate, and thrive in an inclusive workplace.

Senior Data Engineer – PySpark
Experience: 7–8 Years
Location: Pune (Work from office)
Job Summary

We are seeking a highly skilled Senior PySpark Data Engineer with 7–8 years of experience in designing, developing, and optimizing large-scale data engineering solutions. The ideal candidate should have extensive experience with PySpark, Python, SQL, Databricks, cloud platforms, and modern data architectures. The role requires hands-on development, solution design, client interaction, and mentoring junior team members.

Key Responsibilities
  • Design, develop, and maintain scalable ETL/ELT pipelines using PySpark.
  • Build high-performance data pipelines to process large volumes of structured and semi-structured data.
  • Develop reusable frameworks for data ingestion, transformation, validation, and monitoring.
  • Optimize Spark jobs by tuning partitions, joins, caching, and memory configurations.
  • Design and implement Data Lake/Lakehouse architectures using Delta Lake.
  • Write complex SQL queries, stored procedures, CTEs, and window functions.
  • Collaborate with business stakeholders, architects, and data analysts to understand business requirements.
  • Participate in architecture discussions and recommend scalable data solutions.
  • Perform code reviews and enforce coding standards and best practices.
  • Monitor production pipelines, troubleshoot issues, and implement performance improvements.
  • Work with DevOps teams to implement CI/CD for data pipelines.
  • Mentor junior engineers and provide technical leadership.
  • Estimate effort, prepare technical documentation, and participate in Agile ceremonies.
Required Technical Skills
Programming
  • Python (Advanced)
  • PySpark (Advanced)
  • SQL (Advanced)
Big Data Technologies
  • Apache Spark
  • Spark SQL
  • Delta Lake
  • Parquet
  • iceberg
  • IOMETE
Databases
  • SQL Server
  • PostgreSQL
  • Azure SQL
Cloud Platforms (Any One)
  • Azure
  • AWS
Data Engineering Tools
  • Databricks
  • Apache Airflow
Version Control & DevOps
  • Git
  • Azure DevOps / GitHub
  • CI/CD Pipelines


Similar Jobs

6 Days Ago
Hybrid
Junior
Junior
Financial Services
Join an agile Employee Platforms team to design, develop, test, and troubleshoot secure, scalable software. Build ETL pipelines, perform database and ETL testing, write high-quality code, use CI/CD and AI-assisted development tools, and collaborate on SDLC, resiliency, and security practices.
Top Skills: SparkAWSCallidusCi/CdDatabase TestingETLOracle IcmPysparkPythonSQLVaricentXactly
10 Days Ago
In-Office
Pune, Maharashtra, IND
Senior level
Senior level
Artificial Intelligence • Analytics • Business Intelligence • Consulting
Design and maintain scalable data pipelines using Python and PySpark; build FastAPI REST APIs; develop SQL queries and transformations; apply data warehousing, modeling, and ETL/ELT practices; support CI/CD deployments; perform data profiling and quality checks; gather requirements; document solutions; and contribute to architecture and process improvements.
Top Skills: Ci/CdFastapiGitPysparkPythonSQL
20 Days Ago
In-Office
Pune, Maharashtra, IND
Mid level
Mid level
Fintech • Financial Services
Design, build and maintain data pipelines, data warehouses and lakes using PySpark, Snowflake and AWS analytics services. Implement scalable ETL/ELT, collaborate with data scientists for ML deployment, ensure data governance, security and high-quality deliverables while leading or advising team members.
Top Skills: AbinitioAlationApache AirflowAthenaAWSCsvDbtGlueIcebergImmutaJSONLake FormationLambdaNoSQLParquetPl/SqlPysparkRddS3SnowflakeSnowflake TasksSpark DataframesSparksqlSQL

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account