Synthlane Technologies Logo

Synthlane Technologies

Data Engineer – Mid & Senior Level

Posted 2 Days Ago
Remote
Hiring Remotely in IND
Senior level
Remote
Hiring Remotely in IND
Senior level
Design, build, and support production-grade AWS data pipelines that transform operational data into secure, high-quality, AI-ready datasets. Responsibilities include distributed data processing, Parquet curation, privacy-preserving transformations, orchestration, data quality monitoring, schema management, CI/CD, infrastructure as code, metadata and lineage management, troubleshooting, and reliable backfills.
The summary above was generated by AI

This is a remote position.

Role Overview

We are looking for Data Engineers at Senior and Mid-Level to join our team in building a privacy-preserving data platform where data engineering meets production-grade software engineering.

You will work on designing, developing, and maintaining reliable data pipelines that transform operational data into high-quality, secure, and AI-ready datasets.

Key Responsibilities
  • Build and maintain production-grade data pipelines on AWS.
  • Extract, transform, validate, and curate large-scale Parquet datasets.
  • Implement data de-identification, masking, and privacy-preserving transformations.
  • Design and maintain data pipeline orchestration, scheduling, retries, and backfill mechanisms.
  • Implement comprehensive data quality checks, monitoring, and alerting.
  • Work with workflow orchestration tools such as Airflow, Dagster, or AWS Step Functions.
  • Contribute to CI/CD pipelines and Infrastructure as Code (IaC) practices.
  • Manage schema evolution and schema drift across data sources and pipelines.
  • Provide production support, troubleshooting, and root cause analysis for data pipeline issues.
  • Maintain data catalogs, metadata, and data lineage.
  • Follow software engineering best practices including Git, code reviews, automated testing, and maintainable code.
  • Build reliable and idempotent data pipelines capable of handling retries and large-scale backfills.

Required Skills & Experience
  • Strong proficiency in Python and SQL.
  • Hands-on experience with AWS data services and production data pipelines.
  • Experience with Apache Spark or equivalent distributed data processing technologies.
  • Practical experience with Airflow, Dagster, AWS Step Functions, or similar orchestration tools.
  • Strong understanding of data pipeline architecture, ETL/ELT, and data transformation.
  • Experience working with Parquet and large-scale datasets.
  • Understanding of data quality, schema management, monitoring, and alerting.
  • Strong software engineering practices including:
    • Git and version control
    • Code reviews
    • Automated testing
    • Idempotency
    • Error handling
    • Retries and backfills
  • Experience supporting and troubleshooting production data pipelines.
  • Ability to work effectively with cross-functional engineering and data teams.


Requirements Required Skills

Python | SQL | AWS | Data Engineering | Data Pipelines | ETL/ELT | Apache Spark | Parquet | Airflow | Dagster | AWS Step Functions | Data Orchestration | Data Quality | Schema Management | Data Transformation | Production Support | Git | CI/CD | Automated Testing | Data De-identification | Data Lineage | Data Catalog

Good to Have

Debezium | AWS DMS | Apache Iceberg | Delta Lake | Apache Hudi | Data Masking | Data Tokenization | Terraform | CloudFormation | ML/AI Training Data | Privacy-Preserving Data



Similar Jobs

26 Days Ago
Remote
India
Mid level
Mid level
AdTech
As a Data Engineer, you will design, develop, and maintain data integration solutions, optimize ETL/ELT pipelines, and support data workflows for analytics and reporting.
Top Skills: Azure Data FactoryAzure DatabricksMicrosoft FabricPysparkPythonSpark SqlSQLT-Sql
An Hour Ago
Remote or Hybrid
India
Senior level
Senior level
Aerospace • Artificial Intelligence • Cloud • Machine Learning • Software • Cybersecurity • Defense
Designs, develops, integrates, and validates aerospace control systems and embedded solutions for engine and power systems. Responsibilities include requirements engineering, system architecture, control-system analysis, MATLAB/Simulink modeling, SIL/HIL testing, sensor and actuator integration, embedded C development, and aerospace certification support. The role applies systems engineering, model-based development, real-time embedded concepts, and compliance with ARP4754A, ARP4761, DO-178C, DO-254, and DO-160.
Top Skills: AfdxArinc 429Arm CortexArp4754AArp4761C++CanDo-160Do-178CDo-254DoorsDoors NgEmbedded CHilMatlabMbsePowerpcRs232Rs422/485RtosSilSimulinkStateflowSysmlTi C2000
2 Hours Ago
Remote or Hybrid
Pune, Maharashtra, IND
Mid level
Mid level
Artificial Intelligence • Cloud • Information Technology • Sales • Security • Software • Cybersecurity
Build, operate, and scale cloud infrastructure and deployment pipelines for a secure FedRAMP-authorized environment. Develop automation for reliable microservice operations, manage platform systems and runbooks, monitor infrastructure, improve performance and MTTR, participate in on-call and incident response, and contribute to operational architecture and SRE practices.
Top Skills: AnsibleArgocdAtlantisAWSAzureCi/CdDockerFedrampGCPGitopsGrafanaKubernetesLinuxPrometheusTerraform

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account