Greetings !
Role:Senior Data Engineer
Experience:10+ Years
Location: Remote (India)
Key Responsibilities:
- Design, develop, and optimize scalable ETL/ELT pipelines using Python, SQL, and PySpark.
- Build batch and real-time data processing solutions.
- Develop distributed data processing applications using Apache Spark.
- Create and maintain cloud-native data platforms on AWS and/or Azure.
- Implement workflow orchestration using Apache Airflow, AWS MWAA, AWS Glue, and Databricks Workflows.
- Build streaming data solutions using Kafka, Kinesis, or Azure Event Hubs.
- Implement data quality, observability, metadata, and lineage frameworks.
- Develop CI/CD pipelines and automate deployments for data engineering solutions.
- Collaborate with architects, analysts, and data scientists to deliver scalable solutions.
- Mentor junior engineers and participate in architecture reviews.
Mandatory Skills:
- 10+ years of Data Engineering experience
- Python
- SQL
- PySpark
- Apache Spark
- ETL/ELT Development
- Batch & Streaming Data Processing
- AWS and/or Azure
- Apache Airflow / AWS MWAA
- AWS Glue
- Databricks
- CI/CD (GitHub Actions or Azure DevOps)
- Git
- Docker
- Data Lake / Lakehouse Architecture
- Data Quality & Data Governance
- Snowflake or Redshift or Azure Synapse