Job ID: z5G7h3l6a1kMvyS65NP3c90KoA8p0AXa1woVYBCQVOA=Job Code: JPC Location: Peoria, IL, 61654, Plano, TX, 75094, Dallas, TX, 75342, San Jose, CA, 95117, Edison, NJ, 08906, United StatesExperience: 5-10 yearsContact: Vighneshwar GoudaContact Email: are seeking an experienced Data Engineer to design, build, and optimize scalable data pipelines and cloud-based data platforms. The ideal candidate will have strong expertise in Python, SQL, Spark, ETL processes, and cloud technologies to support data-driven decision-making and advanced analytics.Key ResponsibilitiesDesign, develop, and maintain scalable data pipelines and ETL/ELT workflows.Build and optimize data processing solutions using Python, SQL, and Apache Spark (PySpark).Develop batch and real-time data ingestion pipelines.Work with structured and unstructured datasets from multiple data sources.Design and implement cloud-based data solutions using Azure, AWS, or GCP.Develop and maintain data models, data lakes, and enterprise data warehouses.Integrate data from APIs, databases, and streaming platforms.Optimize data pipeline performance and ensure data quality.Collaborate with Data Scientists, Business Analysts, and Software Engineers.Implement CI/CD pipelines and DevOps best practices for data engineering.Troubleshoot production issues and provide ongoing support.Ensure compliance with data governance, security, and privacy standards.Required Skills5+ years of experience as a Data Engineer.Strong programming skills in Python.Excellent SQL skills.Hands-on experience with Apache Spark / PySpark.Experience with ETL/ELT development.Strong knowledge of Data Warehousing concepts.Experience with Hadoop ecosystem (Hive, HDFS).Experience with Databricks.Experience with Apache Kafka.Hands-on experience with Apache Airflow or similar workflow orchestration tools.Experience working on Linux/Unix environments.Knowledge of Git and CI/CD pipelines.Cloud SkillsExperience with one or more of the following:Microsoft Azure (ADF, ADLS, Synapse)Amazon Web Services (AWS)Google Cloud Platform (GCP)Preferred SkillsSnowflakeDelta LakeDockerKubernetesTerraformScala or JavaPower BI or TableauREST APIsAgile/Scrum methodologyExperience working with large-scale enterprise data platformsExperience5-10 yearsSkillsPYTHON, SQL, Apache Spark, PySpark, Hadoop, Hive, ETL, ELT, Data Pipeline Development, Apache Kafka, Databricks, Azure Data Factory (ADF), Azure Synapse, Snowflake, Azure, AWS, GCP, Delta Lake, DATA WAREHOUSING, Data Modeling, LINUX, UNIX, Airflow, REST APIs, Git, CI/CD
Data Engineer in new brunswick at Unknown Company
This position is listed as full time and onsite.