- Job Summary
- We are seeking a Data Engineer with 5+ years of experience in designing, developing, and optimizing scalable data pipelines and cloud-based data platforms. The ideal candidate will have strong expertise in Python, SQL, Databricks, Apache Spark, and ETL/ELT development, along with hands-on experience in at least one cloud platform (AWS, Azure, or Google Cloud Platform).
- Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines.
- Build and optimize data ingestion, transformation, and integration workflows using Databricks and Apache Spark.
- Develop and maintain data lakes and cloud-based data warehouses.
- Create scalable batch and real-time data processing solutions.
- Optimize data pipelines for performance, reliability, and scalability.
- Develop reusable data engineering frameworks and automation solutions.
- Collaborate with data analysts, data scientists, and business stakeholders to deliver high-quality data solutions.
- Implement data quality, governance, monitoring, and security best practices.
- Troubleshoot production issues and continuously improve data platform performance.
- Follow DevOps and CI/CD best practices for data engineering projects.
- Required Qualifications
- Bachelor’s or Master’s degree in Computer Science, Information Technology, Engineering, or a related field.
- 5+ years of experience as a Data Engineer.
- Strong programming skills in Python.
- Advanced SQL proficiency.
- Hands-on experience with Databricks.
- Strong experience with Apache Spark (PySpark preferred).
- Experience building and maintaining ETL/ELT pipelines.
- Experience with Apache Airflow or similar workflow orchestration tools.
- Experience with data warehousing technologies such as Snowflake, Amazon Redshift, Google BigQuery, or Azure Synapse Analytics.
- Hands-on experience with at least one cloud platform (AWS, Azure, or Google Cloud Platform).
- Experience with Git, Docker, and CI/CD pipelines.
- Knowledge of relational and NoSQL databases.
- Preferred Qualifications
- Experience with Kafka or other streaming platforms.
- Experience with Delta Lake.
- Familiarity with Apache Iceberg or Apache Hudi.
- Experience with Infrastructure as Code (Terraform or CloudFormation).
- Knowledge of data governance and data quality frameworks.
- Exposure to DevOps and MLOps practices.
Mention you found this on Data First Jobs — it helps us bring you more roles like this.
Data Engineer
ATC
Similar Engineering Jobs
View all Engineering jobs→University of Washington
Research Scientist/Engineer 3
New
RemoteUSA
Meta
Data Engineer, Product Analytics
New
Fremont, California (USA)$177,000 - $247,000
Seagate Technology
Sr Engineer - Early Career - Experimental Data Storage Physics
New
Bloomington, Minnesota (USA)
First Citizens Bank
AI Platform Director - Data Engineering
New
RemoteUSA
General Dynamics Information Technology
Senior Systems Analyst (Proposal Engineer)
New
Springfield, Virginia (USA)
ST Global Tech LLC
Lead AI Engineer / Data Scientist - Local to CA Only
New
Santa Clara, California (USA)
Like this role? Get carefully selected jobs like it, twice a week, straight to your inbox.
Free, no spam. Unsubscribe anytime.