Data First Jobs

InfoSmart Technologies, Inc

Senior PySpark & Databricks Data Engineer

Contract ยท In Office ยท Atlanta, Georgia (USA)

Posted Sep 9, 2026

Work Options
Seniority Level
Cloud Stack
Industry
Positions
Job Type
Position Group
  • Location: Atlanta, Georgia - Hybrid 3 days/onsite
  • Role Summary :
  • We are looking for a skilled Data Engineer with strong hands-on experience in PySpark and Databricks to design, build, and maintain scalable data pipelines. The ideal candidate has deep expertise in stream
  • processing with Apache Kafka, workflow orchestration with Apache Airflow, data warehousing on Amazon Redshift, and building cloud-native solutions on AWS.

Key Responsibilities

  • - Design and develop scalable ETL/ELT pipelines using PySpark on Databricks
  • - Build and maintain real-time data streaming pipelines using Apache Kafka
  • - Orchestrate and schedule data workflows using Apache Airflow (DAG development, monitoring, and troubleshooting)
  • - Manage and optimize data models and queries in Amazon Redshift
  • - Architect and deploy data solutions on AWS using services such as S3, Glue, Lambda, EMR, IAM, and CloudWatch
  • - Collaborate with analysts and platform teams to deliver high-quality data products
  • - Monitor pipeline performance and implement tuning strategies for large-scale data workloads
  • - Implement data quality checks, observability, and alerting across pipelines
  • - Participate in code reviews and contribute to engineering best practices
  • - Document data flows, architecture decisions, and pipeline logic

Required Skills & Experience

  • ๐Ÿ”น PySpark โ€” 3+ years
  • DataFrame API, Spark SQL, query optimizations and performance tuning
  • ๐Ÿ”น Databricks โ€” 3+ years
  • Notebooks, Jobs, Delta Lake, Unity Catalog
  • ๐Ÿ”น Apache Kafka โ€” 2+ years
  • Producers/consumers, Kafka Streams, schema registry
  • ๐Ÿ”น Apache Airflow โ€” 2+ years
  • DAG authoring, task dependencies, operators, scheduling
  • ๐Ÿ”น Amazon Redshift โ€” 2+ years
  • Data modeling, query tuning, Redshift Spectrum
  • ๐Ÿ”น AWS โ€” 3+ years
  • S3, Glue, Lambda, EMR, IAM, CloudWatch, VPC
  • ๐Ÿ”น Python โ€” 4+ years
  • Strong scripting and engineering fundamentals
  • ๐Ÿ”น SQL โ€” Strong
  • Complex queries, window functions, performance tuning
  • ---
  • Nice to Have
  • - Experience with Delta Lake
  • - Familiarity with dbt for transformation layers
  • - Knowledge of Databricks Workflows alongside Airflow
  • - Exposure to Confluent Platform or AWS MSK (Managed Kafka)
  • - Experience with Terraform or AWS CDK for infrastructure-as-code
  • - Understanding of data governance, security best practices, and lake house architecture
  • ---
  • Qualifications
  • - Bachelor's or Master's degree in Computer Science, Engineering, or a related field (or equivalent practical experience)
  • - 4โ€“7 years of overall experience in data engineering roles
  • - Strong problem-solving skills and ability to work independently in an agile environment

Mention you found this on Data First Jobs โ€” it helps us bring you more roles like this.

Senior PySpark & Databricks Data Engineer

InfoSmart Technologies, Inc

Like this role? Get carefully selected jobs like it, twice a week, straight to your inbox.

Free, no spam. Unsubscribe anytime.