InfoSmart Technologies, Inc
Senior PySpark & Databricks Data Engineer
Contract ยท In Office ยท Atlanta, Georgia (USA)
Posted Sep 9, 2026
- Location: Atlanta, Georgia - Hybrid 3 days/onsite
- Role Summary :
- We are looking for a skilled Data Engineer with strong hands-on experience in PySpark and Databricks to design, build, and maintain scalable data pipelines. The ideal candidate has deep expertise in stream
- processing with Apache Kafka, workflow orchestration with Apache Airflow, data warehousing on Amazon Redshift, and building cloud-native solutions on AWS.
Key Responsibilities
- - Design and develop scalable ETL/ELT pipelines using PySpark on Databricks
- - Build and maintain real-time data streaming pipelines using Apache Kafka
- - Orchestrate and schedule data workflows using Apache Airflow (DAG development, monitoring, and troubleshooting)
- - Manage and optimize data models and queries in Amazon Redshift
- - Architect and deploy data solutions on AWS using services such as S3, Glue, Lambda, EMR, IAM, and CloudWatch
- - Collaborate with analysts and platform teams to deliver high-quality data products
- - Monitor pipeline performance and implement tuning strategies for large-scale data workloads
- - Implement data quality checks, observability, and alerting across pipelines
- - Participate in code reviews and contribute to engineering best practices
- - Document data flows, architecture decisions, and pipeline logic
Required Skills & Experience
- ๐น PySpark โ 3+ years
- DataFrame API, Spark SQL, query optimizations and performance tuning
- ๐น Databricks โ 3+ years
- Notebooks, Jobs, Delta Lake, Unity Catalog
- ๐น Apache Kafka โ 2+ years
- Producers/consumers, Kafka Streams, schema registry
- ๐น Apache Airflow โ 2+ years
- DAG authoring, task dependencies, operators, scheduling
- ๐น Amazon Redshift โ 2+ years
- Data modeling, query tuning, Redshift Spectrum
- ๐น AWS โ 3+ years
- S3, Glue, Lambda, EMR, IAM, CloudWatch, VPC
- ๐น Python โ 4+ years
- Strong scripting and engineering fundamentals
- ๐น SQL โ Strong
- Complex queries, window functions, performance tuning
- ---
- Nice to Have
- - Experience with Delta Lake
- - Familiarity with dbt for transformation layers
- - Knowledge of Databricks Workflows alongside Airflow
- - Exposure to Confluent Platform or AWS MSK (Managed Kafka)
- - Experience with Terraform or AWS CDK for infrastructure-as-code
- - Understanding of data governance, security best practices, and lake house architecture
- ---
- Qualifications
- - Bachelor's or Master's degree in Computer Science, Engineering, or a related field (or equivalent practical experience)
- - 4โ7 years of overall experience in data engineering roles
- - Strong problem-solving skills and ability to work independently in an agile environment
Mention you found this on Data First Jobs โ it helps us bring you more roles like this.
Senior PySpark & Databricks Data Engineer
InfoSmart Technologies, Inc
Similar Engineering Jobs
View all Engineering jobsโAgileGrid Solutions
Machine Learning Developer
New
USA
Nexwave
Senior Data Engineering & Platform Engineer (DBT)
New
RemoteUSA
Arkhya Tech. Inc.
Hadoop Data Engineer - Scottsdale AZ (100% Onsite)
New
Scottsdale, Arizona (USA)
Miracle Software Systems, Inc
Sr. Full Stack Data AI Engineer โ GCP & Agentic AI
New
Novi, Michigan (USA)
Skill
Full-Stack Engineer โ Data Experience, Backstage Portal [SK-17196]
New
RemoteUSA$64,000 - $75,000
Jain Global
Solutions Analyst & Engineer
New
New York, New York (USA)
Like this role? Get carefully selected jobs like it, twice a week, straight to your inbox.
Free, no spam. Unsubscribe anytime.