This is a Fully Remote Job
About Our Client:
The organization operates within the data engineering and analytics space, focusing on delivering reliable data pipelines and curated datasets to support reporting, business decision-making, and artificial intelligence initiatives. Addressing challenges in data accuracy and accessibility, the company enables teams to work with large datasets effectively for analytics and AI applications.
About the Opportunity:
The DATA ENGINEER role is centered on building and maintaining data pipelines from ingestion through transformation and delivery. This position is responsible for ensuring data accuracy and accessibility to support analytics, AI, and business intelligence needs. By collaborating with analysts, data scientists, and engineers, the role significantly contributes to the quality and usability of data assets that drive organizational decisions and AI capabilities.
Responsibilities:
- Build and maintain ETL/ELT pipelines using SQL, Python, dbt, and tools like Apache Spark, PySpark, or Airflow.
- Ingest data from databases, APIs, SaaS tools, and event streams through connectors or custom pipelines.
- Develop tested data models and curated datasets in cloud platforms such as Snowflake, Databricks, BigQuery, or Redshift.
- Collaborate with data scientists and ML engineers to prepare feature datasets for model training and inference.
- Prepare and refresh structured data for AI search, retrieval-augmented generation, and other Generative AI applications.
- Create dashboards and analyses using Tableau, Looker, Power BI, or similar tools when needed.
- Implement data quality checks, monitoring, and documentation to ensure data trustworthiness and early issue detection.
- Optimize query performance and pipeline costs; use Git, code reviews, and CI/CD for safe releases.
- Manage data access, lineage, and sensitive information including personally identifiable information (PII).
Requirements:
- 3 to 6 years of experience in data engineering, analytics engineering, or related data roles with production delivery.
- Strong SQL skills with experience in complex transformations and query performance optimization.
- Hands-on experience with cloud data platforms; preferred platforms include Snowflake, Databricks, BigQuery, or Redshift.
- Production experience with dbt for transformation, testing, and documentation.
- Proficiency in Python and Pandas or PySpark for data processing.
- Experience scheduling pipelines using orchestration tools such as Airflow, Dagster, or Prefect.
- Solid understanding of data modeling and building usable datasets for analysts and business teams.
Preferred Experience:
- Experience with Apache Spark, Databricks, and large-scale data processing.
- Familiarity with data quality or observability tools like Great Expectations, Soda, or Monte Carlo.
- Knowledge of streaming data platforms such as Kafka or Kinesis and ingestion tools like Fivetran or Airbyte.
- Experience preparing data for machine learning features, AI search, embeddings, or retrieval-augmented generation applications.
- Experience with cloud services across AWS, Azure, or Google Cloud and data governance tools such as Unity Catalog or DataHub.
Pay Range and Compensation Package:
The pay range and compensation package for this role will be determined based on the candidate's experience, skills, and other relevant factors.
Benefits & Perks:
- Unlimited paid time off.
- Generous parental leave exceeding industry standards.
- Medical, Dental, and Vision insurance coverage.
- Disability and Life insurance access.
- Mental health and wellbeing support.
- Annual bonus program.
- Employer Stock Purchase Program.
- Yearly team-building experiences.
- Mentorship and sponsorship opportunities.
- Manager resources and support.
Equal Opportunity Statement: Our client is an equal opportunity employer. They celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, or national origin.
Note:
TalentHop is a recruitment partner of this role. Please note that all employment decisions, including candidate assessment, interviews, hiring, compensation, and employment terms, are made exclusively by the hiring employer.
Mention you found this on Data First Jobs — it helps us bring you more roles like this.
Data engineer - USA
TalentHop
Similar Engineering Jobs
View all Engineering jobs→Cambridgeconsultantslimited
Graduate AI Engineer (2027 start)
Cambridgeconsultantslimited
Graduate Human-Centric AI Engineer (2027 start)
Form3
Senior Cloud Security Engineer - AI Resilience & Security Enhancements (Fixed-term contract)
Meta
Data Center Infrastructure Management (DCIM) Engineer
Ent
Senior AI / ML Engineer
Tenova
R&D Engineer- Data Scientist
Like this role? Get carefully selected jobs like it, twice a week, straight to your inbox.
Free, no spam. Unsubscribe anytime.