Data First Jobs

AgileGrid Solutions

Data Engineer

Full Time · In Office · USA

Posted Jun 28, 2026

About The Company

Bright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications. Our commitment to excellence and innovation drives us to deliver high-quality software products that meet the evolving needs of our clients across various industries. At Bright Vision Technologies, we foster a collaborative and inclusive work environment, encouraging continuous learning and professional growth. Our team of talented engineers and developers work together to solve complex problems and push the boundaries of technology, ensuring our clients stay ahead in a competitive landscape. We are passionate about harnessing the power of technology to create meaningful impact and deliver value to our customers worldwide.

About The Role

We are seeking an experienced AI Data Engineer to join our dynamic team. In this role, you will be responsible for designing, building, and maintaining large-scale data systems that support modern AI training and evaluation pipelines. You will work closely with machine learning researchers and engineers to ensure the seamless ingestion, transformation, and delivery of diverse data modalities, including text, images, audio, video, and structured signals. The ideal candidate will have extensive experience operating petabyte-scale data infrastructure, with a strong foundation in software engineering principles and distributed systems. Your work will directly influence the quality, efficiency, and reproducibility of AI models, making this a critical role within our organization. This position offers the flexibility of a fully remote work environment within the continental United States and provides an opportunity to contribute to cutting-edge AI projects in a long-term, multi-year engagement.

Qualifications

The ideal candidate will possess a bachelor’s or master’s degree in Computer Science or a related field, coupled with at least six years of dedicated data engineering experience supporting machine learning or AI workloads. Proficiency in Python is essential, along with experience in at least one JVM or systems programming language. Candidates should have hands-on experience with modern data processing frameworks such as Spark, Ray, or Beam, and a proven track record of managing petabyte-scale storage and pipeline systems. A deep understanding of distributed systems, data modeling, and storage formats is required. Familiarity with dataset versioning, lineage, and reproducibility for ML workflows is highly desirable. Strong software engineering practices, including testing, continuous integration/deployment, and code review, are essential. Excellent communication skills and the ability to collaborate effectively across teams are also important. Preferred qualifications include experience with multimodal datasets, data quality tooling, privacy-preserving data systems, open-source contributions, and experience supporting frontier model training pipelines.

Responsibilities

Design and operate large-scale data pipelines supporting AI training, evaluation, and continuous improvement workflows. Build robust ingestion systems capable of handling diverse modalities such as text, image, audio, video, and structured signals. Implement data cleaning, deduplication, filtering, and quality assurance processes at petabyte scale to ensure data integrity. Develop dataset versioning, lineage, and provenance tracking systems to facilitate reproducible training processes. Build high-throughput data loading systems optimized for GPU utilization during training sessions. Design and implement labeling workflows, active learning pipelines, and human-in-the-loop data improvement systems to enhance dataset quality. Architect storage solutions that balance cost, throughput, and latency across various data tiers. Construct evaluation dataset pipelines with strict controls to prevent contamination and ensure data integrity. Enforce data privacy, redaction, and consent policies throughout the data lifecycle. Collaborate closely with machine learning teams to align data infrastructure with model development needs, ensuring seamless data flow and accessibility. Drive observability initiatives to monitor data quality, detect drift, and maintain pipeline health across the entire data estate. Optimize system performance through compression, format selection, and caching strategies to reduce costs and improve efficiency. Document all data systems, schemas, and operational procedures comprehensively for internal reference and knowledge sharing. Stay updated with the latest research and emerging open-source tools in AI data infrastructure to continuously enhance system capabilities.

Benefits

Bright Vision Technologies offers a competitive salary range of $100,000 to $150,000 per annum, commensurate with experience. As a full-time W2 employee, you will enjoy comprehensive benefits including health, dental, and vision insurance, paid time off, and opportunities for professional development. Our remote work environment provides flexibility and work-life balance, allowing you to collaborate effectively from anywhere within the continental United States. We foster a culture of innovation, inclusion, and continuous learning, encouraging our team members to grow their skills and advance their careers. Additionally, we support a collaborative work environment where your contributions impact cutting-edge AI projects. Our long-term engagement model ensures stability and opportunities for ongoing involvement in transformative technology initiatives. Bright Vision Technologies is committed to maintaining a supportive and inclusive workplace, where diversity and equal opportunity are celebrated.

Equal Opportunity (EEO)

Bright Vision Technologies (BV Teck) is committed to providing equal employment opportunities to all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. We prohibit workplace harassment and discrimination and ensure that all employment decisions are made based on merit and business needs. Our commitment extends to recruitment, hiring, training, promotion, transfer, and all other aspects of employment. We strive to create an inclusive environment where everyone can thrive and contribute their best. Bright Vision Technologies is an equal opportunity employer, including Disability and Veterans.

Mention you found this on Data First Jobs — it helps us bring you more roles like this.

Data Engineer

AgileGrid Solutions

Like this role? Get carefully selected jobs like it, twice a week, straight to your inbox.

Free, no spam. Unsubscribe anytime.