AI Data Engineer
Back to Jobs
BrightvisionGet Smart Job AI Coach in the appFree on iOS and Android

AI Data Engineer
80,000–100,000 / Year
Location
Remote
Experience
Senior
Posted
Jul 30, 2026
Apply by
August 29, 2026
Applicants
0
Early applicantEasy applyFull-timeWork from Home
Job Description
AI Data Engineer – Remote
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: AI Data Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $80,000–$100,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
We are seeking an AI Data Engineer to build and operate the large-scale data systems that power modern AI training and evaluation pipelines. The role combines deep data engineering expertise with a strong understanding of AI workloads, focusing on ingestion, transformation, quality assurance, lineage, and high-throughput delivery of data to training jobs across diverse modalities. The ideal candidate has experience operating petabyte-scale data systems, strong software engineering fundamentals, and clear understanding of how data infrastructure choices propagate into model quality and training efficiency.
Required Qualifications
- Bachelor’s or Master’s degree in Computer Science or a related field.
- Six or more years of data engineering experience, with significant work supporting ML or AI workloads.
- Strong proficiency in Python and at least one JVM or systems language.
- Deep experience with modern data processing frameworks such as Spark, Ray, or Beam.
- Hands-on experience operating petabyte-scale storage and pipeline systems.
- Strong understanding of distributed systems, data modeling, and storage formats.
- Experience with dataset versioning, lineage, and reproducibility for ML workflows.
- Familiarity with high-throughput data loading for accelerator-based training.
- Strong software engineering practices including testing, CI/CD, and code review.
- Excellent communication and cross-functional collaboration skills.
Preferred Qualifications
- Experience with multimodal datasets at large scale.
- Familiarity with data quality tooling and dataset evaluation methodology.
- Exposure to privacy-preserving data systems and regulated data handling.
- Open-source contributions to data infrastructure projects.
- Experience supporting frontier model training pipelines.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [\[email protected\]](/cdn-cgi/l/email-protection). Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Equal Employment Opportunity (EEO) Statement
Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.
BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Key Responsibilities
- Build and operate large-scale data systems for AI training and evaluation pipelines.
- Manage data ingestion, transformation, quality assurance, and lineage.
- Ensure high-throughput delivery of data to training jobs across diverse modalities.
- Operate petabyte-scale storage and pipeline systems.
- Implement dataset versioning, lineage, and reproducibility for ML workflows.
- Support high-throughput data loading for accelerator-based training.
- Apply strong software engineering practices including testing, CI/CD, and code review.
Requirements
- Bachelor's or Master's degree in Computer Science or a related field
Skills Required
PythonJVMSystems LanguageSparkRayBeamDistributed SystemsData ModelingStorage FormatsDataset VersioningLineageCI/CDTestingCommunicationCross-functional collaborationMultimodal DatasetsData Quality ToolingDataset EvaluationPrivacy-Preserving Data SystemsRegulated Data HandlingOpen-Source ContributionsFrontier Model Training Pipelines
App exclusive · Free
Smart Job AI Coach
Your personal interview coach on every job — readiness tips, profile improvements, and role-specific prep. Available only in the Pulse Job app.
Interview readiness
See how prepared you are and what to improve for each role.
Personalized tips
Actionable suggestions based on your profile and the job.
After you apply
Keep coaching momentum from job detail through application success.




