595,900open jobs
29,787companies
85,713added this week
Browse all
Salary
$100k – $150k per year
Location
Remote (United States)
Seniority
Senior · 6+ years exp
Overview
Company
Impact
Profile match
Bright Vision Technologies is an IT consulting, enterprise technology services, and workforce solutions enterprise. Headquartered in Bridgewater, New Jersey, United States, the minority-owned firm specializes in technology staffing, cybersecurity, application management, and digital product engineering. Founded in 2020, the enterprise delivers specialized staffing and IT services alongside proprietary automation and AI software - including its flagship enterprise talent intelligence platform, Lumina - serving clients across information technology, defense, healthcare, government, and manufacturing sectors.

AI Pipeline Engineer- Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title:AI Pipeline Engineer

Location: 100% Remote (U.S.)

PositionType:Full-time, Direct W2

Salary Range:$100,000-$150,000 Annually

ExperienceRequired:6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

We are seeking an AI Pipeline Engineer to build and operate the large-scale data systems that power modern AI training and evaluation pipelines. The role combines deep data engineering expertise with a strong understanding of AI workloads, focusing on ingestion, transformation, quality assurance, lineage, and high-throughput delivery of data to training jobs across diverse modalities. The ideal candidate has experience operating petabyte-scale data systems, strong software engineering fundamentals, and clear understanding of how data infrastructure choices propagate into model quality and training efficiency.

Key Responsibilities

  • Design and operate large-scale data pipelines supporting AI training, evaluation, and continual improvement workflows.
  • Build ingestion systems for diverse modalities including text, image, audio, video, and structured signals.
  • Implement data cleaning, deduplication, filtering, and quality assurance at petabyte scale.
  • Develop dataset versioning, lineage, and provenance tracking systems suitable for reproducible training.
  • Build high-throughput data loading systems that maximize GPU utilization during training.
  • Implement labeling workflows, active learning pipelines, and human-in-the-loop data improvement systems.
  • Design storage architectures balancing cost, throughput, and latency across data tiers.
  • Build evaluation dataset construction pipelines with strict integrity and contamination controls.
  • Implement data privacy, redaction, and consent enforcement throughout the pipeline.
  • Collaborate with ML researchers and engineers to align data systems with model development needs.
  • Drive observability of data quality, drift, and pipeline health across the AI data estate.
  • Optimize cost and performance through compression, format selection, and caching strategies.
  • Document data systems, schemas, and operational procedures for broad internal use.
  • Stay current with AI data infrastructure research and emerging open-source tools.
Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science or a related field.
  • Six or more years of data engineering experience, with significant work supporting ML or AI workloads.
  • Strong proficiency in Python and at least one JVM or systems language.
  • Deep experience with modern data processing frameworks such as Spark, Ray, or Beam.
  • Hands-on experience operating petabyte-scale storage and pipeline systems.
  • Strong understanding of distributed systems, data modeling, and storage formats.
  • Experience with dataset versioning, lineage, and reproducibility for ML workflows.
  • Familiarity with high-throughput data loading for accelerator-based training.
  • Strong software engineering practices including testing, CI/CD, and code review.
  • Excellent communication and cross-functional collaboration skills.
Preferred Qualifications
  • Experience with multimodal datasets at large scale.
  • Familiarity with data quality tooling and dataset evaluation methodology.
  • Exposure to privacy-preserving data systems and regulated data handling.
  • Open-source contributions to data infrastructure projects.
  • Experience supporting frontier model training pipelines.

How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at www.bvteck.com.

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
595,900 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
In office
Python
SQL
Databases
PostgreSQL
Redis
DevOps
Rest API
GitHub Actions
containerd
CI/CD
Jenkins
Git
Docker
GitHub
QA
Selenium
Swagger
Robot Framework
Apply
$140k – $273k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • McLean
Python
Java
SQL
Scala
Databases
MySQL
MongoDB
Snowflake
Cassandra
Apache Kafka
Amazon Redshift
AI/ML
Hadoop
Spark
DevOps
GCP
Azure
AWS
Management
Agile
Apply
$209k – $239k per year • In office • Full-Time • 9+ years exp • Bachelor's Degree • Wilmington
Python
Java
SQL
Scala
Databases
MySQL
MongoDB
Snowflake
Cassandra
Apache Kafka
Amazon Redshift
AI/ML
Hadoop
Spark
DevOps
GCP
Azure
AWS
Management
Agile
Apply
$32k – $72k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Mexico City
Python
SQL
Databases
Snowflake
Google BigQuery
Amazon Redshift
BigQuery
Analytics
Tableau
Power BI
Looker
Management
Linear
Apply
In office • 6+ years exp
Java
Java
Spring Boot
Hibernate
Micronaut
Databases
MySQL
DynamoDB
DevOps
Splunk
CloudFormation
Datadog
CI/CD
AWS
Docker
Kubernetes
AWS Lambda
Amazon S3
Amazon ECS
API Gateway
Apply
$175k – $200k per year • Remote • 8+ years exp • Bachelor's Degree
Go
Java
C#
DevOps
Istio
Envoy
Linkerd
Kubernetes
Platform Engineering
Service Mesh
FinOps
Apply
$175k – $200k per year • Remote • 8+ years exp • Bachelor's Degree
DevOps
GCP
Azure
AWS
FinOps
Management
ServiceNow
Apply
$115k – $175k per year • Remote • 5+ years exp • PhD
Python
SQL
PowerShell
Python
pySpark
Databases
PostgreSQL
Snowflake
Databricks
Apache Kafka
Google BigQuery
Amazon Redshift
Teradata
BigQuery
AI/ML
Spark
Airflow
DevOps
GCP
Azure DevOps
Azure
CI/CD
Jenkins
Git
AWS
Amazon Kinesis
Analytics
Tableau
Power BI
ETL/ELT
Informatica
Talend
SSIS
DataStage
Pentaho
Azure Data Factory
AWS Glue
Data Vault
Dimensional Modeling
Master Data Management
Management
Agile
Scrum
Apply
$140k – $200k per year • Remote • 6+ years exp • PhD
Go
JavaScript
TypeScript
SQL
C#
Node JS
Solidity
Solidity
Truffle
Ganache
Brownie
Databases
PostgreSQL
Redis
AI/ML
Tokenization
Frontend
Vue.js
GraphQL
Angular
React.js
Remix
DevOps
Rest API
Terraform
WebSockets
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Web3
Foundry
Polygon
Hardhat
Remix
Avalanche
Optimism
Arbitrum
MetaMask
Layer 2
Smart Contracts
Ethereum
Hyperledger Fabric
Reown
Token Standards
Coinbase Wallet
Management
Agile
Scrum
Apply
$120k – $180k per year • Remote • 5+ years exp • PhD
JavaScript
Java
TypeScript
Java
Maven
Spring Boot
Spring MVC
Spring Data JPA
Spring Security
Spring Cloud
Databases
Apache Kafka
Frontend
Webpack
GraphQL
Tailwind CSS
Yarn
RxJS
Angular
Bootstrap
npm
PrimeNG
Sass
NgRx
Angular Material
Lighthouse
Mobile
Dependency Injection
State Management
Offline-First
PWA
DevOps
Rest API
GCP
OpenShift
Azure DevOps
GitHub Actions
WebSockets
GitLab CI
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Bitbucket
GitHub
GitLab
API Gateway
Cybersecurity
Keycloak
Okta
Auth0
Design
Figma
Adobe XD
Management
Agile
Scrum
QA
Selenium
Cypress
Playwright
Swagger
Chrome DevTools
Apply
See all jobs
This is one of many
595,900 more open roles from verified company boards, updated every day.