1,431,395open jobs
83,518companies
216,630added this week
Browse all
Salary
$130k – $180k per year
Location
Remote (United States)
Employment
Full-Time

First seen by Alion on Oct 8, 2026.

Overview
Company
Impact
Profile match
Bright Vision Technologies is a minority-owned, product-focused AI technology company headquartered in Bridgewater, New Jersey. Founded in July 2020, we are a pure-play product engineering firm dedicated to transforming the staffing and enterprise automation industries through our flagship platform — Lumina.Lumina is an advanced AI-powered talent intelligence and enterprise automation platform that integrates artificial intelligence, hybrid cloud, and blockchain technologies to revolutionize how organizations source talent and automate complex workflows.What Lumina delivers:- AI-driven candidate sourcing, semantic resume matching, and intelligent ranking- Generative AI and Large Language Model orchestration with LangChain and RAG-based architecture- Automated recruiter assistant and workflow automation- Blockchain-enabled credential verification and tamper-proof digital identity- Advanced optimization for talent supply chain, workforce planning, and resource allocation- Enterprise-grade security with role-based access, encryption, and comprehensive audit logging- Seamless integration with CRM, ERP, HRMS, ATS, and legacy enterprise systemsBuilt on a scalable microservices architectu
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: AI Platform Engineer
Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $100,000 – $150,000 per annum
Experience: 6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary:
We are seeking an AI Platform Engineer to design, build, and operate scalable AI inference platforms for production ML workloads. The ideal candidate will have expertise in distributed systems, LLM serving, GPU optimization, autoscaling, and cloud-native infrastructure, with a strong focus on performance, reliability, and observability.

Key Responsibilities:
  • Design and maintain scalable AI model serving platforms.
  • Optimize inference performance, GPU utilization, and request routing.
  • Build autoscaling, deployment, and monitoring solutions.
  • Implement caching, security, and high-availability strategies.
  • Collaborate with ML teams to deploy and support production AI models.
Required Qualifications:
  • 6+ years of experience in distributed systems, infrastructure, or ML platform engineering.
  • Strong proficiency in Python and Go, Rust, or C++.
  • Experience with LLM inference frameworks (vLLM, TensorRT-LLM), Kubernetes, cloud platforms, and GPU optimization.
Preferred Qualifications:

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job TitleAI Platform Engineer

Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $130,000–$180,000 Annually (based on experience)
Experience Required:10+ Years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

Bright Vision Technologies is seeking a highly experienced AI Platform Engineer with 10+ years of experience in distributed systems, cloud-native infrastructure, and AI platform engineering to design, build, and operate enterprise-scale AI inference and machine learning platforms. The ideal candidate will possess deep expertise in LLM serving, GPU optimization, Kubernetes, cloud infrastructure, distributed systems, and MLOps, with a proven ability to deliver highly scalable, reliable, secure, and cost-efficient AI platforms supporting production machine learning workloads.

Key Responsibilities
  • Design, build, and maintain scalable AI inference and model-serving platforms for enterprise production environments.
  • Architect highly available, cloud-native infrastructure supporting Large Language Models (LLMs), foundation models, and machine learning services.
  • Optimize inference latency, throughput, GPU utilization, memory management, and request scheduling across distributed AI workloads.
  • Design autoscaling, workload orchestration, traffic management, and intelligent request routing strategies for AI services.
  • Implement model deployment, versioning, rollback, and lifecycle management using modern MLOps practices.
  • Develop monitoring, observability, logging, distributed tracing, and alerting solutions to ensure platform reliability and performance.
  • Implement caching strategies, API gateways, security controls, authentication, authorization, and high-availability architectures.
  • Collaborate with AI researchers, ML engineers, DevOps teams, and software engineers to deploy and support production AI models.
  • Drive cloud infrastructure optimization, resource utilization, FinOps initiatives, and operational excellence.
  • Mentor engineering teams, conduct architecture reviews, and establish best practices for AI platform engineering and cloud-native development.
  • Evaluate emerging AI infrastructure technologies, model-serving frameworks, and GPU acceleration techniques to drive continuous innovation.
Required Qualifications
  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Artificial Intelligence, or a related technical discipline.
  • 10+ years of professional experience in distributed systems, infrastructure engineering, cloud platforms, or machine learning platform engineering.
  • Strong programming skills in Python and at least one systems programming language such as Go, Rust, or C++.
  • Extensive experience with Large Language Model (LLM) serving, model inference optimization, and production AI infrastructure.
  • Hands-on experience with vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve, or similar AI serving frameworks.
  • Strong expertise in Kubernetes, container orchestration, Docker, and cloud-native application architectures.
  • Experience optimizing GPU workloads using CUDA, NVIDIA GPU technologies, distributed inference, and high-performance AI infrastructure.
  • Experience with cloud platforms including AWS, Microsoft Azure, or Google Cloud Platform (GCP).
  • Strong understanding of distributed systems, networking, scalability, observability, and security best practices.
  • Excellent analytical, communication, collaboration, and technical leadership skills.
Preferred Qualifications
  • Experience designing and operating multi-region AI platforms and globally distributed inference services.
  • Knowledge of model optimization techniques such as quantization, pruning, compression, speculative decoding, KV cache optimization, and mixed-precision inference.
  • Experience with MLOps, GitOps, Infrastructure as Code (Terraform, Bicep, CloudFormation), and CI/CD automation.
  • Familiarity with service mesh technologies such as Istio or Linkerd, API gateways, and event-driven architectures.
  • Contributions to open-source AI infrastructure projects, technical publications, patents, or conference presentations.
  • Experience implementing FinOps strategies, cloud cost optimization, and enterprise AI governance.
  • Experience with multi-region AI deployments and AI infrastructure.
  • Familiarity with model optimization techniques such as quantization or compression.
  • Open-source contributions or experience supporting large-scale AI APIs.

Interested in this opportunity? Apply today for immediate consideration!

Email your updated resume to
or Text (908) 505-3545 if you have any questions.
Learn more about Bright Vision Technologies at .

We look forward to connecting with talented professionals and helping you take the next step in your career.

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Originally posted on Himalayas

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,431,395 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
≈ $175k – $330k per year (Estimated) • Remote (United States) • 8+ years exp
Python
AI/ML
NLP
LLM Guardrails
Machine Learning
Apply
Remote (likely Germany) • Internship
Python
AI/ML
LLM
RAG
Machine Learning
DevOps
GitHub
Apply
≈ $129k – $270k per year (Estimated) • Remote (likely Brazil) • Full-Time
Python
AI/ML
LangChain
Hadoop
Spark
LlamaIndex
Embeddings
Langfuse
Pandas
CrewAI
RAG
LLMOps
LLM Guardrails
Agno
Machine Learning
DevOps
Datadog
CI/CD
Git
AWS
Docker
Kubernetes
Apply
Python AI Tech Lead 1 hour ago
≈ $60k – $141k per year (Estimated) • Remote (Portugal) • Full-Time • Lisbon
Python
Databases
Weaviate
Databricks
Delta Lake
Milvus
Pinecone
FAISS
AI/ML
LangGraph
LangChain
Spark
MLFlow
Embeddings
Scikit-learn
Prompt Engineering
NLP
NER
Semantic Kernel
TensorFlow
PyTorch
CrewAI
Semantic Search
OpenAI
Anthropic
Hugging Face
Semantic Search
DevOps
GCP
Azure DevOps
GitHub Actions
Prometheus
Azure
AWS
Docker
Kubernetes
Grafana
IAM
Analytics
ETL/ELT
Apply
Remote (location not specified) • Full-Time • PhD
Python
AI/ML
LangGraph
LangChain
MLFlow
Machine Learning
DevOps
Azure
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
Apply
Hybrid • 1+ year exp • Madrid
Python
AI/ML
vLLM
LLM
DevOps
Jenkins
Linux
Management
Scrum
Apply
Ingeniero/a IA 1 day ago
≈ $53k – $125k per year (Estimated) • In office • A Coruña
Python
DevOps
Rest API
Git
Apply
≈ $51k – $121k per year (Estimated) • Hybrid • Madrid
Python
AI/ML
LLM
Machine Learning
Apply
Hybrid • 2+ years exp • Madrid
Python
DevOps
Terraform
Azure DevOps
GitHub Actions
OpenTelemetry
Azure
CI/CD
Git
Docker
Kubernetes
Azure AKS
IAM
Apply
In office • China
C++
Apply
$130k – $180k per year • Remote (United States) • Full-Time
Python
Databases
Weaviate
Milvus
Pinecone
FAISS
AI/ML
LangGraph
LangChain
LlamaIndex
Fine-tuning
Embeddings
JAX
Prompt Engineering
Multimodal AI
Computer Vision
AI Agents
NLP
TensorFlow
PyTorch
LLM
RAG
Ray
Recommender Systems
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
$130k – $180k per year • Remote (likely United States) • Full-Time
Python
AI/ML
LangGraph
AutoGen
LangChain
Claude
LlamaIndex
LoRA
Fine-tuning
Embeddings
Prompt Engineering
AI Agents
PEFT
QLoRA
Semantic Kernel
Llama
Mistral
CrewAI
Gemini
LLM
RAG
Hallucination
OpenAI
Anthropic
LLM Guardrails
Multi-Agent Systems
Tool Use
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Kubernetes
Cybersecurity
PCI DSS
SOC 2
HIPAA
FedRAMP
Apply
$100k – $150k per year • Remote (United States) • Full-Time
Python
AI/ML
Fine-tuning
JAX
Multimodal AI
AI Agents
PyTorch
RAG
Machine Learning
Apply
NLP AI Engineer 2 days ago
$130k – $180k per year • Remote (United States) • Full-Time
Python
AI/ML
LangGraph
LangChain
DeepSpeed
LlamaIndex
LoRA
Fine-tuning
RLHF
Multimodal AI
AI Agents
NLP
PEFT
QLoRA
Transformers
PyTorch
LLM
RAG
Ray
Hallucination
Synthetic Data
DPO
SFT
PPO
FSDP
Knowledge Graph
Machine Learning
DevOps
GCP
Azure
AWS
Docker
Kubernetes
Apply
$100k – $150k per year • Remote (United States) • Full-Time
Python
C++
AI/ML
Quantization
Knowledge Distillation
Federated Learning
Edge AI
LiteRT
ONNX Runtime
Model Distillation
Machine Learning
Mobile
Core ML
Apply
See all jobs
This is one of many
1,431,395 more open roles from verified company boards, updated every day.