1,443,753open jobs
85,355companies
221,329added this week
Browse all
Salary
$100k per year
Location
Remote (United States)
Employment
Full-Time

First seen by Alion on Oct 10, 2026.

Overview
Company
Impact
Profile match
Bright Vision Technologies is a minority-owned, product-focused AI technology company headquartered in Bridgewater, New Jersey. Founded in July 2020, we are a pure-play product engineering firm dedicated to transforming the staffing and enterprise automation industries through our flagship platform — Lumina.Lumina is an advanced AI-powered talent intelligence and enterprise automation platform that integrates artificial intelligence, hybrid cloud, and blockchain technologies to revolutionize how organizations source talent and automate complex workflows.What Lumina delivers:- AI-driven candidate sourcing, semantic resume matching, and intelligent ranking- Generative AI and Large Language Model orchestration with LangChain and RAG-based architecture- Automated recruiter assistant and workflow automation- Blockchain-enabled credential verification and tamper-proof digital identity- Advanced optimization for talent supply chain, workforce planning, and resource allocation- Enterprise-grade security with role-based access, encryption, and comprehensive audit logging- Seamless integration with CRM, ERP, HRMS, ATS, and legacy enterprise systemsBuilt on a scalable microservices architectu
AI Optimization Engineer – Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: AI Optimization Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000 Annually
Experience Required: 6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary
We are seeking an AI Optimization Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. The ideal candidate has demonstrated impact on production AI workloads, with strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production.

Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.
  • Six or more years of experience in performance engineering, ML systems, or HPC.
  • Strong proficiency in Python and C++.
  • Hands-on experience optimizing deep learning workloads on modern GPUs.
  • Deep understanding of distributed training and inference techniques.
  • Experience with profiling tools across CPU, GPU, and distributed systems.
  • Familiarity with model compression techniques and their accuracy implications.
  • Strong grasp of memory hierarchies, communication primitives, and parallelism strategies.
  • Excellent measurement, debugging, and analytical reasoning skills.
  • Strong communication and collaboration skills.
Preferred Qualifications
  • Experience optimizing LLM inference at production scale.
  • Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects.
  • Familiarity with custom kernel authoring in Triton or CUTLASS.
  • Experience with FinOps for AI workloads.
  • Publications or talks on AI systems performance.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to
Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Originally posted on Himalayas

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,443,753 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
≈ $100k – $249k per year (Estimated) • Remote (Europe, Israel) • PhD • Israel • United Kingdom
Python
AI/ML
Post-training
Model Distillation
Machine Learning
Apply
≈ $100k – $249k per year (Estimated) • Remote (Europe) • PhD • Prague
Python
AI/ML
Post-training
Model Distillation
Machine Learning
Apply
≈ $101k – $251k per year (Estimated) • Remote (Europe, Israel) • Bachelor's Degree • Amsterdam • Prague
Python
AI/ML
Flash Attention
Reinforcement Learning
Quantization
AI Agents
LLM
DPO
PPO
Post-training
FSDP
Model Distillation
Reward Modeling
Machine Learning
DevOps
CI/CD
Apply
$195k – $262k per year • Remote (United States) • Palo Alto
Python
AI/ML
DeepSpeed
CUDA Toolkit
RLHF
Function Calling
TRL
Transformers
PyTorch
LLM
Ray
Synthetic Data
CUDA
Triton
DPO
SFT
PPO
GRPO
Post-training
Megatron-LM
FSDP
NCCL
InfiniBand
Agentic Workflows
Tool Use
RLAIF
Reward Modeling
Machine Learning
DevOps
SLURM
Kubernetes
Apply
$195k – $262k per year • Remote (United States) • Palo Alto • San Francisco
Python
AI/ML
Ray Serve
vLLM
Triton Inference Server
Quantization
Knowledge Distillation
Function Calling
AI Agents
VLM
SGLang
AWQ
GPTQ
TensorRT
TensorRT-LLM
PyTorch
LLM
Ray
Mixture of Experts
KServe
TPOT
Triton
Post-training
NCCL
InfiniBand
NVLink
Structured Outputs
Speculative Decoding
KV Cache
Tool Use
Model Distillation
Machine Learning
Apply
See all jobs
This is one of many
1,443,753 more open roles from verified company boards, updated every day.