1,132,385open jobs
65,147companies
198,503added this week
Browse all
Salary
$100k – $150k per year
Location
Remote (United States)
Employment
Full-Time

First seen by Alion on Oct 2, 2026.

Overview
Company
Impact
Profile match
Bright Vision Technologies is a minority-owned, product-focused AI technology company headquartered in Bridgewater, New Jersey. Founded in July 2020, we are a pure-play product engineering firm dedicated to transforming the staffing and enterprise automation industries through our flagship platform — Lumina.Lumina is an advanced AI-powered talent intelligence and enterprise automation platform that integrates artificial intelligence, hybrid cloud, and blockchain technologies to revolutionize how organizations source talent and automate complex workflows.What Lumina delivers:- AI-driven candidate sourcing, semantic resume matching, and intelligent ranking- Generative AI and Large Language Model orchestration with LangChain and RAG-based architecture- Automated recruiter assistant and workflow automation- Blockchain-enabled credential verification and tamper-proof digital identity- Advanced optimization for talent supply chain, workforce planning, and resource allocation- Enterprise-grade security with role-based access, encryption, and comprehensive audit logging- Seamless integration with CRM, ERP, HRMS, ATS, and legacy enterprise systemsBuilt on a scalable microservices architectu
GPU Systems Engineer - Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: GPU Systems Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $100,000–$150,000 Annually
Experience Required: 6+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary
We are seeking a GPU Systems Engineer with deep expertise in CUDA programming, GPU architecture, and high-performance computing to design and optimize compute-intensive workloads on modern accelerator hardware. This role focuses on extracting maximum performance from GPU platforms for AI training, inference, scientific computing, and high-throughput data processing workloads. The ideal candidate combines low-level systems mastery with strong software engineering practices, and has a track record of delivering measurable performance improvements on production GPU systems. In this role you will work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions, and will be expected to raise the bar through code review, design review, and mentorship of more junior engineers. The successful candidate brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production.

Key Responsibilities
  • Design and implement high-performance CUDA kernels for compute-intensive workloads across AI and HPC use cases.
  • Profile and optimize GPU code using tools such as Nsight Systems, Nsight Compute, and CUDA profilers.
  • Tune memory access patterns, occupancy, register usage, and shared memory utilization for peak performance.
  • Develop highly optimized libraries for linear algebra, attention, and other ML primitives.
  • Optimize multi-GPU and multi-node training using NCCL, RDMA, and high-performance networking.
  • Implement custom operators and fused kernels in PyTorch, JAX, or Triton.
  • Collaborate with ML engineers to identify performance bottlenecks in training and inference pipelines.
  • Develop benchmarks and regression tests to safeguard performance over time.
  • Evaluate new GPU architectures and feature sets, and advise on adoption strategy.
  • Contribute to compiler-level optimizations for tensor programs where appropriate, working at the boundary between ML frameworks and underlying accelerator codegen to unlock performance not reachable through framework-level tuning alone.
  • Optimize memory hierarchy usage across HBM, L2, shared memory, and registers.
  • Implement mixed-precision and quantized compute paths that maximize accelerator throughput while preserving numerical fidelity within bounds acceptable for the target workloads.
  • Document performance characteristics, design decisions, and tuning playbooks for internal teams.
  • Stay current with GPU architecture, CUDA evolution, and emerging accelerator technologies.
Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.
  • Six or more years of experience in GPU programming and performance engineering.
  • Deep expertise in CUDA C/C++ and GPU programming models.
  • Strong understanding of modern GPU architectures, memory hierarchies, and execution models.
  • Hands-on experience profiling and optimizing GPU workloads in production.
  • Familiarity with NCCL, MPI, and high-performance interconnect technologies.
  • Experience integrating custom kernels into ML frameworks.
  • Strong C++ skills and familiarity with modern systems programming practices.
  • Solid grounding in linear algebra and numerical methods.
  • Strong communication and collaboration skills with research and engineering teams.
Preferred Qualifications
  • Experience with Triton, CUTLASS, or other GPU kernel authoring frameworks.
  • Familiarity with TensorRT, FasterTransformer, or vLLM internals.
  • Exposure to compiler infrastructure such as LLVM or MLIR.
  • Open-source contributions to GPU or ML performance libraries.
  • Experience with large-scale distributed training infrastructure.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to  or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at .
Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Originally posted on Himalayas

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,132,385 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
AI Architect, Finance 4 months ago
$82k – $102k per year • Remote (Italy) • Full-Time • 10+ years exp • Rome
Python
Clarity
AI/ML
Scikit-learn
Prompt Engineering
AI Agents
Prophet
TensorFlow
PyTorch
LLM
RAG
Statsmodels
Anomaly Detection
Time Series Forecasting
Human-in-the-Loop
LLM Guardrails
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Apply
Remote (location not specified) • Part-Time
Python
Python
FastAPI
AI/ML
LangChain
Claude
LlamaIndex
Model Context Protocol
Prompt Engineering
AI Agents
OpenAI
Anthropic
A2A
DevOps
Rest API
GitHub
Management
Stripe
Apply
Senior AI Scientist 2 days ago
≈ $137k – $245k per year (Estimated) • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • New Orleans
Python
SQL
AI/ML
Function Calling
AI Agents
TensorFlow
PyTorch
LLM
RAG
LLMOps
Voice Agents
Tool Use
Machine Learning
Analytics
A/B Testing
Apply
$130k – $150k per year • Remote (United States) • Full-Time
AI/ML
Embeddings
NLP
LLM
Semantic Search
Semantic Search
Machine Learning
Apply
206303 - AI Engineer 1 month ago
≈ $96k – $243k per year (Estimated) • Remote (Canada) • Toronto
AI/ML
AutoGen
LangChain
Prompt Engineering
AI Agents
LLM
Apply
≈ $20k – $49k per year (Estimated) • Hybrid • 7+ years exp • Chennai
C++
DevOps
CI/CD
Jenkins
Git
Self-Healing
Linux
Windows
Management
Agile
Apply
≈ $65k – $188k per year (Estimated) • In office • Internship • Toronto
Python
C++
AI/ML
Copilot
ChatGPT
LLM
Machine Learning
DevOps
CI/CD
Jenkins
Apply
≈ $134k – $285k per year (Estimated) • Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • Charlotte
Python
Python
pySpark
AI/ML
Spark
AI Agents
NLP
TensorFlow
PyTorch
Machine Learning
DevOps
GCP
Apply
GN&C Engineer 1 day ago
≈ $74k – $139k per year (Estimated) • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Houston
Python
Kotlin
C++
MATLAB
Kotlin
StateFlow
MATLAB
Simulink
AI/ML
Machine Learning
DevOps
CI/CD
Management
Agile
Apply
≈ $106k – $216k per year (Estimated) • Equity • In office • Full-Time • 10+ years exp • PhD • Huntsville
Python
C++
Bash
MATLAB
Fortran
MATLAB
Simulink
DevOps
Linux
Robotics
Sensor Fusion
SpaceTech
GMAT
Apply
$96k – $120k per year • Remote (United States) • Full-Time
Python
AI/ML
RLHF
Reinforcement Learning
AI Agents
Reward Modeling
Machine Learning
Robotics
Reinforcement Learning
Apply
$80k – $100k per year • Remote (United States) • Full-Time
Python
AI/ML
Spark
Multimodal AI
Ray
Machine Learning
DevOps
CI/CD
Apply
Prompt Engineer 1 day ago
$76k – $99k per year • Remote (United States) • Full-Time
Python
AI/ML
Fine-tuning
Embeddings
LLM
RAG
Agentic Workflows
Multi-Agent Systems
Tool Use
Apply
$100k – $150k per year • Remote (United States) • Full-Time
Python
AI/ML
Fine-tuning
LLM
LLM Guardrails
Machine Learning
Cybersecurity
Threat Modeling
Apply
$130k – $180k per year • Remote (United States) • Full-Time • 10+ years exp • Bachelor's Degree
Python
C++
C++
PyTorch C++
AI/ML
DeepSpeed
vLLM
CUDA Toolkit
Triton Inference Server
Quantization
TensorRT-LLM
PyTorch
LLM
TensorBoard
Ray
CUDA
NCCL
ROCm
CUTLASS
Speculative Decoding
KV Cache
Machine Learning
DevOps
GCP
Azure
AWS
FinOps
HPC
Apply
See all jobs
This is one of many
1,132,385 more open roles from verified company boards, updated every day.