428,639open jobs
14,645companies
62,546added this week
Browse all
Salary
$130k – $180k per year
Location
Remote (United States)
Seniority
Staff · 10+ years exp
Overview
Company
Impact
Profile match
Bright Vision Technologies is an IT consulting, enterprise technology services, and workforce solutions enterprise. Headquartered in Bridgewater, New Jersey, United States, the minority-owned firm specializes in technology staffing, cybersecurity, application management, and digital product engineering. Founded in 2020, the enterprise delivers specialized staffing and IT services alongside proprietary automation and AI software - including its flagship enterprise talent intelligence platform, Lumina - serving clients across information technology, defense, healthcare, government, and manufacturing sectors.

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title

AI Systems Performance Specialist

Location: 100% Remote (Continental United States)

Position Type: Full-time, Direct W2

Salary Range: $130,000-$180,000 Annually

Experience Required: 10+ Years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

Bright Vision Technologies is seeking a highly experienced AI Systems Performance Specialist with 10+ years of experience in AI infrastructure, machine learning systems, High-Performance Computing (HPC), and performance engineering. The ideal candidate will optimize AI training and inference workloads for maximum performance, scalability, reliability, and cost efficiency. This role requires deep expertise in GPU optimization, distributed training, Large Language Model (LLM) inference, Python, C++, CUDA, and production AI systems, along with the ability to lead performance optimization initiatives across enterprise-scale AI platforms.

Key Responsibilities

  • Optimize AI training and inference pipelines for maximum throughput, low latency, scalability, and infrastructure efficiency.
  • Analyze and improve GPU utilization, memory management, kernel execution, and multi-GPU performance across production AI workloads.
  • Design and implement optimization techniques including quantization, pruning, mixed precision, batching, caching, speculative decoding, and model parallelism.
  • Profile AI applications using industry-standard performance analysis tools and identify bottlenecks across compute, memory, networking, and storage.
  • Optimize distributed training and inference using NCCL, DeepSpeed, PyTorch Distributed, Ray, MPI, or similar distributed computing frameworks.
  • Collaborate with AI researchers, ML engineers, platform engineers, and infrastructure teams to improve model performance and production reliability.
  • Build automated benchmarking frameworks, performance dashboards, monitoring solutions, and regression testing pipelines.
  • Evaluate emerging AI hardware, GPU architectures, inference frameworks, and optimization technologies to improve enterprise AI capabilities.
  • Drive AI infrastructure cost optimization through efficient resource utilization, cloud optimization, and FinOps best practices.
  • Mentor engineering teams and provide technical leadership on AI systems architecture, GPU optimization, and performance engineering.

Required Qualifications

  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, Artificial Intelligence, or a related technical discipline.
  • 10+ years of professional experience in performance engineering, AI infrastructure, machine learning systems, High-Performance Computing (HPC), or distributed computing.
  • Expert-level programming skills in Python and C++.
  • Extensive experience optimizing GPU-accelerated AI workloads using CUDA, distributed training frameworks, and modern deep learning libraries.
  • Strong knowledge of Large Language Models (LLMs), deep learning frameworks, model serving, and production AI inference.
  • Hands-on experience with profiling tools such as NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, TensorBoard, or similar performance analysis tools.
  • Experience deploying and optimizing AI workloads on AWS, Microsoft Azure, or Google Cloud Platform (GCP).
  • Strong understanding of distributed systems, networking, storage optimization, and AI infrastructure architecture.
  • Excellent analytical, troubleshooting, communication, and technical leadership skills.

Preferred Qualifications

  • Experience optimizing production-scale LLM inference and serving large foundation models.
  • Hands-on experience with vLLM, TensorRT-LLM, DeepSpeed, Triton Inference Server, CUTLASS, FasterTransformer, or similar AI optimization frameworks.
  • Knowledge of model compression, KV cache optimization, speculative decoding, and advanced inference optimization techniques.
  • Experience implementing FinOps strategies for AI infrastructure cost optimization and resource management.
  • Contributions to AI systems research, open-source AI infrastructure projects, patents, or technical publications.
  • Familiarity with emerging AI accelerator technologies, including AMD ROCm, Intel oneAPI, or custom AI hardware.

Interested in this opportunity? Apply today for immediate consideration! Email your updated resume: [email protected] Call or Text: (908) 505-3545 Learn more: www.bvteck.com

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
428,639 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$30k – $76k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • Bengaluru
Python
C++
C++
PyTorch C++
AI/ML
Fine-tuning
Computer Vision
Speech Recognition
Transformers
PyTorch
Edge AI
DevOps
Docker
Kubernetes
Apply
$56k – $134k per year (Estimated) • Equity • In office • Full-Time • Bachelor's Degree • Cambridge
Python
Rust
C++
DevOps
Git
Docker
Apply
$20k – $42k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Manila
ABAP
ABAP
SAP BTP
DevOps
Terraform
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
$112k – $201k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Fayetteville
Python
SQL
Databases
Databricks
AI/ML
Copilot
Copilot Studio
DevOps
CI/CD
Management
Jira
Power Automate
Apply
$171k – $322k per year (Estimated) • Remote • Full-Time • 10+ years exp • Master's Degree • United States
Python
Databases
Databricks
AI/ML
Multimodal AI
TensorFlow
PyTorch
Hugging Face
DevOps
Azure
Apply
$175k – $200k per year • Remote • 8+ years exp • Bachelor's Degree
Go
Java
C#
DevOps
Istio
Envoy
Linkerd
Kubernetes
Platform Engineering
Service Mesh
FinOps
Apply
$175k – $200k per year • Remote • 8+ years exp • Bachelor's Degree
DevOps
GCP
Azure
AWS
FinOps
Management
ServiceNow
Apply
$115k – $175k per year • Remote • 5+ years exp • PhD
Python
SQL
PowerShell
Python
pySpark
Databases
PostgreSQL
Snowflake
Databricks
Apache Kafka
Google BigQuery
Amazon Redshift
Teradata
BigQuery
AI/ML
Spark
Airflow
DevOps
GCP
Azure DevOps
Azure
CI/CD
Jenkins
Git
AWS
Amazon Kinesis
Analytics
Tableau
Power BI
ETL/ELT
Informatica
Talend
SSIS
DataStage
Pentaho
Azure Data Factory
AWS Glue
Data Vault
Dimensional Modeling
Master Data Management
Apply
$140k – $200k per year • Remote • 6+ years exp • PhD
Go
JavaScript
TypeScript
SQL
C#
Node JS
Solidity
Solidity
Truffle
Ganache
Brownie
Databases
PostgreSQL
Redis
AI/ML
Tokenization
Frontend
Vue.js
GraphQL
Angular
React.js
Remix
DevOps
Rest API
Terraform
WebSockets
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Web3
Foundry
Polygon
Hardhat
Remix
Avalanche
Optimism
Arbitrum
MetaMask
Layer 2
Smart Contracts
Ethereum
Hyperledger Fabric
Reown
Token Standards
Coinbase Wallet
Apply
$120k – $180k per year • Remote • 5+ years exp • PhD
JavaScript
Java
TypeScript
Java
Maven
Spring Boot
Spring MVC
Spring Data JPA
Spring Security
Spring Cloud
Databases
Apache Kafka
Frontend
Webpack
GraphQL
Tailwind CSS
Yarn
RxJS
Angular
Bootstrap
npm
PrimeNG
Sass
NgRx
Angular Material
Lighthouse
Mobile
MVC
Dependency Injection
State Management
Offline-First
PWA
DevOps
Rest API
GCP
OpenShift
Azure DevOps
GitHub Actions
WebSockets
GitLab CI
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Bitbucket
GitHub
GitLab
API Gateway
Cybersecurity
Keycloak
Okta
Auth0
Design
Figma
Adobe XD
QA
Selenium
Cypress
Playwright
Swagger
Chrome DevTools
Apply
See all jobs
This is one of many
428,639 more open roles from verified company boards, updated every day.