662,880open jobs
38,667companies
98,736added this week
Browse all
Salary
$130k – $180k per year
Location
Remote (United States)
Seniority
Staff · 10+ years exp
Overview
Company
Impact
Profile match
Bright Vision Technologies is an IT consulting, enterprise technology services, and workforce solutions enterprise. Headquartered in Bridgewater, New Jersey, United States, the minority-owned firm specializes in technology staffing, cybersecurity, application management, and digital product engineering. Founded in 2020, the enterprise delivers specialized staffing and IT services alongside proprietary automation and AI software - including its flagship enterprise talent intelligence platform, Lumina - serving clients across information technology, defense, healthcare, government, and manufacturing sectors.

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title

Parallel Computing Engineer

Location: 100% Remote (U.S.)

Position Type: Full-time, Direct W2

Salary Range: $130,000-$180,000 Annually

Experience Required: 10+ Years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

Bright Vision Technologies is seeking a highly experienced Parallel Computing Engineer with 10+ years of experience in High-Performance Computing (HPC), GPU programming, and parallel computing to optimize AI, machine learning, and scientific computing workloads. The ideal candidate will possess deep expertise in CUDA, GPU architecture, distributed computing, performance optimization, and large-scale AI infrastructure, with a proven track record of designing high-performance computing solutions for enterprise and research environments.

Key Responsibilities

  • Design, develop, and optimize high-performance CUDA kernels for AI, deep learning, and scientific computing applications.
  • Analyze, profile, and optimize GPU workloads using NVIDIA Nsight Systems, Nsight Compute, CUDA Profiler, and related performance analysis tools.
  • Optimize GPU memory management, kernel execution, multi-GPU scaling, and distributed computing performance.
  • Design scalable distributed training and inference architectures using NCCL, MPI, CUDA-aware communication libraries, and high-performance networking technologies.
  • Develop custom GPU operators and optimized kernels for PyTorch, JAX, Triton, TensorFlow, or similar AI frameworks.
  • Improve training and inference performance for large language models (LLMs), deep learning, and high-performance AI workloads.
  • Collaborate with AI researchers, ML engineers, and software architects to accelerate production AI applications.
  • Build automated benchmarking frameworks, performance regression testing, and optimization pipelines.
  • Evaluate emerging GPU technologies, programming models, and accelerator architectures to improve computational efficiency.
  • Mentor engineers and provide technical leadership in GPU optimization, HPC architecture, and parallel programming best practices.

Required Qualifications

  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical discipline.
  • 10+ years of professional experience in GPU programming, High-Performance Computing (HPC), or parallel computing.
  • Expert-level proficiency in CUDA C/C++, GPU architecture, and massively parallel programming techniques.
  • Extensive experience with NCCL, MPI, CUDA-aware MPI, and distributed GPU communication frameworks.
  • Strong understanding of GPU memory hierarchy, kernel optimization, occupancy tuning, and performance analysis.
  • Hands-on experience integrating custom GPU kernels into PyTorch, TensorFlow, JAX, Triton, or other machine learning frameworks.
  • Strong C/C++ programming skills with expertise in debugging, profiling, and performance optimization.
  • Experience developing scalable AI or HPC solutions on cloud platforms or large GPU clusters.
  • Excellent analytical, communication, collaboration, and technical leadership skills.

Preferred Qualifications

  • Experience with Triton, CUTLASS, TensorRT, FasterTransformer, vLLM, DeepSpeed, or similar GPU optimization frameworks.
  • Knowledge of LLVM, MLIR, compiler optimization techniques, or code generation technologies.
  • Experience with large-scale distributed AI training, model parallelism, pipeline parallelism, and inference optimization.
  • Familiarity with cloud-based GPU infrastructure on AWS, Microsoft Azure, or Google Cloud Platform (GCP).
  • Contributions to open-source GPU libraries, research publications, patents, or technical presentations.
  • Experience with emerging accelerator technologies such as AMD ROCm, Intel

How to Apply

Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3545. Learn more about Bright Vision Technologies at www.bvteck.com.

Bright Vision Technologies is an Equal Opportunity Employer.

Equal Employment Opportunity (EEO) Statement

Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall.

BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
662,880 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$160k – $180k per year • Remote • 6+ years exp • Bachelor's Degree
Python
SQL
Databases
Snowflake
Databricks
Apache Kafka
Google BigQuery
BigQuery
AI/ML
Hadoop
Spark
MLFlow
Reinforcement Learning
Scikit-learn
Computer Vision
Kubeflow
TensorFlow
Pandas
NumPy
PyTorch
Explainable AI
Feature Store
Recommender Systems
DevOps
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Analytics
ETL/ELT
Management
Agile
Apply
$245k – $275k per year • Remote • 10+ years exp • Bachelor's Degree
Go
Java
C++
DevOps
FinOps
Apply
$150k – $190k per year • Remote • 10+ years exp • Bachelor's Degree
Python
Go
Java
C#
C++
DevOps
Terraform
GCP
CloudFormation
Azure
CI/CD
AWS
Kubernetes
FinOps
Apply
In office • 10+ years exp • Bachelor's Degree
Python
Verilog
C++
SystemVerilog
VHDL
DevOps
RTOS
CI/CD
HPC
Chips/EDA
UVM
Apply
$145k – $205k per year • Remote • 6+ years exp • PhD
Rust
SQL
C++
Rust
Actix Web
Axum
Rocket
Databases
MySQL
PostgreSQL
Redis
NATS
Cassandra
DynamoDB
RabbitMQ
Apache Kafka
DevOps
gRPC
Terraform
GCP
Istio
OpenTelemetry
Linkerd
Prometheus
Pulumi
WebSockets
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Grafana
Platform Engineering
Service Mesh
API Gateway
Management
Agile
Apply
$138k – $173k per year • Remote/Hybrid • Bachelor's Degree
Apply
In office • 5+ years exp • Associate's Degree
Apply
$101k – $129k per year • Remote/Hybrid • 2+ years exp • Bachelor's Degree
Analytics
Microsoft Excel
Apply
$138k – $173k per year • Remote/Hybrid • Bachelor's Degree
Apply
$115k – $135k per year • Remote • 6+ years exp • Bachelor's Degree
Management
ServiceNow
Apply
See all jobs
This is one of many
662,880 more open roles from verified company boards, updated every day.