706,063open jobs
41,958companies
98,117added this week
Browse all
Salary
$170k – $210k per year
Location
In office (Bethesda)
Seniority
Middle · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Through innovation and optimization, we have the ability to discover and maximize advantages for our Clients.

Our partner, Vexterra Group, is looking for a mid-career Systems Engineer (Mid-Career) - HPC & GPU Infrastructure with a deep understanding of operating systems, hardware, Kubernetes, and NVIDIA GPU products. As a Systems Engineer (Mid-Career) - HPC & GPU Infrastructure, you will play a pivotal role in designing, developing, and optimizing GPU clusters for the IC community customers. This is a 100% on-site position. All work must be performed at the customer site in Bethesda at the Intelligence Community Campus.

Primary Responsibilities

1. HPC and GPU environment engineering: Contribute to the installation and maintenance of GPU

and HPC hardware on-prem and in the cloud, providing insights into hardware performance to

ensure efficient interaction with software components.

2. Performance Optimization: Analyze HPC/GPU cluster performance, identify bottlenecks, and

develop strategies to enhance performance across various applications in Linux, addressing both

hardware and software considerations. Regularly monitor and improve performance.

3. HPC/GPU tooling: Install and configure HPC/GPU job scheduling and workload management

platforms such as Slurm, PBS , Apache Airflow, Kubernetes

4. Power Efficiency: Work on power management techniques to optimize GPU power

consumption, ensuring efficient operation on both mobile and desktop Linux platforms.

Continuously assess and enhance power efficiency strategies.

5. Testing and Validation: Design and execute tests to validate GPU performance and functionality

on Linux, including stress testing, benchmarking, and debugging to ensure robust operation.

Maintain and expand the testing suite.

6. Documentation: Maintain comprehensive technical documentation, including architectural

specifications, code documentation, and Linux-specific best practices for GPU development.

Keep documentation up to date with changes and improvements.

7. Industry Insight: Stay updated on the latest trends, innovations, and competitive landscapes

within the GPU industry, contributing to research efforts and proposing Linux-specific

approaches to GPU design and optimization. Share regular updates and insights with the team.

Basic Qualifications

  • Bachelor's or higher degree in Computer Science, Electrical Engineering, or a related field.
  • Additional years of experience may be considered in lieu of a degree.
  • 4+ years of relevant systems engineering experience
  • Expertise in operating system integration for Linux.
  • Strong understanding of computer hardware architecture, particularly as it relates to Linux systems.
  • Knowledge of parallel computing, graphics algorithms, and real-time rendering in Linux environments.
  • Excellent problem-solving skills and the ability to collaborate within a team.
  • Strong communication skills for conveying technical information in a Linux context.
  • Proficiency with scripting languages such as Python or BASH.
  • Proficiency with automation tools such Ansible, Puppet, Salt, Terraform, etc.
  • Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
  • TS/SCI clearance with Polygraph required or a TS/SCI and willingness to obtain a Polygraph prior to starting.

Preferred Qualifications

  • Knowledge of GPU virtualization, cloud computing, and emerging Linux-based technologies in the field.
  • Experience with container technologies (Docker, Kubernetes)
  • Experience with Prometheus/Grafana for monitoring
  • Knowledge of distributed resource scheduling systems
  • Understanding data center networking hardware and cabling concepts.
  • Understanding of networking technologies such as DHCP, DNS, TCP/IP, VLANs, HSRP, and SNMP.
  • Knowledge of data center networking security principles Firewall ACLs, IPS/IDS, and Policy Based Routing.

All your information will be kept confidential according to EEO guidelines.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
706,063 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bethesda
$130k – $268k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Singapore
Python
AI/ML
Edge AI
Apply
Data Scientist 1 day ago
In office • Bengaluru
Python
AI/ML
XGBoost
NLP
BERT
OCR
LLM Guardrails
Machine Learning
Apply
Research Engineer 1 day ago
$142k – $275k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • New York • San Francisco
Python
AI/ML
Fine-tuning
Reinforcement Learning
Machine Learning
DevOps
OpenTelemetry
Docker
Grafana
Apply
$120k – $165k per year • Remote/Hybrid • Public Trust • 5+ years exp • Bachelor's Degree • Ashburn
Python
JavaScript
Java
TypeScript
COBOL
Java
Spring Boot
Hibernate
COBOL
IBM MQ
Databases
PostgreSQL
ActiveMQ
Apache Kafka
Frontend
Angular
Mobile
JUnit
DevOps
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Bamboo
Linux
Cybersecurity
SonarQube
Management
Jira
Agile
Apply
$121k – $219k per year • Equity • In office • 5+ years exp • Bachelor's Degree
Python
Go
Rust
C++
DevOps
KVM
QEMU
Akamai
Linux
Apply
DevOps Engineer 4 days ago
$120k – $150k per year • In office • Full-Time • 3+ years exp • Chantilly
Python
Rust
SQL
PowerShell
Rust
Rocket
Frontend
WebAssembly
DevOps
Ansible
VMWare
Windows Server
Docker
Ubuntu
GitHub
GitLab
Linux
Management
Agile
ITIL
Apply
$135k – $145k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • Andrews
Python
PowerShell
DevOps
Terraform
Ansible
VMWare
Windows Server
Linux
Apply
$170k – $210k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Bethesda
Python
C
Bash
C
MPI
AI/ML
Kubeflow
DevOps
Terraform
Ansible
GitLab CI
SLURM
CI/CD
AWS
Docker
Kubernetes
Ubuntu
Platform Engineering
Linux
Apply
$80k – $92k per year • In office • Full-Time • 2+ years exp • Associate's Degree • Huntsville
DevOps
Linux
Windows
Wi-Fi
Apply
Senior Linux Engineer 1 month ago
$114k – $221k per year (Estimated) • In office • Full-Time • 10+ years exp • Reston
Python
DevOps
Puppet
Ansible
Linux
Management
Microsoft Office
Apply
$79k – $183k per year (Estimated) • In office • Public Trust • Contractor • Bethesda
Management
Outlook
Microsoft Office
Apply
$64k – $74k per year • In office • Contractor • Bethesda
Apply
$84k – $94k per year • In office • Contractor • Bethesda
Apply
$92k – $167k per year • In office • TS/SCI • Full-Time • 8+ years exp • Bachelor's Degree • Bethesda
DevOps
SLI/SLO/SLA
TCP/IP
DNS
DHCP
BGP
OSPF
Management
ServiceNow
ITSM
Apply
$66k – $119k per year • In office • Public Trust • Full-Time • 7+ years exp • Associate's Degree • Bethesda
AI/ML
Copilot
Claude
ElevenLabs
Runway
Mobile
Lottie
Design
Adobe Photoshop
Adobe Premiere Pro
Figma
Adobe After Effects
Cinema 4D
DaVinci Resolve
Rive
Management
Power Automate
Agile
Microsoft Office
Marketing
YouTube
Apply
See all jobs
This is one of many
706,063 more open roles from verified company boards, updated every day.