552,642open jobs
20,924companies
75,446added this week
Browse all
Salary
$140k – $210k per year
Location
Remote (United States)
Seniority
Staff · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Software Engineer - Backend based in United States.

This role is a senior technical leadership opportunity focused on building the infrastructure and backend platforms that power production machine learning and generative AI workloads. You will design and evolve scalable, reliable services that enable advanced models to operate efficiently at global scale. Working closely with ML Scientists, Data Engineers, Product teams, and engineers, you will turn sophisticated models into robust production systems. The role combines hands-on software engineering with architectural leadership, technical strategy, and mentorship. You will also contribute to observability, GPU infrastructure, workload optimization, and cloud cost efficiency. This is an opportunity to shape modern ML infrastructure while influencing engineering practices across a globally distributed organization.

Accountabilities:

    • Lead the design, architecture, and evolution of infrastructure, platforms, and backend services supporting machine learning and generative AI workloads.
    • Build and operate highly scalable, reliable, and observable distributed systems and microservices for production environments.
    • Deploy and manage ML and generative AI workloads using serving technologies such as vLLM, Triton, TorchServe, SageMaker Endpoints, or comparable frameworks.
    • Develop cloud-native infrastructure and services on AWS, supporting high-availability workloads across services such as ECS, EKS, Lambda, DynamoDB, S3, IAM, and CloudWatch.
    • Establish and improve observability practices using technologies such as OpenTelemetry, Prometheus, Grafana, and CloudWatch.
    • Work with vector databases, feature stores, caching systems such as Valkey/Redis, and infrastructure-as-code technologies including CDK, CloudFormation, or Terraform.
    • Contribute to GPU infrastructure management, workload scheduling, performance optimization, and cloud cost efficiency.
    • Drive complex technical initiatives, influence architectural decisions, and establish engineering standards across distributed teams.
    • Partner closely with ML Scientists, Data Engineers, Product teams, and other stakeholders to translate machine learning capabilities into reliable production solutions.
    • Mentor engineers and serve as a technical leader, helping elevate engineering quality and accelerate software delivery through AI-assisted development tools.
    • Requirements:

      • 7+ years of professional software engineering experience building and operating large-scale, production-grade distributed systems and microservices.
      • Strong proficiency in Python, Go, and/or TypeScript, with deep knowledge of API design, system architecture, scalability, resiliency, and performance optimization.
      • Hands-on experience developing CI/CD pipelines, cloud-native applications, and production infrastructure on AWS.
      • Strong experience with containerization and orchestration technologies such as Docker, Kubernetes, ECS, or comparable platforms.
      • Experience designing and operating highly available production workloads in distributed environments.
      • Experience with machine learning or generative AI infrastructure and model-serving technologies is highly valuable.
      • Familiarity with observability frameworks and tools such as OpenTelemetry, Prometheus, Grafana, and CloudWatch.
      • Experience with vector databases, feature stores, caching technologies, and infrastructure-as-code solutions is preferred.
      • Understanding of GPU infrastructure, workload scheduling, performance tuning, and cloud cost optimization is an advantage.
      • Proven ability to lead complex technical initiatives, influence architecture, and collaborate effectively with diverse technical and business stakeholders.
      • Demonstrated experience mentoring engineers and providing technical leadership within distributed or globally collaborative teams.
      • Experience using AI-assisted development tools to improve engineering productivity and accelerate software delivery is preferred.
      • Strong communication skills and the ability to operate effectively in a fast-paced, evolving environment.
      • Benefits:

        • Remote position within the United States, with occasional visits to an office for team events or meetings.
        • Flexible remote and hybrid work options depending on team and location.
        • Estimated base salary ranging from $140,000-$210,000 USD for most U.S. locations.
        • Estimated base salary of $157,000-$235,000 USD for Austin, D.C. Metro, non-Bay Area California, Hawaii, Illinois, Massachusetts, New Hampshire, Oregon, Virginia, and Washington.
        • Estimated base salary of $166,800-$250,200 USD for the New York City Metro and Kirkland/Seattle areas.
        • Estimated base salary of $182,000-$273,000 USD for the Bay Area and Los Angeles.
        • Potential eligibility for additional compensation, including corporate bonuses and/or equity awards.
        • Medical, dental, and vision insurance.
        • 401(k) retirement savings plan.
        • Paid sick time, flexible paid time off, and paid holidays.
        • Paid parental leave and wellness days.
        • Life insurance, short- and long-term disability, and AD&D insurance.
        • Mental health and Employee Assistance Program resources.
        • Tuition assistance and financial education and advice.
        • Adoption, surrogacy, and fertility benefits.
        • Dependent daycare and backup care benefits.
        • Employee stock purchase plan.
        • Opportunities to work remotely with a globally distributed engineering organization.
        • Benefits and compensation may vary based on geographic location and applicable eligibility requirements.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
552,642 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$197k – $225k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • San Jose • San Francisco • McLean • Cambridge • New York
Python
Go
Java
C#
C++
Scala
C++
PyTorch C++
AI/ML
NeMo Guardrails
PyTorch
LLM
Hugging Face
LLM Guardrails
DevOps
GCP
Azure
AWS
Apply
$20k – $40k per year (Estimated) • Remote • Full-Time • 3+ years exp • Moscow
Python
JavaScript
Python
Flask
SQLAlchemy
FastAPI
Alembic
Databases
PostgreSQL
Redis
RabbitMQ
OpenSearch
Frontend
Vue.js
GraphQL
React.js
DevOps
Rest API
Docker Compose
Prometheus
WebSockets
GitLab CI
CI/CD
Git
Docker
Grafana
QA
Sentry
Apply
Senior Data Engineer 2 hours ago
$24k – $50k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
Snowflake
DynamoDB
Apache Kafka
Amazon Redshift
AI/ML
Copilot
Claude
Spark
ChatGPT
Cline
DevOps
Terraform
CI/CD
Jenkins
AWS
Docker
Kubernetes
AWS Lambda
Amazon EC2
Amazon S3
IAM
Amazon Kinesis
AWS Step Functions
Management
ServiceNow
Apply
QE Platform Engineer 2 hours ago
$150k – $220k per year • Equity • Remote • Full-Time • 5+ years exp
JavaScript
TypeScript
Node JS
Frontend
React.js
Mobile
React Native
DevOps
AWS
Kubernetes
Platform Engineering
QA
Selenium
Cypress
Playwright
Appium
Apply
R&D Engineer 2 hours ago
$101k – $211k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp
Python
AI/ML
CUDA Toolkit
YOLO
Fine-tuning
Quantization
Knowledge Distillation
Computer Vision
PyTorch
CUDA
Edge AI
Model Distillation
Robotics
Localization
Apply
$196k – $270k per year • Equity • Remote • Full-Time • 5+ years exp • Bachelor's Degree
Apply
QE Platform Engineer 2 hours ago
$150k – $220k per year • Equity • Remote • Full-Time • 5+ years exp
JavaScript
TypeScript
Node JS
Frontend
React.js
Mobile
React Native
DevOps
AWS
Kubernetes
Platform Engineering
QA
Selenium
Cypress
Playwright
Appium
Apply
$57k – $65k per year • Remote • Full-Time
Apply
$99k – $223k per year (Estimated) • Remote • Full-Time • 5+ years exp
Marketing
Salesforce
Apply
R&D Engineer 2 hours ago
$101k – $211k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp
Python
AI/ML
CUDA Toolkit
YOLO
Fine-tuning
Quantization
Knowledge Distillation
Computer Vision
PyTorch
CUDA
Edge AI
Model Distillation
Robotics
Localization
Apply
See all jobs
This is one of many
552,642 more open roles from verified company boards, updated every day.