368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$78k – $164k per year (Estimated)
Location
Remote/Hybrid (Washington, United States)
Seniority
Middle · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Palantir Technologies is an American software company founded in 2003 that builds platforms for integrating, analysing and acting on large and fragmented datasets. Its Gotham platform serves defence, intelligence and law enforcement customers, Foundry brings the same data-operating model to commercial industries such as manufacturing, healthcare and energy, and the Artificial Intelligence Platform layers large language models over both. Headquartered in Denver, Colorado and listed on the New York Stock Exchange since 2020, the company is known for deploying forward engineers directly inside customer organisations.

The Role

We’re looking for Site Reliability Engineers who can help us build, operate, and maintain high-performance, scalable, and reliable services for our production infrastructure, across both cloud & on-prem environments. Site Reliability Engineers combine engineering experience and an innate drive to improve existing systems and processes, with the creativity to develop novel solutions to evolving challenges. Our team strives to automate processes wherever possible, using whichever tools are best for the job. You’ll be the experts for the environments that you operate infrastructure in, helping partner teams build & configure their software to operate reliably within.

We strongly believe in engineering teams being responsible for the operations of their services in production. In this role, you’ll work closely with engineers to advocate and participate in sensible, scalable, systems design and share responsibility with them in diagnosing, resolving, and preventing production issues.

Core Responsibilities

  • Maintaining availability of cloud & physical Linux servers that power the Palantir platform in air-gapped production environments
  • Design, deploy, and operate infrastructure to support customer & product requirements via modern orchestration & monitoring platforms.
  • Collaborate closely with product teams on requirements & SLOs for deploying software into air-gapped environments.
  • Identifying, troubleshooting, and solving network & systems issues
  • Scripting to automate away routine operational tasks
  • Provide technical troubleshooting support for production issues, ensuring timely resolution and minimal impact on operations. Participate in a support on-call schedule

What We Value

  • Confidence in troubleshooting complex systems issues independently using stack traces and observability & systems tools
  • Comfort with managing large scale production systems and technologies with configuration management, load balancing, monitoring & alerting infrastructure, and container orchestration
  • Demonstrated ability to continuously learn and work independently, making decisions with minimal supervision while working in secure facilities
  • Experience with containers (Docker/Podman) and orchestration (OpenShift/Kubernetes) at scale is a plus
  • Preferred Certifications: DOD 8570 IAT Level II or greater (CISSP, Sec+), Unix/Linux Computing Environment (e.g Linux+, RHCE)

What We Require

  • Active security clearance
  • 4+ years of experience with Linux system administration (RHEL or equivalent preferred)
  • Experience with cloud-based hosting platforms like AWS, Azure, or GCP and/or experience with hardware-based environments
  • Familiarity with monitoring systems using tools like Prometheus and writing health checks
  • Proficiency with at least one programming language, such as Java, Go, Python, JavaScript, Bash, or similar languages.
  • Strong engineering background, preferred in fields such as Computer Science, Mathematics, Software Engineering, Physics, and Data Science
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Washington
$77k – $193k per year (Estimated) • In office • Full-Time • Singapore
PowerShell
Python
DevOps
Ansible
AWS
Azure
Chef
CI/CD
CloudFormation
Configuration Management
Datadog
Docker
Docker Swarm
Grafana
Helm
Hyper-V
Incident Management
Kubernetes
Kustomize
OpenShift
Platform Engineering
Prometheus
Puppet
Service Mesh
Terraform
VMWare
Apply
$18k – $45k per year (Estimated) • In office • Bachelor's Degree • Moscow
Node JS
TypeScript
JavaScript
Node JS
Nest.JS
Pino
PM2
Prisma
Databases
PostgreSQL
Frontend
ESLint
Husky
npm
Pinia
Prettier
Vite
Vue.js
DevOps
CI/CD
Debian
Docker
GitLab CI
Kubernetes
Nginx
Amazon S3
GitLab
Cybersecurity
Keycloak
QA
Jest
Playwright
Swagger
Vitest
Apply
$22k – $53k per year (Estimated) • In office • Astana
Python
Databases
FAISS
Milvus
pgvector
Qdrant
PostgreSQL
AI/ML
Fine-tuning
LangChain
Langfuse
LangGraph
LlamaIndex
LLM
LoRA
NLP
Ollama
QLoRA
RAG
Triton
vLLM
Whisper
PEFT
TGI
DevOps
Docker
Kubernetes
Apply
$28k – $57k per year (Estimated) • Remote • 6+ years exp • Moscow
Databases
Apache Kafka
ClickHouse
PostgreSQL
Redpanda
TimescaleDB
DevOps
Alertmanager
ArgoCD
cert-manager
CI/CD
External Secrets
GitLab CI
Grafana
Helm
Jsonnet
Kubernetes
Loki
Nginx
Prometheus
VictoriaMetrics
GitLab
Cybersecurity
Trivy
Apply
$135k – $244k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree
Clojure
Go
Clojure
Datomic
Databases
Apache Kafka
Google Cloud Spanner
AI/ML
AI Agents
DevOps
Docker
GitHub Actions
GitHub
Cybersecurity
Cosign
Sigstore
SLSA
Apply
$83k – $187k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Washington
Python
AI/ML
LLM
AI Agents
Red Teaming
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
IAM
Cybersecurity
BloodHound
Burp Suite
Certipy
Impacket
Nuclei
Pacu
ScoutSuite
Apply
$145k – $200k per year • Equity • Remote/Hybrid • Full-Time • 4+ years exp • Washington
Python
AI/ML
LLM
AI Agents
Red Teaming
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
IAM
Cybersecurity
BloodHound
Burp Suite
Certipy
Impacket
Nuclei
Pacu
ScoutSuite
Apply
$85k – $172k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Huntsville
Apply
$83k – $169k per year (Estimated) • In office • Full-Time • 2+ years exp • Washington
Apply
$95k – $191k per year (Estimated) • In office • Full-Time • 2+ years exp • New York
Apply
$81k – $122k per year • Remote/Hybrid • Full-Time • PhD • Atlanta • Washington
JavaScript
SQL
AI/ML
AI Agents
Agentforce
Marketing
Salesforce
Apply
$171k – $273k per year • In office • Full-Time • 8+ years exp • PhD • San Francisco • Washington
AI/ML
A2A
Agentforce
AI Agents
Model Context Protocol
DevOps
AWS
GCP
Marketing
Salesforce
Apply
Data Scientist 2 hours ago
$113k – $188k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Arlington • Washington
Python
Databases
Databricks
DevOps
AWS
Azure
Analytics
ETL/ELT
Power BI
Apply
Data Scientist 2 hours ago
$98k – $163k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Arlington • Washington
SQL
Databases
Databricks
DevOps
AWS
Azure
Analytics
ETL/ELT
Power BI
Apply
$99k – $150k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Washington
JavaScript
AI/ML
Agentforce
AI Agents
Marketing
HubSpot
Marketo
Salesforce
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.