703,239open jobs
41,516companies
100,102added this week
Browse all
Salary
$113k – $209k per year (Estimated)
Location
Remote (United States)
Seniority
Senior · 5+ years exp
Overview
Company
Impact
Profile match
Vannevar Labs is a defense technology company focused on digital intelligence to support non-kinetic missions aimed at deterring and de-escalating conflicts. The company specializes in developing advanced machine learning techniques to ensure access to mission-relevant data, building tools to counter malign influence, and enhancing maritime domain awareness using open-source and targeted data collection techniques. Vannevar Labs is dedicated to quickly iterating with national security partners to meet urgent and dynamic security challenges, positioning its solutions to address the strategic competition for legitimacy and influence in the modern geopolitical landscape.

Vannevar builds AI systems for the Department of War's most consequential missions. We have 125 deployments across every branch and combatant command spanning our core platform and five distinct products. An an example of how we work, when major combat operations started with Iran, we fielded a new product supporting 24/7 operations and 9,000+ users in three months.

How we build is our advantage: we deploy forward with the people who own the mission. Our engineers, product team, and CTO deploy forward to the point of friction, including visiting units in Ukraine. We focus on embedding directly with operational units supporting great power competition with China, combat operations with Iran, and counter-narcotics missions.

About the role

We are looking for an Site Reliability Engineer to own the reliability, health, and deployment automation of the platform at Vannevar Labs. In this role you'll be the person watching the system's pulse - monitoring dashboards, catching health issues before they become incidents, and owning the debugging process from first alert to resolution. Your decisions today will have a large impact on the company's future.

We believe that simple systems are easier to understand, maintain, and scale. You will be making trade-offs as you work to ensure that our systems are prepared to operate reliably in high-side environments at scale. A strong sense of judgment matters here: knowing when to dig deeper into a problem yourself and when to pull in the right people to escalate. Clear, calm communication - during an incident and in day-to-day work - is a must.

What you'll do

  • Monitor dashboards and system telemetry to detect health issues, performance degradation, and reliability risks - often before anyone else notices them.
  • Own the debugging and incident response process end to end, exercising good judgment about when to investigate more deeply and when to escalate.
  • Build logging, monitoring, and observability tooling to visualize the state of the platform and continuously mature our SRE practices.
  • Develop, maintain, and be responsible for overall platform health, scaling, and capacity planning.
  • Understand and help improve the deployment process, and automate build & deployment pipelines.
  • Identify bottlenecks in engineering workflows and drive improvements that make the whole team faster and more reliable.
  • Develop self-service tools and automation to improve engineering efficiency.
  • Play a critical part in implementing a secure, robust, high-availability delivery pipeline.
  • Communicate system status, trade-offs, and post-incident learnings clearly with teammates and stakeholders.

Qualifications

  • 5+ years of experience in SRE, DevOps, or software engineering.
  • Hands-on experience monitoring production systems and responding to incidents - comfortable owning a debugging process and making the call on when to dig in versus escalate.
  • Excellent communication skills, especially the ability to stay clear and organized while troubleshooting live issues.
  • Experience with the PLG stack, Datadog, or other enterprise monitoring/observability tools.
  • Experience participating in an on-call rotation and running or contributing to post-mortems.
  • Knowledge of AWS cloud technologies.
  • Familiarity with infrastructure-as-code technologies such as Terraform and Pulumi.
  • Experience with Python, Bash, or other scripting languages.
  • Experience working in an agile scrum environment, with the ability to work independently.
  • Able to quickly learn new and existing technologies.
  • Strong attention to detail and analytical capabilities.
  • Willingness and ability to work on-site in San Diego, CA.
  • U.S. Citizenship status is required, as this position requires the ability to access U.S.-only data systems and export-controlled data.
  • TS/SCI Clearance required.

Additional Qualifications (Nice to haves)

  • Experience defining and tracking SLOs/SLIs and error budgets.
  • Experience crafting CI/CD processes and automation.
  • Proficient with containerization technologies like Docker.
  • Experience working in AWS GovCloud.
  • Experience with modern web services architectures.
  • Experience with relational database systems, including SQL and relational design.
  • Experience working with Elasticsearch/OpenSearch.
  • Strong collaboration and negotiation skills, with the ability to work on cross-functional projects with internal partner engineering teams.

What we offer

We’re proud to offer competitive benefits that support our employees. Some key highlights of our benefits package include:

  • Health, dental, and vision insurance
  • 100% remote first culture. You can work from anywhere in the US and all full time employees have WeWork access
  • Unlimited PTO including competitive vacation and holiday schedules
  • Lifestyle stipends - Monthly mental health, wellness & fitness stipend, in-home office setup stipend and family planning assistance
  • Salary top-up during military reserve duty
  • Fully paid parental leave
  • Child and pet care reimbursement during travel

Vannevar is an equal opportunity employer, and qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender perception or identity, national origin, age, marital status, protected veteran status, or disability status.

We encourage candidates from all backgrounds to apply, even if you don't feel like you're a perfect fit. If you're passionate about contributing to our mission, we'd love to hear from you!

IMPORTANT NOTICE

We are committed to protecting the privacy of all applicants. Official emails from the company will come from an @vannevarlabs.com domain. Under no circumstances will a legitimate representative from our company contact you to request passwords, financial information, or other sensitive personal data. Please be vigilant of potential scams.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
703,239 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Diego
$102k – $190k per year • Equity • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Frisco • Toronto • Ann Arbor
Python
Go
Rust
Databases
PostgreSQL
AI/ML
Embeddings
AI Agents
LLM
OpenAI
Anthropic
DevOps
Rest API
Terraform
GCP
GitHub Actions
Vercel
Datadog
CI/CD
AWS
Kubernetes
SRE
Platform Engineering
Service Mesh
Amazon ECS
DNS
Cybersecurity
Snyk
SOC 2
QA
Playwright
Apply
$54k – $120k per year (Estimated) • Remote/Hybrid • Full-Time • 14+ years exp • Bachelor's Degree • Bengaluru
Python
Go
DevOps
Splunk
Terraform
Ansible
OpenTelemetry
Prometheus
CI/CD
AWS
Docker
Kubernetes
Grafana
Thanos
Amazon EKS
Alertmanager
Amazon CloudWatch
Linux
Unix
Apply
$30k – $62k per year (Estimated) • In office • Full-Time • 18+ years exp • Bachelor's Degree • Hyderabad
Python
JavaScript
Java
TypeScript
SQL
COBOL
Java
Spring Framework
Spring Boot
COBOL
IBM MQ
Databases
Oracle
Apache Kafka
Frontend
Angular
DevOps
Jenkins
AWS
Apply
$122k – $178k per year • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Austin • Richardson
Python
DevOps
Azure
AWS
Linux
Apply
Data Engineer 10 hours ago
$82k – $152k per year • Equity • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • McLean
Python
Java
SQL
Databases
Snowflake
Databricks
MS SQL
DevOps
Rest API
Azure
AWS
GitHub
Analytics
ETL/ELT
Apply
$172k – $284k per year (Estimated) • Remote • Top Secret • Raleigh
Management
Notion
Apply
$41k – $70k per year (Estimated) • Remote • TS/SCI • Internship • Master's Degree
Apply
$65k – $171k per year (Estimated) • Remote • 3+ years exp • Bachelor's Degree
Apply
$150k – $215k per year • Remote • 5+ years exp
Apply
$169k – $279k per year (Estimated) • Remote • Top Secret • Tampa
AI/ML
AI Agents
Apply
$115k – $184k per year • In office • Full-Time • 5+ years exp • San Diego • San Antonio
Analytics
Power BI
Apply
$56k – $90k per year • In office • Full-Time • 3+ years exp • Associate's Degree • San Diego
DevOps
Windows
Apply
Director of Sales 10 hours ago
$138k – $263k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • San Diego
Apply
$94k – $174k per year • Equity • In office • Full-Time • 4+ years exp • PhD • San Diego
AI/ML
Machine Learning
Apply
$161k – $241k per year • In office • Secret • Full-Time • Bachelor's Degree • San Diego
Apply
See all jobs
This is one of many
703,239 more open roles from verified company boards, updated every day.