812,549open jobs
52,293companies
130,976added this week
Browse all
Salary
$150k – $250k per year
Location
Remote (United States)
Seniority
Staff · 6+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 26, 2026. First seen by Alion on Sep 25, 2026.

Overview
Company
Impact
Profile match
Build, develop, test, and understand mobile applications with cloud compute, live devices, agent-driven verification, and Atlas runtime evidence.

As a Staff Infrastructure Engineer at Revyl, you'll own the technical strategy and systems required to scale the infrastructure that powers our mobile development platform.

Revyl gives developers and AI agents access to cloud iOS and Android environments where they can build, run, inspect, and verify mobile applications. Behind that experience is a distributed compute platform spanning GCP, physical Mac infrastructure, simulators, emulators, build runners, streaming, and test execution.

We're entering a period where our infrastructure needs to support significantly more customers, workloads, and compute. You'll be responsible not only for operating that infrastructure, but for figuring out how it needs to evolve.

You'll dig into how our platform operates today, identify bottlenecks and scaling limits, understand where operational toil and reliability issues are coming from, model future capacity requirements, and develop the technical roadmap that gets us from where we are today to where we need to be.

This isn't a traditional DevOps or SRE role. We're looking for someone who can move fluidly between understanding the current system, designing the next version of it, and getting hands-on to build it.

What you'll work on

  • Define our infrastructure scaling strategy. Understand the architecture and operational characteristics of Revyl's platform today, identify the systems that will break as load increases, and develop a roadmap for scaling them ahead of demand.

  • Capacity planning and modeling. Translate expected customer and partnership growth into compute, storage, network, device, and hardware requirements. Build the models and instrumentation that let us understand when and where we need to add capacity.

  • Scaling our device cloud. Design and build the systems that provision, schedule, manage, and recycle large fleets of iOS simulators, Android emulators, and physical compute.

  • Distributed workload orchestration. Evolve how builds, tests, agent sessions, and interactive workloads are scheduled across heterogeneous pools of compute as concurrency grows.

  • Cloud infrastructure. Scale and evolve the GCP infrastructure behind Revyl's APIs, control plane, workflow execution, storage, networking, and supporting services.

  • Identify and eliminate bottlenecks. Use production data to understand where we're constrained by CPU, memory, disk, network, scheduler throughput, database performance, host capacity, or architecture-and determine the right solution rather than simply adding more machines.

  • Build for step-function growth. Prepare Revyl for customers and partnerships that can introduce significantly more traffic than the platform handles today. Design load tests, failure tests, and capacity plans that give us confidence before traffic arrives.

What we're looking for

We are looking for someone who has taken an infrastructure platform from one stage of scale to the next.

You likely have:

  • Has 5+ years of software engineering, infrastructure, platform engineering, or SRE experience, with significant technical ownership.
  • Has previously helped scale a production infrastructure or compute platform through a meaningful increase in load.
  • Can analyze an existing architecture, identify scaling constraints, and turn those findings into a prioritized technical roadmap.
  • Understands capacity planning and can translate business forecasts and expected workload growth into infrastructure requirements.
  • Is capable of making architectural decisions under uncertainty and knows how to validate assumptions through instrumentation, benchmarking, load testing, and production data.
  • Has built or operated distributed systems at meaningful scale.
  • Is a strong software engineer who can build infrastructure services and control-plane systems, not just configure infrastructure tooling.
  • Understands scheduling, queues, worker pools, concurrency, backpressure, resource allocation, failure recovery, and distributed systems fundamentals.
  • Has deep experience with cloud infrastructure such as GCP, AWS, or Azure.

What success looks like

Within your first year, we'd expect you to have materially changed both Revyl's infrastructure and our understanding of how it scales.

That means:

  • We have a clear model of the capacity and architecture required to support our next stages of growth.
  • We know where our major scaling limits are before customers encounter them.
  • Large new customers and partnerships have concrete capacity and load plans before launch.
  • Revyl can handle an order of magnitude more concurrent compute without requiring an order of magnitude more operational work.
  • Common infrastructure failures are automatically detected and remediated.
  • Device scheduling and capacity management are predictable rather than reactive.
  • We can confidently answer how much additional traffic the platform can support.
  • Infrastructure projects are prioritized based on real scaling constraints rather than whichever fire happened most recently.
  • Our existing SRE and engineering team spend substantially less time firefighting.
  • The Device Cloud becomes a platform we can scale repeatedly rather than an infrastructure stack that needs to be reinvented every time Revyl grows.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
812,549 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
San Francisco
≈ $103k – $198k per year (Estimated) • In office • Full-Time • 5+ years exp • Glendale
Python
PowerShell
AI/ML
AutoGen
LangChain
Model Context Protocol
Prompt Engineering
AI Agents
LLM
Anomaly Detection
Human-in-the-Loop
Frontend
GraphQL
DevOps
Rest API
Terraform
Azure
CI/CD
GitOps
Windows Server
Platform Engineering
Self-Healing
Bicep
Windows
TCP/IP
DNS
DHCP
VPN
Cybersecurity
HIPAA
Microsoft Entra ID
Management
OneDrive
SharePoint
Apply
≈ $115k – $212k per year (Estimated) • Remote (United States) • Full-Time • 5+ years exp • Austin
Python
JavaScript
Node JS
Frontend
React.js
DevOps
Rest API
GCP
Git
Management
Miro
Zapier
Apply
≈ $93k – $206k per year (Estimated) • Equity • Remote (Poland) • Full-Time • 5+ years exp
Python
AI/ML
Anthropic
DevOps
Terraform
GCP
GitHub Actions
Azure
CI/CD
AWS
Platform Engineering
GitHub
IAM
Management
Agile
Apply
≈ $90k – $200k per year (Estimated) • Equity • Remote (Poland) • Full-Time • 3+ years exp
AI/ML
Anthropic
DevOps
GCP
Azure
AWS
Platform Engineering
Management
Power Automate
Power Apps
Agile
Apply
≈ $89k – $199k per year (Estimated) • Equity • Remote (Poland) • Full-Time • 5+ years exp
Python
AI/ML
Anthropic
DevOps
Rest API
Terraform
GCP
Azure DevOps
GitHub Actions
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Platform Engineering
Bicep
GitHub
IAM
Cybersecurity
Microsoft Sentinel
Microsoft Defender
Microsoft Defender for Cloud
Microsoft Entra ID
Management
Agile
QA
Pytest
Apply
$60k – $85k per year • Equity 0.1–0.3% • Remote (Poland, Romania, Ukraine, Lebanon, Lithuania) • Full-Time • 3+ years exp
Python
SQL
Databases
MySQL
PostgreSQL
Redis
DynamoDB
RabbitMQ
ActiveMQ
ElasticSearch
Apache Kafka
OpenSearch
AI/ML
Prompt Engineering
AI Agents
LLM
DevOps
GCP
DigitalOcean
Jenkins
AWS
Docker
Kubernetes
GitLab
Apply
Founding Engineer 1 day ago
$150k – $175k per year • Equity 2.5–5% • In office • Full-Time • New York
Python
JavaScript
TypeScript
Python
FastAPI
Databases
PostgreSQL
AI/ML
dbt
AI Agents
LLM
Frontend
React.js
DevOps
AWS
Cloudflare
Apply
$160k – $220k per year • Equity 0.2–1% • Remote (United States, Thailand) • Full-Time • 6+ years exp • San Francisco
AI/ML
Model Context Protocol
AI Agents
Apply
$120k – $240k per year • Equity 0.2–1% • Remote (United States, Thailand) • Full-Time • 6+ years exp • San Francisco
Rust
TypeScript
AI/ML
Model Context Protocol
AI Agents
Apply
Backend Engineer 1 day ago
$120k – $200k per year • Equity 0.2–0.8% • In office • Full-Time • New York
AI/ML
Embeddings
Multimodal AI
Computer Vision
AI Agents
Context Engineering
LLM Guardrails
Management
Stripe
Apply
GTM Intern 1 day ago
$60k – $120k per year • In office • Internship • San Francisco
Apply
Software Engineer 1 day ago
$125k – $200k per year • In office • Full-Time • 1+ year exp • San Francisco
Kotlin
Dart
Swift
AI/ML
AI Agents
Mobile
Flutter
DevOps
OpenTelemetry
WebSockets
Apply
Engineering Intern 1 day ago
$60k – $120k per year • In office • Internship • San Francisco
Python
Go
TypeScript
Apply
$120k – $180k per year • Equity 0.2–1% • In office • Full-Time • 6+ years exp • San Francisco
DevOps
CI/CD
Apply
Founding Engineer 1 day ago
$110k – $180k per year • Equity 0.1–1% • In office • Full-Time • San Francisco
Python
JavaScript
Node JS
AI/ML
Vertex AI
OpenAI
Anthropic
Frontend
Next.js
React.js
DevOps
Azure
Kubernetes
Apply
$60k – $84k per year • In office • Internship • San Francisco
Python
JavaScript
TypeScript
AI/ML
Copilot
Cursor
Claude
Claude Code
Model Context Protocol
Vertex AI
AI Agents
LLM
OpenAI
Anthropic
LLM Guardrails
Tool Use
Frontend
Next.js
React.js
DevOps
Azure
AWS
Kubernetes
Apply
$80k – $135k per year • Equity 0.1–0.5% • Remote (United States) • Full-Time • 1+ year exp • San Francisco
Management
Zapier
Apply
$40k – $120k per year • Equity 0.1–0.5% • Remote (United States) • Full-Time • San Francisco
Frontend
Next.js
React.js
Apply
See all jobs
This is one of many
812,549 more open roles from verified company boards, updated every day.