1,434,848open jobs
83,826companies
217,404added this week
Browse all
Salary
$135k – $168k per year
Location
Hybrid (San Carlos, United States)
Seniority
Junior · 2+ years exp
Visa
No sponsorship (stated in the posting)
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Sep 14, 2026.

Overview
Company
Impact
Profile match
Beacon AI is an aviation technology company that builds AI-powered pilot assistance systems and flight safety software for commercial, private, and defense aviation. Often described as an "R2-D2 for pilots", the platform acts as an intelligent onboard teammate - integrating directly with cockpit infrastructure to assist aviators with real-time checklists, route optimization, fleet monitoring, and situational awareness to reduce human error and workload.

About Beacon AI

We’re a fast-moving team of aviators, engineers, and operators building an AI platform to make flying safer, more efficient, and more capable. Backed by top investors, we’ve secured a dozen Department of Defense contracts and partnered with major airlines to deliver mission-critical systems. We operate without silos or heavy processes. Small, focused teams own what they build, ship quickly, and learn fast, pushing the boundaries of how humans and AI work together in aviation.

You will ship LLM-powered product features end-to-end. That means designing retrieval and tool-calling flows, writing the services that run them, building evals and guardrails, and watching cost, latency, and quality in production. You’ll partner with the ML/infra teammates on embeddings, indexing, and model hosting, and with the product teammates on user experience and outcomes. We move fast, and we care about reliability in a safety-critical domain.

This role is for engineers early in their LLM/production ML career (2-4 years software engineering experience). You'll implement features within existing service patterns, with architecture and eval design guided by senior engineers.

What you’ll do

Build user-facing LLM features

  • Design and implement retrieval-augmented generation and tool-calling flows using frameworks like LangChain or equivalent primitives, where simpler is better.

  • Deliver robust JSON and schema-bound outputs with validation, retries, and fallbacks.

  • Add function calling to integrate with internal tools, search, routing, and data services.

Own the service layer

  • Ship APIs and workers in Python or TypeScript with clear contracts, streaming, and backoff.

  • Add caching, request shaping, prompt templates, and context packing to control latency and cost.

  • Integrate with AWS Bedrock, OpenAI, Anthropic, or self-hosted endpoints as needed.

Retrieval and data prep

  • Collaborate with infrastructure teammates to develop chunking, embeddings, and indexing capabilities for documents, time series, and multimedia.

  • Choose and tune vector backends such as OpenSearch, pgvector, or Pinecone.

  • Keep knowledge bases fresh with data syncs from S3, Aurora, DynamoDB, and external sources.

Evaluation and quality

  • Create offline evals and golden sets for prompts, retrievers, and tools.

  • Stand up online metrics for task success, hallucination rate, retrieval precision/recall, p95 latency, and cost per request.

  • Run A/B tests and prompt/version rollouts with guardrails and canaries.

Safety, privacy, and compliance

  • Implement content and policy checks, PII detection and redaction, access controls, and auditing.

  • Design human-in-the-loop paths for sensitive actions.

  • Handle aviation data with care and follow internal security standards.

Operate what you build

  • Add tracing, logs, and dashboards for model calls, token usage, errors, and saturation.

  • Debug tricky failures across retrieval, prompts, tools, and providers.

What will make you successful

  • Shipped LLM apps: You’ve put LLM features in front of users and improved them with data.

  • Strong builder: Comfortable writing production code, tests, and docs. You keep things simple and observable.

  • RAG and tools depth: You understand embeddings, chunking, vector search tradeoffs, and function calling.

  • Quality mindset: You design evals, define success metrics, and iterate based on evidence.

  • Cost and latency aware: You track p95, hit SLAs, and reduce cost without hurting quality.

  • Clear communicator: You explain tradeoffs and align partners across product, infra, and security.

  • Growth mindset: You take direction well, ask good questions, and level up fast with feedback.

Nice to have

  • Experience with Bedrock, OpenSearch Serverless, pgvector, Pinecone, or Weaviate.

  • Prompt versioning, guardrails, and provider routing in production.

  • Multimodal work with time series or video.

  • Familiarity with GPU inference, Triton, or TensorRT-LLM.

  • Aviation or other safety-critical domain exposure.

  • DevOps basics for CI/CD, IaC, and secure secrets handling.

Example problems you might tackle in month one

  • Transform an internal knowledge base into a low-latency RAG service, complete with explicit schemas and evaluations.

  • Add tool-calling to automate a repetitive cockpit or ops workflow with guardrails and audit trails.

  • Reduce the cost per request through improved chunking, caching, and prompt refactoring, while maintaining task success rates.

Work Location

This is a hybrid role based in San Carlos, CA, with 3+ days per week onsite and the option to work remotely on remaining days.

Perks & Benefits (Full-Time Employees)

  • Healthcare: 100%* of employee medical premiums covered; 25% for dependents

  • Time Off: 3 weeks PTO plus 13+ paid company holidays

  • 401(k): Offered (no current employer match, but we are committed to enhancing this benefit in the future)

Due to U.S. export control regulations, we can only hire U.S. Persons (U.S. citizens, Green Card holders, lawful permanent residents, or individuals granted asylum or refugee status). We are unable to provide visa sponsorship or support visa transfers. All work must be performed in the United States.

Beacon AI is an equal opportunity employer and does not discriminate based on race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected characteristic. We prohibit harassment or discrimination of any kind in the workplace and comply with all applicable federal, state, and local employment laws.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,434,848 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
San Carlos
$73k – $121k per year • In office • Full-Time • PhD • Lemont
AI/ML
Fine-tuning
JAX
Multimodal AI
PyTorch
GNN
Machine Learning
DevOps
HPC
Robotics
Digital Twin
Apply
≈ $163k – $370k per year (Estimated) • In office • Full-Time • Bachelor's Degree • San Jose
Python
Java
Rust
C++
AI/ML
Claude Code
Model Context Protocol
vLLM
RLHF
Multimodal AI
Knowledge Distillation
AI Agents
NLP
TensorRT
TensorRT-LLM
LLM
RAG
Mixture of Experts
OpenAI Codex
DPO
SFT
Post-training
Context Engineering
LLM Guardrails
Prompt Caching
Tool Use
Model Distillation
RLAIF
Reward Modeling
Apply
≈ $162k – $368k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Seattle
Python
Java
Rust
C++
AI/ML
Claude Code
Model Context Protocol
vLLM
RLHF
Multimodal AI
Knowledge Distillation
AI Agents
NLP
TensorRT
TensorRT-LLM
LLM
RAG
Mixture of Experts
OpenAI Codex
DPO
SFT
Post-training
Context Engineering
LLM Guardrails
Prompt Caching
Tool Use
Model Distillation
RLAIF
Reward Modeling
Apply
≈ $163k – $370k per year (Estimated) • In office • Full-Time • PhD • San Jose
Python
Java
Rust
C++
AI/ML
Claude Code
Model Context Protocol
vLLM
RLHF
Knowledge Distillation
AI Agents
NLP
TensorRT
TensorRT-LLM
LLM
RAG
Mixture of Experts
OpenAI Codex
DPO
SFT
Post-training
Context Engineering
LLM Guardrails
Prompt Caching
Tool Use
Model Distillation
RLAIF
Reward Modeling
Apply
≈ $105k – $208k per year (Estimated) • In office • Contractor • 6+ years exp • Bachelor's Degree • Charlotte
Python
JavaScript
TypeScript
AI/ML
Computer Vision
AI Agents
NLP
AWS Bedrock
LLM
Machine Learning
Frontend
Angular
DevOps
CI/CD
AWS
Apply
≈ $78k – $165k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • San Carlos
Apply
≈ $64k – $130k per year (Estimated) • In office • 5+ years exp • High School Diploma • San Carlos
Apply
$144k – $169k per year • Equity • In office • Full-Time • 5+ years exp • Associate's Degree • San Carlos
Design
AutoCAD
Management
Microsoft Office
Apply
$190k – $230k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Carlos
Management
Linear
Jira
Apply
$80k – $120k per year • Equity • In office • Full-Time • San Carlos
Apply
See all jobs
This is one of many
1,434,848 more open roles from verified company boards, updated every day.