845,403open jobs
53,832companies
143,151added this week
Browse all
Salary
≈ $40k – $103k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Senior · 4+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 27, 2026. First seen by Alion on Jun 24, 2026.

Overview
Company
Impact
Profile match
Honest and Grounded AI infrastructure for every research delivery team. 35+ research agencies use Metaforms for their research ops.

About Metaforms

Market research runs on 30-year-old survey platforms and armies of specialists hand-coding questionnaires in proprietary languages. Metaforms is the agent layer that does that work. Every survey is a program - full of skip logic, piping, quotas, and loops - and a single wrong number in a client report is unrecoverable. Our AI agents write production survey code, QA live deployments, process and clean large structured datasets, configure analysis, and generate client-ready reports, so agencies like Dynata, Savanta, and Borderless Access ship more projects with far less friction.

  • 1,000+ surveys processed monthly

  • Serving Fortune 500 companies across the globe

  • Rapid month-over-month growth

We’re Series A funded and scaling fast, aggressively growing our AI engineering team to build the next generation of production-grade AI agent systems.

The Role

We’re hiring a Senior AI Engineer to own the design, development, and continuous improvement of the AI agent systems that power modern research operations.

This is a high-ownership, high-impact role at the intersection of applied AI and systems engineering. You’ll work on genuinely hard problems: agent reliability at scale, long-context handling, cascading error mitigation, and evaluation infrastructure - like codegen agents that write in proprietary DSLs, computer-use agents that QA live deployments, data agents that clean tabular exports and configure multi-step analysis, and evals for outputs where “correct” is genuinely ambiguous. And you’ll do it on a team that ships fast and treats quality as non-negotiable.

What You’ll Own

Agent Harness and Architecture

  • Own the agent harness our production agents run on - the loop where agents plan, use tools, check their work, and recover from failures

  • Lead research and implementation for long-context handling and cascading-error challenges in multi-step agent pipelines

  • Drive context engineering strategy and experimentation frameworks across the team

Evaluation and Production Monitoring

  • Define structured rubrics for evaluating AI outputs on nuanced, ambiguous research tasks

  • Build continuous monitoring, tracing, and failure-mode analysis for agents in production - including the loop that turns production failures into test cases

  • Create tooling that lets domain experts refine and evolve the skill files, eval sets, and knowledge bases our agents consume

Reliability for High-Stakes Outputs

  • Build eval suites - regression sets, golden datasets, LLM-as-judge pipelines - that catch regressions before deploy

  • Develop evaluation datasets for DSLs, structured data transforms, and computed outputs to systematically find and close model weaknesses

  • Design human-in-the-loop and review workflows for outputs where a single wrong number in a client report is unrecoverable

What We’re Looking For

Must-Have

  • Built and operated agentic systems in production - multi-step pipelines, tool use, codegen, computer-use, or data and reporting agents - not just prototypes

  • 4+ years of engineering experience, with at least 1 year focused on LLM/agent systems in production

  • Deep hands-on experience with frontier model APIs (Anthropic, OpenAI, Gemini), evaluation frameworks, and AI system optimization

  • Strong Python skills; Go or TypeScript a plus

  • Solid grasp of context engineering and evaluation methodology

  • Strong instincts for debugging complex, non-deterministic system failures

  • High ownership: you drive problems to resolution independently and pull others in when it matters

Nice to Have

  • Experience with LLM observability and eval tooling (Braintrust, Langfuse, LangSmith, Weave, promptfoo, or in-house equivalents)

  • Background in semantic parsing, DSLs, or structured-output generation

  • Prior work on computer-use or browser agents

  • Experience with human-in-the-loop agent workflows where proposals are reviewed before apply, or agents over large structured datasets

Why Metaforms

  • Work at the frontier of production AI: systems handling 1,000+ research projects a month, with the reliability bar that implies

  • A small, senior team where your decisions carry real architectural weight

  • Zero-bureaucracy culture: high autonomy, fast feedback loops, direct access to leadership

  • Well-funded and financially stable, with a clear roadmap and the runway to execute on it

Benefits

  • Full family health insurance

  • $1,000 USD annual learning and development budget

  • Dedicated mentor and coaching support

  • Free snacks and dinner at the office

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
845,403 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Bengaluru
Staff AI Developer 6 days ago
≈ $46k – $103k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
Python
Java
C#
C++
C#
.NET
C++
TensorFlow C++
PyTorch C++
Databases
Apache Kafka
AI/ML
Spark
MLFlow
Vertex AI
Scikit-learn
NLP
Kubeflow
TensorFlow
Pandas
NumPy
PyTorch
Amazon SageMaker
Interpretability
Machine Learning
DevOps
Rest API
GCP
Azure
AWS
Docker
Kubernetes
GitHub
Cybersecurity
GDPR
HIPAA
Apply
≈ $40k – $101k per year (Estimated) • In office • Full-Time • 2+ years exp • India
Python
Java
AI/ML
LangGraph
LangChain
Vertex AI
Fine-tuning
Prompt Engineering
AI Agents
LLM
Google ADK
Machine Learning
DevOps
GCP
Google Cloud Run
IAM
Amazon ECS
Apply
≈ $28k – $72k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • India
Python
SQL
AI/ML
Vertex AI
Embeddings
Scikit-learn
Prompt Engineering
NLP
AWS Bedrock
TensorFlow
NumPy
PyTorch
RAG
Synthetic Data
Amazon SageMaker
OCR
Structured Outputs
Knowledge Graph
Tool Use
Machine Learning
DevOps
Azure
Git
AWS
Unix
Analytics
A/B Testing
Apply
≈ $28k – $71k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Bengaluru
Apply
≈ $39k – $99k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
Snowflake
Databricks
Oracle
Delta Lake
Apache Kafka
AI/ML
Spark
MLFlow
AI Agents
LLM
Feature Store
Machine Learning
DevOps
Terraform
GCP
Azure DevOps
GitHub Actions
Azure
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
Incident Management
Analytics
Azure Data Factory
Apply
$180k – $350k per year • In office • 20+ years exp • New York
Python
SQL
AI/ML
AI Agents
Apply
≈ $99k – $214k per year (Estimated) • In office • 20+ years exp • London
Python
SQL
AI/ML
AI Agents
Apply
$110k – $190k per year • In office • 4+ years exp • Princeton
Python
JavaScript
SQL
AI/ML
AI Agents
Analytics
Tableau
QlikSense
Management
Agile
Scrum
Apply
$110k – $190k per year • In office • 4+ years exp • New York
Python
JavaScript
SQL
AI/ML
AI Agents
Analytics
Tableau
QlikSense
Management
Agile
Scrum
Apply
$140k – $295k per year • In office • New York
Python
SQL
AI/ML
AI Agents
Management
Agile
Apply
≈ $38k – $79k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Python
JavaScript
TypeScript
AI/ML
AI Agents
LLM
Context Engineering
Frontend
React.js
DevOps
AWS
Apply
AI / ML Engineer 3 hours ago
≈ $39k – $99k per year (Estimated) • In office • 5+ years exp • Bengaluru
AI/ML
LangGraph
LangChain
Claude
AI Agents
AutoGPT
TensorFlow
CrewAI
Gemini
LLM
LLMOps
GPT-4
LLM Guardrails
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Apply
≈ $30k – $63k per year (Estimated) • In office • 6+ years exp • Bengaluru
Python
AI/ML
LLM
Agentic Workflows
DevOps
Rest API
gRPC
Docker
Kubernetes
Platform Engineering
Apply
In office • 5+ years exp • Bengaluru
JavaScript
Kotlin
TypeScript
Swift
AI/ML
Copilot
Cursor
ChatGPT
LLM
Frontend
GraphQL
React.js
Mobile
React Native
Firebase
Expo
State Management
Crashlytics
DevOps
Rest API
CI/CD
Git
Management
Gmail
Telegram
WhatsApp
QA
Sentry
Apply
≈ $23k – $62k per year (Estimated) • In office • 6+ years exp • Gandhinagar • Mumbai • Pune • Bengaluru
Apply
AWS Data Engineer 5 hours ago
≈ $22k – $47k per year (Estimated) • In office • 6+ years exp • Hyderabad • Bengaluru
Python
DevOps
AWS
AWS Lambda
Analytics
AWS Glue
Apply
See all jobs
This is one of many
845,403 more open roles from verified company boards, updated every day.