368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$29k – $72k per year (Estimated)
Location
Remote (India)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Credo AI is a company founded in 2020 that builds governance software for enterprise artificial intelligence. Its platform registers AI use cases, runs policy and fairness assessments and generates the evidence organisations need for frameworks such as the European Union AI Act and the NIST risk management standard. It works with large enterprises and public bodies that must show responsible AI practice.

About Credo AI

The Credo AI Governance Platform operates as the trust layer for global enterprises focused on transforming their organizations with AI. It is powered by a knowledge graph that is updated by agents, enriched with global regulatory intelligence and curated business context, supported by Forward-Deployed Experts, giving enterprises a clear view of the risk profile behind every AI system. By making that risk visible, we enable teams to move quickly and confidently scale across systems to adopt compliant AI, faster and in a compliant manner.

About Us

Credo AI is a leading AI governance platform helping enterprises implement responsible AI at scale. We are building the next generation of AI-powered governance solutions that enable organizations to manage risk, ensure compliance, and scale their AI deployments with confidence.

About the Role

This is a senior technical role building the AI systems that power Credo AI's governance products. You will design and build intelligent systems that observe, reason about, and act on governance knowledge - from the agent infrastructure that monitors AI behavior in enterprise deployments, to the knowledge systems that make governance intelligence accessible and actionable at scale.

You are someone who cares about how AI systems behave, not just whether they perform. You think carefully about reliability, consistency, and failure modes - and you have the engineering discipline to turn that thinking into production systems. You work across the full stack of modern AI engineering: LLM-based agents, retrieval and knowledge systems, evaluation pipelines, and the behavioral configuration layers that shape how AI acts within defined constraints.

You will write code, ship systems, and own technical direction - while collaborating closely with AI researchers, governance experts, and product teams.

What You'll Build

AI agent systems

  • Design and implement agent architectures that reason about governance policies and take action within defined constraints

  • Build instrumentation that exposes how agents reason and act, making their behavior auditable at the session level

  • Design and build the telemetry, filtering, and analytics infrastructure that lets governance owners empirically verify how agents are behaving at organizational scale

  • Develop behavioral configuration and constraint systems that encode organizational policies into how AI agents operate

  • Architect evaluation frameworks that measure whether agent behavior actually aligns with governance intent

RAG - Governance knowledge systems

  • Develop retrieval and context systems that surface relevant governance knowledge at the right moment, to the right system or user

  • Design hybrid retrieval architectures combining semantic search, structured knowledge traversal, and dynamic context assembly

Platform infrastructure

  • Contribute to the broader AI systems architecture underlying Credo AI's platform

  • Work with data and product teams to translate governance intelligence into reliable, scalable product features

  • Establish engineering standards and best practices for AI system development across the team

About You

You have built LLM-based systems in production and you have encountered their failure modes firsthand - inconsistency, instruction-following breakdowns, unexpected behaviors at distribution edges. You think carefully about how to specify, evaluate, and constrain AI behavior, and you bring engineering rigor to problems that sit at the boundary of ML research and systems design.

You read the literature on agent architectures, evaluation methodology, and alignment-adjacent topics not because your job requires it but because you find them genuinely useful. You move fast, ship things, and iterate - but you think carefully about failure modes before they reach production.

Minimum Qualifications

  • 5+ years building production AI/ML systems, with meaningful experience shipping LLM-based applications

  • Strong experience with agent architectures: tool use, planning, multi-step reasoning, and the failure modes that accompany them

  • Evaluation mindset - you have designed evals, run them, and used results to make systems meaningfully better

  • Experience with behavioral shaping techniques: prompt architecture, output validation, policy-grounded constraints

  • Solid systems engineering: you can design data pipelines, APIs, and distributed systems that hold up in production

  • Experience building monitoring or observability systems for AI applications in production

  • Experience with retrieval-augmented generation and tradeoffs in hybrid retrieval systems.

  • Strong communicator and collaborator

Preferred Qualifications

  • Research background or publications in LLM evaluation, alignment, agent safety, or AI robustness

  • Experience with red-teaming, adversarial evaluation, or automated failure detection

  • Background in multi-agent systems or systems where multiple AI components interact

  • Familiarity with AI governance, compliance, or risk management domains

  • Track record contributing to open-source ML or evaluation infrastructure

We don't expect any single candidate to meet every qualification listed. If you bring deep experience in some of these areas, sharp engineering judgment, and the curiosity to grow into the rest, we want to hear from you.

Why This Role

The AI systems problems at Credo AI are not solved problems. How do you reliably make AI agents behave within governance boundaries? How do you evaluate whether a governance-reasoning AI system is actually getting it right? You will be working on these questions in a domain - enterprise AI governance - where the right architectural decisions are still being made, with close collaboration between engineering, research, and deep domain expertise.

*Supports EST/CST/PST time zone (your choice)

Compensation

The expected base salary range for this position is ₹40-50L. Our salary ranges are determined by role, level, and location. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position in the specified location. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.

Location & Remote Culture

While this is a remote role and we're a fully distributed team, we routinely meet up in-person. We support individual members to coordinate in-person coworking whenever possible, and organize company-wide offsites multiple times a year. At Credo AI we value diversity, equity, and inclusion as core principles in our work environment, and the development of our product offerings, and we have implemented initiatives to foster and support these values.

Credo AI Benefits & Perks

  • Competitive Salary and Equity

  • Health: We offer health, dental, and vision coverage. We also offer an ergonomic benefit to cover the costs of equipment to help staff stay healthy while working, both in the office and at home.

  • Coworking: We will cover the cost of co-working spaces like WeWork and in-person meetups.

  • Unlimited PTO: Credo AI has unlimited time off to support our employees

  • Generous Parental Leave: We offer up to 12 weeks of paid parental leave.

  • 401(k) plan for employees (US only)

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$28k – $53k per year (Estimated) • In office • Full-Time • Rostov-on-Don
AI/ML
LLM
RAG
Apply
$39k – $87k per year (Estimated) • In office • 9+ years exp • Bengaluru
Java
Python
Java
Spring Boot
Databases
ElasticSearch
Google BigQuery
Neo4j
PostgreSQL
AI/ML
Claude
Claude Code
Copilot
Cursor
LLM
AI Agents
Devin
Model Context Protocol
OpenAI Codex
DevOps
AWS
Datadog
GCP
Grafana
New Relic
Prometheus
Management
Jira
Apply
$21k per year • In office • Contractor • Yekaterinburg
Python
SQL
Python
FastAPI
Flask
AI/ML
Claude
Claude Code
Embeddings
Function Calling
LLM
RAG
OpenAI
OpenAI Codex
Structured Outputs
DevOps
Docker
Git
Apply
ML Engineer 1 day ago
Remote/Hybrid • 3+ years exp
Python
SQL
AI/ML
AI Agents
Computer Vision
MLFlow
PyTorch
RAG
A2A
Model Context Protocol
DevOps
AWS
Azure
CI/CD
Docker
Docker Compose
GCP
Kubernetes
Analytics
ETL/ELT
Apply
$140k – $225k per year • Remote • Full-Time • 5+ years exp • Seattle
Python
Lua
Python
FastAPI
Celery
Pydantic
SQLAlchemy
Databases
pgvector
PostgreSQL
Redis
AI/ML
LLM
RAG
Hybrid Search
Reranking
Anthropic
LLM Evaluation
LLM Guardrails
OpenAI
Model Context Protocol
DevOps
OpenTelemetry
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.