1,274,618open jobs
74,028companies
209,134added this week
Browse all
Salary
≈ $53k – $133k per year (Estimated)
Location
Remote (Brazil)
Seniority
Senior

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Oct 6, 2026.

Overview
Company
Impact
Profile match
Codurance is a global software consultancy firm focused on helping businesses modernize their technical capabilities and build sustainable growth. With a dedication to Software Craftsmanship, Codurance emphasizes professionalism and technical excellence in software development. The company offers a range of services, including software modernization, product development, and cloud engineering, tailored to help clients solve complex challenges and achieve their business goals.

Design and extend production-grade LLM applications and agentic workflows using

NestJS, XState v5, and the OpenAI SDK - flows include RAG, intent detection,

clarification, fulfillment, escalation, tool-use, and human-in-the-loop state machines

- Build and maintain the conversation-machine substrate: guard/action registries, flow

validation (ajv), DB-driven flow configs, and design-time tooling in Epicenter admin

- Build and evolve the AI systems behind Epic Support Assistant (ESA), the

player-facing support chatbot, and Agent Support Assistant, the AI copilot used by

customer support agents

- Integrate with MCP servers (Model Context Protocol) for tool-use and agentic behaviors

- Evaluate, benchmark, and tune models across providers including OpenAI, Gemini,

Anthropic, and future providers; own model selection decisions balancing quality,

latency, throughput, reliability, and cost

- Troubleshoot production LLM issues including hallucinations, retrieval failures, prompt

regressions, model drift, token inefficiencies, latency bottlenecks, and provider outages

- Build resilience mechanisms: retries, fallback routing, caching, streaming, rate limiting,

and provider routing

- Instrument and tune model quality using Langfuse (tracing, evals, prompt

management), evaluation datasets, A/B testing, prompt versioning, and production

telemetry

- Manage async workloads via BullMQ and caching with Redis; PostgreSQL persistence

via Kysely

Requirements

Must-Have

- Proven experience building and operating production LLM-powered systems

similar in scope to chatbots, AI assistants, agent copilots, RAG systems, or LLM

orchestration platforms

- Strong TypeScript/Node.js engineering; TypeScript strict-mode fluency

- Production AI experience: prompt engineering, RAG pipelines, agent design, tool

calling, model evaluation, observability, and failure-mode analysis - you've shipped AI

features, not just prototyped them

- Fullstack depth: comfortable moving between NestJS APIs, React UIs, databases,

infrastructure, and production operations; you don't artificially limit yourself to one layer

- Ability to evaluate tradeoffs between model quality, latency, reliability, throughput,

and cost

- Ability to troubleshoot AI systems across prompts, retrieval pipelines, model

configuration, infrastructure, and application code

- State machine thinking - you naturally model complex async workflows; XState or

similar experience is a strong signal

- Solid understanding of REST API design, async patterns (queues, events), and caching

strategies

- Strong testing culture: unit, integration, and contract tests are first-class deliverables, not

afterthoughts

- Experience working in a monorepo with multiple interconnected services

Strong Plus

- Hands-on experience with MCP (Model Context Protocol) or building tool-use agentic

workflows

- Familiarity with Langfuse or other LLM observability/evaluation platforms

- Experience operating AI workloads at scale

- Experience evaluating multiple foundation models and providers

- Experience building AI copilots, assistants, or conversational products

- Experience with semantic search and retrieval architectures

- Experience with AI gateways such as Portkey or similar platforms

- Experience with NestJS specifically: modules, providers, guards, interceptors, DI

patterns

- Background in customer support or player support platforms - you understand the

stakes of getting AI-generated responses wrong

- Experience shipping under low-latency constraints (chatbot response time budgets,

streaming)

- Previous work in gaming or high-volume consumer products

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,274,618 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
In your city
≈ $26k – $56k per year (Estimated) • Remote (likely EAEU) • 3+ years exp • Moscow
Java
Databases
PostgreSQL
Weaviate
ClickHouse
Milvus
pgvector
Qdrant
TimescaleDB
AI/ML
LangGraph
LangChain
Model Context Protocol
vLLM
Embeddings
AI Agents
LLM
RAG
OpenAI
LLM Guardrails
DevOps
Rest API
gRPC
Terraform
Zabbix
OpenTelemetry
Prometheus
CI/CD
Git
Grafana
Management
Confluence
UML
Apply
≈ $109k – $238k per year (Estimated) • Remote (Ukraine)
Python
AI/ML
Weights & Biases
OpenCV
Computer Vision
NumPy
PyTorch
OCR
DevOps
Git
GitHub
Apply
≈ $109k – $238k per year (Estimated) • Remote (Europe)
Python
AI/ML
Weights & Biases
OpenCV
Computer Vision
NumPy
PyTorch
OCR
DevOps
Git
GitHub
Apply
≈ $115k – $251k per year (Estimated) • Equity • Remote (Poland) • Full-Time • 5+ years exp
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
CUDA Toolkit
Fine-tuning
TensorFlow
PyTorch
Gemini
LLM
CUDA
Anthropic
TPU
Machine Learning
DevOps
GCP
Azure
AWS
Platform Engineering
Management
Agile
Apply
≈ $117k – $255k per year (Estimated) • Remote (United States) • Full-Time
AI/ML
LLM
LLM Evaluation
Apply
See all jobs
This is one of many
1,274,618 more open roles from verified company boards, updated every day.