368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$67k – $129k per year (Estimated)
Location
Remote/Hybrid (Bulgaria, Cyprus, Czech Republic, Hungary, Italy, Poland, Romania, Serbia, Spain, Ukraine, Albania, Estonia, Greece, Lithuania, Moldova, North Macedonia, Portugal, Slovakia, Slovenia)
Seniority
Architect · 6+ years exp
Employment
Contractor
Overview
Company
Impact
Profile match
Neurons Lab helps financial institutions move from AI-curious to AI-enabled. We deliver AI training programs and custom AI agents designed for regulated environments — from pilot to production, at scale.

About the project (description, duration, stage)

Hands-on Tech Lead for an AI Companion in an online mahjong game. The client is a social gaming company (web3 element) that scales its product and team. We deliver the AI side of their game as their embedded AI partner.

The AI Companion plays mahjong at a strong level and explains its moves. The core of the role is to build the mahjong-playing algorithm: a dedicated decision-making model (RL, imitation learning, or search-based - trained on the client's hand-history data) with an LLM reasoning layer on top. Key design constraints: a valid-action contract with the game engine (the bridge supplies legal moves), win detection, and a 2-second response budget per move. Explanations run async. Support for more than one rule set (riichi and regional variants) is on the roadmap.

Duration: 3 months, 0.5 FTE.

What you'll actually do (example tasks)

  • Design and build the mahjong-playing algorithm: choose and defend the approach (imitation learning on hand histories, RL / self-play, search with MCTS, or a hybrid), then train, evaluate, and ship it.

  • Own the technical architecture end to end: game model + LLM reasoning layer, valid-action mask, win detection, and the API contract with the client's game bridge.

  • Hit the 2-second response budget: design and measure the inference path, batching, and caching; keep a latency buffer for the client-facing number.

  • Define what data and event names we need from the client (hand histories, event streams); build the training and calibration pipeline on that data.

  • Build and run the evaluation harness: measure play strength against the client's reference points, and validate explanation quality.

  • Stand up LLM observability with Langfuse (async logging, N+1 batch) as an early sprint quick win.

  • Take over context from Vlad Borysenko (0.15-0.2 FTE supervision during ramp-up) and lead the sprint work with the AI Engineer; work with the client's Product Owner in a scrum process.

  • Front the client's CTO and engineers on technical decisions; explain trade-offs in plain language and in depth when asked.

  • Watch the risks the account team flagged: licensing on new training data, engine-bridge capabilities, and multi-rule-set scope.

Skills (hands-on first)

  • Game AI / sequential decision-making: hands-on RL, imitation learning, or search-based agents (MCTS, self-play) - ideally for imperfect-information games (mahjong, poker, card games)

  • Expert Python for ML systems; strong software engineering (APIs, testing, CI)

  • Model training on gameplay data end to end: data → training → evaluation → serving

  • LLM application engineering: reasoning layers, prompt and context design, structured outputs, guardrails

  • Low-latency inference: profiling, batching, caching, model-size trade-offs against a hard time budget

  • LLM observability and evaluation (Langfuse or similar)

  • AWS deployment for ML workloads

  • Technical leadership of a small pod; clear written and spoken communication with client engineers and executives

Knowledge

  • Game theory for imperfect-information games; evaluation of play strength (win rates, Elo-style ratings, baseline agents)

  • Game-engine integration patterns (event streams, action masks, state bridges)

  • Web3 / gaming product context - plus, not required

  • AWS Well-Architected for ML workloads

Experience

Key characteristics (ideally 4/4):

  • Hands-on ML/AI engineering at production scale

  • Shipped an AI system inside a live product with hard latency limits

  • Cloud hyperscaler experience (AWS preferred)

  • Technology consulting / client-facing delivery background

Role-specific characteristics:

  • 6+ years hands-on ML/AI engineering, with real game AI or sequential decision-making work (RL / MCTS / self-play - not only LLM apps)

  • Trained models on user or gameplay data end-to-end (data → training → evaluation → serving)

  • Led small delivery teams while still coding personally

  • Comfortable owning an architecture in front of a technical client CTO

Questions for Applicants

  • Imperfect information: mahjong hides most tiles from each player. How does hidden information change your algorithm choice compared to a perfect-information game like chess?

  • Latency budget: tell us about a system you shipped with a hard response-time limit. How did you design, measure, and defend the budget?

  • LLM + model hybrid: how would you combine a trained game model with an LLM explanation layer so the explanation never contradicts the move?

  • Hands-on + lead: how do you balance personally coding the hard parts with leading an engineer and fronting the client?

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$96k – $218k per year (Estimated) • Equity • In office • Full-Time • 8+ years exp • Toronto
Python
Databases
Databricks
Snowflake
AI/ML
AI Agents
AWS Bedrock
AWS Bedrock AgentCore
LLM
LLM Evaluation
DevOps
AWS
CI/CD
GCP
Apply
$122k – $200k per year • In office • Full-Time • 10+ years exp • Jersey City • Charlotte
Java
Node JS
Python
SQL
JavaScript
Java
Hibernate
Spring Framework
Spring MVC
Databases
Apache Kafka
DynamoDB
Oracle
RabbitMQ
AI/ML
Ray
DevOps
Amazon CloudWatch
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
Ansible
API Gateway
AWS
AWS Lambda
Azure
CI/CD
CloudFormation
GCP
Git
Grafana
IAM
Jenkins
Prometheus
Splunk
Terraform
Kubernetes
Apply
Founding Engineer 6 hours ago
$120k – $200k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • San Francisco
TypeScript
AI/ML
AI Agents
Claude
LLM
DevOps
GitHub
Management
Linear
Slack
Apply
$111k – $200k per year (Estimated) • In office • Full-Time • 1+ year exp • Plano
JavaScript
Python
C#
C#
.NET
DevOps
Azure
CI/CD
Apply
$71k – $112k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Austria
C#
C++
Python
Visual Basic
Apply
$77k – $138k per year (Estimated) • Remote • Contractor • 6+ years exp
Python
TypeScript
AI/ML
AWS Bedrock
Claude
Copilot
Langfuse
LLM
Prompt Engineering
LiveKit
LLM Guardrails
Mobile
Twilio
DevOps
AWS
Chips/EDA
PoC Library
Analytics
A/B Testing
Apply
$52k – $92k per year (Estimated) • Remote • Full-Time • 4+ years exp
Python
SQL
Databases
OpenSearch
pgvector
Pinecone
PostgreSQL
AI/ML
Composio
LLM
RAG
Unstructured.io
Model Context Protocol
DevOps
AWS
GCP
AWS Step Functions
Cybersecurity
GDPR
Management
Google Workspace
Slack
Apply
$39k – $85k per year (Estimated) • Remote • Full-Time
SQL
AI/ML
Knowledge Distillation
LLM
Human-in-the-Loop
Knowledge Graph
Cybersecurity
GDPR
Management
Slack
Apply
$74k – $129k per year (Estimated) • Remote • Full-Time • Madrid
AI/ML
AWS Bedrock
Claude
Claude Code
Cursor
LiteLLM
LLM
OpenRouter
LLM Guardrails
OpenAI Codex
AI Agents
Model Context Protocol
DevOps
AWS
Apply
$62k – $106k per year (Estimated) • Remote • Part-Time • 5+ years exp • Madrid
Databases
Neo4j
AI/ML
AutoGen
LangChain
LLM
Prompt Engineering
AI Agents
EU AI Act
Knowledge Graph
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.