368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$70k – $139k per year
Location
Remote (EAEU)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
All-in-One Ecosystem to Discover, Develop, & Commercialize Innovations.

AI Engineer - LLM Systems // Evals @ Ventora

Location: Remote

Employment Type: Full-time

Level: Mid-level / Senior

About Ventora

Ventora is an AI platform that helps people turn ideas into real products - from business planning and MVP creation to payments, marketing, and market validation.

The core platform is already built and in use. We’re now entering a stage of active growth and are looking for an AI-first engineer to help us make our AI systems smarter, more reliable, and easier to improve.

This is a hands-on product engineering role with real ownership. You’ll build LLM-powered features, improve AI pipelines, and create practical evaluation systems that show us what works and what needs to be fixed.

Your Mission

Help Ventora ship better AI systems - and make their quality measurable.

You’ll work across the product’s AI pipeline: business plan generation, app building, code generation, review, and marketing.

You’ll:

  • Build and improve LLM-powered product features
  • Develop multi-stage AI agents and workflows
  • Integrate and test new models
  • Improve prompts, context, routing, and fallbacks
  • Create automated evals and regression tests
  • Measure quality, cost, speed, and reliability
  • Turn evaluation findings into product improvements
  • Contribute to regular product engineering when needed

This is not a research-only or eval-only position. You’ll ship production code and see your work directly affect the product and its users.

What You’ll Work On

AI Systems & Product Features

Build and improve AI-powered features across Ventora’s product journey.

This may include multi-stage workflows such as plan → build → review, model integrations, structured outputs, tool calling, streaming, caching, prompt improvements, and new AI agents.

You’ll also contribute to production features, bug fixes, and pipeline development in the main codebase.

Evals & Quality

Help us understand whether each product change actually makes the AI better.

You’ll create practical evaluation systems for generated applications, business plans, marketing content, AI reliability, and release regressions. Depending on the task, you may use automated checks, browser tests, LLM judges, human-reviewed datasets, and side-by-side comparisons. When something fails, you’ll help find the cause, improve the prompt or pipeline, ship the fix, and add a regression test so the same issue does not return.

The goal is simple: catch problems before users do and make product decisions based on evidence.

What Success Looks Like

  • New AI features and improvements reach production regularly
  • AI quality becomes visible and measurable
  • Regressions are caught before release
  • Generated products become more functional and reliable
  • Evaluation findings consistently turn into shipped fixes
  • Model quality, cost, and latency stay under control

Over time, you’ll become a key technical partner for the founders, product team, and engineers shaping Ventora’s AI platform.

What You Bring

Required:

  • 3+ years of experience in AI/ML, applied data science, or backend engineering
  • Hands-on experience building production applications with LLMs
  • Experience with model APIs such as OpenAI, Anthropic, or OpenRouter
  • Understanding of prompts, agents, tool calling, structured outputs, and multi-stage AI workflows
  • Ability to turn unclear questions like “Is this version better?” into practical, measurable tests
  • Strong ownership, product thinking, and clear communication
  • Regular use of AI coding tools such as Codex, Claude Code, Cursor, or similar
  • Comfortable working in Linux environments and with Docker

You don’t need experience with every framework or evaluation platform. We care more about strong engineering fundamentals, curiosity, and the ability to l"}

menu

AI Engineer LLM Systems

Прямой работодатель Atlantix( www.atlantix.cc )

Миддл

  • Сеньор

Информационные технологии

  • Разработка
  • SaaS/PaaS

11 августа

Удаленная работа

Опыт работы любой

Работодатель Atlantix

Короткая ссылка: geekjob.ru/hidX

Откликнуться

Описание вакансии

AI Engineer - LLM Systems // Evals @ Ventora

Location: Remote

Employment Type: Full-time

Level: Mid-level / Senior

About Ventora

Ventora is an AI platform that helps people turn ideas into real products - from business planning and MVP creation to payments, marketing, and market validation.

The core platform is already built and in use. We're now entering a stage of active growth and are looking for an AI-first engineer to help us make our AI systems smarter, more reliable, and easier to improve.

This is a hands-on product engineering role with real ownership. You'll build LLM-powered features, improve AI pipelines, and create practical evaluation systems that show us what works and what needs to be fixed.

Your Mission

Help Ventora ship better AI systems - and make their quality measurable.

You'll work across the product's AI pipeline: business plan generation, app building, code generation, review, and marketing.

You'll:

  • Build and improve LLM-powered product features
  • Develop multi-stage AI agents and workflows
  • Integrate and test new models
  • Improve prompts, context, routing, and fallbacks
  • Create automated evals and regression tests
  • Measure quality, cost, speed, and reliability
  • Turn evaluation findings into product improvements
  • Contribute to regular product engineering when needed

This is not a research-only or eval-only position. You'll ship production code and see your work directly affect the product and its users.

What You'll Work On

AI Systems & Product Features

Build and improve AI-powered features across Ventora's product journey.

This may include multi-stage workflows such as plan → build → review, model integrations, structured outputs, tool calling, streaming, caching, prompt improvements, and new AI agents.

You'll also contribute to production features, bug fixes, and pipeline development in the main codebase.

Evals & Quality

Help us understand whether each product change actually makes the AI better.

You'll create practical evaluation systems for generated applications, business plans, marketing content, AI reliability, and release regressions. Depending on the task, you may use automated checks, browser tests, LLM judges, human-reviewed datasets, and side-by-side comparisons. When something fails, you'll help find the cause, improve the prompt or pipeline, ship the fix, and add a regression test so the same issue does not return.

The goal is simple: catch problems before users do and make product decisions based on evidence.

What Success Looks Like

  • New AI features and improvements reach production regularly
  • AI quality becomes visible and measurable
  • Regressions are caught before release
  • Generated products become more functional and reliable
  • Evaluation findings consistently turn into shipped fixes
  • Model quality, cost, and latency stay under control

Over time, you'll become a key technical partner for the founders, product team, and engineers shaping Ventora's AI platform.

What You Bring

Required:

  • 3+ years of experience in AI/ML, applied data science, or backend engineering
  • Hands-on experience building production applications with LLMs
  • Experience with model APIs such as OpenAI, Anthropic, or OpenRouter
  • Understanding of prompts, agents, tool calling, structured outputs, and multi-stage AI workflows
  • Ability to turn unclear questions like "Is this version better?" into practical, measurable tests
  • Strong ownership, product thinking, and clear communication
  • Regular use of AI coding tools such as Codex, Claude Code, Cursor, or similar
  • Comfortable working in Linux environments and with Docker

You don't need experience with every framework or evaluation platform. We care more about strong engineering fundamentals, curiosity, and the ability to learn and ship.

Nice to Have

  • Experience with LLM evals, datasets, scorers, or CI quality gates
  • TypeScript or Node.js experience
  • Browser automation, SQL, or production monitoring experience
  • Familiarity with tools such as promptfoo, DeepEval, Ragas, Braintrust, LangSmith, or Langfuse
  • Experience with agent frameworks, RAG evaluation, vision models, or AI red-teaming
  • A background in mathematics, statistics, physics, machine learning, or another quantitative field
  • Early-stage startup or side-project experience

Why Join Ventora

  • Build real AI systems used in production
  • Influence both the product and its technical foundation
  • Own projects from idea to launch and measurement
  • Work directly with founders, product, and engineering
  • Experiment with new models and AI development workflows
  • Help define quality standards for an AI-native platform
  • Grow into a leading role as the product and team expand

If you enjoy building with LLMs, care about whether AI systems actually work, and want your work to have a direct product impact, we'd be glad to talk.

Join us and help build the next stage of Ventora!

Специализация

Информационные технологии

Разработка

Отрасль и сфера применения

SaaS/PaaS

Уровень должности

Миддл

Сеньор

Откликнуться на вакансию

Быстрый отклик и регистрация/авторизация

Или быстрая регистрация/авторизация через OAuth

* Продолжая регистрацию/авторизацию вы подтверждаете свое согласие на хранение и обработку своих персональных данныхв соответствии с Федеральным законом № 152-ФЗ "О персональных данных", а так же подтверждаете, что прочитали и принимаете Пользовательское соглашение и согласие на обработку персональных данных, Условия использования сайта, Политику Конфиденциальности иРазъяснение использования личных данных,и согласны получать информацию и рассылкиот Geekjob через телекомуникационные средства личной связи.

Вакансии от"Atlantix"

Еще интересные вакансии

  • Москва

    RU

    AI Product Manager в RUFP

    Rufp.ru

    office

    10 августа

  • AR

    Lead Product Analyst / Analytics Engineer - AI & Experimentation

    Artworkout.app

    remote

    7 августа

  • LE

    E-mail/SMM Marketing Manager

    Lev Haolam

    remote

    6 августа

  • LE

    CRM Marketing Manager (Email Social)

    Lev Haolam

    remote

    6 августа

  • London

    500K - 1000K ₽

    CTO / Head of Engineering

    Kanzum

    relocateremote

    5 августа

  • Junior/Middle Full-Stack Developer (AI-First)

    Insinet

    remote

    3 августа

  • Junior+/Middle QA Engineer

    Starplay Games

    remote

    3 августа

  • 500 $

    Junior+ QA Engineer

    ITDEAL GROUP

    remote

    3 августа

  • Новосибирск

    TH

    Senior CV Engineer / Team Lead

    The One (агентство)

    office

    31 июля

  • Москва

    150K ₽

    PL

    DevOps Engineer

    PlaysDev

    remote

    31 июля

Еще...

Эбаут

Ви ар зе дрим тим оф крэзи спешиалистс энд рекрутерс.Ви лов программерс энд ви олвэйс хелп ту файнд зе дрим ворк фор гуд пипл.

Доказательный рекрутинг

NEWHR Medium Blog

Исследования рынка труда

Дайджесты вакансий

Статьи о карьере в IT

NewHR Podcast

Блог нашего FullStack CTO

Контакты

  • Обратная связь

  • Условия использования

  • Конфиденциальность

  • Услуги и цены

Партнеры

Сейчас на сайте 28482 кандидатов и более 10000 компаний.В среднем 16 откликов на вакансию

Ⓒ Geekjob

Мы используем куки, потому что без кук наш сайт не работал бы, другие сайты не работали бы, да и вообще весьинтернет не работал бы

window.Vacancy = {"id":"6a7af7879319f6a4ef049fba","ci":"661cd76bb72af3b82a0a80db","ic":"661cd5b901ba39cb8504f155","lang":"en","position":"AI Engineer LLM Systems

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$113k – $188k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • Tysons
Python
SQL
Python
pySpark
Databases
Amazon Redshift
Apache Iceberg
Databricks
Delta Lake
AI/ML
Airflow
Anomaly Detection
ChatGPT
Copilot
Cursor
GraphRAG
Knowledge Graph
RAG
Spark
DevOps
Amazon Kinesis
Amazon S3
AWS
AWS CDK
AWS Lambda
AWS Step Functions
CI/CD
CloudFormation
Docker
Git
GitHub
GitHub Actions
IAM
Jenkins
Terraform
Vector
Cybersecurity
Least Privilege
Analytics
ETL/ELT
Power BI
Tableau
Apply
Founding Engineer 6 hours ago
$120k – $200k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • San Francisco
TypeScript
AI/ML
AI Agents
Claude
LLM
DevOps
GitHub
Management
Linear
Slack
Apply
$87k – $131k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Tysons
Java
Mobile
JUnit
DevOps
AWS
Azure
Azure DevOps
CI/CD
Datadog
Docker
Dynatrace
GCP
Git
Jenkins
Kubernetes
New Relic
Shift-Left
Splunk
Cybersecurity
Shift-Left Security
SonarQube
Management
Jira
QA
Cucumber
JMeter
Postman
Rest-Assured
Selenium
TestNG
Apply
$220k – $250k per year • Equity 0.5–1% • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
AI/ML
AI Agents
Apply
Data Engineer 3 hours ago
$22k – $55k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad
Databases
Databricks
AI/ML
AI Agents
Analytics
ETL/ELT
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.