612,151open jobs
29,363companies
85,745added this week
Browse all
Salary
$33k – $88k per year (Estimated)
Location
Remote/Hybrid (Bucharest, Romania)
Overview
Company
Impact
Profile match
IONOS is a Montabaur company that is one of the largest web hosting and cloud providers in Europe. It hosts millions of domains and websites for small businesses and offers cloud infrastructure with servers located inside the European Union. The company was carved out of United Internet and listed on the Frankfurt exchange in 2023.

At IONOS, the leading European provider of cloud infrastructure, cloud services and hosting services, you will work together with a wide range of teams. We are characterized by open structures, a friendly working culture and flat hierarchies with a strong team spirit. We firmly believe that work and fun are compatible, and offer you the right environment for this. Our constant growth means that we are always looking for new colleagues. Become part of IONOS and grow with us.

About the team:

Our mission is to  build a modern ecosystem used for all IONOS customer support needs. The tools developed by us are used in over 20 locations, by more than 2.000 users, supporting 8 million customer contracts in 10 markets.

The development team has full responsibility for the development lifecycle. This means we plan, develop, test and deploy our software without any other internal or external dependencies.

Our portfolio revolves around an internally built CRM which is now being enhanced with AI capabilities. 

About the product you will be building:

We are building a next-generation AI platform designed to redefine how our company interacts with customers. This isn't just a chatbot; it's a high-performance, multimodal AI ecosystem powered by state-of-the-art Speech-to-Speech (S2S) models, advanced Large Language Models (LLMs), and intelligent orchestration frameworks. Our platform will understand, reason, and respond across text and voice - while seamlessly executing real-time actions to resolve customer needs.

We are aiming for a hybrid architecture of Open Source LLMs, industry-leading proprietary models, and Model Context Protocol (MCP) to enable contextual reasoning, tool invocation, and seamless orchestration across systems. The goal is not just to talk to the customer, but to act on their needs.

What makes this project unique:

The Voice Frontier: We are building low-latency, emotive speech-to-speech pipelines for a truly natural voice channel experience.

Deep System Integration: Our platform connects directly to the company's core systems via MCPs, allowing the AI to access real-time customer context and execute complex workflows.

Self-Evolving Logic: We are developing  an automated QA and evaluation module that continuously analyzes interactions across channels.By programmatically measuring quality, accuracy, latency, and resolution outcomes, we can close the feedback loop, and adapt system behavior in hours, not weeks.

Hybrid Innovation: You’ll work at the intersection of "build vs. buy," integrating the best of the open-source community with custom-built internal infrastructure.

What's in it for you:

You won't just be shipping code; you’ll be part of making this concept evolve and shift.

You’ll join a friendly, experienced team where your voice matters and your contribution shapes real-world outcomes. You’ll work in a modern environment with technologies and practices that help us ship reliable software efficiently.

Role description:

As an AI Engineer on this team, you will build the core intelligence systems behind our multimodal AI platform.You will be responsible for moving beyond simple chat interfaces to build high-performance, real-time systems that handle complex reasoning, deep context retrieval, LLM orchestration, retrieval-augmented generation (RAG) and seamless voice interactions.

Main responsibilities:

  • Design Agentic Workflows: Design and implement LLM-based systems that go behind response generation - enabling structured tool usage, workflow orchestration, and secure interaction with internal services via MCP (Model Context Protocol).
  • Build and Optimize RAG & CAG: Develop high-performance Retrieval-Augmented Generation and Context-Augmented Generation pipelines to ensure accurate, relevant, and low-latency responses. Continuously improve context management, ranking strategies, and grounding mechanisms to support complex, multi-step interactions.
  • Voice Channel Mastery: Develop and optimize real-time Speech-to-Speech (S2S) pipelines, focusing on streaming architectures, latency reduction  (including Time to First Word - TTFW) and maintaining a natural conversational flow.
  • Evaluation, Quality & Alignment: Build and maintain an automated QA module, including LLM-as-a-judge patterns, to measure accuracy, safety, latency, and resolution quality at scale. Translate evaluation insights into systematic models and prompt improvements..
  • Model Strategy & Hybrid Integration: Integrate and operate both commercial foundation  models (e.g., OpenAI, Anthropic, Google) and open-source alternatives (e.g., Qwen, Kimi, DeepSeek, Moonshot, GLM), selecting and optimizing models based on performance, latency, cost, and use-case requirements.

We are looking for some of:

  • Strong Python and/or Java Engineering Skills: Advanced-level Python development experience, including  asynchronous programming (e.g., FastAPI, asyncio) and building high-performance, production-grade services. Experience with streaming architectures is a strong advantage.
  • LLM Application & Multi-Agent Orchestration Experience: Hands-on experience building LLM-powered systems, including multi-step workflows, stateful agents, and tool invocation. Familiarity with orchestration frameworks such as LangChain, LlamaIndex, or LangGraph, particularly in building stateful, multi-turn agents.
  • Advanced Retrieval & Context Management: Deep understanding of vector databases (e.g., Weaviate, Qdrant, pgvector, Elasticsearch), semantic search, embedding strategies, and re-ranking techniques. Experience designing and optimizing RAG pipelines.
  • Real-Time & Low-Latency Systems: Experience in designing systems that operate under latency constraints, including streaming APIs, event-driven architectures, and performance optimization. Understanding of trade-offs between quality, cost, and response time.
  • Evaluation-Driven Development: Experience in implementing evaluation frameworks for LLM-based systems, including automated QA pipelines and LLM-as-a-judge patterns.
  • Familiar with API Design: knowledge of RESTful API design, OAuth2

What we offer:

  • Access to local/international trainings, development and growth opportunities, including access to e-learning platforms, covering both technical and soft skills areas;
  • Modern technologies, product responsibility;
  • Flexible work schedule;
  • Hybrid work option;
  • Medical services package from one of two private providers;
  • 25 vacation days per year;
  • Substitute days off for public holidays that occur on the weekend;
  • Meal tickets;
  • Internal referral program;
  • Team events, networking events organized to promote a passionate, creative and diverse culture;
  • Summerfest and Winterfest parties;
  • Of course, coffee, soft drinks and fresh fruits are on us in the office.

About IONOS

IONOS is the leading European digitalization partner for small and medium-sized businesses (SMB). The company serves around six million customers and operates across 18 markets in Europe and North America, with its services being accessible worldwide. With its Web Presence & Productivity portfolio, IONOS acts as a 'one-stop shop' for all digitalization needs: from domains and web hosting to classic website builders and do-it-yourself solutions, from e-commerce to online marketing tools. In addition, the company offers Cloud Solutions to enterprises who are looking to move to the cloud as their businesses evolve. 

We value diversity and welcome all applications - regardless of, for example, gender, nationality, ethnic or social origin, religion, disability, age as well as sexual orientation and identity, physical characteristics, marital status or any other irrelevant factor subject to applicable law.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
612,151 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bucharest
$82k per year (net) • In office • 4+ years exp • Master's Degree
Python
JavaScript
Java
Rust
TypeScript
SQL
Python
FastAPI
Django
Java
Maven
Gradle
Databases
Databricks
Apache Iceberg
Apache Kafka
Trino
AI/ML
Copilot
Hadoop
Spark
Airflow
Flink
Frontend
Angular
npm
DevOps
CI/CD
Git
Kubernetes
Cybersecurity
Zero Trust
Analytics
ETL/ELT
Management
Agile
Scrum
Kanban
Apply
$63k – $160k per year (Estimated) • In office • Internship • Master's Degree
Python
C++
Apply
Developer Internship 4 months ago
$39k – $64k per year (Estimated) • In office • Internship • Bachelor's Degree
Java
SQL
C++
Bash
Perl
Apply
Design Engineer II 8 hours ago
$23k – $49k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Nanjing
Python
C++
SystemVerilog
Chips/EDA
Cadence Palladium
Apply
Lead Design Engineer 8 hours ago
$30k – $72k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Nanjing
Python
C++
SystemVerilog
Chips/EDA
Cadence Palladium
Apply
$31k – $51k per year (Estimated) • Remote/Hybrid • Karlsruhe
Management
Confluence
Jira
Google Workspace
Apply
$32k – $53k per year (Estimated) • Remote/Hybrid • Master's Degree • Karlsruhe
Python
JavaScript
AI/ML
Prompt Engineering
AI Agents
LLM
Edge AI
DevOps
Rest API
Management
Confluence
Jira
Google Workspace
Apply
$30k – $81k per year (Estimated) • Remote/Hybrid • Bucharest
Python
C++
Bash
C++
Protobuf
STL
DevOps
Rest API
OpenShift
Debian
Jenkins
Docker
Kubernetes
Cryptography
OpenSSL
Apply
$31k – $51k per year (Estimated) • Remote/Hybrid • PhD • Berlin
Management
Google Workspace
Apply
$31k – $51k per year (Estimated) • Remote/Hybrid • Berlin
Apply
$23k – $65k per year (Estimated) • In office • 1+ year exp • Associate's Degree • Bucharest
Apply
$11k – $25k per year (Estimated) • In office • 1+ year exp • Associate's Degree • Bucharest
Apply
$49k – $102k per year (Estimated) • Remote/Hybrid • Full-Time • Bucharest
Apply
Remote/Hybrid • Bucharest
JavaScript
Node JS
DevOps
CI/CD
Management
Agile
QA
TestRail
Mocha
Apply
Remote/Hybrid • Internship • Master's Degree • Bucharest
Analytics
Microsoft Excel
Apply
See all jobs
This is one of many
612,151 more open roles from verified company boards, updated every day.