{"id":1160309,"url":"https://alion.io/job/lucid-labs-ai-engineer-mfd","title":"AI Engineer (m/f/d)","company":{"id":679888,"name":"Lucid Labs","domain":"lucidlabs.de","url":"https://alion.io/company/lucidlabs","size_band":"51-200","is_staffing_agency":false,"is_intermediary":false,"listed_via":null,"ats_vendor":"Join","truth_index":null},"role":"AI/ML","role_family":"AI/ML","seniority":"middle","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Berlin, Germany"],"countries":["DE"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":67000,"max_usd":180000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":456},"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AI Agents","optional":false},{"name":"Anthropic","optional":false},{"name":"Arize Phoenix","optional":false},{"name":"Claude","optional":false},{"name":"Claude Code","optional":false},{"name":"Cursor","optional":false},{"name":"Docker","optional":false},{"name":"Fine-tuning","optional":false},{"name":"Function Calling","optional":false},{"name":"GDPR","optional":false},{"name":"Gemini","optional":false},{"name":"Human-in-the-Loop","optional":false},{"name":"Kubernetes","optional":false},{"name":"Langfuse","optional":false},{"name":"LLM","optional":false},{"name":"LLM Guardrails","optional":false},{"name":"Machine Learning","optional":false},{"name":"Mistral","optional":false},{"name":"n8n","optional":false},{"name":"Next.js","optional":false},{"name":"OpenAI","optional":false},{"name":"Promptfoo","optional":false},{"name":"React.js","optional":false},{"name":"Slack","optional":false},{"name":"Structured Outputs","optional":false},{"name":"Tool Use","optional":false},{"name":"TypeScript","optional":false},{"name":"Vercel","optional":false},{"name":"Vercel AI SDK","optional":false},{"name":"Agentic Workflows","optional":true},{"name":"GitHub Actions","optional":true},{"name":"JavaScript","optional":true},{"name":"Node JS","optional":true},{"name":"pnpm","optional":true},{"name":"PostgreSQL","optional":true},{"name":"Radix UI","optional":true},{"name":"shadcn/ui","optional":true},{"name":"Tailwind CSS","optional":true}],"status":"live","first_seen_at":"2026-08-24T11:29:59Z","employer_posted_date":"2026-08-24","last_verified_at":"2026-09-24T10:42:56Z","board_verified":true,"closed_at":null,"days_open":31,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":31},"description":"What this role is about\nAI has arrived in the German Mittelstand. The distance between a convincing demo and a system a department trusts every single day is still large - and it gets closed by work that almost nobody does properly: defining what \"good\" concretely means, building the system toward it, measuring whether it hits, and iterating until it holds.\nThat is what you do:\nYou don't start from zero. You usually get a real starting point, often a golden data set from the department of a German Mittelstand company: real cases, each with the result an experienced expert would deliver. Your job is to build a system that reaches that result reliably, to bring it into people's working day through an interface they can actually use, and to be able to prove that it works.\nTasks\nYou typically work on two to three customer projects in parallel.\n80% Building: The core of the role\nDerive a defensible definition of \"correct\" from reference data, and build an eval that measures it\nDesign and build the AI system itself: model and provider choice, context and prompt design, structured outputs, tool calling, agent/workflow logic, retrieval where it earns its place - and deterministic logic where an LLM isn't needed.\nDo error analysis and iterate deliberately instead of guessing - and demonstrate the improvement\nBuild guardrails, escalation paths and monitoring, so the system doesn't fail loudly and confidently\nBuild the production interface that goes with it: review screens, approvals, human-in-the-loop - in Next.js and TypeScript, not a throwaway prototype\nRoll the solution out and keep it running + improve it as AI advances\n20% Communication: Mostly internal\nCapture and pass on your state in a structured way: what is built, what was measured, what is still open - in writing, without anyone having to ask\nExplain complex technical matters so that the project lead and colleagues without an AI background can decide on a sound basis\nRaise questions, blockers and wrong assumptions early instead of collecting them until the next meeting\nOccasionally demonstrate or explain a result at the customer yourself\nYour standard is not \"it's built\". Your standard is: it is demonstrably good, it gets used, and it creates real value for the customer.\nWhat this role is not\nThree boundaries and all three exist to protect your build time.\nYou don't own the customer. The relationship, the workshops and the conversations with management sit with the project lead. You see customers occasionally - when a result needs to be demonstrated or explained - not weekly.\nYou don't run the project - and it does get run. We work with structured project management: scope and deadlines are defined up front and actively managed, so you know what you're delivering and by when, and your work stays plannable instead of absorbing whatever shifted this week. Very little ticket boards, no worklogs, no status lists on your side. What we do need: that your state is legible to everyone else at any time, without anyone having to ask.\nYou don't work inside our customers' infrastructure - no SAP customizing, no system administration, no legacy integration as your core work.\nYour first 90 days\nWeeks 1-3: You work inside two running projects, get to know our stack and our eval practice, and take over your first AI ise case of your own.\nWeeks 4-8: You own a complete use case - from the reference data set through the evals to the production interface - and you record results and measurements so that the project lead can represent them without you.\nAfter three months: You build your use cases independently, decide on approach and scope, and are the person on the project team who translates technical topics for everyone else.\nRequirements\nImportant\nYou have built an AI solution that real users used in production - not just a prototype, not just a concept.\nYou know how to engineer production LLM systems: context and prompt design, model selection, structured outputs, tool calling, failure handling, latency/cost trade-offs - and you know when a deterministic component beats an LLM.\nYou work eval-driven: you define up front what a good result is, you measure systematically, and you don't ship on gut feeling. Whether that runs on Langfuse, Promptfoo, Arize Phoenix or your own spreadsheet is up to you; that you do it at all is not.\nYou have built a system in which AI agents plan and execute tasks: tool calling, structured outputs, human-in-the-loop.\nYou build production frontend: React and Next.js (App Router, Server Components), TypeScript. Not just demos, but error states, permissions, and the edges where prototypes fall apart.\nYou can ship and operate your own solution: Docker, deployment, logging, cost control.\nYou can explain a complex technical topic so that someone without an AI background can decide soundly afterwards - in writing just as well as in conversation.\nYou work in a structured way internally: your status, your measurements and your open points are written down and findable, without anyone having to chase you.\nCoding agents (Claude Code, Cursor) as part of your daily work\nEnglish at working level. The working language on this role is English, meaning Slack, meetings, code, PRs, and the write-ups that go with them.\nYou can work in Germany full-time. A student visa with a day limit doesn't cover this role.\n3+ years of professional experience in delivery projects.\nHighly valuable\nGerman. Our customers are German Mittelstand, and the occasional demo there runs in German. Without it, the project lead covers that part - but it makes life easier, for you and for us.\nExperience with a production AI SDK layer (e.g. Vercel AI SDK), streaming architectures, several providers (OpenAI, Anthropic, Gemini, Mistral) \nEvaluation pipelines and feedback loops as a topic in their own right, not a by-product\nUsing retrieval where it earns its place - and recognising when you don't need it\nExperience with customers or business departments in the German Mittelstand\nConfidence on GDPR and EU hosting questions\nWorkflow orchestration with n8n or something comparable\nNot required\nA computer science degree\nModel training, fine-tuning, classical machine learning\nMLOps, Kubernetes, data engineering pipelines\nTen years of experience with a three-year-old framework ;)\nWhat matters to us more than a perfect CV\nWe're not looking for the most elegant architecture, the most impressive stack, or the solution that demos best. We're looking for the simplest solution that demonstrably solves the use case - and for someone who knows the difference and holds it even when the more elaborate option would be more fun.\nIf your first question on a new project is what infrastructure we should set up, we're probably not a fit. If your first question is what we actually want to measure this thing against: then we are.\nYou'll fit here particularly well if …\n… you'd rather ship something demonstrably useful in two weeks than something complete in three months.\n… you can sink your teeth into an error rate until you understand where it comes from.\n… you say when a use case isn't worth it, even when it already sounds sold.\nBenefits\nWhat you can expect\nDelivery that doesn't end at slides. You see projects from discovery through to go-live, and you build the solution yourself.\nOwnership without micromanagement. Clear goals, no step-by-step instructions. Good arguments change decisions, regardless of who makes them.\nVisible impact. Short paths, flat structure, no sign-off loops for the sake of form. Within a few weeks you see whether something works.\nBecoming genuinely AI-native. Agents in production, not in a notebook - and evaluating what actually holds up, on systems with real users. Full AI stack: Claude, OpenAI, Cursor, Claude Code, n8n, and whatever else your project needs.\nRemote, but not anonymous. Our team works mostly out of Berlin, with a team day every Wednesday at the EDGE Coworking by Hauptbahnhof. Otherwise flexible remote, with occasional appointments at customers.\nWho we are\nLucid Labs is an AI-first studio for applied AI in the German Mittelstand. We don't just advise companies on which use cases to pick - we build and operate the solutions that run in their processes. Founder Marek Janetzke previously co-built Flightright and stayed with it through to the exit.\nWho you'll work with\nYou won't be the only AI engineer here. You work alongside colleagues who build the same kind of systems you do - people to think a hard problem through with, to have an approach reviewed before you commit to it, and to borrow a solution from when someone has already solved it once. Technical decisions get argued out, not handed down.\nOur stack\n Frontend: Next.js (App Router), React, TypeScript, Tailwind CSS, shadcn/ui\n AI: for example Vercel AI SDK, provider-flexible (OpenAI, Anthropic, Gemini, Mistral), Claude Skills, agentic workflows, n8n\nData & deployment: PostgreSQL with Drizzle ORM, Docker, Vercel, Elestio on EU servers\nDay-to-day: Cursor, Claude Code, pnpm, Node.js, GitHub Actions\nHow to apply\nPlease send us two things:\nyour CV\none example, in two paragraphs at most: an AI solution you built that real users used.\nBriefly describe:\nwhat you used to decide the result was good enough\nhow you measured it\nwhat didn't work on the first attempt, and why\nwhat you changed as a result\nPlease don't include confidential information from previous employers or customers. The second part matters more to us than the first. You don't need a cover letter.\nWe're looking forward to hearing from you.","description_format":"text","description_chars":9555,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[{"language":"English","level":"All levels","optional":false}]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Artificial Intelligence","Machine Learning","AI Agents"],"lifecycle":[{"event":"open","at":"2026-09-23T23:30:19Z"}],"liveness":{"score":55,"band":"ok","label":"Likely open","p_open":1,"p_active":0.731,"p_room":0.75,"age_days":30,"expected_fill_days":42,"reasons":["conf:6","win:late"],"computed_at":"2026-09-24T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/lucid-labs-ai-engineer-mfd","json_url":"https://alion.io/job/lucid-labs-ai-engineer-mfd.json","meta":{"generated_at":"2026-09-24T16:50:20Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers"}}