683,239open jobs
39,562companies
98,223added this week
Browse all
Salary
$67k – $154k per year (Estimated)
Location
Remote (France)
Seniority
Architect · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior ML Solutions Architect - Token Factory based in France.

Join a fast-growing AI infrastructure team building a serverless platform for running and customizing open-source LLMs in production.

Help customers move from AI prototypes to scalable, reliable production applications without building their own complex inference stacks.

Design optimized inference workflows and customized LLM solutions across multiple models and modalities.

Work hands-on with inference, fine-tuning, evaluation, prompt engineering, and retrieval-augmented generation.

Partner directly with customers to understand technical challenges and translate them into effective AI architectures.

Collaborate closely with product and engineering teams to turn customer feedback into platform improvements.

Work remotely across Europe in an international environment focused on high-impact AI projects, technical ownership, and continuous innovation.

Accountabilities

    • Optimize LLM inference workflows across different modalities to deliver measurable business value and meet customer requirements.
    • Support customers with supervised and reinforcement-learning-based fine-tuning approaches to improve model quality and performance.
    • Design and implement LLM-powered solutions using serverless inference services and served open-source models.
    • Build production-ready applications using LLM APIs, including multimodal models covering text, vision, audio, and domain-specific use cases.
    • Provide technical guidance on prompt engineering, RAG architectures, model selection, inference optimization, and deployment strategies.
    • Guide customers through the transition from proof of concept to production, with a focus on performance, reliability, scalability, and cost efficiency.
    • Work closely with product and engineering teams to communicate customer needs, identify platform gaps, and contribute to roadmap development.
    • Help customers select appropriate models, inference configurations, and fine-tuning strategies based on their use cases and technical constraints.
    • Contribute to improving the platform and its capabilities by sharing practical insights from customer implementations and production workloads.
    • Requirements:

      • 5+ years of professional experience working with ML/AI systems, including at least 2 years focused specifically on LLMs and generative AI.
      • Deep understanding of the modern LLM ecosystem, including model architectures, inference approaches, and fine-tuning techniques.
      • Hands-on experience running LLMs in production, including deploying and operating inference workloads at scale.
      • Strong practical experience with LLM fine-tuning, including supervised fine-tuning, SFT, LoRA, and data preparation or curation; experience with reinforcement-learning-based fine-tuning is a strong advantage.
      • Experience building LLM evaluation frameworks, including task-specific benchmarks, offline and online evaluation pipelines, and LLM-as-a-judge approaches.
      • Practical experience with modern inference frameworks and ML libraries such as vLLM, SGLang, TensorRT-LLM, or Transformers.
      • Experience deploying LLM-powered applications through APIs from providers such as OpenAI or Anthropic, as well as open-source models.
      • Strong Python programming skills and the ability to develop practical, production-oriented AI solutions.
      • Excellent communication skills, with the ability to explain complex technical concepts clearly to customers, engineers, product teams, and other audiences.
      • Experience working with multimodal AI models, such as vision-language or speech models, is a plus.
      • Familiarity with DevOps technologies including Docker, Kubernetes, and Git is beneficial.
      • Contributions to open-source ML or AI projects are an additional advantage.
      • Familiarity with cloud AI platforms such as AWS SageMaker or Bedrock, Google Vertex AI, or Azure ML is welcome.
      • Benefits:

        • Competitive compensation.
        • Career growth and ongoing learning opportunities.
        • Flexible working arrangements with a high degree of autonomy and ownership.
        • Fully remote work opportunity from Europe.
        • Collaborative and innovative international working environment.
        • Opportunity to work on impactful AI infrastructure and production-grade LLM projects.
        • Exposure to advanced technologies across LLM inference, fine-tuning, evaluation, retrieval, and multimodal AI.
        • Opportunity to work closely with talented AI, engineering, product, and customer-facing teams.
        • Meaningful technical ownership and the opportunity to influence the evolution of an emerging AI platform.
        • Inclusive workplace committed to equal employment opportunities and a diverse working environment.
        • Reasonable accommodations are available during the application process where required.
        • Candidates must be authorized to work in the country where they apply and may need to provide proof of employment eligibility.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
683,239 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$47k – $114k per year (Estimated) • In office • Full-Time • Bengaluru
Python
SQL
Databases
Weaviate
Milvus
Pinecone
FAISS
AI/ML
Weights & Biases
LangChain
Claude
Spark
LlamaIndex
LoRA
MLFlow
Vertex AI
Fine-tuning
Embeddings
Scikit-learn
Multimodal AI
AI Agents
NLP
Gensim
LangSmith
PEFT
Mistral
Transformers
TensorFlow
Pandas
NumPy
Keras
PyTorch
LLM
RAG
NLTK
Hugging Face
Amazon SageMaker
Agentic Workflows
DevOps
Azure
CI/CD
AWS
Docker
Kubernetes
Vector
Apply
Analytics Engineer 3 hours ago
Remote/Hybrid • 3+ years exp
Python
SQL
Databases
Firestore
Google BigQuery
BigQuery
AI/ML
Claude Code
Streamlit
Mobile
Firebase
DevOps
GCP
Management
Trello
Agile
Apply
$149k – $201k per year • Remote • Secret • Full-Time • 10+ years exp
Python
Cybersecurity
Zero Trust
Management
Agile
Apply
Cloud Data Engineer 3 hours ago
$140k – $190k per year • Remote • Full-Time • 8+ years exp • Bachelor's Degree
Python
JavaScript
SQL
Bash
Databases
Snowflake
DevOps
CI/CD
Git
AWS
Bitbucket
AWS Lambda
GitHub
IAM
Analytics
ETL/ELT
Management
Confluence
Jira
Agile
Scrum
Apply
$143k – $251k per year (Estimated) • Remote • Full-Time • New York
Python
Databases
PostgreSQL
DevOps
OpenTelemetry
VMWare
Debian
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
TeamCity
Cybersecurity
Zero Trust
Apply
Billing Manager 3 hours ago
$115k – $135k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree
SQL
Analytics
Microsoft Excel
Apply
HR Manager 3 hours ago
$83k – $169k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree
Analytics
Microsoft Excel
Apply
$91k – $178k per year (Estimated) • Remote • Full-Time • 4+ years exp • Bachelor's Degree
Cybersecurity
Zscaler
MITRE ATT&CK
HIPAA
Apply
$50k per year • Remote • Full-Time • 2+ years exp
Apply
$85k – $161k per year (Estimated) • Remote • Full-Time • 5+ years exp
Analytics
Microsoft Excel
Apply
See all jobs
This is one of many
683,239 more open roles from verified company boards, updated every day.