374,200open jobs
9,699companies
47,772added this week
Browse all
Location
Remote (Pakistan, Egypt)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Soum is a Saudi Arabian recommerce marketplace that enables consumers to buy and sell pre-owned electronics, automobiles, and high-value goods. Founded in 2021, the digital platform manages the entire transaction process by providing quality inspections, secure payment handling, integrated logistics, and buyer warranties. By introducing transparency and trust to secondhand commerce, the company streamlines peer-to-peer trading across Saudi Arabia and the broader Middle East.

Role Name: AI / GenAI Solutions Engineer

Location: Egypt, Pakistan (Remote)

Work Week: Sunday - Thursday

Working Hours: 9:00 AM - 6:00 PM (Saudi Arabia Standard Time)

Overview:

Soum is building an AI native layer across our C2C marketplace, from customer conversations to automated order handling, personalised discovery, recommendations, fraud signals, and seller tooling. We're looking for a Senior GenAI Engineer to design and ship production-grade AI systems that touch millions of users across the buying and selling journey.

You'll work end-to-end: architecting LLM-powered services on our Python backend, integrating them into product surfaces on our React frontend, and partnering with teams across the company to find where AI genuinely moves the needle. This is a senior builder's role with high autonomy, broad scope, and real impact on the business.

What You’ll Do

  • Architect production GenAI systems across multiple domains, including conversational agents, automated order and dispute workflows, personalized discovery and recommendations, content generation, search relevance, and emerging use cases.

  • Own features end-to-end, from problem framing and model selection through backend services (FastAPI, Python), frontend integration (React, TypeScript), evaluation, deployment, and monitoring.

  • Design agentic workflows with tool calling, multi-step reasoning, retrieval augmented generation, and integrations with internal APIs, third-party SaaS, and event-driven systems.

  • Build the retrieval and embeddings stack, including chunking strategies, embedding model selection, vector indexes, hybrid search, reranking, and retrieval evaluation pipelines.

  • Make it reliable and cost-efficient through streaming, prompt caching, latency budgets, token cost optimization, observability for LLM calls, and graceful fallback when models or upstreams misbehave.

  • Establish evaluation rigor with offline and online evals covering response quality, tool call correctness, hallucination rate, retrieval precision, and business KPIs.

  • Drive experimentation and research by evaluating new models, frameworks, and agent patterns, running focused experiments, and bringing what works into production.

  • Mentor and raise the bar for engineers across the team on AI and ML best practices, prompt engineering, and production readiness.

  • Partner cross-functionally with Product, Engineering, Data, Ops, and CX to identify high-leverage AI opportunities and ship them.

Where You'll Have Impact

    A non-exhaustive list of areas we are actively building in or want to:

    • Conversational AI for customer support, dispute resolution, and seller assistance

    • Automated order handling, escalation routing, and workflow orchestration

    • Personalized discovery, recommendations, and search relevance

    • Listing quality, including auto-generated titles, descriptions, categorization, and image understanding

    • Trust and safety, including fraud signals, anomaly detection, and content moderation

    • Internal agent tooling for ops and CX teams

Qualifications

    Required

    • 5+ years building production software, with at least 2 years shipping LLM and ML-powered features at scale.
    • Strong Python for backend services, scripting, data pipelines, and ML tooling.
    • Working ability in TypeScript and React for integrating AI features directly into product surfaces.
    • Production experience with major LLM providers (Gemini, Claude, GPT) covering tool and function calling, structured outputs, streaming, prompt caching, and cost control.
    • Deep understanding of Retrieval Augmented Generation (RAG): document ingestion, chunking, embedding generation, vector databases (pgvector, Pinecone, Weaviate, or similar), hybrid retrieval, and reranking.
    • Solid grounding in embeddings and vector search: dense vs. sparse representations, similarity metrics, indexing strategies (HNSW, IVF), and dimensionality tradeoffs.
    • Strong ML and NLP fundamentals: transformer architectures, tokenization, fine-tuning vs. prompting tradeoffs, classification, ranking, and evaluation methodology.
    • Experience with scripting and automation for data preparation, model evaluation harnesses, and offline analysis.
    • Comfortable with relational databases (PostgreSQL), caching layers (Redis), REST, SSE, WebSockets, and event-driven architectures.
    • Production experience with observability, A/B testing, and rolling out model changes safely.
    • Nice to Have

      • Experience with recommendation systems, learning to rank, or search relevance at scale.
      • Marketplace or C2C background covering buyer and seller dynamics, disputes, fraud, and payouts.
      • Multimodal model experience, including vision and image understanding for listings.
      • Arabic NLP or bilingual product experience.
      • Experience with agent frameworks (LangGraph, custom orchestrators) and the judgment to know when to use them.
      • Fine-tuning, LoRA, distillation, or hosting open weight models in production.
      • Open source contributions to LLM tooling, eval frameworks, or retrieval libraries.

What We Care About

    • Ship over the architect. Lean code, no premature abstractions, no half-finished frameworks.

    • Measurement-driven development. Features ship with evals and metrics, not vibes.

    • Ownership. From idea to deployment to monitoring the first real users.

    • Curiosity and range. This role spans many problem domains, and we want someone energized by that.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
374,200 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$25k – $62k per year (Estimated) • In office • Contractor • 10+ years exp • Bengaluru
AI/ML
Claude
Claude Code
Copilot
Cursor
AI Agents
Human-in-the-Loop
DevOps
AIOps
SLI/SLO/SLA
GitHub
Apply
$100k – $206k per year (Estimated) • Remote • 15+ years exp
AI/ML
Claude
DevOps
AWS
Management
Jira
Miro
Marketing
Amplitude
Apply
In office • Full-Time • 5+ years exp • Malaysia
AI/ML
Anomaly Detection
Hallucination
NLP
RAG
Speech Recognition
DevOps
AWS
Azure
GCP
Apply
$217k – $440k per year (Estimated) • In office • Full-Time • 8+ years exp • Mountain View
Python
AI/ML
AI Agents
Fine-tuning
Knowledge Distillation
LLM
Prompt Engineering
RLHF
LLM Guardrails
SFT
Management
ServiceNow
Apply
$95k – $214k per year (Estimated) • Equity • Remote • Contractor
C#
Go
Python
C#
.NET
Python
FastAPI
DevOps
Amazon EKS
AWS
Azure
CI/CD
GCP
GitOps
Helm
Kubernetes
Kustomize
Nginx
Platform Engineering
Pulumi
Terraform
Apply
See all jobs
This is one of many
374,200 more open roles from verified company boards, updated every day.