412,433open jobs
14,112companies
70,830added this week
Browse all
Salary
$48k – $122k per year (Estimated)
Location
Remote (Portugal)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Sales Engineer - Token Factory based in Portugal.

This is a foundational technical role supporting customers building and scaling performance-sensitive AI inference workloads.

You will bridge customer requirements, commercial objectives, and engineering realities from initial discovery through production validation.

The role combines deep technical architecture with customer engagement, helping ensure that proposed solutions are scalable, efficient, and economically viable.

You will work closely with Sales and Engineering to shape strategic opportunities and make informed technical decisions before significant resources are committed.

You will also identify recurring workload patterns and translate customer needs into actionable insights for product and platform evolution.

The environment is fast-moving, international, engineering-led, and focused on solving complex challenges at the forefront of AI infrastructure.

Success means building customer trust while improving deal quality, PoC-to-production conversion, and the effective use of engineering capacity.

Accountabilities

    • Lead in-depth technical discovery with engineering teams, technical founders, and other customer stakeholders to understand models, traffic expectations, latency requirements, GPU economics, and system dependencies.
    • Translate customer objectives into production-ready architectures while identifying technical risks, scalability constraints, and hidden dependencies early.
    • Partner closely with Sales on strategic opportunities, using architectural clarity to influence deal strategy and prevent technically misaligned commitments.
    • Define measurable PoC success criteria covering latency, time to first token (TTFT), throughput, cost, and other relevant performance indicators.
    • Assess workload complexity and determine the appropriate level of optimization, technical support, engineering involvement, and GPU capacity required.
    • Drive structured Go/No-Go decisions and help ensure PoCs remain appropriately scoped, economically justified, and free from uncontrolled customization or hidden R&D requirements.
    • Identify recurring customer configuration and workload patterns, quantify demand for advanced inference optimizations, and communicate structured insights to Product and Engineering.
    • Contribute to the evolution of platform capabilities by turning real-world workload data and customer requirements into actionable product opportunities.
    • Help ensure strategic deals are technically sound before engineering engagement, engineering resources are allocated predictably, and customers receive scalable solutions.
    • Requirements

      • Deep understanding of AI inference systems and GPU-backed infrastructure, with strong knowledge of performance-sensitive environments.
      • Hands-on experience working with LLM workloads and the architectural trade-offs involved in deploying inference systems at scale.
      • Practical experience with inference frameworks and libraries such as vLLM, SGLang, or TensorRT-LLM.
      • Strong ability to reason about latency, throughput, cost, scalability, resource utilization, and overall architecture trade-offs.
      • Experience working directly with engineering-led organizations, technical founders, developers, or highly technical customer teams.
      • Strong customer-facing and communication skills, with the confidence to challenge assumptions and push back constructively when necessary.
      • Commercial awareness and an understanding that engineering capacity is a strategic resource that must be allocated thoughtfully.
      • Strong Python skills and familiarity with modern AI, API, MLOps, DevOps, and cloud technologies.
      • Experience with technologies such as OpenAI or Anthropic SDKs, LangChain, LangSmith, smolagents, FastAPI, Flask, Kubernetes, Docker, and Git is highly valuable.
      • Familiarity with major cloud AI platforms such as AWS SageMaker and Bedrock, Google Cloud Vertex AI, or Azure Machine Learning is preferred.
      • Ability to operate effectively in a fast-moving, international environment with a high degree of ownership and autonomy.
      • Benefits

        • Competitive compensation.
        • Flexible remote work opportunities within Europe.
        • High level of ownership, autonomy, and flexibility.
        • Career growth and continuous learning opportunities.
        • Opportunity to work on impactful, technically challenging AI infrastructure projects.
        • Collaborative, innovative, and engineering-driven working environment.
        • International team with highly experienced professionals across AI, software, and infrastructure.
        • Opportunity to influence platform evolution and shape solutions for real-world AI workloads.
        • Fast-paced environment with meaningful impact and significant opportunities for professional growth.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
412,433 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$23k – $62k per year (Estimated) • In office • Full-Time • 4+ years exp • Chennai
Python
Bash
Databases
Snowflake
Databricks
DynamoDB
Apache Kafka
Kafka
AI/ML
LangChain
LlamaIndex
Model Context Protocol
MLFlow
dbt
Embeddings
Prompt Engineering
AI Agents
AWS Bedrock
LLM
RAG
Hugging Face
Amazon SageMaker
LLMOps
AWS Strands Agents
LLM Guardrails
DevOps
Rest API
Terraform
Ansible
Helm
GitHub Actions
CloudFormation
Datadog
GitLab CI
CI/CD
GitOps
ArgoCD
Jenkins
Git
AWS
Docker
Kubernetes
Grafana
Platform Engineering
Self-Healing
Bitbucket
Amazon EKS
AWS Lambda
Amazon EC2
Vector
GitHub
GitLab
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
Amazon EventBridge
Cybersecurity
Least Privilege
Apply
Full Stack Engineer 4 hours ago
$23k – $107k per year (Estimated) • In office • Full-Time • Chennai
Python
JavaScript
Java
TypeScript
C++
Node JS
Python
Flask
Django
Databases
RabbitMQ
Frontend
Vue.js
Angular
Bootstrap
React.js
DevOps
GitHub Actions
CI/CD
GitOps
Jenkins
Docker
Kubernetes
GitHub
Apply
$150k – $250k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • New York
Python
SQL
Python
FastAPI
Dask
Databases
PostgreSQL
Google BigQuery
BigQuery
AI/ML
Cursor
Claude
Claude Code
Model Context Protocol
Dagster
AI Agents
OpenAI
Anthropic
Tool Use
DevOps
Terraform
GCP
Docker
Analytics
ETL/ELT
Apply
Backend Engineer 4 hours ago
$150k – $250k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • New York
Python
Python
FastAPI
Dask
Databases
PostgreSQL
Google BigQuery
BigQuery
AI/ML
Cursor
Claude
Claude Code
Model Context Protocol
Dagster
Function Calling
AI Agents
LLM
RAG
OpenAI
Anthropic
Agentic Workflows
Tool Use
DevOps
Terraform
GCP
Docker
Analytics
ETL/ELT
Apply
Fullstack Engineer 4 hours ago
$150k – $250k per year • Equity 0.5–1% • In office • Full-Time • 3+ years exp • New York
Python
JavaScript
TypeScript
Python
FastAPI
Databases
PostgreSQL
Google BigQuery
BigQuery
AI/ML
Cursor
Claude
Claude Code
Model Context Protocol
Function Calling
AI Agents
RAG
OpenAI
Anthropic
Agentic Workflows
Tool Use
Frontend
Next.js
React.js
Mobile
React Native
State Management
DevOps
Terraform
GCP
Docker
Analytics
ETL/ELT
Apply
$111k – $199k per year (Estimated) • Remote • Full-Time • 4+ years exp
Python
JavaScript
Databases
PostgreSQL
AI/ML
Copilot
Cursor
Claude
Claude Code
OpenAI Codex
Frontend
Next.js
React.js
DevOps
GitHub
Apply
$57k – $112k per year (Estimated) • Remote • Full-Time • 4+ years exp
Python
JavaScript
Databases
PostgreSQL
AI/ML
Copilot
Cursor
Claude
Claude Code
OpenAI Codex
Frontend
Next.js
React.js
DevOps
GitHub
Apply
$81k – $166k per year (Estimated) • Remote • Full-Time • 4+ years exp
Python
JavaScript
Databases
PostgreSQL
AI/ML
Copilot
Cursor
Claude
Claude Code
OpenAI Codex
Frontend
Next.js
React.js
DevOps
GitHub
Apply
$72k – $140k per year (Estimated) • Remote • Full-Time • 4+ years exp
Python
JavaScript
Databases
PostgreSQL
AI/ML
Copilot
Cursor
Claude
Claude Code
OpenAI Codex
Frontend
Next.js
React.js
DevOps
GitHub
Apply
$99k – $192k per year (Estimated) • Remote • Full-Time • 4+ years exp
Python
JavaScript
Databases
PostgreSQL
AI/ML
Copilot
Cursor
Claude
Claude Code
OpenAI Codex
Frontend
Next.js
React.js
DevOps
GitHub
Apply
See all jobs
This is one of many
412,433 more open roles from verified company boards, updated every day.