663,523open jobs
38,718companies
99,026added this week
Browse all
Salary
$18k per year
Location
Remote (India)
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI/ML Evaluator - English based in India.

As an AI/ML Evaluator, you will use analytical reasoning and technical expertise to evaluate complex AI system behavior.

You will analyze system outputs, telemetry, event data, and other technical signals to identify meaningful behavioral patterns.

A key focus will be distinguishing legitimate user activity from sophisticated automated or bot activity using available evidence and project guidelines.

You will work with complex and sometimes ambiguous information, applying AI/ML knowledge to make accurate and consistent assessments.

Your evaluations will generate high-quality, human-verified data that supports the training and improvement of AI systems.

This freelance opportunity offers the potential to transition into a full-time role while providing hands-on experience in AI evaluation and technical data analysis.

Accountabilities:

    • Review and evaluate AI system behavior, outputs, and relevant technical data according to defined project guidelines.

    • Analyze system telemetry, event data, signals, and other technical information to identify meaningful patterns and behavioral indicators.

    • Evaluate signals that may help distinguish legitimate user activity from automated or bot activity.

    • Apply AI and machine learning knowledge, analytical reasoning, and expert human judgment to complex or ambiguous cases.

    • Identify unusual patterns, inconsistencies, anomalies, or behaviors that may require additional review.

    • Evaluate complex cases using available evidence, contextual information, and defined project requirements.

    • Classify, label, and annotate assigned data accurately and consistently.

    • Provide high-quality human evaluations to support AI system training, testing, evaluation, and continuous improvement.

    • Clearly document decisions and supporting reasoning when required.

    • Identify unclear, unusual, or particularly complex cases and flag them through established project processes.

    • Apply detailed project guidelines consistently across all assigned tasks.

    • Maintain high standards of accuracy, quality, consistency, and attention to detail.

    • Requirements:

      • Educational background or equivalent professional experience in Artificial Intelligence, Machine Learning, Computer Science, Data Science, Statistics, Engineering, or a related field.

      • Experience working with AI/ML systems, machine learning models, AI evaluation, data analysis, or technical system analysis.

      • Strong understanding of core artificial intelligence and machine learning concepts.

      • Ability to analyze complex technical data, system behavior, and behavioral patterns.

      • Experience evaluating AI systems, models, outputs, or datasets is preferred.

      • Familiarity with system telemetry, event data, logs, or other technical signals is an advantage.

      • Strong analytical, critical-thinking, and problem-solving skills.

      • Ability to identify meaningful patterns and make informed decisions using complex, incomplete, or ambiguous information.

      • Strong attention to detail and ability to consistently apply detailed evaluation guidelines.

      • Experience with data annotation, AI evaluation, human feedback, model evaluation, or structured data review is preferred.

      • Familiarity with bot detection, automated abuse detection, anomaly detection, fraud detection, or related areas is advantageous.

      • Strong written English comprehension and documentation skills.

      • Ability to work independently while maintaining consistent quality and accuracy across assigned evaluations.

      • Benefits:

        • Compensation: US$9 per hour.

        • Contract type: Freelance, with potential opportunity to transition to a full-time role.

        • Location: India.

        • Language: English.

        • Flexibility: Freelance structure offering flexibility in managing assigned work.

        • Professional exposure: Hands-on experience evaluating AI systems, technical signals, and machine learning behavior.

        • Impact: Contribute to improving the accuracy and reliability of AI systems used to identify automated abuse while helping reduce incorrect classification of legitimate activity.

        • Skill development: Opportunity to deepen expertise in AI/ML evaluation, technical data analysis, anomaly detection, and human-in-the-loop AI systems.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
663,523 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
Sr. AI Engineer 11 hours ago
$146k – $234k per year • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Cambridge
Python
JavaScript
TypeScript
SQL
Node JS
Scala
Databases
Databricks
Delta Lake
AI/ML
MLFlow
Embeddings
Function Calling
LLM
RAG
Anomaly Detection
LLMOps
Feature Store
Human-in-the-Loop
LLM Guardrails
Agentic Workflows
Tool Use
Frontend
Vue.js
Svelte
GraphQL
Angular
React.js
DevOps
Terraform
GCP
Azure
CI/CD
AWS
Vector
IAM
Apply
$93k – $204k per year (Estimated) • In office • 5+ years exp • Singapore
AI/ML
AI Agents
Human-in-the-Loop
DevOps
SLI/SLO/SLA
Apply
$180k – $250k per year • Equity • Remote • Full-Time
Python
AI/ML
Function Calling
AI Agents
LLM
RAG
Human-in-the-Loop
LLM Evaluation
Edge AI
Agentic Workflows
Multi-Agent Systems
Tool Use
Apply
$63k – $182k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Auckland
Python
C++
AI/ML
Human-in-the-Loop
DevOps
CI/CD
Git
Design
SolidWorks
Management
Jira
Apply
$54k – $98k per year (Estimated) • In office • Full-Time • 1+ year exp • Bengaluru
AI/ML
Human-in-the-Loop
Recommender Systems
Agentic Workflows
Design
Figma
Management
Agile
Apply
Graphic Designer 4 hours ago
Remote • Full-Time • 1+ year exp
PHP
PHP
WordPress
Design
Adobe Photoshop
Adobe Premiere Pro
Adobe Illustrator
Adobe InDesign
Apply
$20k per year • Remote • Full-Time
Apply
Full Stack Engineer 4 hours ago
$25k – $57k per year (Estimated) • Remote • Full-Time • 3+ years exp
Python
JavaScript
TypeScript
SQL
Node JS
Python
Django
Frontend
Vue.js
Angular
React.js
DevOps
Git
Apply
$26k per year • Remote
AI/ML
Red Teaming
Apply
$30k – $73k per year (Estimated) • Remote • Full-Time
Apply
See all jobs
This is one of many
663,523 more open roles from verified company boards, updated every day.