698,009open jobs
40,743companies
105,967added this week
Browse all
Salary
$136k – $274k per year (Estimated)
Location
Remote (United States)
Employment
Contractor
Overview
Company
Impact
Profile match
Terac is an AI-powered market research and human data platform headquartered in the United States. Founded by Zachary Baker and Jack Blair, the enterprise operates an infrastructure layer connecting AI agents with verified human domain experts and research participants. Operating on a B2B SaaS and usage-based API business model, Terac automates qualitative research by using conversational voice and text AI agents to recruit participants, moderate interviews, collect feedback, and synthesize structured data insights in real time for product, marketing, and frontier AI research teams.

What We're Researching

We're running a paid study on the quality and realism of coding environments designed to test AI agents. We are developing a comprehensive suite of programming tasks and evaluation harnesses to measure AI performance accurately. Your feedback directly shapes how these agents are benchmarked against real-world engineering standards.

How It Works

During this remote session, you will review several programming tasks and their corresponding evaluation harnesses. You will assess the technical accuracy, complexity, and realism of each coding challenge. We will ask you to walk through the logic of the evaluation environments and identify any potential flaws. Finally, you will provide feedback on how these environments compare to standard industry practices.

Who This Is For

We are looking for software engineers who have hands-on experience building, reviewing, or testing realistic programming tasks. We welcome full-stack developers, backend engineers, test automation engineers, and systems architects who are familiar with evaluation harnesses. Ideal candidates understand what makes a coding challenge robust, verifiable, and technically sound.

What You'll Do

  • Review and assess the quality of realistic programming tasks

  • Evaluate coding environments and technical harnesses for AI testing

  • Walk us through potential flaws in the provided code structures

  • Provide feedback on the difficulty and realism of the challenges

Who Should Apply

  • Professional experience as a software engineer or developer

  • Familiarity with building or verifying programming tasks

  • Experience with code review and evaluation harnesses

  • Comfortable discussing technical architecture and testing methodologies

Compensation

$75 per hour

Ready to participate?

Start your paid interview now

About Terac

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Learn more at terac.com or on YouTube at @jointerac.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
698,009 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$34k – $81k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Shanghai
Python
TypeScript
AI/ML
Model Context Protocol
Prompt Engineering
Function Calling
AI Agents
RAG
LLM Guardrails
DevOps
Azure
CI/CD
Platform Engineering
Management
Agile
Apply
$292k – $383k per year • In office • Full-Time • 10+ years exp • Menlo Park
Databases
Snowflake
AI/ML
Claude Code
dbt
Prompt Engineering
AI Agents
Streamlit
Agentic Workflows
DevOps
CI/CD
Analytics
Tableau
Master Data Management
Management
Confluence
Jira
Agile
Apply
$145k – $218k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • Mississauga
Python
Java
Kotlin
SQL
Databases
Weaviate
Pinecone
Apache Kafka
AI/ML
Copilot
AutoGen
LangChain
Fine-tuning
Prompt Engineering
AI Agents
Flink
LLM
RAG
Anomaly Detection
Feature Store
DevOps
CI/CD
AWS
Docker
Kubernetes
Trunk-Based Development
Chaos Engineering
Progressive Delivery
AIOps
SLI/SLO/SLA
Cybersecurity
Threat Modeling
Management
Agile
Apply
$67k – $145k per year (Estimated) • Remote/Hybrid • 8+ years exp • Bachelor's Degree • Mexico City
AI/ML
AI Agents
Agentic Workflows
Apply
$92k – $126k per year • In office • Full-Time • 2+ years exp • Berlin
Go
Dart
AI/ML
Claude Code
Model Context Protocol
AI Agents
OpenAI Codex
Mobile
Flutter
State Management
DevOps
CI/CD
Apply
$40k – $99k per year (Estimated) • Remote • Contractor
Apply
$69k – $177k per year (Estimated) • Remote • Contractor
Management
QuickBooks
Apply
$87k – $169k per year (Estimated) • Remote • Contractor
Apply
Remote • Contractor
Marketing
YouTube
Apply
See all jobs
This is one of many
698,009 more open roles from verified company boards, updated every day.