{"id":1292954,"url":"https://alion.io/job/axiom-bio-agent-harness-engineer","title":"Agent Harness Engineer","company":{"id":2041615,"name":"Axiom Bio","domain":"axi.om","url":"https://alion.io/company/axi-om","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Ashby","truth_index":null},"role":"Industrial Engineering","role_family":"Industrial Engineering","seniority":null,"employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["San Francisco, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":88000,"max_usd":223000,"period":"year","method":"role_country_seniority_unknown","sample_n":6763},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AI Agents","optional":false},{"name":"Azure Data Factory","optional":false},{"name":"Docker","optional":false},{"name":"DuckDB","optional":false},{"name":"FastAPI","optional":false},{"name":"Function Calling","optional":false},{"name":"LLM","optional":false},{"name":"LLM Guardrails","optional":false},{"name":"Python","optional":false},{"name":"Svelte","optional":false},{"name":"SvelteKit","optional":false},{"name":"Terraform","optional":false},{"name":"Tool Use","optional":false},{"name":"JavaScript","optional":true}],"status":"live","first_seen_at":"2026-07-14T04:05:24Z","employer_posted_date":"2026-07-14","last_verified_at":"2026-10-01T11:14:50Z","board_verified":true,"closed_at":null,"days_open":79,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":79},"description":"About Axiom:\nAxiom is building the closed-loop scientific AI system required to replace animal testing and, over time, much of human safety testing. We start with pharma’s hardest drug development toxicology problems. Those problems define the proprietary human biological data we generate through Axiom’s Data Factory. We use that data to train scientific AI, partnering with leading AI labs to improve frontier models while building our own specialist agentic harness to deploy the improved frontier models back into pharma. Each deployment reveals the next capabilities to build, creating a compounding loop across data, models, and drug development. Today, liver toxicity is our proving ground. Axiom is already helping leading pharmaceutical companies understand toxicity, identify its mechanism, and design safer drugs. Over time, we will expand across the major organ systems and build the experimental and agentic system of record for translational drug development. Our goal is to dramatically reduce the risk of testing new molecules in humans, enabling high throughput evaluation of efficacy in humans.\nWhat you will do:\nOwn the harness: the scaffolding, tooling, and infrastructure that turn frontier models into agents that do long-horizon scientific analysis\n\nBuild the data backbone that gets everything to the right place: pipelines, storage, and systems for runtime context, agent trajectories, eval results, and training data\n\nBuild sandboxed execution environments with instant spin-up/tear-down, reproducible and deterministic enough to trust for evals and RL\n\nDesign and run the eval systems: offline suites, test cases on production traces, LLM-as-judge pipelines, regression gates\n\nWork closely with domain experts and encode their taste into rubrics, golden sets, and review workflows, turning \"I know it when I see it\" into something measurable\n\nBuild the tools the agent needs to do better work, and stay tuned in to the state of the art for new methods, protocols, and patterns worth adopting\n\nMake every agent run observable and replayable: trace every model call, tool call, and state transition, and build the debugging tooling to make sense of it\n\nEngineer the context: memory, compaction, retrieval, and recovery so long-horizon agent runs stay coherent across hours and crashes\n\nOwn the loop itself: retries, budget caps, stop conditions, output verification, permissions, and guardrails\n\nSupport ML research with environments, reward instrumentation, and rollout infra for RL on agentic tasks\n\nExpertise:\nPython, Modal, DuckDB, FastAPI, Docker, Containerization, Terraform\n\nEngineers who've built with LLM APIs and shipped agentic systems: tool use, loops, and the debugging scars to prove it\n\nBuilt bespoke evaluation, monitoring, and RL env observability tooling (SvelteKit, Svelte 5, React)\n\nWhat we look for:\nCan tackle deep technical challenges and own/ship simple, clean, maintainable code\n\nHigh ownership: owns outcomes end to end, not tickets, and doesn't wait for a spec to start moving\n\nAllergic to complexity: reaches for the simplest system that works and keeps it that way as it scales\n\nStrong software engineer first, with infrastructure, platform, data, or devtools depth and production systems they're proud of\n\nInstinctively asks \"how would we know if this is working?\" and builds the measurement alongside the feature\n\nReads an agent failure trace the way other engineers read a stack trace\n\nCares about reliability because they know environments that break silently poison evals and training data\n\nHas a knack for surfacing the important questions about what the agent actually needs to do better work\n\nSharp and confident\nkeeps up with a really technical & sophisticated crowd\n\ncomfortable getting in over their head and figuring it out as they go\n\nthrives in a discipline with no playbook, because it's being invented right now\n\nCurious about how things work: an engineering/tinkering mindset, good at scavenging the state of the art\n\nPassion for learning what \"good\" looks like from deep domain experts and turning it into systems","description_format":"text","description_chars":4086,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Artificial Intelligence","Biotechnology"],"lifecycle":[{"event":"open","at":"2026-09-26T08:32:57Z"}],"liveness":{"score":9,"band":"cold","label":"Long shot","p_open":1,"p_active":0.329,"p_room":0.28,"age_days":79,"expected_fill_days":21,"reasons":["conf:14","win:tail","crowd:"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/axiom-bio-agent-harness-engineer","json_url":"https://alion.io/job/axiom-bio-agent-harness-engineer.json","meta":{"generated_at":"2026-10-01T12:27:37Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":3870,"day_limit":5000,"remaining_today":1130,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}