{"id":1532844,"url":"https://alion.io/job/ixigo-fellowship-agent-intelligence-evaluations","title":"Fellowship : (Agent Intelligence & Evaluations)","company":{"id":4076,"name":"Ixigo","domain":"ixigo.com","url":"https://alion.io/company/ixigo","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"SmartRecruiters","truth_index":null},"role":"Science & Research","role_family":"Science & Research","seniority":null,"employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["New Delhi, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":15500,"max_usd":44000,"period":"year","method":null,"sample_n":2420},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Arize Phoenix","optional":false},{"name":"Fine-tuning","optional":false},{"name":"Langfuse","optional":false},{"name":"LLM","optional":false},{"name":"NLP","optional":false},{"name":"OpenTelemetry","optional":false},{"name":"Python","optional":false},{"name":"Self-Healing","optional":false},{"name":"Speech Recognition","optional":false},{"name":"Text-to-Speech","optional":false},{"name":"Tool Use","optional":false},{"name":"Voice Agents","optional":false},{"name":"Whisper","optional":false},{"name":"Interpretability","optional":true}],"status":"closed","first_seen_at":"2026-09-30T10:53:53Z","employer_posted_date":"2026-09-30","last_verified_at":"2026-10-05T06:06:58Z","board_verified":false,"closed_at":"2026-10-05T06:06:58Z","days_open":4,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":4},"description":"We're building self-healing voice agents for enterprise customer support within ixigo. The system has to know when it's failing, why it's failing, and how to fix itself before a human notices. This fellowship sits at the intelligence layer behind that work.\n Voice agents fail in ways traditional software doesn't. An ASR confidence drop on a regional accent misfires a tool call, an LLM hallucinates a policy because upstream latency broke turn-taking, and support teams roll these agents back within a week without anyone able to explain what went wrong.\nWhat you'll work on\nEvaluation frameworks. Text-only evals miss most of what matters in voice: barge-in, prosody, latency-induced errors, cross-turn context loss. You'll design audio-native metrics, generate adversarial conversational datasets across accents and edge cases, and build LLM-as-judge rubrics for task completion, empathy, and recovery from tool failures.\nEnd-to-end observability. Tracing a failed interaction means correlating audio packets, STT hypotheses, LLM reasoning traces, tool calls, and TTS output back to a single conversation ID. You'll help shape the schema and analysis layer that makes cascade failures visible across the stack.\nSelf-improvement systems. Once you can measure and trace, the interesting work is closing the loop: mining production traces for failure patterns, generating targeted fine-tuning data or prompt updates, and validating that fixes hold under adversarial replay.\nWho we're looking for\nSomeone who cares about the research questions for their own sake, and equally cares whether the work ships. Papers at Interspeech, ACL, NeurIPS, or EMNLP on speech, dialogue systems, agent evaluation, or human-AI interaction are directly relevant.\nComfortable in Python, and familiar with at least one of: speech models (Whisper, Conformer variants), LLM tool-use and agent frameworks, or observability stacks (OpenTelemetry, Langfuse, Arize, Hamming).\nCurrent PhD students in ML, NLP, or speech are the strong default; exceptional MS students or research engineers with a publication track record are welcome to apply.\nNice to have\nPrior work on evaluation methodology, dataset synthesis, or interpretability. Experience with real-time systems, telephony, or streaming pipelines. A blog, repo, or workshop paper that shows how you think in public.\nThis is a full-time role with a competitive salary and ESOPs.\n Our Culture: ixigo is proud to have built an entrepreneurial culture that has become a folk-lore in the startup ecosystem. One in every four ixigems has gone on to build successful startups and companies. Our cultural values of integrity, empathy, ingenuity, awesomeness, and resilience have stood the tests of time and we’ve built a fun, flexible and creative work environment that is driven by people with a high degree of ownership. You will get to work with some of the smartest folks in the Indian startup ecosystem, and solve some of the toughest problems for the next billion users by using bleeding-edge technologies. Oh, and we have an awesome “play” area, great chai/coffee, free lunches (yes, they exist!) and a workspace you will fall in love with.","description_format":"text","description_chars":3170,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"phd","optional":false},"security_clearance":false,"languages":[{"language":"English","level":"All levels","optional":false}]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Travel & Tourism"],"lifecycle":[{"event":"open","at":"2026-09-30T17:03:05Z"},{"event":"close","at":"2026-10-05T06:06:58Z"}],"visa":[],"liveness":null,"pay":null,"html_url":"https://alion.io/job/ixigo-fellowship-agent-intelligence-evaluations","json_url":"https://alion.io/job/ixigo-fellowship-agent-intelligence-evaluations.json","meta":{"generated_at":"2026-10-06T21:07:37Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":467,"day_limit":5000,"remaining_today":4533,"minute_limit":60,"resets_at":"2026-10-07T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":4076},"rest":"https://alion.io/mcp/rest/get_company?id=4076"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fixigo-fellowship-agent-intelligence-evaluations"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fixigo-fellowship-agent-intelligence-evaluations"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fixigo-fellowship-agent-intelligence-evaluations"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/ixigo-fellowship-agent-intelligence-evaluations\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fixigo-fellowship-agent-intelligence-evaluations"}]}