{"id":1172093,"url":"https://alion.io/job/weekday-ai-engineer-ai-labs","title":"AI Engineer (AI labs)","company":{"id":7097,"name":"Weekday","domain":"weekday.works","url":"https://alion.io/company/weekday","size_band":"51-200","is_staffing_agency":true,"employer_type":"agency","is_intermediary":false,"listed_via":null,"ats_vendor":"Workable","truth_index":{"grade":"B","score":81,"open_postings":86,"ghost_share":0,"stale_share":0.977,"repost_share":0,"time_to_fill_p50_days":5,"computed_at":"2026-09-26T05:45:00Z"}},"role":"AI/ML","role_family":"AI/ML","seniority":null,"employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"explicit","locations":["Bengaluru, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":3000000,"max":20000000,"currency":"INR","period":"year","gross":null,"usd_annual":31446},"salary_estimate":null,"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AI Agents","optional":false},{"name":"Computer Use","optional":false},{"name":"Function Calling","optional":false},{"name":"LLM","optional":false},{"name":"Python","optional":false},{"name":"Structured Outputs","optional":false},{"name":"Tool Use","optional":false}],"status":"live","first_seen_at":"2026-09-24T08:24:25Z","employer_posted_date":"2026-09-24","last_verified_at":"2026-09-27T00:51:42Z","board_verified":true,"closed_at":null,"days_open":2,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":2},"description":"This role is for one of Weekday’s clients\nSalary range: Rs 3000000 - Rs 20000000 (ie INR 30 - 200 LPA)\nMin Experience: 3+ years\nLocation: Bengaluru, Karnataka, India\nJobType: full-time\nThis role involves building and improving AI agents that generate and validate real-world tasks. You will be responsible for enhancing these agents using traces and evaluation results, turning enterprise workflows into executable tasks for AI training and evaluation, and ensuring their quality and behavior as frontier models evolve.\nRequirements\nKey Responsibilities\nBuild and improve agents, prompts, tool interfaces, context handling, and orchestration across task generation and validation.\nInspect model calls, tool use, trajectories, screenshots, state changes, and grader outputs for trace-driven debugging.\nTurn recurring failures into targeted code and agent changes.\nReview pipeline outputs, form hypotheses, implement fixes, and test on fresh runs and regression suites before rolling out improvements.\nChoose and instrument tracing, debugging, and experiment tools.\nTrack quality, latency, token usage, and cost by agent and pipeline stage to improve model and tool choices.\nImprove validation logic and reward criteria for task and grader quality, ensuring shipped tasks are solvable and evaluations reflect genuine completion.\nPartner with Platform Engineers to turn agent improvements into modular components that can be evaluated and released reliably.\nOwn a workflow from output review through debugging, evaluation, and release to find the actual failure.\nBuild repeatable checks to distinguish agent failures from task or grader defects, making evaluation signals trustworthy.\nFix recurring failure patterns and prove gains across task families, fresh batches, and regression suites.\nEstablish a repeatable way to compare quality, latency, and cost across versions for fast iteration and measurable improvements.\nQualifications\nStrong software engineering fundamentals and hands-on Python experience.\nAbility to read unfamiliar code and debug a system across component boundaries.\nExperience building LLM agents or multi-step AI workflows with tool use, structured outputs, and state management.\nComfort designing evaluations and using traces and data to distinguish a real improvement from a lucky run.\nDemonstrated ownership through implementation and verification, with clear communication about failures, tradeoffs, and evidence-supported decisions.\nUseful additional experience includes computer-use agents, RL environments, reward design, browser automation, or tracing and evaluation platforms.\nMust-have skills\nLLM Agents, Multi-step AI workflows, Reinforced Learning environments\nGood-to-have skills\nBrowser automation, State management, Reward design","description_format":"text","description_chars":2765,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"India","iso":"IN","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Artificial Intelligence"],"lifecycle":[{"event":"open","at":"2026-09-24T08:24:25Z"}],"liveness":{"score":30,"band":"fade","label":"Fading","p_open":1,"p_active":0.332,"p_room":0.9,"age_days":1,"expected_fill_days":5,"reasons":["conf:0","agency","stale_co","velocity","win:mid","comp:brand"],"computed_at":"2026-09-26T05:45:00Z"},"pay":{"stated_usd_annual":31446,"is_top_pay":false},"html_url":"https://alion.io/job/weekday-ai-engineer-ai-labs","json_url":"https://alion.io/job/weekday-ai-engineer-ai-labs.json","meta":{"generated_at":"2026-09-27T01:30:19Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1358,"day_limit":5000,"remaining_today":3642,"minute_limit":60,"resets_at":"2026-09-28T00:00:00Z"}}}