{"id":827767,"url":"https://alion.io/job/radixark-technical-program-manager","title":"Technical Program Manager","company":{"id":686458,"name":"RadixArk","domain":"radixark.com","url":"https://alion.io/company/radixark","size_band":null,"is_staffing_agency":false,"is_intermediary":false,"ats_vendor":"Greenhouse","truth_index":{"grade":"C","score":55,"open_postings":20,"ghost_share":0.75,"stale_share":0,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-09-23T05:45:00Z"}},"role":"Product","role_family":"Product","seniority":"middle","employment_type":null,"work_mode":"on_site","remote_scope":null,"hiring_geo_confidence":"structured","locations":["Palo Alto, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":160000,"max":300000,"currency":"USD","period":"year","gross":null,"usd_annual":300000},"salary_estimate":null,"experience_years_min":4,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"AWS Trainium","optional":false},{"name":"C++","optional":false},{"name":"CI/CD","optional":false},{"name":"CUDA","optional":false},{"name":"CUDA Toolkit","optional":false},{"name":"GitHub","optional":false},{"name":"LLM","optional":false},{"name":"Python","optional":false},{"name":"SGLang","optional":false},{"name":"TPU","optional":false}],"status":"live","first_seen_at":"2026-04-30T02:00:08Z","employer_posted_date":"2026-09-16","last_verified_at":"2026-09-23T20:05:20Z","board_verified":true,"closed_at":null,"days_open":146,"trust":{"level":"ghost","repost_count":0,"flags":["stale","company_stale"],"days_open":146},"description":"About the Role\nAs a Technical Program Manager at RadixArk, you'll drive the execution of complex, cross-functional programs across our inference and training infrastructure. You'll partner closely with Product Management, Research, and Engineering to turn ambitious technical roadmaps into shipped reality, coordinating across kernel teams, distributed systems engineers, and external partners to deliver infrastructure that serves billions of tokens daily and coordinates 10,000+ GPU training runs.\nThis role is for someone who thrives at the intersection of deep technical understanding and rigorous program execution. You'll own the \"how\" and \"when\" of our most critical initiatives.\nKey Responsibilities\nProgram Execution & Delivery\nDrive end-to-end execution of large-scale, cross-functional programs spanning inference engines (e.g., SGLang), training frameworks (e.g., Miles), and hardware integration efforts.\n\nDefine program structure, including milestones, dependencies, critical paths, risks, and success criteria. Maintain a clear source of truth for status across all stakeholders.\n\nRun design reviews, sprint planning, release readiness reviews, and post-mortems. Ensure decisions are documented and follow-ups are closed out.\n\nIdentify and unblock cross-team dependencies across kernel, runtime, scheduler, networking, and model teams before they become release blockers.\n\nDrive release management for major versions, including changelog ownership, compatibility validation, partner rollout sequencing, and rollback planning.\n\nTechnical Coordination\nPartner with Product Management to translate roadmap priorities into executable program plans, with clear scope, staffing, and timelines.\n\nWork shoulder-to-shoulder with engineering leads on technical trade-off decisions; understand the architecture deeply enough to ask the right questions and surface hidden risks.\n\nCoordinate hardware enablement programs with partners like Nvidia, Google, and AWS, including new accelerator bring-up, kernel co-development, and benchmark validation.\n\nManage integration programs with frontier AI labs and early adopters, ensuring technical requirements, SLAs, and feedback loops are well-defined.\n\nOperational Excellence\nBuild and improve the engineering operating cadence, including standups, planning rituals, OKR tracking, dashboards, and reporting to leadership.\n\nEstablish metrics and instrumentation for program health such as velocity, defect rates, benchmark regressions, and customer-reported issues, and drive accountability against them.\n\nLead incident response coordination for production issues affecting partners; own root-cause review and corrective-action tracking.\n\nImprove developer productivity by identifying and removing systemic friction in our build, test, and release pipelines.\n\nStakeholder Communication\nServe as the connective tissue between engineering, product, GTM, and external partners, ensuring everyone has the right information at the right altitude.\n\nProduce clear, concise written updates for leadership and partners. Translate engineering progress into business-relevant signals.\n\nRepresent program status honestly, including risks and slips, with concrete mitigation plans.\n\nQualifications\nMinimum Requirements\nBachelor's or Master's degree in Computer Science, Engineering, or a related technical field.\n\n4+ years of direct experience in Technical Program Management, Engineering Management, or a senior engineering role with significant program ownership, in a software or infrastructure company.\n\nStrong technical fluency in systems software, distributed systems, or AI/ML infrastructure; able to read code, follow architecture discussions, and challenge technical assumptions productively.\n\nDemonstrated track record shipping complex, multi-team programs on time, including managing dependencies, risks, and scope changes.\n\nExcellent written and verbal communication skills; able to drive alignment across engineers, executives, and external partners.\n\nPreferred (Bonus) Qualifications\nDirect experience shipping AI/ML infrastructure such as inference engines, training frameworks, GPU kernels, distributed schedulers, or model serving platforms.\n\nHands-on coding background (Python, C++, CUDA) and comfort working in engineering codebases, including reading PRs, running benchmarks, and reproducing issues.\n\nExperience coordinating with hardware vendors (Nvidia, AMD, Google TPU, AWS Trainium/Inferentia) on enablement or co-engineering programs.\n\nExperience driving open-source release programs or working in OSS communities, including issue triage, RFC processes, and contributor coordination.\n\nFamiliarity with release engineering, CI/CD systems, and observability tooling for large-scale distributed systems.\n\nExperience supporting B2B or developer-facing products with enterprise SLAs.\n\nAbout RadixArk\nRadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Founded by AI infrastructure veterans from xAI and NVIDIA, we're on a mission to democratize frontier-level AI infrastructure by building world-class open systems for inference and training. Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.\nCompensation\nDepending on background, skills, and experience, the expected annual salary range for this position is $160,000 - $300,000 USD + equity.\nEqual Opportunity\nRadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.","description_format":"text","description_chars":5936,"description_truncated":false,"requirements":{"experience_years_min":4,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[]},"benefits":["Equity"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["AI Infrastructure"],"lifecycle":[{"event":"open","at":"2026-09-12T15:26:28Z"}],"liveness":{"score":5,"band":"cold","label":"Long shot","p_open":1,"p_active":0.18,"p_room":0.28,"age_days":146,"expected_fill_days":19,"reasons":["conf:0","stale_co","ghost","win:tail","crowd:"],"computed_at":"2026-09-23T05:45:00Z"},"pay":{"stated_usd_annual":300000,"is_top_pay":true},"html_url":"https://alion.io/job/radixark-technical-program-manager","json_url":"https://alion.io/job/radixark-technical-program-manager.json","meta":{"generated_at":"2026-09-24T00:46:05Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers"}}