{"id":1228668,"url":"https://alion.io/job/parsewave-python-developer-llm-post-training-remote-san-francisco","title":"Python Developer – LLM Post-Training (Remote | San Francisco)","company":{"id":1317,"name":"Parsewave","domain":"parsewave.ai","url":"https://alion.io/company/parsewave","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"AI/ML","role_family":"AI/ML","seniority":null,"employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"inferred_payroll_markers","remote_working_hours":null,"hiring_geo_confidence":"inferred","locations":[],"countries":[],"hiring_countries":["IN"],"hiring_countries_total":1,"salary":null,"salary_estimate":null,"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Git","optional":false},{"name":"Linux","optional":false},{"name":"LLM","optional":false},{"name":"Post-training","optional":false},{"name":"Python","optional":false},{"name":"Reinforcement Learning","optional":false},{"name":"RLHF","optional":false},{"name":"SFT","optional":false},{"name":"Unix","optional":false},{"name":"AI Agents","optional":true},{"name":"AWS","optional":true},{"name":"Bash","optional":true},{"name":"Docker","optional":true},{"name":"FAISS","optional":true},{"name":"FastAPI","optional":true},{"name":"Fine-tuning","optional":true},{"name":"Flask","optional":true},{"name":"GCP","optional":true},{"name":"Hugging Face","optional":true},{"name":"LangChain","optional":true},{"name":"LangGraph","optional":true},{"name":"LlamaIndex","optional":true},{"name":"LoRA","optional":true},{"name":"Milvus","optional":true},{"name":"PEFT","optional":true},{"name":"Pinecone","optional":true},{"name":"Qdrant","optional":true},{"name":"QLoRA","optional":true},{"name":"RAG","optional":true},{"name":"Transformers","optional":true},{"name":"Weaviate","optional":true}],"status":"live","first_seen_at":"2026-09-06T15:48:14Z","employer_posted_date":null,"last_verified_at":"2026-09-06T15:48:14Z","board_verified":false,"closed_at":null,"days_open":27,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":27},"description":"We are a San Francisco-based AI infrastructure company working with leading frontier AI labs to build post-training data and evaluation infrastructure for foundation models. We are hiring a Python Developer to create high-quality datasets, reinforcement learning environments, and benchmarking pipelines used to improve and evaluate state-of-the-art LLMs. This is a remote role with flexible working hours.\n\nResponsibilities\n\n* Create and curate datasets for LLM post-training (SFT, RLHF, RL, preference optimization).\n* Build and maintain RL environments for agent evaluation.\n* Develop Python tooling for dataset generation, validation, and transformation.\n* Evaluate models on custom benchmarks and testing pipelines.\n* Collaborate with research and engineering teams to deliver client-specific post-training datasets.\n* Work with terminal-first development workflows and cloud infrastructure.\n\nRequired Skills\n\n* Strong Python programming skills.\n* Understanding of LLM fundamentals and post-training concepts (SFT, RLHF, RL).\n* Experience working with structured data (JSON, CSV, YAML).\n* Git, Linux/Unix command line, and solid software engineering fundamentals.\n\nGood to Have\n\nExperience with RAG, agentic AI systems, Hugging Face Transformers, LoRA/PEFT, LangChain or LlamaIndex, vector databases (FAISS, Qdrant, Milvus, Pinecone, Weaviate, ChromaDB), Docker, AWS/GCP, FastAPI/Flask, Bash, CLI tooling, model evaluation frameworks, benchmarking, and AI infrastructure.\n\nCompensation\n\nBase Salary: USD $1,250/month\nEquity: ESOP/Equity package included.\nPerformance Bonuses: Up to USD $4,000/month (in addition to base salary).\n\nLocation\n\nRemote (Worldwide)\n\nWork Hours\n\nFlexible, remote-first, asynchronous work environment.\n\nHow to Apply\n\nApply here: https://tally.so/r/wLReJG\nPlease complete the application form and submit the required details. Only shortlisted candidates will be contacted.\nSkills\nPython, Large Language Models (LLM), Fine-tuning LLMs, PEFT (Parameter-Efficient Fine-Tuning), Amazon Web Services (AWS), Google Cloud Platform (GCP), Docker, Retrieval Augmented Generation (RAG), Vector database, FastAPI, Agentic AI, Hugging Face Transformers, LoRA / QLoRA, LangChain, LangGraph, LlamaIndex, Bash","description_format":"text","description_chars":2223,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Equity","Flexible schedule"],"hiring_locations":[{"name":"India","iso":"IN","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["AI Training Data & Annotation"],"lifecycle":[{"event":"open","at":"2026-09-25T14:00:00Z"}],"visa":[],"liveness":{"score":30,"band":"fade","label":"Fading","p_open":0.85,"p_active":0.645,"p_room":0.55,"age_days":26,"expected_fill_days":26,"reasons":["seen:26","win:tail"],"computed_at":"2026-10-03T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/parsewave-python-developer-llm-post-training-remote-san-francisco","json_url":"https://alion.io/job/parsewave-python-developer-llm-post-training-remote-san-francisco.json","meta":{"generated_at":"2026-10-04T01:07:22Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1421,"day_limit":5000,"remaining_today":3579,"minute_limit":60,"resets_at":"2026-10-05T00:00:00Z"}}}