{"id":1303800,"url":"https://alion.io/job/millennium-management-senior-full-stack-data-platform-engineer","title":"Senior Full Stack Data Platform Engineer","company":{"id":704283,"name":"Millennium Management","domain":"mlp.com","url":"https://alion.io/company/millennium-management","size_band":"11-50","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Eightfold","truth_index":{"grade":"B","score":75,"open_postings":80,"ghost_share":0,"stale_share":1,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-09-27T05:45:00Z"}},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":null,"work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Bengaluru, India"],"countries":["IN"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":27000,"max_usd":57000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":9},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Airflow","optional":false},{"name":"Apache Iceberg","optional":false},{"name":"Apache Kafka","optional":false},{"name":"Apache Solr","optional":false},{"name":"AWS","optional":false},{"name":"C++","optional":false},{"name":"ElasticSearch","optional":false},{"name":"ETL/ELT","optional":false},{"name":"GCP","optional":false},{"name":"Java","optional":false},{"name":"NLP","optional":false},{"name":"OCR","optional":false},{"name":"Python","optional":false},{"name":"Redis","optional":false},{"name":"Semantic Search","optional":false},{"name":"Semantic Search","optional":false},{"name":"SQL","optional":false},{"name":"Unstructured.io","optional":false},{"name":"AG Grid","optional":true},{"name":"Angular","optional":true},{"name":"D3.js","optional":true},{"name":"Highcharts","optional":true},{"name":"Hugging Face","optional":true},{"name":"JavaScript","optional":true},{"name":"Milvus","optional":true},{"name":"Pinecone","optional":true},{"name":"React.js","optional":true},{"name":"spaCy","optional":true},{"name":"TypeScript","optional":true},{"name":"Vue.js","optional":true},{"name":"Weaviate","optional":true}],"status":"live","first_seen_at":"2026-01-05T00:00:00Z","employer_posted_date":"2026-01-05","last_verified_at":"2026-09-27T16:10:00Z","board_verified":true,"closed_at":null,"days_open":266,"trust":{"level":"stale","repost_count":0,"flags":["stale"],"days_open":265},"description":"Senior Full Stack Data Platform EngineerFounded in 1989, Millennium is a global alternative investment management firm. Millennium seeks to pursue a diverse array of investment strategies across industry sectors, asset classes and geographies. The firm’s primary investment areas are Fundamental Equity, Equity Arbitrage, Fixed Income, Commodities and Quantitative Strategies. We solve hard and interesting problems at the intersection of computer science, finance, and mathematics. We are focused on innovating and rapidly applying innovations to real world scenarios. This enables engineers to work on interesting problems, learn quickly and have deep impact to the firm and the business.\nAt Millennium, we are redefining how investment decisions are made. We don't just look at balance sheets; we harness the chaos of the real world. By analyzing vast amounts of unstructured data-from news briefings and earnings call audio to regulatory documents-we provide our Portfolio Managers (PMs) with the \"informational edge\" (Alpha) they need to outperform the market.\nThe Role\nWe are seeking a Senior Full Stack Software Engineer with deep expertise in building high-throughput data platforms.\nIn this role, you will architect scalable data platforms using Python, Java, C++, build robust APIs, and enable processing of data using genAI techniques. You will build and optimize a config-driven, plugin-enabled data platform that will allow the construction of DAGs for data processing. You will then apply the platform to build reusable components and pipelines that will ingest gigabytes of unstructured text, audio, and video. You will enable a variety of rich data consumption use-cases by building the right abstractions and APIs for data consumers.\nYou will be the bridge between complex ML research and real-time trading decisions, working in a poly-language environment (Python, Java, C++) where performance is paramount.\nKey Responsibilities\nHigh-Performance Data Pipelines: Architect low-latency, high-throughput platform that enables rapid development of pipelines to ingest and normalize unstructured data (PDFs, news feeds, audio streams).\n\nAI & ML Integration: Build the infrastructure that wraps and serves NLP and ML models. You will ensure that model inference happens in real-time within the data stream.\n\nBackend Microservices: Develop robust backend services to handle metadata management, search, and retrieval of processed alternative data.\n\nSystem Optimization: Tune the platform for speed. In financial markets, milliseconds matter; you will optimize database queries, serialization, and network calls to ensure data reaches the PMs instantly.\n\nData Strategy: Implement storage strategies for unstructured data, utilizing Vector Databases for semantic search and Distributed File Systems for raw storage.\n\nRequired Qualifications\nData Platform Experience: Minimum 5+ years of software engineering experience, preferably building data platforms.\n\nCore Languages: Strong proficiency in both Python and Java/C++ is required. You should be comfortable switching between these languages for different use cases (e.g., Python for data processing, Java or C++ for high-concurrency, scalable services).\n\nData Engineering: Proven experience building data pipelines, ETL processes, or working with big data frameworks (e.g., Kafka, Airflow, Apache Parquet, Arrow, Iceberg , KDB etc).\n\nUnstructured Data Expertise: Proven experience working with unstructured data types (Text, Audio, Documents). Familiarity with techniques such as OCR, transcription normalization, text extraction.\n\nDatabase Knowledge: Proficiency in SQL and significant experience with search/NoSQL engines (Elasticsearch, Redis, Solr, MongoDB or equivalent).\n\nCloud Native: Experience building serverless data lakes or processing pipelines on AWS/GCP, etc\n\nPreferred Qualifications\nAI/NLP Exposure: Experience working with Large Language Models (LLMs), Vector Databases (Pinecone, Milvus, Weaviate), or NLP libraries (Hugging Face, spaCy) or similar\n\nFrontend Competence: Solid experience with modern frontend frameworks (React, Vue, or Angular) and data visualization libraries (e.g., D3.js, Highcharts, or AG Grid).\n\nFinancial Knowledge: Understanding of financial instruments (Equities, Fixed Income) or the investment lifecycle.\n\nDocument Processing: Familiarity with parsing complex document structures (Earnings calls transcripts, 10-K/10-Q filings, Broker Research, Sector and Industry Reports, Central Bank documents, news wires, social media, etc).","description_format":"text","description_chars":4536,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Equity"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Hedge Funds & Alternative Investments"],"lifecycle":[{"event":"open","at":"2026-09-26T12:48:45Z"}],"liveness":{"score":5,"band":"cold","label":"Long shot","p_open":1,"p_active":0.18,"p_room":0.28,"age_days":265,"expected_fill_days":31,"reasons":["conf:0","stale_co","velocity","win:tail","crowd:brand"],"computed_at":"2026-09-27T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/millennium-management-senior-full-stack-data-platform-engineer","json_url":"https://alion.io/job/millennium-management-senior-full-stack-data-platform-engineer.json","meta":{"generated_at":"2026-09-28T04:34:54Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2873,"day_limit":5000,"remaining_today":2127,"minute_limit":60,"resets_at":"2026-09-29T00:00:00Z"}}}