{"id":1424913,"url":"https://alion.io/job/clarium-staff-data-engineer","title":"Staff Data Engineer","company":{"id":5320,"name":"Clarium","domain":"clarium.com","url":"https://alion.io/company/clarium","size_band":"51-200","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Ashby","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"staff","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":[],"countries":[],"hiring_countries":["US"],"hiring_countries_total":1,"salary":{"min":170000,"max":210000,"currency":"USD","period":"year","gross":null,"usd_annual":210000},"salary_estimate":null,"experience_years_min":8,"visa_sponsorship":false,"relocation_package":false,"has_equity":true,"technologies":[{"name":"Airflow","optional":false},{"name":"Amazon ECS","optional":false},{"name":"Amazon S3","optional":false},{"name":"Apache Kafka","optional":false},{"name":"AWS","optional":false},{"name":"AWS Lambda","optional":false},{"name":"CI/CD","optional":false},{"name":"Dagster","optional":false},{"name":"dbt","optional":false},{"name":"Fivetran","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Prefect","optional":false},{"name":"Python","optional":false},{"name":"Snowflake","optional":false},{"name":"Spark","optional":false},{"name":"SQL","optional":false},{"name":"Amazon EKS","optional":true},{"name":"GitHub Actions","optional":true},{"name":"HIPAA","optional":true},{"name":"Kubernetes","optional":true},{"name":"Terraform","optional":true}],"status":"live","first_seen_at":"2026-09-28T14:52:07Z","employer_posted_date":"2026-09-28","last_verified_at":"2026-09-30T17:21:50Z","board_verified":true,"closed_at":null,"days_open":2,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":2},"description":"Why Clarium?\nThe healthcare industry overspends on its supply chain by over $25B each year, the result of fragmented data, inefficient workflows, and wasted supplies. Clarium is fixing that. Our AI-powered platform, Astra OS, gives hospitals end-to-end visibility into their supply chain operations, automating workflows and surfacing actionable insights so supply chain teams can focus on what matters most: patient care. We're trusted by some of the world's leading health systems, including Yale New Haven Health, Stanford, Geisinger, and Kaiser Permanente.\nFounded in 2020, Clarium has raised $43M in total funding. Our Series A was led by Northzone, with participation from General Catalyst, AlleyCorp, Kaiser Permanente Ventures, Texas Medical Center Ventures, and 1984 Ventures.\nThe Opportunity\nThis is a staff-level individual contributor role for an engineer who sets technical direction for our data platform rather than working ticket to ticket. You'll decide how data gets modeled, moved, and trusted across Clarium, and you'll leave both the systems and the engineers around you measurably better.\nYou'll work across our transactional databases and analytical warehouse, own the pipelines that connect them, and partner directly with the analytics, product, and engineering teams who depend on that data. Much of the work is supply chain data arriving from hospital ERP systems, where schemas and semantics vary by source and are rarely well documented, so a large part of the job is deciding what should be built, not only building it.\nIn This Role You Will\nOwn the architecture and technical roadmap for core data infrastructure on AWS, spanning ingestion, transformation, storage, and serving layers\n\nDesign, build, and operate reliable batch and near-real-time pipelines with clear SLAs and the observability to back them up\n\nModel supply chain data from external ERP systems into a coherent, reusable warehouse model, and lead the migration of legacy assets toward it\n\nTune performance and cost across Postgres and Snowflake, including query plans, indexing, partitioning, warehouse sizing, and storage strategy\n\nEstablish engineering standards for data work (testing, code review, CI/CD, data quality checks, documentation) and mentor engineers through design reviews, pairing, and honest technical feedback\n\nPartner with analytics, product, and engineering teams to turn ambiguous questions into durable data models instead of ad hoc extracts\n\nLead governance for sensitive data, including lineage, access controls, retention, auditability, and de-identification where required\n\nWhat You'll Bring\n8+ years of data or software engineering experience, including significant time owning systems end to end in production\n\nExpert-level SQL: you write complex analytical queries as a matter of course and can read a query plan and explain why something is slow\n\nStrong Python for production data work, including pipeline code, transformation logic, testing, and tooling\n\nDeep experience with relational databases, including schema design, normalization tradeoffs, transactions, and performance tuning (Postgres and Snowflake strongly preferred)\n\nHands-on experience with data pipeline and orchestration tooling such as Airflow, Dagster, Prefect, dbt, Fivetran, Spark, or Kafka; we care more about depth and judgment than an exact stack match\n\nProduction experience running data workloads on AWS (e.g., S3, RDS, Lambda, ECS), with an understanding of the cost, security, and networking implications of how you build\n\nA track record of leading multi-quarter, cross-team initiatives that depended on teams you don't manage, and comfort with the ambiguity of deciding what should be built\n\nNice to Have\nExperience with healthcare data (claims, EHR/EMR, HL7 or FHIR, ICD-10, CPT) or working under HIPAA with PHI, de-identification, and audit requirements\n\nFamiliarity with supply chain data from ERP systems such as Oracle, Workday, or Lawson, including where it tends to be unreliable\n\nInfrastructure-as-code and deployment automation (Terraform, CDK, CI/CD for data infrastructure)\n\nStreaming and event-driven architectures, or data quality and observability tooling\n\nSkills & Tools You'll Use\nNeed to Know: SQL · Python · Postgres · Snowflake · AWS (S3, RDS, Lambda, ECS)\nNice to Know: Terraform · Github Actions · EKS / Kubernetes · dbt · Prefect · Apache Kafka\nWhat You Get at Clarium\nTarget Base Salary Range: $170K - $210K\nThe base salary Clarium offers may vary depending upon the ultimate scope and responsibilities of the position and on the candidate's job-related knowledge, skills, and experience. The total package will include equity, in addition to a full range of medical and/or other benefits, depending on the position offered. Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation or benefit plans.\nIncentive Stock Options proportionate to your salary\n\nFully remote, with a NYC co-working space available; distributed team across multiple time zones with opportunities for in-person time\n\nUnlimited PTO\n\nTop-tier health, vision, and dental benefits\n\n401K\n\nThe opportunity to build on a strong foundational team with deep data and engineering roots at a stage where your work genuinely shapes the product\n\nA fast-paced, high-growth environment where we move quickly to solve critical healthcare challenges\n\nThe chance to empower frontline healthcare workers by automating the administrative friction in their day-to-day\n\nCollaborative work alongside a high-caliber team of clinicians, engineers, and revenue leaders\n\nEqual Opportunity Statement\nClarium is committed to promoting an inclusive work environment free of discrimination and harassment. We value a diverse and balanced team where everyone can belong.","description_format":"text","description_chars":5794,"description_truncated":false,"requirements":{"experience_years_min":8,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["401k plan","Equity","Stock options","Unlimited PTO"],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Information Technology"],"lifecycle":[{"event":"open","at":"2026-09-28T23:25:46Z"}],"liveness":{"score":86,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.86,"p_room":1,"age_days":1,"expected_fill_days":42,"reasons":["conf:7","win:early"],"computed_at":"2026-09-30T05:45:00Z"},"pay":{"stated_usd_annual":210000,"is_top_pay":true},"html_url":"https://alion.io/job/clarium-staff-data-engineer","json_url":"https://alion.io/job/clarium-staff-data-engineer.json","meta":{"generated_at":"2026-10-01T03:25:46Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2883,"day_limit":5000,"remaining_today":2117,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}