{"id":1627302,"url":"https://alion.io/job/spacenexus-data-platform-engineer","title":"Data Platform Engineer","company":{"id":2086894,"name":"SpaceNexus","domain":"spacenexus.us","url":"https://alion.io/company/spacenexus","size_band":"1-10","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Schema","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":null,"work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["San Francisco, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":153000,"max":250000,"currency":"USD","period":"year","gross":null,"usd_annual":250000},"salary_estimate":null,"experience_years_min":7,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AI Agents","optional":false},{"name":"Amazon Kinesis","optional":false},{"name":"Amazon Redshift","optional":false},{"name":"Anomaly Detection","optional":false},{"name":"Apache Iceberg","optional":false},{"name":"Apache Kafka","optional":false},{"name":"dbt","optional":false},{"name":"Dimensional Modeling","optional":false},{"name":"Fivetran","optional":false},{"name":"LLM Guardrails","optional":false},{"name":"Metabase","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Python","optional":false},{"name":"Redpanda","optional":false},{"name":"Scala","optional":false},{"name":"SOC 2","optional":false},{"name":"SQL","optional":false},{"name":"Stripe","optional":false}],"status":"live","first_seen_at":"2026-09-22T23:11:43Z","employer_posted_date":"2026-09-22","last_verified_at":"2026-10-01T21:33:12Z","board_verified":true,"closed_at":null,"days_open":9,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":9},"description":"Hubble Network was founded with the intention of delivering on the promise of what Internet-of-Things (IoT) was supposed to be. We're building a global Bluetooth® network dedicated to machine-to-machine connectivity. We differentiate ourselves as the first modem-less and gateway-less, direct-to-satellite network from off-the-shelf Bluetooth® Low Energy chips. Hubble is ideal for applications in logistics, AgTech, and maritime where economies of scale for volume consumer and enterprise asset tracking is a priority. Our goal is to be the first billion-endpoint-connected network in the world.\nHubble is an early-stage, venture-backed startup supported by some of the best investors in the world. In their previous lives, the founding team has been successful in raising $100s of millions in venture funding, developing the Amazon Sidewalk network, launching billions of dollars of space assets, and leading their teams to successful exits, both through acquisition and IPO. We are now looking to bring on talented team members who are the best at what they do to help us make Hubble a reality for the world.\nAbout the Role\nThis role will be based in San Francisco, CA\nThis is the founding data hire. You will build Hubble's data platform from the warehouse layer up, make decisions on architecture, determine processes for managing schema and data models, and work with many stakeholders.\nWe know what we want: a layered lakehouse with Landing, Staging, Warehouse, Mart stages. Each layer has an owner, a purpose, and a quality bar. You’ll work with the Infrastructure team to manage the pipes: Fivetran, Redshift, orchestration, monitoring. You own everything downstream of raw: the transformations, the models, the metric definitions, the quality gates, and the performance bar.\nThe guiding principle is data democratization. Every person at Hubble should be able to answer their own questions. Success is not how many queries you run for other people, it's how few they need you to run.\nYou'll be a peer to Platform Engineering and to the team building our internal operations platform. You'll be an embedded consultant to Finance, Product, and Customer Success. You’ll think about how to present a strong story from data.\nWe expect you to use AI agents as part of how you work, and data work rewards this more than most: schema exploration, model scaffolding, test generation, and documentation are all places where agentic tooling earns its keep if you review the output critically.\nKey Responsibilities\nOwn the transformation layer end-to-end: Defined models across Staging, Warehouse, and Mart. Star-schema design, SCD2, and data contracts enforced by tests at every layer.\nBuild the data layer our internal tools run on: mart tables designed for specific operational workflows, sitting behind a generated internal metrics API. You own the mart and the API's performance characteristics; Engineering owns the framework and the frontend. Hold the line at p95 under 5 seconds on mart queries and under 500ms on metric endpoints.\nMake freshness a commitment, not an aspiration: core data and critical operational metrics streamed in real time with Apache Iceberg, with monitoring that catches breakage before a stakeholder does.\nInstrument the funnels: Design event collection across our different product surfaces, unified into a customer 360 spanning authentication, usage, and billing. Acquisition, activation, engagement, retention, revenue.\nData in space: We collect large amounts of telemetry from our satellites which is critical to mission operations - help our Mission Operations group land this telemetry in dashboards and monitoring systems so they can take action on data immediately.\nAutomate revenue reconciliation: Daily comparison of Stripe Platform billing against Hubble usage, with discrepancies flagged before month-end and drill-down to the device and packet level. Turn a multi-day manual close into a few hours.\nDefine metrics once: Active devices, packet volume, MRR, contract utilization. Every number has a canonical definition, an owner, and visible lineage from source to dashboard. “Which revenue number is correct?” should stop being a question anyone asks.\nImplement classification in the pipeline: Security defines the tiers and the guardrails. You tag datasets by sensitivity, enforce retention, masking, and anonymization, and keep lineage and access logs auditable. Hubble does not store PII, and the pipeline is where that commitment either holds or quietly fails.\nBuild for self-service: Semantic layers that let business users work with customers and revenue rather than join keys, curated datasets, a data catalog, and documentation good enough that people stop asking you.\nChoose boring technology: Fivetran, dbt, Redshift, Metabase are all options - we lean towards buy over build. Bring a strong opinion about which is which, and be able to defend it.\nTechnical Requirements\nRequired\n7+ years building production data systems, with 2+ years at senior or staff scope owning architecture rather than executing someone else's.\nDeep SQL and dimensional modeling: star schemas, slowly changing dimensions, grain discipline. You can explain fact tables and grains.\ndbt (or similar) in production: model organization, macros, incremental strategies, tests as contracts, and CI that blocks bad merges.\nCloud data warehouse depth, ideally Redshift (or similar): distribution and sort keys, vacuum and analyze behavior, workload management, and the ability to take a slow query apart and make it fast.\nPython or Scala for the specialized pipelines: custom connectors against proprietary APIs, backfills, and reconciliation jobs.\nEvent and product analytics instrumentation: you've defined a tracking plan, argued about event taxonomy, and built funnel tables that survive contact with a changing product.\nData quality as engineering: anomaly detection, freshness monitoring, and incident response with real escalation paths. You've been paged for a broken pipeline and you've fixed the class of bug, not the instance.\nFluency with AI agents in your own workflow: you already use agentic coding tools to ship real work, and you have a point of view on where they help and where they quietly introduce errors that only show up three models downstream.\nArchitecture & Scale\nExperience with high-volume machine-generated data, telemetry, events, or IoT at billions of rows, where your partitioning and incremental strategy created a cost-efficient and performant data product.\nSemantic layer and metric definition work: you've built the abstraction that keeps one number meaning one thing across every tool.\nStreaming or near-real-time ingestion (Kinesis, Kafka, Redpanda) with Apache Iceberg or Pinot and a clear view of when it's warranted and when a scheduled batch is the honest answer.\nOwnership of a BI tool as a platform: permissions, curated collections, and the discipline that keeps a Metabase instance from becoming an unmaintained graveyard.\nPreferred\nInterest or experience in IoT, satellite systems, telemetry, or geospatial data.\nBilling, usage-based pricing, or revenue reconciliation and recognition workflows.\nSOC 2 or similar compliance work: classification frameworks, access reviews, audit evidence.\nOpen source contributions or public technical writing.\nWhat We Look For\nCritical Thinking\nYou interrogate a metric request before building it, because the question behind it is often different from the question asked. You spot the difference between a data quality problem and a source system problem, and you fix upstream when you can. You make decisions with incomplete information and revise them as you learn.\nPartnership\nYou work as a peer to Platform and Engineering, with clear ownership boundaries and no throwing work over the wall. You sit with Finance and Product long enough to understand what they're actually trying to decide. You disagree constructively and commit fully once a decision is made.\nOwnership & Pragmatism\nYou've shipped a data model you later had to tear out, and you learned something specific from it. You explain lineage and trade-offs to people who don't write SQL. You know when to ship the imperfect table that unblocks a decision this week and when to slow down and get the grain right.\nCompensation & Benefits\nSalary: $153000-$250,000 (commensurate with experience )\nComprehensive benefits: Health, Dental, Vision, HSA/FSA options\nUnlimited PTO\nCommuter benefits (if working from HQ)\nLearning & Development allowance\nHealth & Wellness stipend\nSabbatical program\nWork on state-of-the-art satellite systems\nThe posted compensation range reflects multiple levels within Hubble Network's Talent architecture. Final level placement and corresponding compensation will be determined through the interview and assessment process, taking into consideration factors such as the candidate's relevant experience, qualifications, skills, and geographical location.\nITAR REQUIREMENTS\nHubble is required by the U.S. Government to comply with various space technology export regulations including the International Traffic in Arms Regulations (ITAR). All applicants must be U.S. citizens or lawful permanent residents (“green card holders”) as defined by ITAR (22 CFR §120.15). More information on ITAR can be found here .\nHubble is committed to creating a diverse environment and is proud to be an equal-opportunity employer. Each individual has the right to work in a professional environment that promotes equal employment opportunity and prohibits discriminatory practices, including harassment. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.\nSeattle Washington\nWashington Pay\n$153,000 - $250,000 USD","description_format":"text","description_chars":9866,"description_truncated":false,"requirements":{"experience_years_min":7,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Unlimited PTO"],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Space & Aerospace"],"lifecycle":[{"event":"open","at":"2026-10-01T21:05:57Z"}],"liveness":{"score":74,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.824,"p_room":0.9,"age_days":9,"expected_fill_days":25,"reasons":["conf:4","win:mid","comp:brand"],"computed_at":"2026-10-02T02:30:19Z"},"pay":{"stated_usd_annual":250000,"is_top_pay":true},"html_url":"https://alion.io/job/spacenexus-data-platform-engineer","json_url":"https://alion.io/job/spacenexus-data-platform-engineer.json","meta":{"generated_at":"2026-10-02T02:30:19Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":3359,"day_limit":5000,"remaining_today":1641,"minute_limit":60,"resets_at":"2026-10-03T00:00:00Z"}}}