1,287,628open jobs
74,529companies
207,347added this week
Browse all
Salary
$160k – $200k per year
Location
In office (New York)
Seniority
Senior · 4+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 7, 2026. First seen by Alion on Oct 6, 2026. Distributed Spectrum │ scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Gain actionable insights from the radio spectrum in real-time with small, cheap and efficient sensors from Distributed Spectrum.

DS creates systems that power the next generation of radio spectrum intelligence. We collect radio data from all over the world, train neural networks to decipher it, and run them on the smallest chips we can. We’re solving a new, technically hard problem where nothing from other fields works out of the box, and along the way, we’ve built our own stack from scratch, including entirely new embedding model architectures, custom GPU kernels, and much more.

Joining DS means owning major parts of a fast-growing AI research organization, joining a collaborative, talent-dense team with decades of experience in probabilistic ML, accelerated computing, embedded systems, and signal theory, and growing your career in the areas that interest you. You’ll fit in if you want to come to work for the problem itself and don’t want to choose between technical rigor, business value, and real-world impact.

We work with high ownership and trust.

The Role

DS collects radio data at a scale few organizations ever see - continuous, high-rate streams from sensors around the globe. We're hiring a Senior Data Engineer to design the platform that turns that firehose into an asset: a data lakehouse that serves researchers training models, production systems running inference, and agents retrieving context in real time.

You'll own the architecture from ingestion through storage, cataloging, vector search, and access, and you'll design it explicitly for AI/ML workloads, not just analytics.

What you'll do

  • Architect and build our data lakehouse: ingestion pipelines, open table formats, partitioning and compaction strategies, cataloging, and governance.

  • Design and operate vector database infrastructure for embedding storage, similarity search, and retrieval at scale - choosing, tuning, and evolving the right systems for our workloads.

  • Build large-scale data management systems: lifecycle and retention, lineage, versioning of datasets for reproducible training, quality monitoring, and cost management across petabyte-class storage.

  • Design data platform capabilities that directly support ML use cases - feature and embedding pipelines, training-set assembly, evaluation datasets, and low-latency retrieval for agents.

  • Establish schema and data-contract standards across teams, and build tooling that lets researchers and engineers self-serve.

  • Own reliability and performance of data pipelines in production, including observability and failure handling.

  • Mentor engineers and lead technical design across the data domain.

What we're looking for

  • 4-6+ years of software / data engineering experience, including several years designing and operating large-scale data platforms in production.

  • Deep experience with lakehouse architectures and technologies - e.g., Apache Iceberg, Delta Lake, or Hudi; Parquet; Spark, Flink, Trino, DuckDB, or similar engines.

  • Hands-on experience with vector databases and embedding retrieval (e.g., pgvector, Milvus, Qdrant, Weaviate, Pinecone, LanceDB, or FAISS-based systems), including indexing trade-offs and scaling.

  • Strong experience with AWS or another major cloud provider - object storage, managed data services, compute, and cost optimization at scale.

  • Fluency in Python and SQL; experience with orchestration tools (Airflow, Dagster, Prefect, Step Functions, etc.).

  • Demonstrated experience designing data architectures specifically for ML / AI workloads: training data pipelines, feature stores, dataset versioning, or retrieval systems.

  • Strong grasp of data modeling, consistency, and the trade-offs between batch and streaming.

Nice to have

  • Experience with streaming ingestion at high volume (Kafka, Kinesis, Pulsar).

  • Experience with time-series, geospatial, or signal / sensor data.

  • Familiarity with data governance and security requirements in regulated environments.

  • Experience with Rust, Go, or C++ for performance-critical data paths.

Who Thrives at Distributed Spectrum

  • Fast learners over specific backgrounds - We care more about how quickly you can pick up new skills than where you’ve worked before.

  • Intellectual honesty - The right answer matters more than being right. You challenge assumptions, test ideas, and pivot when needed.

  • Adaptability - We’re organized, but sometimes things change quickly. You find a way to make it work and balance short-term deliverables with long-term goals.

  • Ownership of outcomes - You optimize your own time, focus on what matters to deliver quickly, and cut out inefficiencies.

  • Not building in a vacuum - You stay connected to the rest of our teams and our customers to make sure all the pieces fit together.

What We Offer

  • Above-market salary, equity, and benefits package.

  • Early Series A Equity

  • Excellent health, dental, and vision coverage

  • 401(k) match - up to 4% of your salary

  • Flexible PTO

  • Daily office lunches in NYC

ITAR Requirements

To conform to U.S. Government technology export regulations, including the International Traffic in Arms Regulations (ITAR) you must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,287,628 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
New York
$197k – $225k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Richmond • McLean
Python
Java
SQL
Scala
Databases
Snowflake
Databricks
Cassandra
DynamoDB
Amazon Redshift
AI/ML
Spark
Dagster
Machine Learning
DevOps
Splunk
GCP
Azure
AWS
Management
Agile
Apply
AI Data Engineer 1 day ago
$129k – $161k per year • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • United States
Python
SQL
Python
pySpark
Databases
Snowflake
Neo4j
Databricks
Delta Lake
Google BigQuery
BigQuery
AI/ML
Spark
Great Expectations
LLM
Knowledge Graph
DevOps
Terraform
CI/CD
Git
AWS
SLI/SLO/SLA
Amazon S3
IAM
Analytics
ETL/ELT
Apply
Data Engineer 1 hour ago
$110k – $135k per year • In office • 3+ years exp • Bachelor's Degree • Columbus
Python
SQL
Python
pySpark
Databases
Databricks
Microsoft Fabric
AI/ML
Spark
Machine Learning
DevOps
Azure DevOps
Azure
Git
Analytics
Power BI
ETL/ELT
Azure Data Factory
Dimensional Modeling
Management
SharePoint
Apply
Data Engineer 1 hour ago
$125k – $150k per year • In office • 3+ years exp • Bachelor's Degree • Los Angeles
Python
SQL
Python
pySpark
Databases
Databricks
Microsoft Fabric
AI/ML
Spark
Machine Learning
DevOps
Azure DevOps
Azure
Git
Analytics
Power BI
ETL/ELT
Azure Data Factory
Dimensional Modeling
Management
SharePoint
Apply
Data Engineer 1 hour ago
$110k – $135k per year • In office • 3+ years exp • Bachelor's Degree • Spokane
Python
SQL
Python
pySpark
Databases
Databricks
Microsoft Fabric
AI/ML
Spark
Machine Learning
DevOps
Azure DevOps
Azure
Git
Analytics
Power BI
ETL/ELT
Azure Data Factory
Dimensional Modeling
Management
SharePoint
Apply
$110k – $190k per year • In office • 3+ years exp • Bachelor's Degree • New York
Python
SQL
AI/ML
AI Agents
NLP
DevOps
Platform Engineering
Analytics
Tableau
Power BI
QlikSense
ETL/ELT
Apply
$191k – $248k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • New York • Pittsburgh • Boston
Marketing
Salesforce
Apply
≈ $101k – $214k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Malvern • New York • Boston • Charlotte • Washington
Apply
$132k – $212k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • New York
Management
Microsoft Office
Apply
$174k – $286k per year • In office • Full-Time • New York
Apply
See all jobs
This is one of many
1,287,628 more open roles from verified company boards, updated every day.