Skip to content

Open nowPosted 16 hours ago

Senior Software Engineer, Data Engineering

Distributed Spectrum14 open roles

Pay
$160,000 – $200,000 a year
Where
New York City
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior Software Engineer, Data EngineeringDistributed Spectrum · New York City
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Distributed Spectrum's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.1% of postings close within 7 days. Measured by our own scanner across the market. Distributed Spectrum postings stay open a median of 51 days.

Share of postings closed within
  1. 1.7%1 day
  2. 3.6%3 days
  3. 8.1%7 days
  4. 15.1%14 days
  5. 34.0%30 days
This job: posted 16 hours ago

Distributed Spectrum median: 51 days open

The posting

DS creates systems that power the next generation of radio spectrum intelligence. We collect radio data from all over the world, train neural networks to decipher it, and run them on the smallest chips we can. We’re solving a new, technically hard problem where nothing from other fields works out of the box, and along the way, we’ve built our own stack from scratch, including entirely new embedding model architectures, custom GPU kernels, and much more.

Joining DS means owning major parts of a fast-growing AI research organization, joining a collaborative, talent-dense team with decades of experience in probabilistic ML, accelerated computing, embedded systems, and signal theory, and growing your career in the areas that interest you. You’ll fit in if you want to come to work for the problem itself and don’t want to choose between technical rigor, business value, and real-world impact.

We work with high ownership and trust.

THE ROLE

DS collects radio data at a scale few organizations ever see — continuous, high-rate streams from sensors around the globe. We're hiring a Senior Data Engineer to design the platform that turns that firehose into an asset: a data lakehouse that serves researchers training models, production systems running inference, and agents retrieving context in real time.

You'll own the architecture from ingestion through storage, cataloging, vector search, and access, and you'll design it explicitly for AI/ML workloads, not just analytics.

WHAT YOU'LL DO

- Architect and build our data lakehouse: ingestion pipelines, open table formats, partitioning and compaction strategies, cataloging, and governance.

- Design and operate vector database infrastructure for embedding storage, similarity search, and retrieval at scale — choosing, tuning, and evolving the right systems for our workloads.

- Build large-scale data management systems: lifecycle and retention, lineage, versioning of datasets for reproducible training, quality monitoring, and cost management across petabyte-class storage.

- Design data platform capabilities that directly support ML use cases — feature and embedding pipelines, training-set assembly, evaluation datasets, and low-latency retrieval for agents.

- Establish schema and data-contract standards across teams, and build tooling that lets researchers and engineers self-serve.

- Own reliability and performance of data pipelines in production, including observability and failure handling.

- Mentor engineers and lead technical design across the data domain.

WHAT WE'RE LOOKING FOR

- 4–6+ years of software / data engineering experience, including several years designing and operating large-scale data platforms in production.

- Deep experience with lakehouse architectures and technologies — e.g., Apache Iceberg, Delta Lake, or Hudi; Parquet; Spark, Flink, Trino, DuckDB, or similar engines.

- Hands-on experience with vector databases and embedding retrieval (e.g., pgvector, Milvus, Qdrant, Weaviate, Pinecone, LanceDB, or FAISS-based systems), including indexing trade-offs and scaling.

- Strong experience with AWS or another major cloud provider — object storage, managed data services, compute, and cost optimization at scale.

- Fluency in Python and SQL; experience with orchestration tools (Airflow, Dagster, Prefect, Step Functions, etc.).

- Demonstrated experience designing data architectures specifically for ML / AI workloads: training data pipelines, feature stores, dataset versioning, or retrieval systems.

- Strong grasp of data modeling, consistency, and the trade-offs between batch and streaming.

NICE TO HAVE

- Experience with streaming ingestion at high volume (Kafka, Kinesis, Pulsar).

- Experience with time-series, geospatial, or signal / sensor data.

- Familiarity with data governance and security requirements in regulated environments.

- Experience with Rust, Go, or C++ for performance-critical data paths.

WHO THRIVES AT DISTRIBUTED SPECTRUM

- Fast learners over specific backgrounds – We care more about how quickly you can pick up new skills than where you’ve worked before.

- Intellectual honesty – The right answer matters more than being right. You challenge assumptions, test ideas, and pivot when needed.

- Adaptability – We’re organized, but sometimes things change quickly. You find a way to make it work and balance short-term deliverables with long-term goals.

- Ownership of outcomes – You optimize your own time, focus on what matters to deliver quickly, and cut out inefficiencies.

- Not building in a vacuum – You stay connected to the rest of our teams and our customers to make sure all the pieces fit together.

WHAT WE OFFER

- Above-market salary, equity, and benefits package.

- Early Series A Equity

- Excellent health, dental, and vision coverage

- 401(k) match - up to 4% of your salary

- Flexible PTO

- Daily office lunches in NYC

ITAR Requirements

To conform to U.S. Government technology export regulations, including the International Traffic in Arms Regulations (ITAR) you must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State. Learn more about the ITAR here https://www.pmddtc.state.gov/?id=ddtc_kb_article_page&sys_id=24d528fddbfc930044f9ff621f961987.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Distributed Spectrum's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Distributed Spectrum's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Distributed Spectrum's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.