Skip to content

Open nowPosted 66 days ago

Senior Data Engineer

Workable (global search)108,016 open roles

Where
Washington, DC, United States
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior Data EngineerWorkable (global search) · Washington, DC, United States
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Workable (global search)'s own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market. Workable (global search) postings stay open a median of 7 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 7.9%7 days
  4. 14.9%14 days
  5. 34.0%30 days
This job: posted 66 days ago

Workable (global search) median: 7 days open

The posting

Senior Data Engineer

Location: Herndon, VA (Remote Work)

Must have an Public Trust Clearance

KEY RESPONSIBILITIES

• Provide authoritative expertise on data engineering methods and best practices, including code first development approaches and modern pipeline design patterns.

• Design, implement, and maintain the data architecture that supports products and end users, with all assets managed under source control.

• Design, implement, and maintain ELT and ETL pipelines for efficient processing of source data in Azure Synapse and Azure Machine Learning, using both SDK V1 and SDK V2.

• Migrate source data identified by SBA OIG into Azure Data Lake Storage.

• Normalize entity attributes such as addresses, phone numbers, and other common fields.

• Review, maintain, and improve existing architecture and pipelines, including periodic audits addressing bottlenecks, deprecated dependencies, and architecture drift.

• Establish quality controls across all pipelines and introduce error handling, logging mechanisms, and validation checks.

• Incorporate source control across all pipelines and analytics codebases so code can evolve iteratively without destabilizing the architecture.

• Optimize ingestion, processing, and storage across a wide variety of datasets and data types, including modern columnar formats such as Parquet.

• Develop self service capabilities that let SBA OIG analysts query and export data for investigations and audits.

• Author robust standard operating procedures governing the authoring, development, validation, publishing, execution, and monitoring of all data pipelines and assets in the Azure environment.

• Produce detailed documentation of the data architecture, including data dictionaries, entity relationship diagrams, and pipeline process maps.

• Maintain and expand the environment with additional datasets and services on request, following a defined intake and testing process before production deployment.

• Stay current with emerging AI tooling relevant to data engineering and contribute to exploratory work evaluating automation and language model assisted capabilities.

Requirements

Education

Bachelor's degree in data engineering, computer science, data science, machine learning, mathematics, or a related field. Alternatively, five years of applied work experience in any of the same fields.

  • 5 years - Maintaining SQL databases and conducting advanced operations in SQL and T-SQL.
  • 5 years - Designing, implementing, and maintaining ELT and ETL processes in cloud based data analytics environments.
  • 3 years - Working in Azure Synapse and Azure Machine Learning with the modern data stack. Certifications preferred, DP-203 or equivalent.
  • 3 years -Manipulating data in Python. Pandas is required. PySpark and Polars preferred. Experience developing reusable, modular code preferred.

PREFERRED QUALIFICATIONS

  • DP-203, Microsoft Certified Azure Data Engineer Associate, or an equivalent current certification.
  • Implementing pipelines and infrastructure using code first approaches: Python SDK, CLI, REST APIs, or infrastructure as code tooling such as Terraform or Bicep.
  • Implementing source control and continuous integration and delivery workflows for data assets.
  • Demonstrated familiarity with AI coding assistants and large language model integration patterns.
  • PySpark or Polars at production scale.
  • Entity resolution and attribute normalization across records with inconsistent addresses, names, and identifiers.
  • Building self service analytic access for non engineering users.

Benefits

We are proud to offer competitive compensation and benefits packages to include

  • Medical
  • Dental
  • Vision
  • Basic Life
  • Health Saving Account
  • 401K matching
  • Three weeks of PTO/Sick
  • 11 Paid Holidays
  • Pre-Approved Online Training
From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Workable (global search)'s own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Workable (global search)'s form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Workable (global search)'s answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.