Skip to content

Open nowPosted 2 days ago

Staff Data Engineer

bdx150 open roles

Pay
$161,400 – $258,200 a year
Where
USA CA - Irvine Laguna Canyon
Work mode
On site
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff Data Engineerbdx · USA CA - Irvine Laguna Canyon
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on bdx's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market. bdx postings stay open a median of 1 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.4%3 days
  3. 7.9%7 days
  4. 14.2%14 days
  5. 34.1%30 days
This job: posted 2 days ago

bdx median: 1 days open

The posting





We are the people who give possibilities purpose

BD is one of the largest global medical technology companies in the world. Advancing the world of health™ is our Purpose, and it’s no small feat. It takes the imagination and passion of all of us—from design and engineering to the manufacturing and marketing of our billions of MedTech products per year—to look at the impossible and find transformative solutions that turn dreams into possibilities.





Job Description

Summary:

As a Staff Data Engineer, you will lead the design and evolution of the data platform that powers BD Connected Care's AI initiatives. You will architect and build scalable data pipelines, models, and governance solutions that deliver trusted, high-quality data to Machine Learning engineers and other consumers across the organization.

This is a hands-on technical leadership role that combines software development, data architecture, and platform strategy. You will establish engineering standards, drive data quality and governance, and make key architectural decisions that shape the long-term direction of the AI data platform. Working cross-functionally with engineering, product, and clinical teams, you will deliver scalable data solutions that advance BD's mission of improving patient outcomes through innovation and data-driven healthcare.

The ideal candidate can communicate complex data design concepts in a clear, concise manner and confidently explain trade-offs, risks, and limitations to both technical and non-technical audiences, including executive, clinical, and engineering stakeholders.

Key Responsibilities:

  • Architecture and Technical Direction: Own the data architecture for the team's AI workflows, including the warehouse, lake, and serving layers. Set the technical direction and standards that engineers across the team build against, make long-term AI data platform decisions , and design for source diversity so a new device, EMR, or supply chain source can be onboarded without redefining the model.
  • Ingestion and Integration: Propose, design, and implement ingestion from BD product, operational, and clinical-adjacent sources into the analytics and ML data layers, spanning device transactions, inventory, ordering, purchasing, and supply signals, in batch, streaming, and change-data-capture modes.
  • Data Modeling: Design relational and analytical schemas in the cloud data warehouse and the S3-based lake for forecasting, analytics, and feature stores. Author and tune complex SQL, build dimensional and event-based models, manage incremental loads, slowly changing dimensions, and late-arriving data, and deliver analytics-ready models for downstream consumers.
  • Reconciled Histories: Build continuous, reconciled records that align events across systems with different refresh cadences and timestamps, covering procurement, storage, distribution, dispense, administration, waste, and return. Capture the observation time of each value alongside the value itself.
  • Data Quality and Contracts: Own data quality as an outcome. Negotiate and enforce data contracts with producing teams, establish automated validation across completeness, accuracy, consistency, and timeliness, apply confidence scoring to feeds that business and clinical claims depend on, and detect missing, stale, or implausible records upstream of the models.
  • Reliability and Operations: Own the operational health of the data platform. Define SLAs and runbooks, instrument monitoring and alerting, and lead incident response and post-incident review.
  • Privacy, Compliance, and Lineage: Implement access controls, masking, de-identification, and PHI handling in line with HIPAA and BD policy, partnering with information security, legal, and regulatory on row and column-level security. Maintain lineage, cataloging, and audit trails sufficient for clinical, regulatory, and customer security review, and support change-control and validation requirements for regulated products.
  • Technical Leadership: Mentor data engineers and review designs and code. Partner with ML engineers on feature engineering, dataset versioning, and reproducibility, translating operational and clinical context into the features and training datasets models need, and document architecture, schemas, data contracts, and operational runbooks in Confluence.

Required Qualifications:

  • Bachelor's in Computer Science, Data Engineering, or related field.
  • 7+ years of data engineering experience with increasing technical ownership, including end-to-end ownership of production pipelines and data models.
  • Demonstrated architectural ownership beyond a single team: setting standards and driving technical decisions that other engineers build on.
  • Expert SQL: query authoring, optimization, partitioning, distribution keys, window functions.
  • Strong hands-on experience with a cloud data warehouse at production scale, including schema design, performance tuning, workload management, and lake integration. Amazon Redshift (cluster or Serverless, with S3 via Spectrum and COPY/UNLOAD) is our current stack.
  • Production experience with AWS data services: S3, Glue, Lambda, Step Functions, Athena, and streaming or CDC via Kinesis, MSK, or DMS.
  • Python proficiency for pipeline code, with strong software-engineering habits: tests, code review, modular design, CI.
  • Solid grasp of data modeling (Kimball/Inmon, event modeling), governance, lineage, and data quality practice.
  • Hands-on feature engineering experience for ML: building, validating, and maintaining feature and training datasets in partnership with ML engineers, including point-in-time correctness and avoiding leakage.
  • Ability to learn a business domain deeply enough to translate operational and clinical processes into the right data inputs for modeling, including which signals matter, how they are generated, and where they are unreliable.
  • Demonstrated ownership of data quality as an outcome, including validation frameworks, anomaly detection on production feeds, and agreements with upstream producers.
  • Practical knowledge of HIPAA and PHI handling, including access control, masking, and de-identification in production environments.
  • Operational maturity: SLAs, monitoring, on-call, incident response, and post-incident improvement.
  • Working in Agile/Scrum and documenting in Confluence.

Preferred Qualifications:

  • Master's in Computer Science, Data Engineering, or related field.
  • Healthcare or medical-device data experience (HL7/FHIR, device telemetry, operational and clinical data, ERP/SCM feeds).
  • Experience in regulated environments with formal change control, computer system validation, audit trail, and data integrity requirements (21 CFR Part 11, GxP, or equivalent).
  • Experience with dbt for transformations, and with semantic or metrics layers.
  • Experience implementing lakehouse patterns on S3 (Iceberg, Delta, or Hudi).
  • Experience supporting time-series and multi-series ML workloads, including temporal feature structures and dataset versioning for reproducibility.
  • Familiarity with feature stores or vector stores (Feast, Tecton) for ML use cases.
  • Experience building data standardization and entity resolution pipelines using third-party ontologies, NLP, or LLM-assisted mapping (Bedrock or similar).
  • Experience with agentic AI pipelines and services, and with building AI-powered workflows and evals.
  • Professional use of AI-assisted development tools (Claude Code, Copilot, Cursor) with strong judgment around correctness, security, licensing, and review discipline.
  • AWS Data Analytics Specialty certification.

Team Culture:

We are building a high-ownership, mission-driven team, energized by the opportunity to solve hard problems that make a difference in patients' lives.

  • High ownership - You take full responsibility for outcomes. You identify problems early and drive them to resolution, taking initiative rather than waiting for direction.
  • Entrepreneurial spirit - You think and act like an owner: resourceful, creative, willing to challenge assumptions, and consistently focused on impact. You find practical paths forward through difficult problems.
  • Mission-driven work ethic - You are motivated by the work we are doing to advance the world of health and improve patient outcomes, and you bring genuine energy and commitment to that mission. You are driven by a real sense of purpose in your work.
  • Creative problem-solving - You are comfortable working through ambiguity. You dig in, experiment, iterate, and deliver, and you are energized by the challenge of figuring things out and finding the right solution.
  • AI-native mindset - You actively use the latest AI tools to accelerate your own work across coding, research, design, and documentation, and you are eager to keep advancing what is possible with AI in your craft.

Why Join Us?

To find purpose in the possibilities, we need people who can see the bigger picture, who understand the human story that underpins everything we do. We welcome people with the imagination and drive to help us reinvent the future of healthcare. At BD, you’ll discover a culture in which you can learn, grow and thrive.

We believe that when people connect in person, we learn faster, collaborate more deeply, and build a stronger culture. Join us and enjoy a culture where face-to-face collaboration supports your learning, your progress, and your success.

To learn more about BD visit https://bd.com/careers.

Becton, Dickinson, and Company is an Equal Opportunity Employer. We evaluate applicants without regard to race, color, religion, age, sex, creed, national origin, ancestry, citizenship status, marital or domestic or civil union status, familial status, affectional or sexual orientation, gender identity or expression, genetics, disability, military eligibility or veteran status, and other legally protected characteristics.

Required Skills

Optional Skills

.

Primary Work Location

USA CA - Irvine Laguna Canyon

Additional Locations

Work Shift





At BD, we reward, support and develop our associates through our comprehensive Total Rewards program. We are committed to attracting and retaining high quality talent by providing reward and recognition opportunities that promote a performance-based culture, as well as a competitive package of compensation and benefits programs. You can learn more on our career site under "Our Commitment to You."

Our salary or hourly rate ranges reward associates fairly and competitively. We regularly review these ranges and factors, such as location, contribute to the range displayed.

Our pay is based on the role and the necessary skills and education to perform it successfully. The salary or hourly rate offered is determined by the role's specific requirements, including any applicable step rate pay system at the work location. Salary or hourly pay ranges are influenced by labor laws and Collective Bargaining Agreement (CBA) requirements applicable to the work location which may also affect the workplace arrangement of the role.

Salary Range Information

$161,400.00 - $258,200.00 USD Annual

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against bdx's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on bdx's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    bdx's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.