Skip to content

Open nowPosted 161 days ago

Staff Scientist – Post-Training and Reinforcement Learning for AI for Science

argonne35 open roles

Pay
$94,486 – $147,399 a year
Where
Lemont, IL USA
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff Scientist – Post-Training and Reinforcement Learning for AI for Scienceargonne · Lemont, IL USA
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on argonne's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.0% of postings close within 7 days. Measured by our own scanner across the market. argonne postings stay open a median of 3 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 8.0%7 days
  4. 15.0%14 days
  5. 34.2%30 days
This job: posted 161 days ago

argonne median: 3 days open

The posting

The Argonne Leadership Computing Facility (ALCF) is seeking a Staff Scientist in Post-Training and Reinforcement Learning for AI for Science to help advance the next generation of foundation models and learning systems for scientific discovery.

This is an opportunity to work at the frontier of AI for science and the Department of Energy Genesis mission, where large-scale machine learning, scientific data, simulation, and leadership-class supercomputers come together to enable new modes of discovery across physics, materials science, chemistry, biology, climate, energy, and related fields. We are looking for a creative and collaborative scientist who is excited to develop, scale, and evaluate post-training methods, including reinforcement learning, preference optimization, adaptation, and alignment techniques, for scientific AI models and workflows.

The successful candidate will conduct research on methods that improve the usefulness, reliability, and scientific performance of large-scale AI models after pretraining, while also advancing the systems and software needed to run these methods efficiently on cutting-edge supercomputers and emerging AI platforms. This role offers the opportunity to contribute both fundamental advances in machine learning and high-impact scientific applications while working in a multidisciplinary environment with experts in AI, simulation, computer science, applied mathematics, and domain science.

You will join the AI group - a highly collaborative, multidisciplinary environment and work alongside experts in AI, simulation, computer science, applied mathematics, and domain science. This role offers the chance to contribute both foundational advances and real-world scientific outcomes, with opportunities to publish in leading journals and conferences, engage with national and international collaborators, and influence AI and HPC for scientific research.

In this role you will:

  • Conduct research and development aligned with Argonne’s strategic mission in computation, AI, and scientific discovery.
  • Develop, scale, and optimize post-training methods for scientific foundation models, including reinforcement learning, preference-based optimization, fine-tuning, alignment, and related approaches.
  • Advance techniques that improve the performance, controllability, reliability, and scientific utility of AI models for science applications.
  • Design and evaluate methods for applying reinforcement learning and post-training pipelines to large-scale scientific and data-intensive environments.
  • Develop and optimize workflows for training and post-training on leadership-class supercomputers and emerging AI-oriented architectures.
  • Partner with computational scientists, applied mathematicians, and domain researchers to apply foundation models and adaptive learning systems to challenging scientific problems with high impact.
  • Address algorithmic, systems, and data challenges associated with large-scale training and post-training, including performance, scalability, robustness, and usability.
  • Conduct original research in computational science and AI at scale, and communicate findings through publications, conference presentations, software, reports, and other research outputs.
  • Work closely with colleagues across national laboratories, universities, industry, and supercomputing centers on current and future systems for the AI for science mission.
  • Contribute to a team culture that values scientific excellence, collaboration, innovation, and inclusive professional growth.

This position qualifies as “Hybrid Remote Work - Mostly Onsite”: which applies to employees regularly scheduled for some onsite and some remote days, with employees typically working up to 40% of their time remotely.

Position Requirements

Required Qualifications:

  • RD2: Bachelor's degree and 5+ years of experience, or a Masters and 3+ years of experience, or a PhD, or equivalent
  • Education in computer science, applied mathematics, statistics, computational science, or a related field
  • Demonstrated advanced knowledge in one or more of the following areas: machine learning, reinforcement learning, large-scale model training, post-training, optimization, data mining, or statistics
  • Strong background in mathematical optimization, linear algebra, or numerical methods
  • Advanced knowledge of and significant programming experience in one or more languages such as Python, C, or C++
  • Significant experience with machine learning frameworks such as PyTorch or JAX
  • Experience with large-scale training, distributed learning systems, or post-training workflows
  • Experience with software development practices and techniques for computational science and machine learning systems
  • Ability to work effectively in interdisciplinary teams involving mathematicians, computer scientists, and application scientists
  • Effective written and verbal communication skills
  • Ability to model Argonne’s core values of impact, safety, respect, integrity, and teamwork

Preferred Qualifications:

  • Experience with reinforcement learning, policy optimization, bandits, preference learning, or related methods
  • Experience with post-training methods for large models, including supervised fine-tuning, reinforcement learning from feedback, direct preference optimization, reward modeling, or model adaptation
  • Experience with distributed training, large-scale optimization, and multi-node or multi-accelerator execution

Job Family

Research Development (RD)

Job Profile

Computer Science 2

Worker Type

Regular

Time Type

Full time

The expected hiring range for this position is $94,486.00 - $147,398.94.

Please note that the pay range information is a general guideline only. The pay offered to a selected candidate will be determined based on factors such as, but not limited to, the scope and responsibilities of the position, the qualifications of the selected candidate, business considerations, internal equity, and external market pay for comparable jobs. Additionally, comprehensive benefits are part of the total rewards package.

Click here to view Argonne employee benefits!

As an equal employment opportunity employer, and in accordance with our core values of impact, safety, respect, integrity and teamwork, Argonne National Laboratory is committed to a safe and welcoming workplace that fosters collaborative scientific discovery and innovation. Argonne encourages everyone to apply for employment. Argonne is committed to nondiscrimination and considers all qualified applicants for employment without regard to any characteristic protected by law.

Argonne employees, and certain guest researchers and contractors, are subject to particular restrictions related to participation in Foreign Government Sponsored or Affiliated Activities, as defined and detailed in United States Department of Energy Order 486.1A. You will be asked to disclose any such participation in the application phase for review by Argonne's Legal Department.

All Argonne offers of employment are contingent upon a background check that includes an assessment of criminal conviction history conducted on an individualized and case-by-case basis. Please be advised that Argonne positions require upon hire (or may require in the future) for the individual be to obtain a government access authorization that involves additional background check requirements. Failure to obtain or maintain such government access authorization could result in the withdrawal of a job offer or future termination of employment.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against argonne's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on argonne's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    argonne's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.