Skip to content

Open nowPosted 257 days ago

Senior Staff AI Engineer

SoFi55 open roles

Where
CA - San Francisco
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowSenior Staff AI EngineerSoFi · CA - San Francisco
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on SoFi's own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

8.0% of postings close within 7 days. Measured by our own scanner across the market. SoFi postings stay open a median of 28 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.5%3 days
  3. 8.0%7 days
  4. 14.9%14 days
  5. 34.1%30 days
This job: posted 257 days ago

SoFi median: 28 days open

The posting

Employee Applicant Privacy Notice

Who we are:

Shape a brighter financial future with us.

Together with our members, we’re changing the way people think about and interact with personal finance.

We’re a next-generation financial services company and national bank using innovative, mobile-first technology to help our millions of members reach their goals. The industry is going through an unprecedented transformation, and we’re at the forefront. We’re proud to come to work every day knowing that what we do has a direct impact on people’s lives, with our core values guiding us every step of the way. Join us to invest in yourself, your career, and the financial world.

The role:

SoFi’s Senior Staff AI Engineer is a hands-on AI engineering role in SoFi’s growing independent risk organization. This is a critical, senior role responsible for setting the technical direction, driving execution, and ensuring the successful delivery of our most complex, production-level AI initiatives. This role will be instrumental in conceptualizing, prototyping and implementing best-in-class AI-based solutions to meet risk management and compliance requirements.

This hands-on role will work closely with the Director of Risk Analytics, and will leverage your deep expertise to solve our hardest problems, mentor the next generation of engineers, and directly connect technical innovation to major business success. This is a crucial role for the independent risk function as we execute our mission to help more members get their money right.

What you’ll do:

  • Architecture and Strategy: Define the long-term technical architecture and strategy for our next-generation AI platform, particularly focusing on robust, scalable agentic frameworks and LLM deployment patterns.
  • Advanced LLM Orchestration: Architect and standardize the use of graph-based LLM orchestration, leveraging expert-level mastery of LangGraph to solve highly complex, multi-stage reasoning problems at scale.
  • Distributed Agent Memory & State: Develop robust, persistent infrastructure for agentic state management, ensuring that long-running agent workflows maintain context and reliability across distributed nodes and regional failovers
  • Deep Model Optimization: Pioneer and institutionalize advanced parameter-efficient fine-tuning (PEFT) and compression techniques to maximize model performance and minimize operational costs across the organization.
  • Model Serving Infrastructure: Support the development of a unified model serving platform designed to host internally fine-tuned and custom-trained models to ensure high-throughput, low-latency inference across diverse hardware footprints.
  • Operational Excellence: Define and enforce high standards for AI operationalization, requiring mastery in designing and deploying comprehensive AI observability solutions and advanced tracing/testing frameworks that guarantee production quality, compliance, and reliability.
  • Mentorship: Mentor senior and junior AI Engineers, elevating the overall engineering quality
  • Cross Functional Collaboration: Coordinate with cross-functional teams to distill specific requirements, project roadmaps, and ensure accurate and on-time project deliveries
  • AI Innovation: Stay up-to-date with the latest trends and advancements in GenAI, LLMs, and NLP, evaluating and experimenting with new techniques and tools to push the boundaries of AI innovation in the banking sector.

What you’ll need:

  • Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Machine Learning, or a related field. PhD is a plus.
  • 8+ years software development experience, with 3+ years of hands-on experience in developing and successfully deploying production-level AI applications that have been used by real customers or internal stakeholders.
  • Expert-level experience with LangGraph to model and orchestrate complex, stateful multi-step reasoning and control flow in LLM applications.
  • Expert-level proficiency in developing sophisticated agentic solutions, with a portfolio demonstrating advanced use of planning, memory management, tool integration, and control flow.
  • Deep understanding of Large Language Model (LLM) architectures, prompt engineering, retrieval-augmented generation (RAG), and advanced text generation techniques.
  • Proven experience implementing parameter-efficient fine-tuning (PEFT) techniques (e.g., LoRA) to customize and optimize pre-trained models for specific tasks with minimal computational overhead.
  • Deep expertise in building or extending inference engines (e.g., vLLM, NVIDIA Triton, or TGI) and managing the underlying Kubernetes/GPU orchestration for custom model deployments.
  • Deep experience designing and institutionalizing AI observability solutions (e.g., LangSmith, Arize, Deepchecks) and advanced tracing and testing methodologies for LLM and agentic systems.
  • Experience with cloud platforms (AWS, Azure, or GCP) and containerization technologies (Docker, Kubernetes).
  • Expert level Python is required.
  • React is strongly preferred.
  • Experience with large-scale data handling, including unstructured and structured data pipelines, with a strong preference for Snowflake and DynamoDB.
  • Experience developing and integrating AI-powered APIs and microservices architecture into banking applications.
  • Experience with vector databases and retrieval-augmented generation (RAG) techniques using systems like Elasticsearch, Pinecone, or FAISS for enhancing LLM performance.
  • Exceptional ability to communicate complex technical concepts, drive consensus among senior technical leaders, and influence organizational AI strategy.
  • Strong analytical and problem-solving skills with attention to detail and an ability to work with complex, large-scale systems.
  • Strong collaboration skills, with experience working in agile, cross-functional teams.

Nice to have:

  • Familiarity with regulatory frameworks and ethical considerations in AI within the banking industry (e.g., GDPR, data privacy, model explainability).
  • Experience in banking or financial services use cases such as conversational AI for customer service, intelligent document processing for loan applications, fraud detection, or risk analysis.

Compensation and Benefits

The base pay range for this role is listed below. Final base pay offer will be determined based on individual factors such as the candidate’s experience, skills, and location.

To view all of our comprehensive and competitive benefits, visit our Benefits at SoFi page!

SoFi provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion (including religious dress and grooming practices), sex (including pregnancy, childbirth and related medical conditions, breastfeeding, and conditions related to breastfeeding), gender, gender identity, gender expression, national origin, ancestry, age (40 or over), physical or medical disability, medical condition, marital status, registered domestic partner status, sexual orientation, genetic information, military and/or veteran status, or any other basis prohibited by applicable state or federal law.

The Company hires the best qualified candidate for the job, without regard to protected characteristics.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

New York applicants: Notice of Employee Rights

SoFi is committed to an inclusive culture. As part of this commitment, SoFi offers reasonable accommodations to candidates with physical or mental disabilities. If you need accommodations to participate in the job application or interview process, please let your recruiter know or email [email protected].

We are unable to accommodate remote work from Hawaii, Alaska or Puerto Rico at this time.

Internal Employees

If you are a current employee, do not apply here - please navigate to our Internal Job Board in Greenhouse to apply to our open roles.

From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against SoFi's own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on SoFi's form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    SoFi's answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.