Skip to content

Open nowPosted 4 days ago

Staff Engineer (Core & MLOps) - Remote

Workable (global search)108,016 open roles

Where
Lisbon, Portugal
Work mode
Remote
Get the CV for this job

From $25 per CV, paid once. No subscription.

Your applicationOpen nowStaff Engineer (Core & MLOps) - RemoteWorkable (global search) · Lisbon, Portugal
  1. YouYes, apply to this one.

  2. CV RocketCV written for this posting.

  3. 25 readersRecruiter, hiring manager, skeptic. Round after round.

  4. CV RocketApplied on Workable (global search)'s own form.

The reply lands in your private mailbox

3×more interviews than doing it yourself with ChatGPT.

The clock on this job

Early applications get read.

7.9% of postings close within 7 days. Measured by our own scanner across the market. Workable (global search) postings stay open a median of 7 days.

Share of postings closed within
  1. 1.6%1 day
  2. 3.6%3 days
  3. 7.9%7 days
  4. 14.9%14 days
  5. 34.0%30 days
This job: posted 4 days ago

Workable (global search) median: 7 days open

The posting

Why Zyte?

At Zyte, we don’t just collect web data—we solve complex web data challenges at scale. We are a globally distributed team that is bold, curious, and dedicated to building innovative ways to deliver clean, reliable web data from across the internet.

  • Data is our passion: We unlock access to open web data, empowering organizations—from startups to global enterprises—to power competitive intelligence, analytics, and AI pipelines.
  • Remote-first culture: Work from anywhere in the world. With over 260 Zytans across 33 countries, our strength comes from diverse backgrounds, perspectives, and skills.
  • Engineering excellence: We love diving deep into complex code bases, evaluating emerging technologies, and pushing technical boundaries. If you thrive on creative problem-solving and reliably delivering production-grade solutions, you will fit right in.
  • Established system design framework: We systematically analyze functional and product requirements alongside architectural trade-offs to make informed, sustainable design decisions.
  • Real-world distributed systems challenges: Tackle rare engineering problems in custom networking, high-throughput streaming, and large-scale orchestration under real-world operational constraints.
  • Strong product ownership: Work on a profitable, market-leading platform with dedicated time and resources allocated for architectural research and strategic feature engineering.

Why This Role?

Zyte API powers a suite of Java and Python microservices running across multiple cloud providers and data centers. Every engineering team—spanning Core API, Browser, Edge, Antiban, AI, and Agentic products—relies on foundational infrastructure managed by our team: Kubernetes clusters, high-throughput Kafka pipelines, billing engines processing per-request metrics, and robust inter-service contracts.

We are evolving Zyte API into an automated software factory: a platform where developers and AI agents seamlessly create, validate, deploy, and maintain data extractors and workflows through governed interfaces. Two foundational layers drive this evolution. The control plane (service/schema registries, health-aware routing, automated canary releases) is actively under deployment. The context plane, which synthesizes observational signals to guide automated repair and optimizations, represents our next major design effort. As a Staff Engineer, you will own both architectures, setting the engineering standard for how services at Zyte are designed, built, and operated. You will be the technical lead in the Core & MLOps squad, collaborating closely with the Team Lead, Senior Engineers, DevOps, and QA, while directly influencing the release velocity and architecture of five neighboring squads.

Requirements

What You’ll Do

  • Architect the control and context planes. Advance the control plane (service registry, schema registry, SLO enforcement, CLI tooling) into a robust production substrate. Design and build the context plane to aggregate operational signals—such as domain extraction histories, IP reputation scores, and BigQuery cost/performance telemetry—creating automated feedback loops between intent, execution, and self-healing maintenance.
  • Own the service chassis and golden path. Maintain and refine multi-language client libraries (Java and Python), standardized workload specifications, Helm charts, and deployment pipelines, driving seamless adoption across all product teams.
  • Define inter-service contracts. Establish gRPC and Protocol Buffer definitions, API gateway transcoding, versioning policies, and schema evolution rules to enable autonomous, safe deployments across teams.
  • Operate the platform substrate. Partner with infrastructure engineers to run Kubernetes (across OCI, Hetzner, Servers.com, and GCP), Terraform, HAProxy/Nginx ingresses, Confluent Kafka, real-time event-billing pipelines, Valkey, and ongoing database modernizations (MySQL to PostgreSQL).
  • Drive architectural strategy via RFDs. Lead Requests for Discussion (RFDs) on critical platform initiatives, including durable workflow orchestration (Temporal/DBOS), per-request gateway orchestration, multi-cluster routing, and automated failover.
  • Institutionalize reliability engineering. Establish clear SLOs and error budgets for core capabilities, implementing health-aware traffic isolation and automated weighted canary deployments to eliminate manual release bottlenecks.
  • Share production ownership. Participate in the shared infrastructure on-call rotation, lead incident post-mortems, and translate operational lessons into platform improvements.
  • Elevate engineering standards. Mentor engineers across squads on distributed systems design, review system proposals, and establish intuitive best practices that make building reliable software effortless.

Who You Are

  • 10+ years of experience building scalable distributed backend systems, with a proven track record of authoring internal platforms or core libraries widely adopted by engineering teams.
  • Expertise in Java (using reactive frameworks like Vert.x or Netty) alongside strong proficiency in Python.
  • Deep gRPC & Protobuf experience, with demonstrated success managing schema evolution and backward compatibility in mission-critical environments.
  • Hands-on infrastructure expertise with production Kubernetes at scale, Terraform, and event streaming with Kafka.
  • Feedback loop designer: Experience designing automated telemetry pipelines, materialized views, or feature stores that dynamically adapt system behavior based on production data.
  • Reliability mindset: Proven success defining SLOs/SLIs, analyzing blast radius, and designing fault-tolerant systems built around rigorous service contracts.
  • Technical leadership: Exceptional technical writing skills with the ability to articulate complex designs clearly and drive technical alignment across multi-functional teams.
  • Async communication: Strong interpersonal and written communication skills tailored for a globally distributed, remote-first environment.
  • Curious problem solver: Passionate about continuous learning and evaluating novel tools, architectures, and techniques.

Our Tech Stack

  • Languages: Java 21, Python.
  • Frameworks & Tools: Vert.x, Netty, gRPC and protobuf, Kubernetes, Helm, Terraform, CircleCI.
  • Infrastructure: Multi-cloud (OCI, GCP) and data centres (Hetzner, Servers.com), blue/green deployments everywhere.
  • Data & Messaging: Confluent Kafka, BigQuery, Valkey, MySQL (Postgres next), HBase, Google Pub/Sub.
  • Monitoring & Observability: Prometheus, Grafana, Loki, OpenTelemetry.

Bonus Points

  • Durable execution engines: Hands-on experience with Temporal, DBOS, or similar workflow platforms.
  • MLOps: Experience in model serving, performance monitoring, and drift detection in production environments.
  • Zero-trust networking & service meshes: Familiarity with SPIRE, mTLS, Cilium, Istio, or Envoy.
  • Developer tooling: Track record of building CLIs, SDKs, or project generators adopted across an engineering organization.
  • Web scraping & crawling experience: Familiarity with the technical challenges of large-scale web data extraction.
  • Open source contributions: History of contributing to or maintaining open-source projects in distributed systems or data extraction.

Why Join Us?

  • Impact at scale: Architect core infrastructure powering web-scale data pipelines for top-tier global enterprises.
  • Autonomy and flexibility: Enjoy a genuine remote-first culture with flexible working hours and high organizational trust.
  • Continuous innovation: Solve emerging, complex technical challenges as AI and web data extraction rapidly evolve.
  • Global community: Collaborate with a passionate, diverse team of distributed systems engineers and data specialists across the world.

If you are a Staff Engineer excited by the prospect of building core platform substrates that empower multiple engineering teams, we would love to hear from you. Join us in shaping the control plane behind web data extraction at Zyte.

Benefits

By joining the Zyte team, you will:

  • Become part of a self-motivated, progressive, multi-cultural team
  • Have the freedom and flexibility to work from where you do your best work
  • Attend conferences and meet with team members from across the globe
  • Work with cutting-edge open source technologies and tools
From $25, paid onceGet the CV for this job

What happens when you press

One press. We do the rest.

  1. A CV for this posting

    Written against Workable (global search)'s own wording, from every piece of relevant proof in your profile.

  2. 25 readers review it

    Recruiter, hiring manager, skeptic and more read every draft, round after round. You get the best round.

    The review screen in CV Rocket: how each CV was read, round by round.
  3. We apply on Workable (global search)'s form

    Our application engine gets through the hardest forms there are. Where a question needs you, AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

    An application in CV Rocket: every answer filled in on the employer's form.
  4. Every reply, sorted

    Workable (global search)'s answer lands in your private mailbox, and we classify it on arrival: interview, question, rejection.

    The CV Rocket inbox: each employer reply classified as an interview, an action or a rejection.
  5. Reply with AI

    AI helps you write the email, checks it and sends it. We show you whether the recruiter read it.

  6. The interview in your calendar

    Full integration with your calendar. The invitation goes straight in.

    An interview invitation in the CV Rocket inbox, added to the candidate's calendar.
Get the CV for this job

From $25 per CV, paid once. No subscription.

Why it works

3×

more interviews than doing it yourself with ChatGPT.

ChatGPT writes a CV and never learns what happened to it. We see every reply. For each CV we know:

  • How it was written, and how the review scored it
  • When we applied, and how long after the posting went up
  • Which posting, which company, which city
  • Who got the interview, and who heard nothing

That is how we know which CVs get called.

Get the CV for this job

From $25 per CV, paid once. No subscription.

The numbers game

More applications. More interviews.

Every application goes out with its own CV, written for that posting and paid once. Send enough of them and the law of large numbers finds you the job.

By hand5–10
With CV Rocket100
applications a day

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

Before you press

Straight answers

Get the CV for this job

From $25 per CV, paid once. No subscription.

What if my background isn't good enough?

We make the most of the background you have. The CV uses every piece of relevant proof your profile holds, and one of the 25 readers reads your whole profile and flags what the CV left out.

Do you really apply for me?

Yes, on the employer's own form, the hardest ones included. Where a question needs you, you answer it right there and AI suggests the best answer. Don't want us applying from our IP addresses? Use our Chrome extension: we apply straight from your own browser.

Is it a subscription?

No. You pay once per CV, from $25. Every application goes out with its own CV, written for that posting.

One job. One CV.
Paid once.

Pick the posting you want. We write for it, apply for you and catch the reply.

Get the CV for this job

From $25 per CV, paid once. No subscription.