← All roles
sieve logosieveFinance and Accounting

Member of Technical Staff, Reliability

GoRustPythonSwiftTerraformSF · Staff · Seed

About Us

Sieve is a multi-modal lab curating the world's highest-quality training datasets — spanning video, audio, images, text, and 3D. We combine exabyte-scale data infrastructure and novel multimodal understanding techniques that push the frontier of foundation models. Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling medium powering creativity, communication, gaming, AR/VR, and robotics. Sieve exists to solve the biggest bottleneck in the growth of these applications: high-quality training data.


We partner with top AI labs and did $XXM last quarter alone, as a team of ~30 people. We also raised our Series A from Tier 1 firms such as Matrix Partners, Swift Ventures, Y Combinator, and AI Grant.

Why Now

Sieve is one of the most capital-efficient teams in AI — roughly 30 people serving the world's leading AI labs across every major data modality. You'll join early, own problems end-to-end, and watch your work ship directly into the models defining the frontier.


About the Role

We process petabytes of video across thousands of nodes and multiple cloud environments, and as we scale, reliability, observability, and security become existential. We're hiring our first engineer fully dedicated to Sieve's infrastructure foundation — a high-ownership role working directly with our CTO and founding engineers to build the core tooling that powers all of engineering. You'll design and validate the infrastructure behind PB-scale workloads, own incident response, harden systems against failure, and build the monitoring, security, and CI/CD tooling the whole team relies on.

This role is ideal for someone who thinks deeply about reliability, throughput, observability, and security — the kind of engineer who anticipates failure modes, eliminates operational risk, and designs systems that don't break. If something goes down, you take it personally, and you thrive in that level of responsibility.

Requirements

  • 3+ years building internal infrastructure at scale

  • Experience on-call for Sev 0 / Sev 1 production incidents (L3 preferred)

  • Strong cloud experience (GCP, AWS, Oracle, Cloudflare, etc.)

  • Deep Infrastructure-as-Code experience (Terraform preferred)

  • Familiarity with Argo, Helm, Kustomize, or similar deployment tools

  • Experience operating observability systems (Prometheus, OTel, VictoriaMetrics)

  • Backend fundamentals in Python, Go, Rust, or C++

  • Strong networking + security intuition, including SSO implementation

  • In-person at our SF HQ

  • Bonus: Experience building lightweight internal tooling (APIs, dashboards, Svelte)

  • Bonus: Familiarity with object storage systems ("buckets")

  • Bonus: Active GitHub or portfolio projects

Benefits

  • 401k + Full Health Insurance

  • Breakfast, Lunch, and Dinner covered and your choice of snacks

  • Ubers covered home

*all roles at Sieve require you to be onsite in San Francisco 5 days per week

AI

Check your CV against this role

Drop your CV. You get a 0-100 fit score against the actual job description, plus the read a senior engineering lead would write. Private to you.

Your CV joins the pool too, so roles that fit can find you. No spam, and nothing reaches a company without your go-ahead.

Score this once, or every future role

Start the candidate journey and every new role on the board gets scored against you.

Five minutes. Tell us what you’re after, drop your CV once, pick how we should reach out. You get a candid read back and you only hear from us when a role fits.

More at sieve