Synced from Ashby · Feb 5

Member of Technical Staff, Reliability

SieveSan FranciscoPosted Feb 5, 2026
Site Reliability EngineerOn-siteSenior
Apply now - freeSave & get alerts

Mirrored from Sieve's own Ashby careers system · refreshed hourly

9
Other open Sieve roles
Feb 5
Posted
Ashby
Applicant system
Job description

About Us

Sieve is a multi-modal lab curating the world's highest-quality training datasets — spanning video, audio, images, text, and 3D. We combine exabyte-scale data infrastructure and novel multimodal understanding techniques that push the frontier of foundation models. Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling medium powering creativity, communication, gaming, AR/VR, and robotics. Sieve exists to solve the biggest bottleneck in the growth of these applications: high-quality training data.


We partner with top AI labs and did $XXM last quarter alone, as a team of ~30 people. We also raised our Series A from Tier 1 firms such as Matrix Partners, Swift Ventures, Y Combinator, and AI Grant.

 

Why Now

Sieve is one of the most capital-efficient teams in AI — roughly 30 people serving the world's leading AI labs across every major data modality. You'll join early, own problems end-to-end, and watch your work ship directly into the models defining the frontier.


About the Role

We process petabytes of video across thousands of nodes and multiple cloud environments, and as we scale, reliability, observability, and security become existential. We're hiring our first engineer fully dedicated to Sieve's infrastructure foundation — a high-ownership role working directly with our CTO and founding engineers to build the core tooling that powers all of engineering. You'll design and validate the infrastructure behind PB-scale workloads, own incident response, harden systems against failure, and build the monitoring, security, and CI/CD tooling the whole team relies on.

This role is ideal for someone who thinks deeply about reliability, throughput, observability, and security — the kind of engineer who anticipates failure modes, eliminates operational risk, and designs systems that don't break. If something goes down, you take it personally, and you thrive in that level of responsibility.

Requirements

  • 3+ years building internal infrastructure at scale

  • Experience on-call for Sev 0 / Sev 1 production incidents (L3 preferred)

  • Strong cloud experience (GCP, AWS, Oracle, Cloudflare, etc.)

  • Deep Infrastructure-as-Code experience (Terraform preferred)

  • Familiarity with Argo, Helm, Kustomize, or similar deployment tools

  • Experience operating observability systems (Prometheus, OTel, VictoriaMetrics)

  • Backend fundamentals in Python, Go, Rust, or C++

  • Strong networking + security intuition, including SSO implementation

  • In-person at our SF HQ

  • Bonus: Experience building lightweight internal tooling (APIs, dashboards, Svelte)

  • Bonus: Familiarity with object storage systems ("buckets")

  • Bonus: Active GitHub or portfolio projects

Benefits

  • 401k + Full Health Insurance

  • Breakfast, Lunch, and Dinner covered and your choice of snacks

  • Ubers covered home

*all roles at Sieve require you to be onsite in San Francisco 5 days per week

View original posting on Ashby

Land this one early - before the req fills.

LandEarly tailors your resume and screening answers to each posting, submits within minutes of a role going live, and tracks every application in one place.

Free to start · No credit card · Cancel anytime

Keep exploring

What this role pays, where else it is open, and how to write the application.

Member of Technical Staff, Reliability
Sieve · San Francisco
Apply