Sieve

Sieve

Member of Technical Staff, Applied Research

San Francisco · Staff+

Sponsorship not specifiedDetected 452 days ago
PythonSwiftGitMachine LearningPyTorchNLPComputer VisionAR/VRResearchCommunication

About the role

  • Often this involves working on ambiguous research problems and finding clever techniques to solve them.
  • You will be working in the computer vision, audio processing, and text processing domains.
  • You're likely a good fit if you're comfortable working with models + APIs and squeezing every drop of performance out of them through clever pre/post-processing, parallelism, pipelining, inference optimization, and occasionally fine-tuning.

Responsibilities

  • We partner with top AI labs and did $XXM last quarter alone, as a team of ~30 people.
  • We also raised our Series A from Tier 1 firms such as Matrix Partners https://matrix.vc/, Swift Ventures https://www.swift.vc/, Y Combinator https://www.ycombinator.com/, and AI Grant https://aigrant.com/.
  • You'll join early, own problems end-to-end, and watch your work ship directly into the models defining the frontier.
  • As an applied research engineer at Sieve https://www.sievedata.com/, you'll build high performance building blocks and large scale pipelines to understand video with high precision at internet scale.

Requirements

  • Strong Python developer with hands-on experience in PyTorch or similar ML frameworks

Nice to have

  • Active contributor to open source projects
  • Experience as an early hire at a startup

Benefits

  • 401k + Full Health Insurance

Company info

  • Sieve is a multi-modal lab curating the world's highest-quality training datasets - spanning video, audio, images, text, and 3D. We combine exabyte-scale data infrastructure and novel multimodal understanding techniques that push the frontier of foundation models. Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling medium powering creativity, communication, gaming, AR/VR, and robotics. Sieve exists to solve the biggest bottleneck in the growth of these applications: high-quality training data.
  • We partner with top AI labs and did $XXM last quarter alone, as a team of ~30 people. We also raised our Series A from Tier 1 firms such as Matrix Partners https://matrix.vc/, Swift Ventures https://www.swift.vc/, Y Combinator https://www.ycombinator.com/, and AI Grant https://aigrant.com/.
  • Sieve is one of the most capital-efficient teams in AI - roughly 30 people serving the world's leading AI labs across every major data modality. You'll join early, own problems end-to-end, and watch your work ship directly into the models defining the frontier.
  • Sieve is a multi-modal lab curating the world's highest-quality training datasets - spanning video, audio, images, text, and 3D.
  • We combine exabyte-scale data infrastructure and novel multimodal understanding techniques that push the frontier of foundation models.
  • Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling medium powering creativity, communication, gaming, AR/VR, and robotics.
  • Sieve exists to solve the biggest bottleneck in the growth of these applications: high-quality training data.
  • Sieve is one of the most capital-efficient teams in AI - roughly 30 people serving the world's leading AI labs across every major data modality.

This listing is sourced directly from Sieve's careers page and normalized into a canonical job model.