Base Salary
$150k - $350k/yr
Responsibilities
- Work directly with customers and internal teams to determine required data and translate ambiguous requirements into technical systems
- Build and ship production systems that find, generate, filter, transform, evaluate, and package high-quality video datasets at scale
- Develop custom algorithms, models, workflows, and large-scale data pipelines
- Work across computer vision, audio processing, text processing, metadata analysis, model adaptation, and quality evaluation
- Improve system performance through preprocessing, post-processing, parallelism, inference optimization, fine-tuning, and evaluation loops
- Deliver reliable end-to-end production outcomes from research prototypes
Requirements
- Comfort working directly with customers or external teams to translate ambiguous needs into concrete technical systems
- Strong Python development skills and hands-on experience with PyTorch or similar machine learning frameworks
- Experience building custom algorithms, model workflows, or large-scale data pipelines
- Strong intuition for dataset quality, filtering, labeling, evaluation, and edge cases
- Ability to decompose customer goals into models, heuristics, infrastructure, and quality-assurance steps
- Ability to write clean, maintainable code and move quickly without creating brittle systems
- Passion for video, media technologies, and frontier AI applications
- Motivation to deliver end-to-end outcomes rather than only training models or writing research code
- Experience with large-scale video, audio, or multimodal data processing is a bonus
- Active open-source contribution is a bonus
- Experience as an early startup hire is a bonus
Benefits
- 401k and full health insurance
- Breakfast, lunch, dinner, and snacks provided
- Ubers home covered
- In-person work at the San Francisco headquarters
Categories
About Sieve
Sieve builds the data and environments frontier AI labs use to train the next generation of multimodal systems. AI is moving beyond chatbots into video, audio, images, software, robotics, and interactive worlds. The next generation of models will need to understand how the world looks, sounds, moves, responds, and changes over time. Progress is bottlenecked by one thing: high-quality data. Sieve brings together exabyte-scale infrastructure, novel multimodal understanding techniques, large-scale sourcing, and deep research partnerships to create datasets and environments with unmatched precision, quality, and speed. This has earned the trust of frontier AI labs, Fortune 100 companies, and fast-growing AI startups working on generative media, robotics, computer use, world models, and agentic systems.
