1 year ago
Base Salary
$160k - $220k/yr
Responsibilities
- Train and fine-tune OCR, layout, table, and vision-language models.
- Build evaluation, data curation, and active learning pipelines.
- Optimize GPU inference, batching, and quantization.
- Productionize models with clear SLAs and rollback plans.
- Write internal notes that inform model and product roadmaps.
Requirements
- At least 3 years of applied machine learning or research experience, or a strong open-source record.
- Experience with PyTorch or JAX and modern vision or multimodal architectures.
- Strong engineering discipline and focus on metrics.
- Eagerness to learn and adapt quickly.
- Prior startup or founding experience is a plus.
- Preferred experience includes Triton Inference Server, TensorRT, ONNX, and distributed training.
Benefits
- Competitive base salary plus equity and a performance-based bonus.
- Relocation assistance for Bay Area moves.
- Daily meal stipend.
- Medical, vision, and dental coverage.
- Five days per week in the San Francisco office.
