
Member of Technical Staff, Applied Research
LlamaIndex2 months ago
Base Salary
$180k - $250k/yr
Responsibilities
- Develop and train vision-language models for document processing and understanding.
- Build data pipelines for data curation, synthetic data generation, labeling, and benchmark creation.
- Evaluate base models and perform post-training or fine-tuning to meet performance targets.
- Improve model accuracy, latency, and cost-effectiveness across real-world document workflows.
- Design and maintain benchmarks for extraction quality, layout understanding, OCR performance, reasoning accuracy, and end-to-end reliability.
- Work with PDFs, scanned documents, tables, charts, forms, and multi-page enterprise documents.
- Collaborate with engineering to move successful research prototypes into production.
- Work with customers when needed to translate product requirements into benchmarks, experiments, and model improvements.
- Stay current with research in vision-language models, document AI, post-training, synthetic data, and agentic systems.
Requirements
- 3–7 years of experience in machine learning engineering, applied research, or research engineering.
- Strong machine learning foundations with hands-on experience benchmarking and training models.
- Strong Python skills and experience with modern machine learning tooling, especially PyTorch.
- Experience with computer vision, vision-language models, NLP, document AI, OCR, extraction, or agentic AI systems.
- Ability to design experiments, evaluate results, and iterate toward measurable performance improvements.
- Strong engineering judgment and ability to write clean, production-quality code.
- Experience with synthetic data generation, post-training, fine-tuning, or benchmark design is preferred.
- Startup, founder, document processing, open-source AI infrastructure, or developer tools experience is preferred.
- Strong technical writing and communication skills.
Benefits
- Work on frontier vision-language models and document AI infrastructure.
- Collaborate directly with technical founders, the CTO, and an ambitious engineering team.
- Have ownership over model quality, product capability, and technical direction.
- Join a fast-growing startup with open-source adoption and commercial traction.
Categories
AI ResearchML Engineering
About LlamaIndex
LlamaParse is the most accurate agentic OCR platform for production AI — purpose-built for the documents agents actually encounter in the real world. Unlike general-purpose models that guess at structure, LlamaParse is engineered for complex layouts, dense tables, handwritten annotations, and scanned pages. Every page is automatically routed to the optimal model, so accuracy and cost are optimized without manual configuration. Trusted by teams at Lovable, 8am, Tabs, KPMG, and others running document-intensive workflows across legal, finance, healthcare, and more.