Member of Technical Staff, ML Product Engineering
The Inception Company6 months ago
San Mateo, CA, USASenior
Responsibilities
- Design, develop, and optimize diffusion and large language models for production use cases.
- Partner with customers to understand requirements and translate them into technical solutions.
- Implement post-training approaches for generative AI models, including agentic workflows.
- Build data preprocessing pipelines, model evaluation processes, and alignment methods for enterprise use cases.
- Deploy and maintain models in production environments at scale.
- Collaborate with product teams to design and implement customer-facing ML features.
Requirements
- BS, MS, or PhD in Computer Science, Machine Learning, or a related field, or equivalent experience.
- At least 5 years of experience working on ML projects using PyTorch or an equivalent technology.
- Excellent familiarity with transformers and core LLM concepts, including autoregressive pretraining, instruction tuning, in-context learning, LoRA, and KV caching.
- Experience training and fine-tuning LLMs.
- Familiarity with large-scale systems and high-performance computing, including GPU and TPU utilization.
- Experience with Git and Docker.
- Strong communication skills for explaining technical concepts to non-technical stakeholders.
- Preferred: expertise in data engineering and synthetic data generation for LLMs.
- Preferred: knowledge of MLOps and production deployment workflows.
- Preferred: experience with vLLM, SGLang, or TensorRT for LLM serving.
- Preferred: experience with AWS, GCP, or Azure.
- Preferred: experience with model quantization and optimization techniques.