The Inception Company

Member of Technical Staff, ML Product Engineering

The Inception Company
Apply
6 months ago
San Mateo, CA, USASenior

Responsibilities

  • Design, develop, and optimize diffusion and large language models for production use cases.
  • Partner with customers to understand requirements and translate them into technical solutions.
  • Implement post-training approaches for generative AI models, including agentic workflows.
  • Build data preprocessing pipelines, model evaluation processes, and alignment methods for enterprise use cases.
  • Deploy and maintain models in production environments at scale.
  • Collaborate with product teams to design and implement customer-facing ML features.

Requirements

  • BS, MS, or PhD in Computer Science, Machine Learning, or a related field, or equivalent experience.
  • At least 5 years of experience working on ML projects using PyTorch or an equivalent technology.
  • Excellent familiarity with transformers and core LLM concepts, including autoregressive pretraining, instruction tuning, in-context learning, LoRA, and KV caching.
  • Experience training and fine-tuning LLMs.
  • Familiarity with large-scale systems and high-performance computing, including GPU and TPU utilization.
  • Experience with Git and Docker.
  • Strong communication skills for explaining technical concepts to non-technical stakeholders.
  • Preferred: expertise in data engineering and synthetic data generation for LLMs.
  • Preferred: knowledge of MLOps and production deployment workflows.
  • Preferred: experience with vLLM, SGLang, or TensorRT for LLM serving.
  • Preferred: experience with AWS, GCP, or Azure.
  • Preferred: experience with model quantization and optimization techniques.
The Inception Company

About The Inception Company

51-200 employees
Contact me