5 months ago
San Francisco, CA, USAMid Level / Senior
Responsibilities
- Build and maintain production-grade LLM pipelines and agentic workflows.
- Design and optimize scalable RAG architectures using Pinecone, FAISS, and Weaviate.
- Implement agentic systems with tool use, multi-agent coordination, and reasoning loops using LangGraph, LlamaIndex, or equivalent technologies.
- Own prompt engineering, model versioning, evaluation, and LLMOps instrumentation.
- Integrate AI features into large-scale data pipelines and maintain production observability and guardrails.
Requirements
- BS or MS in Computer Science, Machine Learning, or a related field.
- 3–5 years of AI/ML engineering experience, including at least 2 years building LLM-powered systems shipped to production.
- Strong Python skills and experience with PyTorch or Hugging Face Transformers.
- Experience with AWS or GCP and Docker/Kubernetes.
- A required portfolio of shipped AI work, such as agentic pipelines, RAG systems, or fine-tuned models.
- Must be authorized to work in the US without current or future employer sponsorship.
