
AI Engineer 4 (AI Foundations, LLM Core and Agentic AI)
Capital One2 days ago
Cambridge, MA, USA +2 moreStaff+
Base Salary
$215k - $246k/yr
Responsibilities
- Partner with engineers, research scientists, technical program managers, and product managers to deliver AI-powered products.
- Design, develop, test, deploy, and support AI software components for foundation-model training, LLM inference, agents, multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Develop foundation-model optimization techniques to improve scalability, cost, latency, throughput, and production performance.
- Contribute to the technical vision and long-term roadmap for foundational AI systems.
- Own end-to-end architecture for complex AI systems with maintainability, observability, and ethical alignment.
- Define and maintain service-level objectives for AI reliability, latency, uptime, and model-performance drift.
- Optimize GPU/TPU utilization and accelerate model-inference pipelines with infrastructure engineering.
- Lead technical reviews for AI deployments covering security, data governance, and compliance.
- Mentor Principal and Senior Associates on scalable design, performance tuning, and research-to-production translation.
Requirements
- Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least 4 years of experience developing AI and ML algorithms or technologies, or a relevant master’s degree plus at least 2 years of that experience.
- At least 4 years of programming experience with Python, Go, Scala, CUDA, or Java.
- Experience leading AI-system development with tradeoffs involving cost, latency, throughput, and accuracy.
- Preferred experience deploying scalable and responsible AI solutions on AWS, Google Cloud, Azure, or equivalent private cloud, including 6 years of such experience.
- Experience designing, developing, delivering, and supporting AI services.
- Experience with LLM inference, similarity search, VectorDBs, guardrails, memory, and AI/ML technologies using Python, C++, C#, Java, CUDA, or Golang.
- Experience optimizing training and inference software for hardware utilization, latency, throughput, and cost.
- Experience building agentic AI systems and agentic workflows.
- Proficiency designing distributed systems for model training, evaluation, and online inference at petabyte scale.
- Experience defining AI model-governance processes, including producibility, lineage tracking, and automated retraining schedules.
- Ability to influence architectural decisions across multiple AI product lines or platforms.
- Ability to apply current AI research and novel techniques in production.
Benefits
- Full-time annual base salary ranges by location from $197,300-$225,100 in Cambridge, MA and McLean, VA, and from $215,200-$245,600 in New York, NY and San Jose, CA; other locations use the applicable local range.
- Eligible for performance-based incentive compensation, including cash bonuses and/or long-term incentives.
- Comprehensive health, financial, and other benefits supporting employee well-being, with eligibility varying by employment status and management level.
- Applications are expected to remain open for a minimum of 5 business days.
- Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.
Categories
About Capital One
Capital One is a U.S. consumer and commercial bank that offers credit cards, checking and savings accounts, auto financing, and lending to small and large businesses. It earns revenue from interest income and interchange/fees across its card and banking products, and develops cloud-based digital services. Founded in 1994, Capital One is publicly traded on the NYSE (COF) and is headquartered in McLean, Virginia.