
AI Engineer 5 (Gen AI Platform Services: Agentic AI, Guardrails, Evaluations)
Capital One1 day ago
Cambridge, MA, USA +4 moreStaff+
Base Salary
$251k - $286k/yr
Responsibilities
- Partner with engineers, research scientists, technical program managers, and product managers to deliver AI-powered products.
- Design, develop, test, deploy, and support AI software for foundation model training, LLM inference, agents, multi-agent workflows, similarity search, guardrails, evaluation, experimentation, governance, and observability.
- Develop optimization techniques for production AI systems to improve scalability, cost, latency, throughput, and hardware utilization.
- Design and optimize multi-model orchestration pipelines integrating LLMs, vector search, and domain-specific models.
- Lead cost-performance governance reviews covering GPU utilization, model throughput, and inference cost efficiency.
- Contribute to the technical vision and long-term roadmap for foundational AI systems.
- Lead design councils and review boards to promote technical consistency and compliance with AI engineering standards.
- Mentor Principal- and Manager-level AI engineers and foster cross-domain technical learning.
- Define and enforce standards for ethical AI deployment, including explainability, fairness, and human-in-the-loop review processes.
Requirements
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least six years of experience developing AI and ML algorithms or technologies, or a master's degree in one of these fields plus at least four years of such experience.
- At least six years of programming experience with Python, Go, Scala, CUDA, or Java.
- Experience leading AI system development with tradeoffs involving cost, latency, throughput, and accuracy.
- Preferred experience deploying scalable and responsible AI solutions on cloud platforms, including seven or more years of experience preferred.
- Experience designing, developing, delivering, and supporting complex AI systems.
- Experience with LLM inference, similarity search, vector databases, guardrails, and memory using languages such as Python, C++, C#, Java, CUDA, or Go.
- Experience optimizing training and inference software and improving hardware utilization, latency, throughput, and cost.
- Experience building agentic AI systems and agentic workflows.
- Experience architecting heterogeneous AI systems, including rule-based, retrieval-augmented, and generative components, into unified production pipelines.
- Experience balancing model performance and operational cost through dynamic inference strategies and model compression.
- Experience right-sizing models, instance counts, and hardware types based on requirements such as context length and token inputs and outputs.
- Strong AI systems, engineering, mathematics, research application, communication, and presentation skills.
Benefits
- Capital One offers health, financial, and other benefits supporting employee well-being, with eligibility varying by employment status and management level.
- The role is eligible for performance-based incentive compensation, which may include cash bonuses or long-term incentives.
- The posting lists full-time work locations in Cambridge, McLean, New York, San Francisco, and San Jose, with applications accepted for a minimum of five business days.
Categories
About Capital One
Capital One is a U.S. consumer and commercial bank that offers credit cards, checking and savings accounts, auto financing, and lending to small and large businesses. It earns revenue from interest income and interchange/fees across its card and banking products, and develops cloud-based digital services. Founded in 1994, Capital One is publicly traded on the NYSE (COF) and is headquartered in McLean, Virginia.