Capital One

AI Engineer 4 (AI Foundations)

Capital One
Apply
1 day ago
Cambridge, MA, USA +4 moreStaff+

Base Salary

$215k - $246k/yr

Responsibilities

  • Partner with engineers, research scientists, technical program managers, and product managers to deliver AI-powered products.
  • Design, develop, test, deploy, and support AI software components including foundation-model training, LLM inference, agents, multi-agent workflows, similarity search, guardrails, evaluation, experimentation, governance, and observability.
  • Develop optimization techniques for large-scale production AI systems to improve scalability, cost, latency, and throughput.
  • Own the end-to-end architecture and long-term technical roadmap for foundational AI systems.
  • Define and maintain service-level objectives for AI reliability, including latency, uptime, and model-performance drift.
  • Optimize GPU/TPU utilization and model-inference pipelines with infrastructure engineering.
  • Lead technical reviews for AI deployments, ensuring security, data governance, and compliance standards are met.
  • Mentor Principal and Senior Associates on scalable design, performance tuning, and research-to-production translation.

Requirements

  • Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least four years of experience developing AI and ML algorithms or technologies, or a master's degree in one of those fields plus at least two years of such experience.
  • At least four years of programming experience with Python, Go, Scala, CUDA, or Java.
  • Preferred experience leading AI-system development with tradeoffs involving cost, latency, throughput, and accuracy.
  • Preferred six years of experience deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, Azure, or equivalent private cloud.
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience with AI and ML technologies such as LLM inference, similarity search, VectorDBs, guardrails, and memory using Python, C++, C#, Java, CUDA, or Golang.
  • Experience optimizing training and inference software for hardware utilization, latency, throughput, and cost.
  • Experience building agentic AI systems and workflows.
  • Proficiency designing distributed systems for model training, evaluation, and online inference at petabyte scale.
  • Experience defining AI model-governance processes, including producibility, lineage tracking, and automated retraining schedules.
  • Ability to influence architectural decisions across multiple AI product lines or platforms.

Benefits

  • Capital One offers health, financial, and other benefits supporting total well-being, with eligibility varying by full- or part-time status, exempt or non-exempt status, and management level.
  • The role is eligible for performance-based incentive compensation, including potential cash bonuses and/or long-term incentives.
  • The posting is expected to accept applications for a minimum of five business days.

Tech Stack

Categories

Capital One

About Capital One

10,000+ employees

Capital One is a U.S. consumer and commercial bank that offers credit cards, checking and savings accounts, auto financing, and lending to small and large businesses. It earns revenue from interest income and interchange/fees across its card and banking products, and develops cloud-based digital services. Founded in 1994, Capital One is publicly traded on the NYSE (COF) and is headquartered in McLean, Virginia.

Contact me