
AI Engineer 4 (AI Foundations, LLM Core and Agentic AI)
Capital One1 day ago
Cambridge, MA, USA +4 moreStaff+
Base Salary
$215k - $246k/yr
Responsibilities
- Partner with engineers, research scientists, technical program managers, and product managers to deliver AI-powered products.
- Design, develop, test, deploy, and support foundation-model training, LLM inference, agentic workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability components.
- Optimize foundation-model training and inference for scalability, cost, latency, throughput, and hardware utilization.
- Own end-to-end architecture for complex AI systems and contribute to the long-term roadmap for foundational AI systems.
- Define service-level objectives for AI reliability, including latency, uptime, and model-performance drift.
- Collaborate with infrastructure engineering to optimize GPU/TPU utilization and inference pipelines.
- Lead technical reviews for AI deployments covering security, data governance, compliance, and ethical alignment.
- Mentor Principal and Senior Associates on scalable design, performance tuning, and research-to-production translation.
Requirements
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least 4 years of experience developing AI and ML algorithms or technologies, or a master's degree in one of these fields plus at least 2 years of relevant experience.
- At least 4 years of programming experience with Python, Go, Scala, CUDA, or Java.
- Ability to understand scientific publications and apply novel AI techniques in production.
- Strong engineering, mathematics, hardware, software, and AI foundations.
- Preferred: experience leading AI-system development with cost, latency, throughput, and accuracy tradeoffs.
- Preferred: 6 years of experience deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, Azure, or equivalent private cloud.
- Preferred: experience developing and supporting AI services, agentic AI systems, distributed systems for model training and inference, and AI governance processes.
- Preferred: experience with LLM inference, similarity search, VectorDBs, guardrails, memory, and training/inference optimization using Python, C++, C#, Java, CUDA, or Golang.
- Preferred: demonstrated ability to influence architectural decisions across multiple AI product lines or platforms.
Benefits
- Comprehensive health, financial, and other benefits supporting total well-being, with eligibility varying by employment status and management level.
- Eligible for performance-based incentive compensation, including cash bonuses and/or long-term incentives.
- Role is offered in listed locations including Cambridge, McLean, New York, San Francisco, and San Jose; applications are accepted for a minimum of 5 business days.
Categories
About Capital One
Capital One is a U.S. consumer and commercial bank that offers credit cards, checking and savings accounts, auto financing, and lending to small and large businesses. It earns revenue from interest income and interchange/fees across its card and banking products, and develops cloud-based digital services. Founded in 1994, Capital One is publicly traded on the NYSE (COF) and is headquartered in McLean, Virginia.