Capital One is hiring an AI Engineer 4 for the Intelligent Foundations and Experiences (IFX) team. The role designs, develops, deploys, and supports responsible, scalable AI systems and foundation model capabilities, with a focus on AI software components and performance optimization across cost, latency, throughput, and reliability.
Responsibilities
- Partner with cross-functional stakeholders, including engineers, research scientists, technical program managers, and product managers, to deliver AI-powered products for how associates work and how customers interact with Capital One.
- Design, develop, test, deploy, and support AI software components such as foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
- Apply a broad stack of Open Source and SaaS AI technologies, including AWS Ultraclusters, Huggingface, VectorDBs, and PyTorch.
- Develop state-of-the-art foundation model optimization techniques to improve performance, including scalability, cost, latency, and throughput, in large-scale production AI systems.
- Contribute to the technical vision and long-term roadmap for foundational AI systems at Capital One.
- Own end-to-end architecture for complex AI systems, ensuring maintainability, observability, and ethical alignment.
- Define and maintain service-level objectives (SLOs) for AI reliability, including latency, uptime, and model performance drift.
- Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines.
- Lead cross-functional technical reviews for new AI system deployments, including security, data governance, and compliance standards.
- Mentor Principal and Senior Associates on scalable design, performance tuning, and research-to-production translation.
Required Qualifications
- Bachelor's degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master's degree plus at least 2 years of experience developing AI and ML algorithms or technologies.
- At least 4 years of experience programming with Python, Go, Scala, CUDA, or Java.
Technologies
Open Source AI technologies, AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, Python, Go, Scala, CUDA, Java, GPU, TPU, foundation model training, large language model inference, agents, multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, observability.
Team Overview
- The Intelligent Foundations and Experiences (IFX) team is central to realizing Capital One’s vision for AI.
- The team works with partners across the company to advance the state of the art in science and AI engineering and builds and deploys proprietary solutions that deliver value to millions of customers.
- Its AI models and platforms help teams enhance products using AI in responsible and scalable ways for high-leverage impact.
Preferred Qualifications
- Experience leading development of AI systems with tradeoff decisions around cost, latency, throughput, and accuracy.
- 6+ years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud).
- Experience designing, developing, delivering, and supporting AI services.
- Experience developing AI and ML algorithms or technologies (e.g., LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, CUDA, or Golang.
- Experience applying state-of-the-art techniques to optimize training and inference software for improved hardware utilization, latency, throughput, and cost.
- Experience building agentic AI systems and agentic workflows.
- Passion for staying current with AI research and applying novel techniques judiciously in production.
- Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale.
- Experience defining AI model governance processes, including producibility, lineage tracking, and automated retaining schedules.
- Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms.
Salary and Incentives
- Salary range: USD 197,300 - 225,100 per year.
- This role may also be eligible for performance-based incentive compensation, including cash bonus(es) and/or long-term incentives (LTI).
- Location salary ranges provided by Capital One:
- Cambridge, MA: $197,300 - $225,100
- McLean, VA: $197,300 - $225,100
- New York, NY: $215,200 - $245,600
- San Jose, CA: $215,200 - $245,600
Additional Information
- No agencies please.
- Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws.
- Capital One promotes a drug-free workplace.
- Capital One will consider qualified applicants with a criminal history in a manner consistent with applicable laws.
- If you require an accommodation for the application process, contact Capital One Recruiting at 1-800-304-9102 or via email at RecruitingAccommodation@capitalone.com.
- For technical support or questions about the recruiting process, email Careers@capitalone.com.
- Capital One does not provide, endorse, or guarantee third-party products, services, educational tools, or other information available through this site.
Location: Cambridge, MA (onsite)