EngineerJobs.io
← Back to all jobs

Job Description

Capital One is hiring an AI Engineer 4 on the Intelligent Foundations and Experiences (IFX) team to build and deploy responsible, scalable AI systems and platform services. This onsite role in San Jose focuses on foundation model training, LLM inference, agentic workflows, evaluation and governance, observability, and optimization for performance and cost.

Responsibilities

  • Work with cross-functional partners including engineers, research scientists, technical program managers, and product managers to deliver AI-powered products for associates and customers.
  • Design, develop, test, deploy, and support AI software components such as foundation model training, large language model inference, agents and multi-agent workflows, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
  • Use a stack of Open Source and SaaS AI technologies including AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, and additional tools.
  • Develop and introduce foundation model optimization techniques to improve scalability, cost, latency, and throughput for production AI systems.
  • Help shape the technical vision and long-term roadmap for foundational AI systems at Capital One.
  • Own end-to-end architecture for complex AI systems, with focus on maintainability, observability, and ethical alignment.
  • Define and maintain service-level objectives (SLOs) for AI reliability across latency, uptime, and model performance drift.
  • Collaborate with infrastructure engineering to optimize GPU/TPU utilization and accelerate model inference pipelines.
  • Lead cross-functional technical reviews for new AI system deployments, ensuring security, data governance, and compliance standards are addressed.
  • Mentor Principal and Senior Associates on scalable design, performance tuning, and research-to-production translation.

Requirements

  • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least 4 years of experience developing AI and ML algorithms or technologies, or
    Master’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least 2 years of experience developing AI and ML algorithms or technologies.
  • At least 4 years of programming experience with Python, Go, Scala, CUDA, or Java.

Technologies

  • AWS Ultraclusters, Huggingface, VectorDBs, PyTorch
  • AWS, Google Cloud, Azure
  • Python, Go, Scala, CUDA, Java, C++, C#, C# (listed as provided), Golang
  • GPU, TPU

Benefits

  • Performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI).
  • Comprehensive, competitive, and inclusive set of health, financial and other benefits.

Preferred Qualifications

  • Experience leading development AI systems with tradeoff decisions across cost, latency, throughput, and accuracy.
  • 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud).
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience developing AI and ML algorithms or technologies such as LLM Inference, Similarity Search and VectorDBs, Guardrails, and Memory using Python, C++, C#, Java, CUDA, or Golang.
  • Experience developing and applying state-of-the-art optimization techniques for training and inference software to improve hardware utilization, latency, throughput, and cost.
  • Experience building agentic AI systems and agentic workflows.
  • Passion for staying abreast of the latest AI research and systems, and applying novel techniques in production.
  • Proficiency in designing distributed systems for model training, evaluation, and online inference at petabyte scale.
  • Experience defining AI model governance processes including producibility, lineage tracking, and automated retention schedules.
  • Demonstrated ability to influence architectural decisions across multiple AI product lines or platforms.

Team Description

  • The Intelligent Foundations and Experiences (IFX) team is at the center of bringing Capital One’s AI vision to life.
  • The team partners across the company to advance the state of the art in science and AI engineering, building and deploying proprietary solutions central to the business.
  • AI models and platforms empower teams across Capital One to enhance products with responsible and scalable AI for high-leverage impact.

Other Pay and Posting Notes

  • Capital One will consider sponsoring a new qualified applicant for employment authorization for this position.
  • Salary range for San Jose, CA: USD 215,200 - 245,600 per year (AI Engineer 4).
  • This role is eligible to earn performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI).
  • This role is expected to accept applications for a minimum of 5 business days.

Similar Jobs