EngineerJobs.io
← Back to all jobs

Job Description

Join Capital One’s Intelligent Foundations and Experiences (IFX) team and help build responsible, reliable foundational AI systems that improve how associates work and how customers interact with Capital One. You will design, develop, deploy, and optimize core AI components spanning LLM training and inference, similarity search, guardrails, evaluation, governance, and observability, while contributing to the long-term technical roadmap.

What you’ll do

  • Collaborate with engineers, research scientists, technical program managers, and product managers to deliver AI-powered products.
  • Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
  • Apply a mix of Open Source and SaaS AI technologies, including AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, and PyTorch.
  • Invent and introduce state-of-the-art LLM optimization techniques to improve scalability, cost, latency, and throughput for large-scale production AI systems.
  • Contribute to the technical vision and long-term roadmap for foundational AI systems at Capital One.

What you’ll bring

  • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or related fields plus at least 4 years of experience developing AI/ML algorithms or technologies, or Master’s plus at least 2 years.
  • At least 4 years of programming experience with Python, Go, Scala, or Java.
  • Strong engineering quality standards, along with a commitment to doing the right thing.
  • Ability to stay current with AI research and apply novel techniques appropriately in production.
  • Comfort bringing clarity to large, undefined problems and adapting quickly.
  • Deep technical foundation across engineering and mathematics, with expertise spanning hardware, software, and AI to identify optimization opportunities.
  • Resilience and initiative to forge new paths to achieve business goals when the route is not yet defined.

Helpful experience (preferred)

  • 6+ years deploying scalable and responsible AI solutions on cloud platforms (for example, AWS, Google Cloud, Azure, or equivalent private cloud).
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience building AI and ML algorithms or technologies such as LLM inference, similarity search and VectorDBs, guardrails, and memory using Python, C++, C#, Java, or Golang.
  • Experience applying state-of-the-art optimization techniques to improve hardware utilization, latency, throughput, and cost for training and inference software.

Technologies you may work with

  • AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, AWS, Google Cloud, Azure
  • Python, Go, Scala, Java, C++, C#, Golang

Compensation and benefits

Salary range: USD 215,200 - 245,600 per year.

  • Health insurance
  • Financial and other benefits supporting your total well-being
  • Performance-based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI)

Additional information

  • Location: New York, NY (onsite)
  • Capital One will consider sponsoring a new qualified applicant for employment authorization.
  • This role is expected to accept applications for a minimum of 5 business days.
  • No agencies please.
  • Capital One is an equal opportunity employer and promotes a drug-free workplace.
  • Capital One may consider qualified applicants with a criminal history in accordance with applicable laws.
  • For technical support or questions about Capital One’s recruiting process: [email protected]

Similar Jobs