EngineerJobs.io
← Back to all jobs

Job Description

Capital One is building responsible, reliable AI systems that can support how associates work and how customers interact with its services. In this Lead AI Engineer role based in New York, you will help deliver foundation model training and large language model inference, along with the surrounding infrastructure needed for safe evaluation, experimentation, and observability.

This position focuses on scalable, high-performance AI infrastructure and the application of state-of-the-art LLM optimization methods to improve production outcomes such as scalability, cost, latency, and throughput. You will collaborate closely across engineering, research, technical program management, and product to advance the long-term roadmap for foundational AI systems.

Responsibilities

  • Partner with a cross-functional team of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products.
  • Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability.
  • Use a broad stack of Open Source and SaaS AI technologies, including AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch, and more.
  • Develop and introduce state-of-the-art LLM optimization techniques to improve performance of production AI systems, targeting scalability, cost, latency, and throughput.
  • Contribute to the technical vision and long-term roadmap of foundational AI systems at Capital One.

Requirements

  • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master’s degree plus at least 2 years of such experience.
  • At least 4 years of programming experience with Python, Go, Scala, or Java.

Technologies

  • AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch
  • AWS, Google Cloud, Azure
  • Python, Go, Scala, Java, C++, C#, Golang

Benefits

  • Comprehensive, competitive, and inclusive health, financial, and other benefits that support total well-being.
  • Eligible to earn performance-based incentive compensation, which may include cash bonuses and/or long term incentives (LTI).

Preferred Qualifications

  • 6 years of experience deploying scalable and responsible AI solutions on cloud platforms such as AWS, Google Cloud, Azure, or equivalent private cloud.
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience developing AI and ML algorithms or technologies (for example: LLM Inference, Similarity Search and VectorDBs, Guardrails, Memory) using Python, C++, C#, Java, or Golang.
  • Experience developing and applying state-of-the-art techniques for optimizing training and inference software to improve hardware utilization, latency, throughput, and cost.
  • Passion for staying abreast of the latest AI research and AI systems, with the ability to judiciously apply novel techniques in production.

Salary: USD 215,200 - 245,600 per year. Location: New York, NY (onsite). Minimum experience: 4 years.

Capital One expects to accept applications for at least 5 business days. No agencies please. Capital One is an equal opportunity employer (EOE, including disability/vet) committed to non-discrimination in compliance with applicable federal, state, and local laws. Capital One promotes a drug-free workplace.

If you need an accommodation to apply due to a disability, contact Capital One Recruiting at 1-800-304-9102 or [email protected]. For technical support or questions about the recruiting process, email [email protected]. Capital One does not provide, endorse, or guarantee third-party products, services, educational tools, or other information through this site.

Similar Jobs