EngineerJobs.io
← Back to all jobs

Job Description

Lead AI Engineer role focused on vision model customization and VLM, building responsible, reliable AI systems end to end.

Responsibilities

  • Work with a cross-functional group of engineers, research scientists, technical program managers, and product managers to deliver AI-powered products for associates and customers.
  • Design, develop, test, deploy, and support AI software components across production workflows, including:
    • Foundation model training
    • Large language model inference
    • Similarity search and vector database applications
    • Guardrails
    • Model evaluation
    • Experimentation, governance, and observability
    • Additional related AI components
  • Build using a mix of open source and SaaS AI technology, including AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, and PyTorch.
  • Develop and introduce state-of-the-art LLM optimization methods to improve scalability, cost, latency, and throughput for large-scale production AI systems.
  • Help shape the technical direction and long-term roadmap for foundational AI systems.

Requirements

  • Bachelor’s degree in Computer Science, AI, Electrical Engineering, Computer Engineering, or a related field plus at least 4 years of experience developing AI and ML algorithms or technologies, or a Master’s degree plus at least 2 years of experience developing AI and ML algorithms or technologies.
  • At least 4 years of programming experience with Python, Go, Scala, or Java.

Technologies

  • AWS Ultraclusters
  • Huggingface
  • VectorDBs
  • Nemo Guardrails
  • PyTorch
  • Python, Go, Scala, Java
  • C++, C#, Golang
  • AWS, Google Cloud, Azure

Benefits

  • Performance-based incentive compensation, potentially including cash bonus(es) and/or long-term incentives (LTI).
  • Comprehensive, competitive, and inclusive health, financial, and other benefits supporting total well-being.

Preferred Qualifications

  • 6 years of experience deploying scalable and responsible AI solutions on cloud platforms (e.g., AWS, Google Cloud, Azure, or equivalent private cloud).
  • Experience designing, developing, delivering, and supporting AI services.
  • Experience developing AI and ML algorithms or technologies (e.g., LLM inference, similarity search and VectorDBs, guardrails, memory) using Python, C++, C#, Java, or Golang.
  • Experience developing and applying state-of-the-art optimization techniques for training and inference software to improve hardware utilization, latency, throughput, and cost.
  • Interest in staying current with AI research and AI systems, with a focus on applying novel techniques in production.

Location: McLean, VA 22101 (onsite)

Compensation: USD 197,300 - 225,100 per year

Similar Jobs