EngineerJobs.io
← Back to all jobs

Job Description

Glint Tech Solutions LLC builds large-scale, production-focused computer vision for enterprise applications, and this role sits on an AI team delivering real-world systems. The position emphasizes NVIDIA DeepStream, GPU performance optimization, and distributed computing, with the goal of shipping production-grade vision pipelines and microservices across GPU-accelerated environments.

What you’ll work on

You will design, develop, and optimize computer vision solutions intended for real-world production environments. Work centers on building scalable AI inference pipelines using NVIDIA DeepStream, with a strong focus on high-throughput, low-latency performance. You will also help deploy distributed computer vision applications across large-scale infrastructure, and create microservices that support AI and computer vision platforms.

  • Design, develop, and optimize computer vision solutions for production use
  • Build scalable AI inference pipelines using NVIDIA DeepStream
  • Optimize GPU performance for high-throughput, low-latency workloads
  • Develop and deploy distributed computer vision applications across large-scale infrastructure
  • Design and implement microservices supporting AI and computer vision platforms
  • Collaborate with AI researchers, software engineers, and infrastructure teams to deliver production-ready solutions
  • Improve system performance, scalability, and reliability across GPU-accelerated environments
  • Troubleshoot performance bottlenecks and optimize inference pipelines

Required skills

  • NVIDIA DeepStream
  • GPU Optimization
  • Computer Vision
  • Microservices Architecture
  • Distributed Computing

Requirements

  • Strong experience with Computer Vision technologies and production deployments
  • Hands-on experience with NVIDIA DeepStream (Required)
  • Strong understanding of GPU optimization and performance tuning
  • Experience building microservices architecture
  • Experience with distributed computing systems
  • Strong programming skills in Python and/or C++
  • Experience deploying scalable AI applications in production
  • Excellent problem-solving and performance optimization skills

Technologies you may work with

  • NVIDIA DeepStream, NVIDIA CUDA, TensorRT, NVIDIA SDKs
  • Python, C++, OpenCV, GStreamer
  • Microservices architecture, GPU optimization
  • Kubernetes, Docker, cloud platforms (AWS, Azure, GCP)
  • Machine learning models, deep learning models

Preferred qualifications

  • Experience with NVIDIA CUDA or TensorRT
  • Experience with OpenCV, GStreamer, or NVIDIA SDKs
  • Familiarity with Kubernetes and Docker
  • Experience with cloud platforms (AWS, Azure, or GCP)
  • Experience deploying machine learning or deep learning models in production
  • Knowledge of real-time streaming and edge AI solutions

Location: Sunnyvale, CA (onsite)

Similar Jobs