NVIDIA logo
NVIDIAPosted 1 week ago

Senior Software Engineer, Metropolis Vision AI

$224,000–$356,500 year

On-siteSanta Clara, California, United States

Full TimeSenior LevelDoctorate Or Professional DegreeEnterprise

Job Summary

Craft and implement high-performance Vision AI pipelines for real-time and streaming scenarios using new computer vision and deep learning models. Develop large-scale distributed services to process video, image, and 3D data in edge and cloud settings, while building multi-modal perception capabilities that combine 2D, 3D, and temporal information. Profile and tune GPU-accelerated inference pipelines to meet strict latency, efficiency, and reliability targets, and use simulation tools to validate perception algorithms at scale. Collaborate with partner teams to translate requirements into technical builds, drive code quality reviews, and mentor engineers on Vision AI systems development.

Required Qualifications

  • BS, MS, or PhD in Computer Science, Electrical/Computer Engineering, or a related field, or equivalent experience
  • 12+ years of professional software development experience using modern C++ (14/17/20) and Python on Linux
  • Strong computer science fundamentals, including algorithms, data structures, concurrency, and distributed systems concepts
  • Demonstrated expertise in computer vision and deep learning, with a history of deploying production systems in these fields with a focus on Vision-Language Model (VLM)
  • Experience building and debugging high-performance, concurrent systems, including multi-threading, asynchronous I/O, and efficient memory management
  • Proficiency working in Linux-based environments with containers and microservices, integrating AI components into scalable back-end services
  • Ability to rapidly prototype vision models and pipelines, then evolve them into production-quality services
  • Practical experience with PyTorch in training, fine-tuning, and deploying models for vision tasks
  • Strong analytical and problem-solving skills, with a data-driven approach to performance optimization and system build
  • Excellent written and verbal communication skills, with demonstrated success collaborating across time zones and functions

Desired Qualifications

  • Proven experience delivering end-to-end computer vision applications in production, such as video analytics, smart cities, autonomous systems, retail analytics, industrial inspection, or digital twins
  • Practical experience with GPU acceleration (such as CUDA, TensorRT, or comparable technologies) and low-level optimization for inference and pre/post-processing
  • Experience in simulation and synthetic data creation employing tools such as Omniverse, Unreal Engine, Unity, or similar digital-twin platforms
  • Background in vision-language models or related multi-modal AI, including integrating these models into real products
  • Background in multimedia, including video-centric processing and delivery (such as codecs, video pipelines, or media frameworks) and integrating vision models into multimedia workflows

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce