Senior Software Engineer, Metropolis Vision AI
$224,000–$356,500 year
On-siteSanta Clara, California, United States
Job Summary
Craft and implement high-performance Vision AI pipelines for real-time and streaming scenarios using new computer vision and deep learning models. Develop large-scale distributed services to process video, image, and 3D data in edge and cloud settings, while building multi-modal perception capabilities that combine 2D, 3D, and temporal information. Profile and tune GPU-accelerated inference pipelines to meet strict latency, efficiency, and reliability targets, and use simulation tools to validate perception algorithms at scale. Collaborate with partner teams to translate requirements into technical builds, drive code quality reviews, and mentor engineers on Vision AI systems development.
Required Qualifications
- BS, MS, or PhD in Computer Science, Electrical/Computer Engineering, or a related field, or equivalent experience
- 12+ years of professional software development experience using modern C++ (14/17/20) and Python on Linux
- Strong computer science fundamentals, including algorithms, data structures, concurrency, and distributed systems concepts
- Demonstrated expertise in computer vision and deep learning, with a history of deploying production systems in these fields with a focus on Vision-Language Model (VLM)
- Experience building and debugging high-performance, concurrent systems, including multi-threading, asynchronous I/O, and efficient memory management
- Proficiency working in Linux-based environments with containers and microservices, integrating AI components into scalable back-end services
- Ability to rapidly prototype vision models and pipelines, then evolve them into production-quality services
- Practical experience with PyTorch in training, fine-tuning, and deploying models for vision tasks
- Strong analytical and problem-solving skills, with a data-driven approach to performance optimization and system build
- Excellent written and verbal communication skills, with demonstrated success collaborating across time zones and functions
Desired Qualifications
- Proven experience delivering end-to-end computer vision applications in production, such as video analytics, smart cities, autonomous systems, retail analytics, industrial inspection, or digital twins
- Practical experience with GPU acceleration (such as CUDA, TensorRT, or comparable technologies) and low-level optimization for inference and pre/post-processing
- Experience in simulation and synthetic data creation employing tools such as Omniverse, Unreal Engine, Unity, or similar digital-twin platforms
- Background in vision-language models or related multi-modal AI, including integrating these models into real products
- Background in multimedia, including video-centric processing and delivery (such as codecs, video pipelines, or media frameworks) and integrating vision models into multimedia workflows
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.