Cerence AI logo
Cerence AIPosted 1 month ago

Principal Software Engineer – Robot Applications & Voice AI

$165,550–$165,550 year

RemoteUnited States

Full TimeSenior LevelMediumAI Services

Job Summary

Implement high-level application software and robotic behavioral state machines to translate human speech, gestures, and environmental cues into deterministic tasks. Build orchestration workflows combining LLMs, VLMs, and memory systems while integrating vision pipelines, object detection, and spatial tracking for contextual awareness. Develop robust voice applications managing STT, NLU, dialog management, and TTS interfaces, then optimize applications to balance low-latency local execution with cloud-based AI services. Consume and orchestrate navigation, manipulation, and device-control services to deploy software in real-world robotic environments.

Required Qualifications

  • Strong proficiency in BehaviorTree.CPP, state machines, workflow orchestration engines, or task-planning architectures
  • Hands-on experience building applications using LLMs, VLM, tool-calling architectures, agent frameworks, RAG systems, or semantic memory platforms
  • 5+ years of professional software engineering experience using Python and/or C++
  • Experience designing scalable application-layer software, API-driven systems, distributed services, and event-driven architectures
  • Experience optimizing software on resource-constrained edge hardware under intermittent connectivity conditions
  • Experience integrating cloud AI services, model serving platforms, and containerized deployments
  • Automated testing, debugging, monitoring, observability, and CI/CD practices
  • Experience designing natural, context-aware human-machine experiences
  • Bachelor's or Master's degree in Computer Science, Robotics, Software Engineering, AI, or a related discipline
  • Basic knowledge of information security and data privacy requirements
  • Demonstrative knowledge of information security through internal training programs

Desired Qualifications

  • Voice & Speech Stack: Cerence, Whisper, Deepgram, Azure Speech, ElevenLabs, AWS Speech Services, STT/TTS/NLU platforms
  • ROS2 application nodes, Services, Actions, Topics, navigation and manipulation APIs
  • Object detection, scene understanding, visual grounding, and spatial reasoning systems
  • Beamforming, Acoustic Echo Cancellation (AEC), microphone arrays, and noise suppression
  • Experience building software for humanoids, service robots, warehouse automation, or embodied AI systems
  • Telemetry, analytics, evaluation systems, retraining pipelines, and AI feedback loops
  • Docker, Kubernetes, OTA updates, and edge deployment environments
  • Comfortable owning solutions from concept through deployment in a fast-moving environment

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce