Red Hat logo
Red HatPosted 3 weeks ago

Principal Forward Deployed Engineer - AI PlatformPrincipal Forward Deployed Engineer - AI Platform/Kubernetes/Pytorch

On-siteSingapore, Singapore

Full TimeSenior LevelEnterpriseTECH

Job Summary

Embed with strategic APAC customers to resolve deep-seated engineering blockers, architect future-proof solutions, and establish robust testing practices for large-scale AI systems. Rapidly prototype custom, secure platform integrations using real-world data, validating complex distributed inference systems, advanced RAG pipelines, and autonomous AI agent architectures. Harden initial proofs-of-concept into scalable production environments while acting as the technical bridge between regional market requirements and global engineering teams. Commit upstream code directly to core open-source projects under the cloud-native AI, model serving, and MLOps umbrella, refusing bespoke isolated code in favor of standardized regional solutions. Drive the evolution of the Software Development Life Cycle and mentor senior technical staff across the organization.

Required Qualifications

  • Exceptional proficiency in C/C++, Go, and Python
  • Must have a proven track record of shipping production-ready, highly optimized, and robustly tested code
  • Demonstrated status as an active upstream contributor or maintainer in key open-source communities (such as deep learning engines, distributed computing schedulers, or container orchestration runtimes)
  • Hands-on experience with deep learning frameworks, model fine-tuning (LoRA, QLoRA, SFT), large-scale model serving architectures, and LLM orchestration (e.g., LangChain, LlamaIndex)
  • Strong experience architecting on Kubernetes or enterprise container platforms, including writing and managing Custom Resource Definitions (CRDs), custom operators, and multi-subsystem integrations
  • Familiarity with hardware-level optimizations, CUDA, ROCm, driver compilation, and GPU/NPU operator configuration
  • Minimum of 8+ years of experience in system engineering, distributed computing, platform engineering, or AI/ML software development
  • Demonstrated experience mentoring senior technical staff and driving modern software delivery practices (SDLC, CI/CD, and robust testing frameworks)
  • High capacity to navigate ambiguous, rapid-velocity environments within a large enterprise structure
  • Excellent communication and consultative skills; ability to engage with enterprise customer architects and C-suite technologists while remaining highly mindful of customer data privacy, compliance policies, and relevant regulatory frameworks
  • Deep capability in at least one of the following two tracks: Hardware Heterogeneity & Sovereign AI Focus OR AI Platform Primitives & Distributed Systems

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce