Principal Platform Engineer, AI & Infrastructure
$255,000–$355,000 year
HybridDenver, Colorado, United States
Job Summary
Shape the technical vision and multi-quarter roadmap for how True Anomaly builds and operates with AI, partnering with leadership to steer the direction of the company's systems in service of space superiority. Parachute into the hardest, most ambiguous technical problems across the enterprise where the business need is critical but the implementation is undefined, getting hands-on designing and delivering the platforms that solve them. Set high-level system design and architecture direction across the platform spanning AI, infrastructure, and data, bringing the enterprise to the cutting edge of agentic engineering in concert with teams and Staff engineers. Play a leading role in DevOps, CI/CD, developer experience, and compute infrastructure decisions, helping teams establish paved-road patterns that scale across the enterprise. Guide technical evaluation and selection of AI providers and tooling, making build-vs-buy decisions based on capability, compliance, and speed to value. Serve as the company's most trusted AI technologist, advising engineering, product, and business functions on how to apply AI and setting the vision for its use across the enterprise. Partner with engineering, security, and IT teams to ensure deployments meet government security and compliance requirements, while acting as the technical anchor and mentor for the engineers on the team.
Required Qualifications
- 12+ years of software engineering, DevOps, SRE, or cloud engineering experience
- Track record of setting technical direction across multiple teams or domains
- Experience owning roadmaps
- Experience making architectural decisions that others build on
- Deep expertise in AI platforms and agentic engineering
- Strong, hands-on command of cloud infrastructure, DevOps, CI/CD, developer experience, and compute
- Ability to set direction and course-correct across the platform
- Track record of taking ambiguous, high-stakes problems from zero to a delivered system
- Ability to frame the problem when requirements and implementation are undefined
- Strong general software background
- Real experience building, deploying, and operating production systems on cloud
- Understanding of a professional software development lifecycle end to end
- Understanding of what constitutes maintainable code
- Hands-on experience with large language model APIs
- Experience with prompt engineering
- Experience with tool use
- Experience with retrieval-augmented generation (RAG)
- Experience with agent frameworks (e.g. CrewAI, Pydantic AI)
- Point of view on where agentic engineering is headed
- Experience with infrastructure-as-code (e.g. Terraform)
- Experience with CI/CD and deployment standards
- Experience with developer platforms and shared tooling
- Experience with cloud networking and identity
- Experience deploying and operating production workloads across cloud environments (AWS, Azure, GCP)
- Experience with networking, identity, and CI/CD foundations
- Experience using AI-powered development tools to deliver production code
- Ability to leverage AI as a force multiplier for yourself and the teams around you
- Demonstrated ability to work across a broad technical surface area
- Ability to context-switch between AI platforms, infrastructure, application development, and platform architecture
- Ability to work across teams to understand requirements and seek consensus
- Strong communication skills
- Ability to translate technical AI concepts for non-technical audiences
- Ability to influence executive leadership on technical direction
- Self-directed and comfortable operating with high autonomy
- Comfortable operating in a fast-paced environment where priorities shift
- Comfortable operating in an environment where the playbook is being written in real time
- US Citizenship
- Ability to obtain and maintain a Top-Secret security clearance
- Work Location — Denver, CO, Long Beach, CA, or Washington, DC
- Onsite availability of at least 3 days per week
- Ability to travel to government sites, classified facilities, launch facilities, or partner locations
Desired Qualifications
- Experience with data platforms, pipelines, or big-data infrastructure
- Ability to help set technical direction for the data team
- Experience deploying AI or ML systems in government, defense, or regulated environments
- Experience with security and compliance constraints
- Familiarity with government cloud environments (AWS GovCloud, Azure Government)
- Familiarity with authorization frameworks (FedRAMP, NIST 800-53, IL4/IL5)
- General experience with various AI technologies, like model harnesses, agentic coding tools, skills, MCPs, RAG, embedding pipelines, etc.
- Experience with LLM eval frameworks
- Experience with OpenTelemetry tracing
- Experience with other model measurement and observability frameworks
- Experience with containers and orchestration
- Experience with service mesh
- Experience with artifact/registry management
- Experience building knowledge retrieval systems
- Experience with embedding pipelines
- Experience with enterprise search infrastructure
- Background in developer tooling, platform engineering, or internal platform teams at high-growth technology companies
- Active U.S. Secret or Top-Secret security clearance
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.