Obvious logo
ObviousPosted 1 month ago

Member of Technical Staff, Infrastructure

$220,000–$280,000 year

RemoteUnited States

Full TimeSenior Level

Job Summary

Own CI/CD pipelines by optimizing build times, improving caching, and reducing flakiness while evolving Kubernetes (EKS) deployment strategies for reliability. Build and harden infrastructure behind model serving, inference, and agent tooling, extending telemetry with instrumentation, sampling, and actionable dashboards including eval pipelines and LLM-ops guardrails. Construct alerting systems that catch actual problems while ignoring noise, and eliminate toil through thoughtful automation to make deployments boring. Improve preview environments, local dev tooling, and testing infrastructure to ensure the feedback loop from code to production remains as fast as possible.

Required Qualifications

  • Infrastructure as the product itself
  • Build and deploy pipelines where rolling back and forth is trivial
  • Model-serving and inference infrastructure that holds up under real AI workloads
  • Observability to know what's actually happening in a system that's non-deterministic by nature
  • Real technical depth in distributed systems
  • Rust or Go
  • Storage engines
  • Control planes
  • Ceph
  • RDMA
  • eBPF
  • Bare-metal automation
  • Kubernetes internals
  • Terraform skills
  • Hands-on experience with observability tools
  • OpenTelemetry
  • Datadog
  • Braintrust
  • Distributed tracing
  • Metrics
  • Structured logging
  • On-call experience
  • Built systems that made on-call better
  • Think like a product manager for internal tools

Desired Qualifications

  • Come from a company where infrastructure was the product itself—not infrastructure work done in service of someone else's product
  • Ideally as an early hire or in a role with real ownership
  • Applied that infra background specifically to AI/LLM workloads
  • Model serving
  • Inference infrastructure
  • Agent tooling
  • Eval pipelines
  • LLM-ops guardrails
  • A personal or self-directed infra track record
  • Side projects
  • Homelabs
  • Open-source infra tooling
  • Published writing or talks
  • Security chops
  • IAM
  • Zero-trust
  • Secrets management
  • SRE practices
  • SLOs
  • SLIs
  • Error budgets
  • Chaos engineering
  • Cost optimization for cloud infrastructure
  • Based in Atlanta
  • You love the talk Simple Made Easy

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce