Principal Platform Software Engineer
On-siteNashville, Tennessee, United States
Job Summary
Design, develop, test, deploy, and operate software provisioning containers across OCI regions using Linux virtualization, hypervisor development, and imaging. Manage containerized cloud workloads with expertise in QEMU/KVM, container runtimes, networking, storage, systemd, and OS image creation. Own end-to-end reliability, scalability, and customer experience for OCI Container Instances, including on-call rotations, incident response, and operational excellence. Leverage Terraform for infrastructure as code and AI-assisted tools to maintain high standards for code quality and security. Collaborate across teams to solve challenging engineering problems at the intersection of Linux, virtualization, and cloud computing.
Required Qualifications
- BS/MS in Computer Science, Engineering, or equivalent practical experience
- 6+ years of experience building and operating production software systems, preferably distributed systems or cloud services
- Strong system administration experience in Linux, including bash/shell scripting
- Strong programming experience in Java or Go
- Strong fundamentals in data structures, algorithms, operating systems, networking, and distributed systems
- Experience working with containers, Docker, Kubernetes, or related cloud-native technologies
- Experience with CI/CD systems, automated testing, and modern software development practices
- Experience using AI-assisted software development tools to improve engineering productivity across coding, testing, debugging, and documentation while maintaining high standards for code quality, security, and engineering judgment
- Experience using Terraform for Infrastructure as Code
- Strong debugging, troubleshooting, and problem-solving skills
- Excellent communication skills with a strong sense of ownership and the ability to collaborate effectively across teams
- Experience operating highly available production services, including on-call rotations, incident response, and driving operational excellence
Desired Qualifications
- Experience with QEMU/KVM: launching and managing VMs, working with virtual devices like virtio-net and virtio-iscsi, performance profiling, troubleshooting
- Experience with container technologies like Docker or OCI containers and container runtimes such as crio, containerd/runc
- Strong knowledge of Linux internals: namespaces, cgroups (v1 & v2), networking, and storage subsystems
- Hands-on experience with systemd: writing and managing unit files, controlling resource allocation, and understanding service lifecycles
- Ability to set up Linux bridges, veth pairs, and configure custom network namespaces and routing
- Experience with block storage: LUKS encryption, LVM logical volumes, and virtual disk provisioning (e.g., virtio-iscsi)
- Familiarity with security best practices for virtual machines and containers, including disk encryption and secure bootstrapping
- Experience in debugging with tools like strace, tcpdump, or wireshark, and analyzing logs
- Ability to analyze core dumps with gdb
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.