Staff Engineer - Systems
On-siteChennai, Tamil Nadu, India
Job Summary
Architect, pioneer, and own next-generation, multi-tenant cloud-native distributed infrastructure from zero to scale, establishing design patterns that serve as the blueprint across the entire engineering organization. Drive the long-term roadmap for core systems, transforming ambiguous business visions into concrete, scalable technical realities while leading deep-dive architectural refactoring and performance tuning across the entire stack. Set the technical standard for mission-critical systems by engineering multi-region fault tolerance, advanced disaster recovery, and deep observability to guarantee 99.999% availability. Guide and grow senior technical talent through foundational design reviews and steering committee contributions, partnering with leadership to translate product roadmaps into long-term infrastructure capacity.
Required Qualifications
- 9+ years of progressive experience in software systems engineering
- Proven track record of advancing into staff-level or principal-level architectural responsibilities within high-growth product organizations
- Verifiable history of designing and shipping high-scale, enterprise-grade SaaS infrastructure that handles massive transactional throughput
- History of successfully surviving major architectural evolutionary cycles
- Master's or Bachelor's degree in Computer Science, Systems Engineering, or a highly quantitative technical field
Desired Qualifications
- Master-level fluency in backend systems languages (e.g., Go, Java, Rust, C++)
- Experience with concurrent programming
- Experience with asynchronous execution frameworks
- Experience with low-latency microservices design
- Deep understanding of distributed systems theory (e.g., CAP theorem, consensus protocols like Raft/Paxos, replication, partitioning, and sharding)
- Expert-level knowledge of data modeling
- Expert-level knowledge of transaction isolation levels
- Expert-level knowledge of tuning distributed storage systems (Relational, NoSQL, NewSQL)
- Expert-level knowledge of distributed caches like Redis
- Expert-level knowledge of message buses like Kafka/Pulsar
- Hands-on experience architecting vector databases
- Experience optimizing hardware accelerators (GPUs/TPUs)
- Experience building backend infrastructure optimized for AI model hosting and routing
- Expert capability in designing structured logging
- Expert capability in designing distributed tracing
- Expert capability in designing metric aggregation models across vast distributed fleets
- Exceptional communication skills with a proven ability to explain complex technical trade-offs to non-technical executives
- Proven ability to inspire and align diverse engineering groups
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.