Oracle logo
OraclePosted 1 month ago

Senior Core Infrastructure Engineer (OCI Object Storage)

On-siteNashville, Tennessee, United States

Full TimeSenior LevelSmallCloud Services

Job Summary

Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability. Delivers features and load/performance tests; leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing; and reviews peers' implementations for scalability compliance. Builds fault-tolerant paths (redundancy, replication, automatic failover), applies recovery-oriented principles, and implements retries, circuit breakers, and timeouts. Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry; authors runbooks and participates in incident response and RCAs. Implements standard replication and synchronization, develops automation/IaC for troubleshooting and maintenance, and applies advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met.

Required Qualifications

  • Nashville, TN location
  • hands-on engineers with expertise and passion in solving difficult problems in distributed systems
  • large scale storage
  • highly available services
  • expertise and passion in solving difficult problems in distributed systems
  • large scale storage
  • highly available services
  • familiarity of distributed systems
  • value simplicity and scale
  • work comfortably in a collaborative, agile environment
  • be excited to learn
  • rock-solid coder
  • Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability
  • Delivers features and load/performance tests
  • leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing
  • reviews peers' implementations for scalability compliance
  • Builds fault-tolerant paths (redundancy, replication, automatic failover)
  • applies recovery-oriented principles
  • implements retries, circuit breakers, and timeouts
  • Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry
  • authors runbooks
  • participates in incident response and RCAs
  • Implements standard replication and synchronization
  • develops automation/IaC for troubleshooting and maintenance
  • applies advanced security controls (encryption, access, remediation)
  • ensuring change, compliance, and documentation standards are met

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce