More from this employer
Harwood, USA
Nashville, USA
Oracle Cloud Infrastructure (OCI) is seeking a highly motivated Software Developer 4 to join the Infrastructure Planning and Capacity Management organization. This team develops the platforms, services, workflows, and op
Nashville, USA
Oracle Cloud Infrastructure (OCI) is seeking a motivated Software Developer 3 to join the Infrastructure Planning and Capacity Management organization. This team develops the platforms, services, workflows, and operation
Santa Clara, USA
Oracle Cloud Infrastructure (OCI) is seeking a hands-on Materials Supply Chain Management, Principal (IC4) to support materials planning, manufacturing coordination, and supply execution across OCI’s global cloud infrast
, USA
, USA
Phoenix, USA
Harwood, USA
, USA
Nashville, USA
Nashville, USA
Santa Clara, USA
Base Career helps you apply smarter for this job.
Within OCI, the Technical Strategy & Oversight organization builds foundational systems for OCI’s most demanding services.
One of its boldest initiatives is Autonomous OCI: a greenfield effort to build a cloud platform that can operate and scale with far less manual work.
Today, every new region and service feature adds operational cost and complexity.
Autonomous OCI is designed to change that by making services fast, resilient, and self-healing by default—helping OCI scale to thousands of regions without growing operations teams at the same rate.
As a Senior Member of Technical Staff, you will help build Lightweight Infrastructure (LWI), the next-generation runtime and platform layer for compute, storage, networking, identity, observability, and governance.
Designs, implements, and optimizes components in distributed systems with an emphasis on scalability, resiliency, and operability.
Delivers features and load/performance tests; leverages data plane platforms and distributed state tools for high-volume retrieval, storage, and processing; and reviews peers’ implementations for scalability compliance.
Builds fault-tolerant paths (redundancy, replication, automatic failover), applies recovery‐oriented principles, and implements retries, circuit breakers, and timeouts.
Proactively detects and mitigates issues via tests, alarms, dashboards, and telemetry; authors runbooks and participates in incident response and RCAs.
Implements standard replication and synchronization, develops automation/IaC for troubleshooting and maintenance, and applies advanced security controls (encryption, access, remediation) while ensuring change, compliance, and documentation standards are met.
Cloud infrastructure and enterprise software solutions provider.
Visit company websiteJobs and hiring trendsSkip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career