Oracle Cloud Infrastructure (OCI) delivers mission-critical applications for leading enterprises worldwide. Our cloud offers hyperscale, multi-tenant services deployed across more than 50 regions globally. OCI continues to expand beyond traditional public-cloud boundaries to support dedicated, hybrid, and multicloud solutions, edge computing, and more.
As a Senior Core Infrastructure Engineer, you will build and operate the foundational distributed systems behind OCI. You will develop scalable, resilient components for high-volume data retrieval, storage, and processing, with a focus on correctness, security, observability, and dependable operation at global scale.
This position is office-based and requires onsite presence in Nashville, Tennessee. Relocation assistance may be available in accordance with Oracle's relocation policies.
Internal Responsibilities
What You'll Do
- Design, implement, test, and optimize components of large-scale distributed systems and data-plane services.
- Build for horizontal and vertical scale, using distributed state, replication, synchronization, and high-volume data-processing patterns.
- Strengthen service reliability through redundancy, automatic failover, recovery-oriented design, retries, circuit breakers, and timeouts.
- Develop and execute performance, load, and fault-injection tests to validate scalability, correctness, availability, and operational readiness.
- Create dashboards, alarms, telemetry, runbooks, and automation that proactively detect issues and accelerate recovery.
- Diagnose production issues, participate in on-call rotations and incident response, and contribute to root-cause analyses and durable corrective actions.
- Develop infrastructure automation and Infrastructure as Code (IaC) for safe maintenance, troubleshooting, patching, updates, and rollbacks.
- Apply strong security controls—including encryption, access controls, and vulnerability remediation—in multi-tenant environments.
What You'll Bring
- Bachelor's or master's degree in Computer Science, Computer Engineering, or a related field, or equivalent practical experience.
- 4+ years of professional software-development experience; advanced degree holders may have fewer years of experience.
- Proficiency in at least one object-oriented or systems programming language, such as Java, C++, C#, or Go.
- Experience building or operating scalable, highly available distributed systems or cloud infrastructure.
- Experience with system-level testing and automation, including performance, load, reliability, or fault-injection testing.
- Understanding of distributed-systems concepts, data structures, algorithms, operating systems, networking, and secure software-development practices.
- Strong debugging, problem-solving, communication, and cross-functional collaboration skills.
Preferred Qualifications
- Experience with cloud platforms, such as Oracle Cloud, AWS, Azure, or Google Cloud.
- Experience with data-plane platforms, distributed storage, microservices, state management, replication, or large-scale data processing.
- Experience with observability, incident management, Infrastructure as Code, and production service operations.
- Familiarity with compliance requirements and security controls for cloud infrastructure.
External Responsibilities
What You'll Do
- Design, implement, test, and optimize components of large-scale distributed systems and data-plane services.
- Build for horizontal and vertical scale, using distributed state, replication, synchronization, and high-volume data-processing patterns.
- Strengthen service reliability through redundancy, automatic failover, recovery-oriented design, retries, circuit breakers, and timeouts.
- Develop and execute performance, load, and fault-injection tests to validate scalability, correctness, availability, and operational readiness.
- Create dashboards, alarms, telemetry, runbooks, and automation that proactively detect issues and accelerate recovery.
- Diagnose production issues, participate in on-call rotations and incident response, and contribute to root-cause analyses and durable corrective actions.
- Develop infrastructure automation and Infrastructure as Code (IaC) for safe maintenance, troubleshooting, patching, updates, and rollbacks.
- Apply strong security controls—including encryption, access controls, and vulnerability remediation—in multi-tenant environments.
What You'll Bring
- Bachelor's or master's degree in Computer Science, Computer Engineering, or a related field, or equivalent practical experience.
- 4+ years of professional software-development experience; advanced degree holders may have fewer years of experience.
- Proficiency in at least one object-oriented or systems programming language, such as Java, C++, C#, or Go.
- Experience building or operating scalable, highly available distributed systems or cloud infrastructure.
- Experience with system-level testing and automation, including performance, load, reliability, or fault-injection testing.
- Understanding of distributed-systems concepts, data structures, algorithms, operating systems, networking, and secure software-development practices.
- Strong debugging, problem-solving, communication, and cross-functional collaboration skills.
Preferred Qualifications
- Experience with cloud platforms, such as Oracle Cloud, AWS, Azure, or Google Cloud.
- Experience with data-plane platforms, distributed storage, microservices, state management, replication, or large-scale data processing.
- Experience with observability, incident management, Infrastructure as Code, and production service operations.
- Familiarity with compliance requirements and security controls for cloud infrastructure.