Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Core Infrastructure Engineer

$114.6k - $234.6k

Oracle Corporation

Job Description

We are looking for experienced engineers to join an incubation team responsible for initiating, prototyping, and transitioning new projects with broad impact across OCI.

One of our key long-term initiatives is the development of a new platform designed to power OCI services at cloud scale. The platform encompasses low-level execution runtimes, application and lifecycle management, and high-level change-management workflows , providing OCI developers with the infrastructure and abstractions they need to build and operate services efficiently.

Our goal is simple: enable OCI developers to focus on building innovative services while the platform provides the scalability, reliability, security, and operational capabilities required to run them at cloud scale.

As part of this team, you will challenge existing engineering assumptions, explore new architectural approaches, and apply your expertise in high-performance and reliable systems to help evolve OCI's infrastructure.

What You'll Do

You will lead the development and begin architecting key components of scalable, elastic distributed systems , taking ownership of their performance, reliability, operability, and security.

You will:
  • Design and build distributed systems at cloud scale , defining and enforcing scalability requirements for the components you own.
  • Optimize critical code and data paths for high-throughput, hyperscale workloads, leveraging data-plane platforms for large-scale retrieval, storage, and processing.
  • Architect resilient and fault-tolerant systems using redundancy, replication, failover, and well-defined policies for handling network and infrastructure partitions.
  • Design systems that support in-service upgrades, safe patching, updates, and rollbacks while minimizing customer impact.
  • Apply load shedding, throttling, rate limiting, backpressure, and other resilience mechanisms to maintain service objectives under overload, partial failures, and unreliable network conditions.
  • Define and maintain Service Level Objectives (SLOs), Key Performance Indicators (KPIs), telemetry, dashboards, and proactive alerting to provide clear visibility into system health and performance.
  • Develop sophisticated validation strategies-including fault injection, brownout testing, failure simulation, and resilience testing -to verify system behavior under adverse conditions.
  • Design and improve replication and synchronization mechanisms that preserve correctness, consistency, and durability across distributed components.
  • Proactively investigate and resolve complex production issues, performing deep technical analysis across system boundaries to identify root causes and long-term improvements.
  • Drive operational readiness , ensuring services are observable, supportable, recoverable, and prepared for production at scale.
  • Implement robust security controls and remediation strategies , while maintaining the documentation and processes necessary to satisfy security and compliance requirements.
  • Build Infrastructure as Code (IaC), automation, and deployment tooling that enables repeatable, reliable, and safe infrastructure and service changes.
  • Contribute to architectural direction, technical standards, and engineering best practices while mentoring peers and raising the technical bar across the team .
  • Prototype and evaluate new technologies and architectural approaches, helping transition successful incubations into production systems and broader OCI adoption.
Responsibilities

As a member of the OCI Technical Strategy and Oversight organization, you will help design, build, and operate highly scalable, reliable, and secure distributed systems that support OCI services at hyperscale. You will take ownership of critical components from architecture and implementation through production operations, while contributing to technical direction, engineering excellence, and the growth of the broader team.

System Design & Architecture

Scalability & Performance
  • Lead the design, development, and implementation of components for scalable, elastic distributed systems , supporting both horizontal and vertical scaling as workload demands evolve.
  • Define scalability and performance requirements for owned components and ensure those requirements are reflected throughout design, implementation, testing, and production operation.
  • Optimize critical code paths and system architectures for high-throughput, large-scale data processing and hyperscale workloads .
  • Design systems with elasticity in mind, enabling resources to scale efficiently both up and down based on demand.
  • Leverage distributed state-management technologies and data-plane platforms to support large-scale data retrieval, storage, processing, and coordination.
  • Develop comprehensive performance, scalability, capacity, and load-testing strategies to validate system behavior under expected and extreme workloads.

    Reliability & Resilience
  • Design and build fault-tolerant, highly available systems capable of remaining operational during failures, maintenance, and in-service updates.
  • Implement redundancy, replication, automatic failover, and recovery mechanisms that minimize disruption and protect critical workloads.
  • Design systems to operate predictably during infrastructure and network failures, including partitions and partial dependency failures, while making appropriate tradeoffs among consistency, availability, and partition tolerance .
  • Implement and optimize resilience mechanisms including load shedding, throttling, rate limiting, backpressure, retries, and graceful degradation .
  • Establish appropriate availability and durability expectations through clearly defined Service Level Objectives (SLOs) and use them to guide architectural and operational decisions.
  • Design systems and components to support upgrades and maintenance with minimal or no customer-visible downtime.

    Observability & System Performance
  • Define meaningful Key Performance Indicators (KPIs), telemetry, and health signals to measure system performance, reliability, capacity, and operational health.
  • Build and customize dashboards, telemetry pipelines, monitoring systems, and alerting mechanisms that proactively identify degradation and emerging issues.
  • Use production telemetry and performance data to identify bottlenecks, capacity constraints, and opportunities for architectural improvement.

    Correctness, Durability & Availability
  • Define and implement functional and correctness requirements for complex features, components, and distributed systems.
  • Design sophisticated validation strategies-including fault injection, brownout testing, failure simulation, and resilience testing -to verify system behavior under adverse conditions.
  • Develop data replication and synchronization mechanisms that maintain correctness, integrity, consistency, durability, and availability across distributed components.
  • Identify complex failure modes and incorporate appropriate safeguards into system architecture and implementation.

    Operational Excellence & Incident Management
  • Take a proactive role in diagnosing, debugging, and resolving complex issues across production components and distributed systems.
  • Maintain deep technical expertise in owned systems to support effective troubleshooting, performance optimization, and production operations.
  • Design and implement strategies that enable zero- or minimal-downtime maintenance , reducing or eliminating the need for customer-facing maintenance windows.
  • Ensure systems meet operational-readiness requirements before entering production, including observability, capacity planning, recovery procedures, documentation, and failure handling.
  • Participate in operational support rotations and provide technical leadership during incident response, mitigation, and recovery.
  • Lead or contribute to root cause investigations , identify systemic improvements, and ensure lessons from incidents are incorporated into future designs.
  • Mentor engineers in debugging, incident response, operational practices, and distributed-systems troubleshooting.

    Security & Compliance
  • Design and implement robust security controls for applications and infrastructure operating in multi-tenant cloud environments .
  • Apply appropriate encryption, authentication, authorization, access-control, and data-protection mechanisms.
  • Identify security gaps and execute remediation plans to reduce risk and strengthen system security.
  • Ensure infrastructure and services meet applicable security, compliance, and regulatory requirements.
  • Maintain accurate security and compliance documentation and incorporate security considerations throughout the development lifecycle.

    Automation & Change Management
  • Develop and maintain automation, tooling, and Infrastructure as Code (IaC) to provision, configure, operate, and manage cloud infrastructure reliably at scale.
  • Establish and follow change-management practices for application and infrastructure patching, upgrades, deployments, and rollbacks.
  • Design systems and components that enable these processes to become increasingly automated, repeatable, observable, and safe.
  • Improve deployment and operational tooling to reduce manual intervention, minimize risk, and accelerate recovery when changes do not behave as expected.

    Core Responsibilities

    Planning & Execution
  • Lead and coordinate moderately complex engineering initiatives, managing priorities, dependencies, timelines, and deliverables to ensure successful execution.
  • Provide technical oversight across multiple workstreams while balancing short-term delivery with long-term architectural objectives.
  • Prioritize and delegate work effectively, monitor progress, identify risks early, and adjust execution plans as resources, requirements, or timelines evolve.
  • Drive projects toward completion while maintaining high standards for engineering quality, reliability, security, and operational readiness.

    Collaboration & Partnership
  • Collaborate across engineering teams and organizational boundaries to align technical direction, expectations, dependencies, and shared objectives.
  • Develop a strong understanding of the needs of business leaders, stakeholders, customers, and partner teams to ensure proposed solutions address meaningful requirements.
  • Communicate technical decisions, tradeoffs, risks, and recommendations clearly to both technical and non-technical stakeholders.
  • Foster an inclusive engineering environment by actively seeking diverse perspectives, encouraging constructive discussion, and ensuring team members feel heard and respected.

    Problem Solving & Technical Judgment
  • Analyze complex technical problems using data, system behavior, telemetry, experimentation, and engineering judgment to identify effective solutions.
  • Investigate issues across component and organizational boundaries rather than limiting analysis to individual services.
  • Proactively escalate critical or unresolved issues with a clear assessment of impact, risks, alternatives, and recommended solutions.
  • Document problem-solving approaches, architectural decisions, tradeoffs, and lessons learned to improve organizational knowledge and future decision-making.

    Continuous Learning & Mentorship
  • Continuously expand expertise in distributed systems, cloud infrastructure, reliability engineering, security, automation, and emerging technologies.
  • Stay current with relevant industry trends, technologies, architectural patterns, and engineering best practices.
  • Actively seek and incorporate feedback to strengthen technical and leadership capabilities.
  • Coach and mentor engineers, sharing technical knowledge and helping others develop stronger design, implementation, debugging, and operational skills.
  • Promote knowledge sharing within and across teams.

    Continuous Improvement
  • Identify opportunities to simplify and improve engineering processes, architectures, tools, protocols, and operational workflows.
  • Develop and recommend improvements that increase engineering velocity, reliability, scalability, security, and operational efficiency .
  • Collaborate with partner teams to implement improvements that span organizational or system boundaries.
  • Evaluate the impact of proposed changes on customers, developers, operators, and other stakeholders.
  • Solicit feedback and continuously explore alternative approaches to improve technical and organizational effectiveness.

    Team & Talent Development
  • Contribute to building and strengthening the engineering organization through technical mentorship and knowledge sharing.
  • Participate in candidate interviews, assess technical and problem-solving capabilities, and provide thoughtful hiring recommendations.
  • Help maintain a high engineering bar while supporting the development and success of existing and incoming team members.
Qualifications

Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $114,600 to $234,600 per annum. May be eligible for bonus, equity, and compensation deferral.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC4

About Us

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing or by calling 1- in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Principal Core Infrastructure Engineer in Nashville, TN vacancy
  • $114.6k - $234.6k

     ...IaC and automation that enable safe patching, updates, and rollbacks within change-management plans. We are seeking a Core Infrastructure Engineer to design, build, and operate the distributed systems that power reliable, secure, and scalable cloud services. This... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    8 hours ago
  • $114.6k - $234.6k

     ...Job Description As a Principal Core Infrastructure Engineer within the Networking organization, you will spearhead the development of innovative services in the Networking Automation & Health domain to enable AI super clusters. This initiative involves building a robust... 
    Principal
    Temporary work
    Long distance
    Flexible hours

    Oracle Corporation

    Nashville, TN
    4 days ago
  • $146.3k - $306.4k

     ...innovation in data plane platforms. Engineers and oversees fault-tolerant,...  ...Oracle Cloud Infrastructure (OCI) is building the next generation...  .... We are seeking a Lead Principal Software Engineer to help...  ...automation of these processes. Core Responsibilities Planning... 
    Principal
    Temporary work
    Flexible hours
    Shift work

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $114.6k - $234.6k

     ...Job Description As a Principal Member of Technical Staff, you will own the software...  ...for major components of Oracle's Cloud Infrastructure. You should be both a rock-solid lead developer...  ...generalist and/or skilled Linux engineer with Systems triage experiance able to... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    1 day ago
  • $114.6k - $234.6k

     ...OCI (Oracle Cloud Infrastructure) AI Infrastructure is at the forefront of building a cutting-...  ...What We’re Looking For: Adaptable Engineers: Self-motivated individuals with a...  ...allow for automation of these processes. Core Responsibilities Planning & Execution... 
    Principal
    Full time
    Temporary work
    Flexible hours
    Shift work

    Oracle

    Nashville, TN
    14 days ago
  • $114.6k - $234.6k

     ...masters degree in Computer Science, Computer Engineering, or a related discipline, or equivalent...  ...-scale distributed systems or cloud infrastructure. We need strong experience designing...  ..., and continuous improvement of core distributed systems and data-plane services... 
    Principal
    Full time
    Work at office
    Worldwide
    Relocation
    Relocation package
    Flexible hours

    Oracle

    Nashville, TN
    11 days ago
  • $114.6k - $234.6k

     ...remediation plans to address identified security gaps. -Ensure cloud infrastructure is in compliance with industry standards and regulations and...  ...and components to allow for automation of these processes. Core Responsibilities Planning & Execution: -Manages and... 
    Principal
    Temporary work
    Flexible hours
    Shift work

    Oracle Corporation

    Nashville, TN
    3 days ago
  • $114.6k - $234.6k

     ...networking technologies (routing, switching, overlays, traffic engineering) to architect resilient, scalable, and highly available...  ...Level - IC4 About Us Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from... 
    Principal
    Temporary work
    Worldwide
    Flexible hours

    Oracle Corporation

    Nashville, TN
    4 days ago
  • $114.6k - $234.6k

     ...Job Description Oracle Cloud Infrastructure (OCI) delivers mission-critical applications for leading enterprises worldwide...  ...solutions, edge computing, and more. As a Principal Core Infrastructure Engineer, you will lead the design and evolution of foundational... 
    Principal
    Temporary work
    Work at office
    Worldwide
    Relocation
    Relocation package
    Flexible hours

    Oracle Corporation

    Nashville, TN
    1 day ago
  • $79.2k - $209.5k

     ...while ensuring change, compliance, and documentation standards are met. Responsibilities Responsibilities As a Core Infrastructure Software Engineer, you will: Design, develop, and operate highly available, scalable, and fault-tolerant microservices across... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    8 hours ago
  • $70k - $166.1k

     ...procedures (encryption, access controls, remediation plans) while escalating complex issues to senior engineers. Responsibilities Responsibilities As a Core Infrastructure Software Engineer, you will: Design, develop, and operate highly available, scalable,... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    1 day ago
  • $146.3k - $306.4k

     ...Job Description Join the Oracle Cloud Infrastructure Team! We are seeking a Senior Software Development Manager to join our OCI AI...  ...Oracle Cloud Infrastructure (OCI). Leveraging your expertise in engineering leadership, hardware and software architecture, distributed... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    1 day ago
  • $79.2k - $209.5k

     ...improve security. -Collaborates with the team to ensure cloud infrastructure complies with relevant industry standards and regulations and...  ...for patching, updating, and rolling back applications. Core Responsibilities Planning & Execution: -Track timelines with... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    3 days ago
  • $70k - $166.1k

     ...while escalating complex issues to senior engineers. Responsibilities Key...  ...updating of documentation to ensure cloud infrastructure is in compliance with relevant industry...  ...rolling back applications, under guidance Core Responsibilities Planning & Execution... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $169.8k - $355.4k

     ...security measures. Drives documentation efforts and ensures cloud infrastructure compliance with industry standards and regulations....  ...that system designs allow for automation of these processes. Core Responsibilities Planning & Execution: Oversees and guides... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $114.6k - $234.6k

     ...building and operating large-scale, highly distributed service infrastructure. You should have experience in an operational environment...  ...lead technical direction on high-impact initiatives, mentor engineers, and shape design reviews with simplicity and resilience in mind... 
    Principal
    Full time
    Work at office
    Relocation
    Flexible hours

    Oracle

    Nashville, TN
    6 days ago
  • $79.2k - $209.5k

     ...Staff, you will help build Lightweight Infrastructure (LWI), the next-generation runtime and...  ...experience. ~3-5+ years of software engineering experience building reliable services,...  ...updating, and rolling back applications. Core Responsibilities Planning &... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $79.2k - $209.5k

     ...Job Description Join Oracle Cloud Infrastructure (OCI) as a Senior Core Infrastructure Engineer and play a pivotal role in shaping the future of cloud computing. In this role, you will lead the design, development, and operation of compute operability solutions, ensuring... 
    Temporary work
    Work at office
    Worldwide
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $146.3k - $306.4k

     ...The OCI AI Infrastructure Network Operations team operates and improves the high-performance RDMA/RoCE network fabrics powering OCI’s...  ...telemetry, and performance troubleshooting with strong software engineering and people leadership. You will drive automation,... 
    Full time
    Temporary work
    Flexible hours
    Night shift

    Oracle

    Nashville, TN
    4 days ago
  • $114.6k - $234.6k

     ...automation scripts, tooling, and IaC for troubleshooting and cloud-infrastructure management. Participates in operational-support rotations...  ...Partners with product managers, architects, and engineering teams to translate requirements into technical solutions.... 
    Principal
    Full time
    Temporary work
    Flexible hours

    Oracle

    Nashville, TN
    14 days ago
  • $193.6k - $414.4k

     ...Are you passionate about building cloud infrastructure that powers millions of customers? Do you enjoy leading high-performing engineering organizations that design and operate mission...  ...engineering teams responsible for core virtual networking Infrastructure services... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $169.8k - $355.4k

     ...Job Description Oracle Cloud Infrastructure (OCI) Networking is the foundation that enables...  ...team. Our team develops the core infrastructure that powers OCI Networking...  ...Responsibilities As a Director of Software Engineering, you will lead teams that design,... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $79.2k - $209.5k

     ...Enterprises on Oracle Cloud! Oracle Cloud Infrastructure (OCI) FastConnect is a mission-critical,...  ...at scale! As a Senior Software Engineer, you should have the ability to understand...  ...updating, and rolling back applications. Core Responsibilities Planning & Execution... 
    Temporary work
    Flexible hours
    Shift work

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $79.2k - $209.5k

     ...Job Description The Oracle Cloud Infrastructure (OCI) team builds and manages a suite of massive scale, integrated cloud services in a...  ...become a part of Virtual Networking Control Plane organization, a core OCI team that own and manage networking resources in virtual... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $114.6k - $234.6k

     ...Description Are you interested in building large-scale distributed infrastructure for the cloud? Oracle's Cloud Infrastructure (OCI) team is...  .... ;br Who are we looking for? We are looking for engineers with distributed systems experience. You should have... 
    Principal
    Full time
    Temporary work
    Work at office
    Relocation
    Long distance
    Flexible hours

    Oracle Corporation

    Nashville, TN
    4 days ago
  • $146.3k - $306.4k

     ...leader with expertise and passion leading engineering teams, solving difficult problems in...  ...This is a new ground-up effort to build Infrastructure-as-a-Service that operates at a high...  ...lead a team whose mission is to build core network infrastructure that is highly performant... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    2 days ago
  • $79.2k - $209.5k

     ...synchronization, develops automation and Infrastructure as Code (IaC) for troubleshooting and...  ...with product managers, architects, and engineering teams to translate requirements into technical...  ...and mentors junior team members. Core Responsibilities Tracks... 
    Full time
    Temporary work
    Flexible hours

    Oracle

    Nashville, TN
    12 days ago
  • $79.2k - $209.5k

     ...This role is based out of Nashville,TN The Oracle Cloud Infrastructure (OCI) team builds and manages a suite of massive scale, integrated...  ...a part of Virtual Networking Control Plane organization, a core OCI team that own and manage networking resources in virtual and... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    4 days ago
  • $104.5k - $234.6k

     ...Description We are seeking a software engineer to build and evolve a large-scale...  ...streaming data pipelines, telemetry agents, infrastructure, query capabilities, and user experience...  ...Responsibilities Develop and maintain core APM repositories, shared platform capabilities... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Nashville, TN
    3 days ago
  • $79.2k - $209.5k

     ...reliability. We collaborate with Product Management and partner engineering teams. We participate in code reviews and provide...  ...with Oracle to connect exceptional professionals with this Core Infrastructure Engineer opportunity on the Identity and Access Management... 
    Full time
    Relocation package
    Flexible hours

    Oracle

    Nashville, TN
    11 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Core Infrastructure Engineer. Be the first to apply!