Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Lead Principal Software Engineer, Core Infrastructure

$135.2k - $306.4k

Oracle Corporation

Job Description

The Oracle Cloud Infrastructure (OCI) team offers the opportunity to build and operate massive-scale, integrated cloud services in a broadly distributed, multi-tenant cloud environment. OCI builds cloud products for customers who are tackling some of the world's largest technical and business challenges.

Oracle Kubernetes Engine (OKE) is OCI's managed Kubernetes service. OKE enables customers to create, run, scale, secure, and operate Kubernetes clusters on OCI, integrating Kubernetes with OCI compute, networking, storage, identity, observability, security, and automation. The OKE team owns a highly available 24x7 cloud service and is expanding the platform to support larger clusters, higher scale, improved operability, deeper OCI integrations, and increasingly demanding cloud native, AI, and GPU workloads.

We are looking for a senior IC5 software engineer with deep Kubernetes expertise, required cloud infrastructure experience, and a strong distributed systems background. This is a high-impact technical leadership role for an engineer who can define architecture, drive cross-team execution, solve ambiguous production and platform problems, and deliver durable systems that improve both customer experience and operational excellence.

You will work on core OKE platform capabilities including cluster lifecycle management, orchestration, scalability, reliability, performance, automation, observability, security, and integration with OCI infrastructure services. The ideal candidate has hands-on experience designing, building, operating, or deeply debugging production cloud services, infrastructure platforms, or Kubernetes-based systems at meaningful scale.

This role requires advanced Kubernetes experience, including Kubernetes control plane behavior, controllers and operators, scheduling, autoscaling, networking, storage, service discovery, container runtimes, node lifecycle, Kubernetes APIs, and etcd. Experience with Kubernetes networking and storage technologies such as CNI, Cilium, Calico, Flannel, other container networking implementations, CSI drivers, and cloud provider integrations is highly relevant.

OKE is also expanding to support demanding AI and accelerated computing use cases. Experience with AI/ML infrastructure, multi-node GPU clusters, accelerated compute, model training or inference platforms, GPU scheduling, device plugins, Karpenter, cluster autoscaling, CUDA, NCCL, RoCE, InfiniBand, RDMA, SmartNIC/DPU offload, or high-performance AI/HPC networking is a significant plus.

This role also requires an engineer who is ready to use modern agentic engineering practices responsibly. We expect senior engineers to apply AI-assisted and agentic workflows to accelerate design exploration, implementation, testing, debugging, documentation, operational analysis, and developer productivity while maintaining strong ownership, security judgment, code quality, and production accountability.

Responsibilities

As a member of the software engineering division, you will take an active role in defining and evolving standard practices and procedures. You will define specifications for significant new projects and specify, design, develop, troubleshoot, and debug software for OCI's managed Kubernetes service.

Responsibilities include:
  • Provide technical leadership for major OKE platform initiatives from architecture through implementation, launch, and production operation.
  • Design and build distributed systems that create, update, scale, repair, and operate Kubernetes clusters across OCI regions.
  • Improve OKE reliability, scalability, performance, upgrade safety, lifecycle management, observability, automation, and operational tooling.
  • Work deeply with Kubernetes technologies, including control plane components, controllers/operators, scheduling, autoscaling, Kubernetes APIs, container runtimes, node behavior, and etcd.
  • Design, debug, and improve Kubernetes networking and storage integrations, including CNI-based networking, Cilium, Calico, Flannel, other container networking implementations, CSI drivers, and OCI infrastructure integrations.
  • Build automation for cluster validation, health checks, readiness testing, failure detection, remote recovery, and reduction of post-deployment operational issues.
  • Lead technical design reviews, code reviews, incident reviews, and production readiness reviews for complex service changes.
  • Debug difficult production issues across service boundaries, including Kubernetes, Linux, networking, compute, storage, identity, telemetry, and OCI infrastructure dependencies.
  • Apply performance engineering practices including profiling, tracing, latency analysis, throughput optimization, and production diagnostics across distributed systems.
  • Build automation that reduces manual operations, improves fleet health, accelerates diagnosis, and raises the quality bar for OKE engineering.
  • Partner with OCI service teams to deliver end-to-end platform capabilities regardless of organizational boundaries.
  • Apply AI-assisted and agentic engineering workflows to improve engineering velocity, test coverage, debugging, operational analysis, and documentation while ensuring correctness, security, and maintainability.
  • Mentor engineers, influence technical direction, and help establish patterns that scale across the OKE organization.
  • Participate in operating a 24x7 cloud service and use customer feedback, production data, and operational experience to prioritize improvements.
Required qualifications:
  • 10+ years of software engineering experience, or equivalent experience building and operating production software systems.
  • Hands-on cloud infrastructure experience is required, ideally designing, building, operating, or debugging production services or platforms on OCI, AWS, Azure, Google Cloud Platform, or a large-scale private cloud.
  • Strong hands-on Kubernetes expertise is required, including Kubernetes architecture, APIs, control plane behavior, controllers/operators, scheduling, autoscaling, networking, storage, nodes, cluster lifecycle management, or production cluster operations.
  • Advanced Kubernetes knowledge, including CNI, CSI, etcd, service discovery, container runtimes, node lifecycle, and Kubernetes failure modes.
  • Experience with Kubernetes networking technologies such as Cilium, Calico, Flannel, or other CNI implementations.
  • Experience with Kubernetes storage integrations, including CSI drivers or cloud storage integrations.
  • Strong distributed systems fundamentals, including availability, failure handling, performance, scalability, and operational tradeoffs.
  • Experience building highly available infrastructure services, platform services, or cloud native systems used in production.
  • Strong development experience in both Go/Golang and Java is required.
  • Strong Linux, networking, debugging, and production operations skills.
  • Demonstrated ability to lead ambiguous technical projects, influence across teams, and deliver through other engineers without relying on formal authority.
  • Strong communication skills, ownership, judgment, and ability to make pragmatic tradeoffs in production systems.
Preferred qualifications:
  • Experience with AI/ML infrastructure, GPU workloads, multi-node GPU clusters, accelerated compute, model training or inference platforms, GPU scheduling, device plugins, Karpenter, cluster autoscaling, CUDA, NCCL, high-performance networking, or distributed training systems.
  • Experience with eBPF-based networking, Kubernetes network policy, service mesh, ingress, load balancing, overlays/underlays, BGP, VXLAN, SmartNIC/DPU offload, RoCE, InfiniBand, RDMA, or multi-cluster networking.
  • Experience with infrastructure as code and cloud provisioning tools such as Terraform, Packer, cloud-init, IAM, VCN/VPC networking, VPN, FastConnect/Direct Connect, or equivalent cloud primitives.
  • Experience building developer productivity, operational automation, or responsible AI-assisted and agentic engineering workflows.
  • Experience with observability systems, incident response, safe deployment practices, canary analysis, rollback strategies, service health automation, and large fleet operations.
  • Open-source or upstream contribution experience in Kubernetes, cloud native infrastructure, observability, networking, or related systems.
Qualifications

Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $135,200 to $306,400 per annum. May be eligible for bonus, equity, and compensation deferral.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC5

About Us

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing or by calling 1- in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Lead Principal Software Engineer, Core Infrastructure in Seattle, WA vacancy
  • $96.8k - $306.4k

     ...Description At Oracle Cloud Infrastructure (OCI) , we are building the...  ...digital twins, simulation, software lifecycle management, and...  ...integration from a manual engineering effort into a scalable...  ...autonomous robot fleets. As a Lead Principal Platform Software Engineer... 
    Suggested
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    7 hours ago
  • $135.2k - $306.4k

     ...Job Description As a Sr. Principal Software Development Engineer in the Oracle Cloud Infrastructure (OCI) Core Platform division, you will play a critical leadership role...  ...availability, and compliance . You will lead technical initiatives across all layers of... 
    Suggested
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    3 days ago
  • $96.8k - $306.4k

    Oracle Cloud Infrastructure (OCI) Networking is the foundation...  .... We're looking for engineers with strong computer...  ...team develops the core infrastructure that...  ...at a company leading the way in AI and cloud...  ...ResponsibilitiesAs a Lead Principal Platform Software Engineer, you will:... 
    Suggested
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    4 days ago
  • $114.6k - $234.6k

    At Oracle Cloud Infrastructure (OCI), we are building the future...  ...of one of the world’s leading enterprise technology...  ...challenges. Engineers at OCI have deep technical...  ...service that supports core OCI services across regions...  ...services. As a Software Engineer on this team,... 
    Suggested
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    4 days ago
  • $135.2k - $306.4k

     ...Future of Cloud at Oracle Cloud Infrastructure (OCI) At Oracle Cloud...  ...What You'll Be Building and Leading End-to-End Solutions:...  ...Frameworks: Spearhead the engineering of new container runtimes and...  ...standards and drive adoption of core data-plane components and... 
    Suggested
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    1 day ago
  • $13 per hour

     ...keeping Salesforce's core values at the heart of...  ...at the company leading workforce transformation...  ...experienced Senior, Lead, and Principal Engineers to serve as a key...  ...ETL pipeline infrastructure to support a daily, self...  ...ship production-grade software using modern engineering... 
    Full time

    Salesforce

    Seattle, WA
    3 days ago
  • $251k - $352k

    Senior Principal Engineer, Infrastructure Base pay range $251,000.00/yr - $352,000.00/yr At Docker, we make...  ...Cross‑Company Technical Leadership Lead large cross‑company programs that...  ...Required Technical Expertise 15+ years of software engineering experience with... 
    Full time
    Contract work
    Immediate start
    Remote work

    Docker

    Seattle, WA
    4 days ago
  •  ...working for one of the world's leading financial institutions, you’ve come to the right place. As a Principal Engineer at JPMorgan Chase on the Core AI Infrastructure Platform team, you will design...  ...skillsFormal training or certification on software engineering concepts and 10+... 

    JP Morgan Chase

    Seattle, WA
    7 hours ago
  •  ...Principal Platform Software Engineer As OCI customers increasingly adopt micro services...  ...than ever. Kafka, the leading message broker of choice...  ...support, and reliability. Core Responsibilities:...  ...multi-tenant, virtualized infrastructure a strong plus • Design, develop... 
    Temporary work
    Shift work

    Hackajob

    Seattle, WA
    3 days ago
  •  ...Principal Engineer, Developer Platform Engineering If you...  ...at one of the world's leading financial institutions...  ..., including Core Web Vitals, bundle budgets...  ...of tools within the Software Development Life Cycle...  ...autoscaling strategies, infrastructure as code, and cost-per... 
    Early shift

    Chase

    Seattle, WA
    2 days ago
  • $104.5k - $234.6k

     ...Responsibilites: ~ Lead engineering and operational management for...  ...Responsibilities Platform Software Development: Lead cross-team...  ..., and reliability. Core Responsibilities Planning...  ...Oracle brings together the data, infrastructure, applications, and expertise... 
    Temporary work
    Flexible hours
    Shift work

    Oracle Corporation

    Seattle, WA
    4 days ago
  • $61k - $101k

     ...formal training or certification in software engineering concepts, along with 5+ years of applied...  .... We require experience with Infrastructure as Code. We need a deep understanding...  ...We require demonstrated experience leading the effective use of enterprise-approved... 
    Full time
    For contractors

    J.P. Morgan

    Seattle, WA
    4 days ago
  •  ...Lead Software Engineer We have an opportunity to impact your career and provide an adventure...  ...Engineer at JPMorgan Chase within the Core Foundational Platforms team, you are...  ...capabilities, and skills Experience in Infrastructure Architecture designs Experience in... 

    Chase

    Seattle, WA
    2 days ago
  • Lead Software Engineer We have an opportunity to impact your career and provide an adventure where...  ..., stable, and scalable way. As a core technical contributor, you are responsible...  ...experience In-depth knowledge of Infrastructure as Code such as Terraform Demonstrated... 

    Chase

    Seattle, WA
    5 days ago
  • $172.6k - $259k

     ...develop our worlds. We are seeking a Principal Software Engineer who can guide product ideas through every...  ...through data-driven insights. Lead product ideas as testable hypotheses....  ...just review it. Strong systems and infrastructure foundation. AWS at scale, containerization... 
    Full time
    Work experience placement

    Hasbro, Inc.

    Renton, WA
    more than 2 months ago
  • $200k - $240k

     ...orbit. Backed by $450M from leading investors including...  ...different class of spacecraft. Engineered to survive the harshest radiation...  ...apply.   The Role The software team at K2 strives to blur the...  ...Build and maintain infrastructure to increase reliability when... 
    Permanent employment
    Full time
    Shift work

    K2 Space Corporation

    Seattle, WA
    more than 2 months ago
  • $163.9k - $270.88k

     ..., your work matters—and so do you. Principal Software Engineer-ENG: This Principal Software Engineer...  ..., and regulatory compliance. - Lead architecture and design reviews; set standards...  ...such as service decomposition, .NET Core enablement, build modernization, and... 
    Full time
    Immediate start

    Ukg

    Seattle, WA
    a month ago
  •  ...as a high-trust collaborator that is core to how you solve problems and accelerate...  ..., and data marketplace. AS A PRINCIPAL SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Solve real business...  ...AT SNOWFLAKE? Build an industry-leading Cloud Data and AI Platform. Solve challenging... 
    Full time

    Snowflake

    Bellevue, WA
    more than 2 months ago
  • $188.7k - $258.39k

     ...committed to building breakthrough software with a spark of magic. We believe a...  ...the Role We are looking for a Principal Software Engineer to join our Applied AI team. This is...  ...Build evaluation and observability infrastructure for non-deterministic systems – including... 
    Full time
    Immediate start
    Flexible hours

    Highspot

    Seattle, WA
    more than 2 months ago
  • $114.6k - $234.6k

     ...defines integration contracts across software, firmware, and hardware. Leads deep-dive investigations for...  ...provides guidance and coaching to engineers to drive improvements. Utilizes...  ...health, support, and reliability. Core Responsibilities Planning & Execution... 
    Temporary work
    Shift work

    Hackajob

    Seattle, WA
    3 days ago
  • Lead Software Engineer Are you passionate about building resilient, scalable systems that power the...  ...— with operational excellence at the core. Job responsibilities Design and...  ...observability, resilience, security controls, infrastructure management, and cost optimization... 

    Chase

    Seattle, WA
    5 days ago
  • Senior Lead Software Engineer, Cloud Platform Are you ready to shape the future of technology at a global financial leader? Join us and make...  ...and development of secure, scalable, and reliable cloud infrastructure and platform tools. Drive adoption of modern DevEx (Developer... 
    Work at office
    Shift work

    Chase

    Seattle, WA
    5 days ago
  • Senior Lead Software Engineer Be an integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top...  ...Engineer at JPMorgan Chase within the Corporate Sector, Infrastructure Platforms team, you are an integral part of an agile team... 
    For contractors

    Hackajob

    Seattle, WA
    5 days ago
  • Senior Lead Software Engineer Are you ready to shape the future of AI-driven reliability engineering? At JPMorganChase, we're building intelligent...  ...across telemetry pipelines, application stacks, and cloud infrastructure — driving engineering productivity, system reliability,... 

    Hackajob

    Seattle, WA
    5 days ago
  • $191k - $253k

     ...partnerships, we’re not just building software—we’re revolutionizing how...  ...for a Senior Software Engineer to join our rapidly growing...  ...integration process Develop infrastructure to simplify the exposure of...  ...engineering teams and leading technical design reviews.... 
    Full time
    Work experience placement
    Immediate start
    Worldwide

    Anduril Industries

    Seattle, WA
    more than 2 months ago
  •  ...like us, we can offer you the ultimate career opportunity that will light a fire within you. Lead Software Engineer The Lead Software Engineer designs and evolves core services within the NICE CXone platform, building scalable cloud systems that support enterprise... 
    Full time
    Worldwide

    Nice

    Seattle, WA
    more than 2 months ago
  • $163.9k - $235.55k

     ...We are searching for an exceptionally skilled and visionary Principal Software Engineer with deep expertise in building, deploying, and scaling...  ...term mission. The Principal Software Engineer, AI/ML will lead the technical vision, strategy, and architectural direction... 
    Full time
    Local area

    Ukg

    Seattle, WA
    a month ago
  •  ...the tools that define how software gets built and delivered....  ...verified images, and secure infrastructure that make autonomous...  ...We’re looking for a Principal Backend Engineer who thrives at the intersection...  ...developer experience. You’ll lead the technical direction of... 
    Full time
    Temporary work
    Remote work
    Home office
    Shift work

    Docker

    Seattle, WA
    more than 2 months ago
  • $79.2k - $209.5k

     ...standards are met. Oracle Cloud Infrastructure Workflow is a Tier 0 service...  ...a fault tolerant manner. An engineer on this team is responsible...  ...back applications. Core Responsibilities Planning...  ...your potential at a company leading the way in AI and cloud solutions... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    7 hours ago
  • $79.2k - $209.5k

     ...will help build Lightweight Infrastructure (LWI), the next-generation...  ...experience. ~3-5+ years of software engineering experience building...  ...rolling back applications. Core Responsibilities Planning...  ...your potential at a company leading the way in AI and cloud solutions... 
    Temporary work
    Flexible hours

    Oracle Corporation

    Seattle, WA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Lead Principal Software Engineer, Core Infrastructure. Be the first to apply!