Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Infrastructure Engineer

$146.3k - $306.4k

Oracle

Job Description Mentors teams and leads the architecture of highly scalable, interdependent distributed systems. Identifies and removes performance/scalability bottlenecks for hyper‑scale workloads; defines scalability requirements with stakeholders; and designs elastic, high‑impact systems while advancing innovation in data plane platforms. Engineers and oversees fault‑tolerant, in‑service‑upgradable designs; optimizes resilience mechanisms (load‑shedding, throttling, rate‑limiting); and sets SLO‑aligned durability and availability standards across dependent services. Establishes KPIs and advanced telemetry; applies formal verification for complex features; and develops robust replication/synchronization strategies. Advises and leads resolution of complex production issues, sets operational readiness and SOP standards, and directs incident response and RCAs. Architects advanced security controls, drives remediation and compliance, and delivers enterprise‑level automation (IaC) and change strategies enabling safe, automated patching, updates, and rollbacks. Mentors teams and leads the architecture of highly scalable, interdependent distributed systems. Identifies and removes performance/scalability bottlenecks for hyper‑scale workloads; defines scalability requirements with stakeholders; and designs elastic, high‑impact systems while advancing innovation in data plane platforms. Engineers and oversees fault‑tolerant, in‑service‑upgradable designs; optimizes resilience mechanisms (load‑shedding, throttling, rate‑limiting); and sets SLO‑aligned durability and availability standards across dependent services. Establishes KPIs and advanced telemetry; applies formal verification for complex features; and develops robust replication/synchronization strategies. Advises and leads resolution of complex production issues, sets operational readiness and SOP standards, and directs incident response and RCAs. Architects advanced security controls, drives remediation and compliance, and delivers enterprise‑level automation (IaC) and change strategies enabling safe, automated patching, updates, and rollbacks. Responsibilities Key Responsibilities System Design & Architecture - System Scalability: Mentor the team in the architecture and design of highly scalable, interdependent distributed systems, ensuring horizontal and vertical scalability and overall performance, including leveraging distributed state management tools. Lead the identification of performance and scalability bottlenecks and recommend solutions to optimize code and/or systems for large-scale data processing and high-throughput requirements to improve performance for hyper‑scale systems. Lead collaboration with stakeholders to define system scalability requirements, ensuring the defined requirements meet customer expectations. Leverage deep expertise to design high-impact, interdependent systems to scale with elasticity (e.g., effectively scaling both up and down). Drive innovation in the use of data plane platforms. Evaluate whether systems are meeting nonfunctional scalability requirements, and proactively anticipate growing business needs within the business unit. System Design & Architecture - System Reliability Design: Design and oversee the implementation of fault‑tolerant, interdependent systems capable of withstanding in‑service updates by implementing sophisticated redundancy, replication, and automatic failover capabilities. Lead the design and implementation of systems that effectively handle service disruptions (e.g., network partitions) by prioritizing consistency, availability, or partition tolerance. Guide the optimization of advanced mechanisms to handle network unreliability, including load‑shedding, throttling, and rate‑limiting. Design interdependent systems that are durable and adhere to service level objectives (SLOs), driving standards for availability and durability of other computing services within the organization System Design & Architecture - System Reliability Performance: Define key performance indicators (KPIs) and telemetry to identify risks, gaps, or cyclical dependencies in running, interdependent systems. Drive the creation and customization of highly complex dashboards, telemetry systems, and alerting mechanisms, proactively ensuring system health and reliability. System Design & Architecture - Correctness / Availability: Maintain expertise in industry standards for verifying correctness and apply existing techniques to interdependent systems. Formally verify complex features (e.g., via TLA+) to ensure system design correctness for various interdependent systems. Develop advanced strategies for data replication and synchronization, ensuring robust data integrity and availability Compliance & Security: Architect advanced security measures to protect data and applications in multi‑tenant environments, and lead initiatives to enhance data and application protection. Guide the execution of comprehensive remediation plans to address identified security vulnerabilities. Ensure cloud infrastructure is in compliance with industry standards and regulations, and guide documentation efforts across projects. Automation & Change Management: Develop enterprise‑level automation tools and strategies (e.g., Infrastructure as Code (IaC)) and oversee their implementation. Drive alignment of change management plans and organizational initiatives for patching, updating, and rolling back applications, and design interdependent systems to allow for automation of these processes. Minimum Qualifications Bachelor’s degree in Computer Science or equivalent proven experience 10+ years of experience building and operating large scale, highly available, cloud based distributed systems Specialist skill in a modern programming language such as Java, C, C++, C#, Go, or Python, with proficiency in additional languages preferred Validated understanding of operating system fundamentals Strong understanding of data models and distributed persistence technologies Thorough understanding of the latest security principles, techniques, and protocols Strong troubleshooting and performance tuning skills Proficiency in network, distributed, asynchronous, and concurrent programming Knowledge of professional software engineering standard methodologies for the full software development process Experience building and operating scalable infrastructure software or distributed systems Proven track record to achieve stretch goals in a highly innovative and fast-paced environment Passion for technical leadership and mentoring Strong verbal and written communication skills Strong analytical skills, with excellent problem‑solving abilities Preferred Qualifications Experience in Agile/SCRUM enterprise‑scale software development Experience building automated network and security solutions Knowledge of Machine Learning fundamentals Working familiarity with networking protocols (TCP/IP, and standard network architectures Working familiarity with storage principles, protocols and practices Working familiarity with building secure software using modern security principles Qualifications Disclaimer: Certain U.S. based or U.S. customer or client‑facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements. Range and benefit information provided in this posting are specific to the stated locations only US: Hiring Range in USD from: $146,300 - $306,400 per year. May be eligible for bonus, equity, and compensation deferral. Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following: Medical, dental, and vision insurance, including expert medical opinion Short term disability and long term disability Life insurance and AD&D Supplemental life insurance (Employee/Spouse/Child) Health care and dependent care Flexible Spending Accounts Pre‑tax commuter and parking benefits 401(k) Savings and Investment Plan with company match Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non‑overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation. 11 paid holidays Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours. Paid parental leave Adoption assistance Employee Stock Purchase Plan Financial planning and group legal Voluntary benefits including auto, homeowner and pet insurance The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted. Career Level - IC5 About Us Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life‑saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation‑View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States. Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law. #J-18808-Ljbffr

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Infrastructure Engineer in Santa Clara, CA vacancy
  • $257.4k

     ...Responsibilities As the head of Infrastructure, you will own the vision, execution, and operational excellence for the infrastructure...  ...platforms. You will lead multiple teams spanning platform engineering, SRE, networking/traffic, storage and databases, data infrastructure... 
    Suggested
    Temporary work
    Local area
    Shift work

    Socket.dev

    San Jose, CA
    3 days ago
  • $140k - $224.25k

    The NVIDIA Experience (NVEX) Solutions Engineering team is looking for a senior Computer or Software Engineer who is ready to become...  ...Spectrum-X network systems that interconnect GPUs and AI compute infrastructure.The level of work we do requires a software engineering... 
    Suggested
    Full time
    Weekend work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...jobs every day, improving the efficiency of NVIDIA's software engineers worldwide. We collaborate with teams such as Graphics...  ...systems, and petabytes of storage.Address exciting challenges in infrastructure such as Kubernetes, job scheduling, multi-region services, resource... 
    Suggested
    Full time
    Remote work
    Worldwide

    Nvidia

    Santa Clara, CA
    1 day ago
  • $196k - $310.5k

     ...intelligence. Make the choice, join our diverse team today.We are now looking for a highly motivated and dedicated Senior DFT Infrastructure Engineer to join our DFX group. You will join this multifaceted and innovative DFX team to develop the next generation of scalable,... 
    Suggested
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...This is our life’s work, to amplify human inventiveness and intelligence.NVIDIA is seeking top-tier Senior Deep Learning Infrastructure Engineers. In this role, you will play a crucial part in building the deep learning infrastructure that powers our AI agents for VLSI... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $120k - $155k

    DescriptionNote- This role focuses on IT infrastructure and systems automation, not software or QA test automation. Candidates with backgrounds...  ...will not be considered.We are seeking a Senior IT Automation Engineer to lead enterprise automation initiatives and improve... 

    Omnivision Technologies

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    We are looking for a senior systems software engineer to improve the operation and user experience of distributed system infrastructure using AI. We are passionate about the opportunity to shape the future of the NVIDIA software platform! We want to develop next-generation... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

    NVIDIA is looking for a top-tier Software Test Engineer to join the NVIDIA-Cumulus Linux Verification Engineering Team! You will play an exciting role that allows you to lead verification of groundbreaking features of NVIDIA-Cumulus Linux and take full ownership of tasks... 
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $146.7k - $339.3k

    What you can expectThis role leads global network infrastructure and data center operations supporting over 10,000 employees and AI/ML training...  ...manager oversees a team of 12 globally distributed network engineers and data center technicians. The team is operating in a... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    8 hours ago
  • Company DescriptionSonsoft , Inc. is a USA based corporation duly organized under the laws of the Commonwealth of Georgia. Sonsoft Inc. is growing at a steady pace specializing in the fields of Software Development, Software Consultancy and Information Technology Enabled...
    Full time

    Sonsoft

    Cupertino, CA
    4 days ago
  • $224k - $356.5k

     ...more details, see We are looking for a Senior System Software Engineer for Cloud who sees the big picture of Cloud Computing and is...  ...and drive best practices in Kubernetes, observability, and infrastructure automation.What we need to see:BS or MS in Computer Science or... 
    Full time
    Local area

    Nvidia

    Santa Clara, CA
    8 hours ago
  •  ...seeking an experienced and bilingual (Spanish/English) Deployment Engineer specializing in network deployments with strong expertise in...  ..., ensuring reliable and optimized performance across client infrastructures. The ideal candidate will have hands-on experience with... 

    Right Hire Consulting LLC

    Santa Clara, CA
    3 days ago
  • $200k - $220k

     ...technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:...  ...lead the design and development of next-generation network infrastructure solutions optimized for AI workloads. This role requires deep... 
    Worldwide

    Supermicro

    San Jose, CA
    2 days ago
  • Partners with customers, sales, engineering and product teams to design, demonstrate and deploy Oracle Cloud architectures that address...  ...design and deployment.Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from... 
    Flexible hours

    Oracle Corporation

    San Jose, CA
    1 day ago
  • $152k - $241.5k

    NVIDIA is building the next generation of production workflow infrastructure for large-scale chip engineering. This platform turns intent, layered configuration, generated files, tool execution, distributed jobs, validation checks, and shared project state into observable... 
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $168k - $264.5k

     ...people work and play. NVIDIA is seeking a Senior Network Deployment Engineer to help build and scale our global network. In this role, you’ll be responsible for sourcing, procuring and delivering infrastructure and circuits to support Point-of-Presence (PoP) deployments.... 
    Full time
    Contract work
    Work experience placement
    Remote work
    Shift work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

     ...Autonomous Vehicles Platform team is seeking a Senior System Software Engineer to help bring NVIDIA's autonomous vehicle platform to new...  ...be doing:Drive the expansion of hardware-in-the-loop (HIL) infrastructure to support the robust deployment and lifecycle management of... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...hardware design and review HW architecture & schematics.What we need to see:A Bachelor of Science Degree (or higher) in Electrical Engineering or Computer Science or equivalent experience.8+ years of experience.Domain expertise in BMC Firmware development on X86 or ARM... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

    Our Autonomous Vehicles Platform team is searching for engineers to develop and bring NVIDIA's automotive platform out to the world. You will participate in a focused effort to develop and productize ground-breaking solutions that will revolutionize the world of transportation... 
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  • $184k - $287.5k

     ...Vehicles Platform team is looking for a hands-on System Software Engineer. As part of our team, you will work on our Autonomous Driving...  ...the stack, from platform and embedded software to cloud infrastructure, underpinned by safety and performance. It extends an opportunity... 
    Full time

    Nvidia

    Santa Clara, CA
    8 hours ago
  •  ...we shape the future of AI and beyond. Together, we advance your career. THE ROLE: We are looking for a dynamic, upbeat system test engineer to join our growing team within NTSG - Network Technology Solutions Group. As a key contributor you will be part of a leading team... 

    AMD

    Santa Clara, CA
    1 day ago
  •  ...global scale, come make a difference at Fiserv.Job TitleSenior Infrastructure Security EngineerJob TitleSenior Infrastructure Security EngineerWhat does a successful Senior Infrastructure Security Engineer do at Fiserv?Excited about shaping the future of fintech? Join... 
    Full time
    Temporary work
    H1b

    Fiserv

    Sunnyvale, CA
    1 day ago
  • $130k - $148k

    Staff Infrastructure Engineer OverviewThe IT Infrastructure Engineer is responsible for designing, implementing, maintaining, and supporting core enterprise infrastructure services across on‑premises and cloud environments. This role ensures high availability, security,... 

    Legence Holdings

    San Jose, CA
    1 day ago
  • $134.5k - $193.5k

     ...challenges of the 21st century.  We are looking for a  Senior Network Engineer to join our team in one of today’s most exciting technologies...  ..., and field environments.Operate and maintain network infrastructure from major vendors, ensuring high availability and... 
    Full time
    Work at office
    Worldwide

    Bloom Energy

    San Jose, CA
    8 hours ago
  • Sr. Network Engineer (Serviceability)This role has been designed as ‘’Onsite’ with an expectation that you will primarily work from an HPE office.Who We Are:Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies... 
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    Remote work

    Hewlett Packard Enterprise

    San Jose, CA
    2 days ago
  •  ...as outside of his/her domainWorks closely with other network engineers, review, execute technical designs and provide technical direction...  ...with key business initiatives.Informed and involved in all infrastructure disruption with relevancy in Data Center to deliver a Global... 
    Work experience placement

    NTT DATA

    San Jose, CA
    8 hours ago
  • $105.3k - $175.21k

     ...Information Security organization is seeking a Network Security Engineer. The candidate chosen for this role will assist senior...  ...F5.o Configuring implementing networks and Network security infrastructure, with network design and troubleshooting, Software Defined Network... 
    Full time
    Internship
    Work at office
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

     ...to advance and implement these technologies into our future offerings, our Compiler team is growing and seeking top-tier compiler engineers who want an exciting and engaging role that bridges compute and networking while leading the charge to even greater accomplishments... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  •  ...RoleWe are looking for a Network Architect to join our Cluster Engineering Team and help shape the front-end datacenter and...  ...fabricAutomate the deployment, configuration, and validation of network infrastructure using Python, including topology provisioning, fabric bring-... 

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  •  ...Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range...  ...be part of day2 operations and on-call rotation for Network Engineering teamYouHave 10+ years of experience in IT and networking spaceHave... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Infrastructure Engineer. Be the first to apply!