Senior Platform Engineer, Observability and AIOps
Synopsys Inc
We Are
Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis solutions, and design services. We partner closely with our customers across a wide range of industries to maximize their R&D capability and productivity, powering innovation today that ignites the ingenuity of tomorrow.
You Are
You are a strong platform engineer with a passion for building platforms and services that improve how complex infrastructure is observed, understood, and operated. You bring experience developing solutions for observability, automation, and operational intelligence in large-scale enterprise environments. You are comfortable working across software, systems, and operations domains, and you enjoy solving difficult technical problems at scale.
In this role, you will design and develop software and platform capabilities that advance observability, operational analytics, and intelligent automation across Synopsys’ infrastructure ecosystem. You will help improve service visibility, accelerate incident response, reduce operational complexity, and increase infrastructure reliability across environments that support critical engineering workloads.
What You'll Be Doing
- Design, develop, and enhance software solutions that support observability, operational analytics, and intelligent automation across infrastructure and platform services.
- Build scalable, reliable, high-performance systems and services for telemetry collection, processing, searching, correlation, analysis, and visualization.
- Develop tools, APIs, and integrations that enhance monitoring, alerting, incident management, and operational workflow automation.
- Create software capabilities that improve visibility across operating systems, orchestration platforms, compute infrastructure, storage, networking, cloud services, and business-critical enterprise platforms.
- Partner with infrastructure, SRE, platform engineering, and operations teams to identify observability gaps and implement scalable solutions.
- Apply Infrastructure as Code practices to deploy, configure, and maintain observability components in a consistent and repeatable way.
- Apply data-driven techniques, AI-assisted methods, or intelligent analytics to improve signal quality, anomaly detection, alert prioritization, and root cause analysis.
- Document technical designs, implementation patterns, and operating procedures to boost teamwork productivity and efficiency within the organization.
The Impact You Will Have
- Enable faster incident response and resolution across a global hybrid-cloud infrastructure environment that supports mission-critical engineering and business workflows.
- Reduce operational complexity and alert fatigue by building intelligent systems that surface actionable signals instead of noise
- Improve infrastructure reliability and uptime by making it easier for teams to see, understand, and act on what is happening in real time
- Accelerate troubleshooting and root cause analysis by correlating telemetry across compute, storage, networking, and cloud platforms
- Increase operational efficiency by automating repetitive triage, escalation, and remediation workflows
- Empower SRE and platform teams with better tooling, better visibility, and better data so they can focus on high-value work instead of firefighting
- Contribute to a culture of operational excellence where observability and intelligent automation are first-class engineering priorities
What You'll Need
- 8-10 years of experience in software engineering, platform engineering, site reliability engineering, or infrastructure engineering, including substantial experience building observability capabilities.
- Proven experience working in large-scale infrastructure environments with thousands of high-performance compute nodes and/or petabyte-scale storage.
- Strong hands-on experience designing, implementing, and operating observability platforms using technologies such as Elastic, Grafana, Kafka, Logstash, OpenTelemetry, and Prometheus.
- Strong scripting and programming skills in Python, Ruby, or Bash for custom tool and ETL process development.
- Solid working knowledge of Linux systems, Kubernetes, and containerized application environments.
- Experience with Infrastructure as Code and configuration management tools such as Ansible, experience with incident management platforms like ServiceNow, Rootly, or PagerDuty is a plus
- Practical knowledge of AI technologies, including machine learning, generative AI, LLM-based tools, or intelligent analytics, with experience applying them to observability, incident response, automation workflows, or operational decision-making.
- Bachelor's or Master's degree in Computer Science, Information Technology, or a related engineering field.
Who You Are
- You can enter into a room of SREs drowning in alerts and leave with a tooling plan that changes how they understand the problem, not just how they respond to it.
- You do not wait for perfect requirements or a fully defined roadmap, you work with what you have, ask the right questions, and start building.
- You think about the person on call when you design a system, if your alerting logic wakes someone up at 3 AM, it better be for something they can actually fix.
- You are comfortable working across domains, you can debug a Kafka pipeline, tune a Grafana query, write a Python exporter, and still understand the operational workflow it all supports.
- You have a point of view on what good observability looks like and you push back when a solution is too noisy, too fragile, or too hard to maintain.
- You care about making your work reusable and understandable, you document your decisions, write maintainable code, and leave things better than you found them
The Team You'll Be Part Of
You will join the Observability Platform Services team in the Enterprise Intelligence and Cloud Platform (EI&CP) organization, a team focused on building scalable observability, operational intelligence, and automation capabilities across Synopsys' global infrastructure. This team works closely with infrastructure, SRE, platform engineering, and operations teams to improve service visibility, reduce operational complexity, and increase reliability across hybrid cloud environments, GPU-enabled compute farms, enterprise networking, storage platforms, and business-critical systems.
$200k - $322k
We are looking for a highly skilled Senior Software Engineer to design and develop AIOps & Observability platforms at NVIDIA. The platforms are used by internal teams to monitor, diagnose, and optimize the products, millions of assets and services in cloud, on-prem, data...SeniorFull time- ...We Are Synopsys is the leader in engineering solutions from silicon to systems, enabling... ...-because you know that the best platforms are the ones that make the right thing... ...things better than you found them-more observable, more resilient, and easier to operate-...SeniorWork at officeRelocation
- ...seeking a world‑class Principal Engineer (Sr Manager‑equivalent) to... ...Cloud Infrastructure and Platform Engineering (CIPE) organization... ..., testing, deployment, and observability. Your work will directly... ...cloud platforms, mentoring senior engineers and infusing industry...SeniorFull timeWork at office3 days per week
$118k - $212k
...is the Content Intelligence platform shaping the future content economy... ...that agents, tooling, and engineers can all operate them safely.... ..., build in first-class observability — and land it incrementally... ...default. ~ Bonus: LLM-agent / AIOps exposure. The US base...SeniorFull timeLocal areaWork from home$155k - $230k
...safe. Our unified data security platform addresses vulnerabilities in hybrid... ...role: We are looking for a Senior/Staff Infrastructure & Platform Engineer to help architect, build, and evolve... ...systems with strong reliability, observability, resilience, security, and...SeniorTemporary workH1bWorldwide$148k - $235.75k
...impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable... ...Ops.Proven ownership of reliability for an observability/AIOps platform: SLOs/SLIs, on-call, addressing incidents...SeniorFull time- ...Palo Alto Networks, Inc. is seeking a Senior Principal Software Engineer to lead development of tools, platforms, and infrastructure enabling engineering excellence at scale. This role combines deep software engineering expertise with strategic thinking and cross-team...Senior
$170k - $277k
...Job Summary We are seeking a Senior Principal Software Engineer who is first and foremost a software... ...passion for building innovative tools, platforms, and infrastructure that enable... ..., Kubernetes, distributed systems, observability, automation, and platform engineering...SeniorFull timeWork at officeVisa sponsorshipWork visa- ...foundation for physical AI — a unified platform that combines high-quality robotic... .... The Role We are looking for a Senior AI Engineer to design, build, and ship AI-powered... ...definitions, memory, feedback loops, observability, and lifecycle control) that makes agents...SeniorFull time
$201.3k - $352.3k
...It all started when engineer Fred Luddy wrote code that automated... ...reinvention. Our ServiceNow AI platform brings together any AI, any... .... About the Role — Senior Staff Data Platform Software... ...principles including reliability, observability, and production readiness...SeniorFull timeWork experience placementWork at officeImmediate startRemote workFlexible hours2 days per week- ...Senior Platform Software Engineer Role Description: The team builds and operates the platform behind a number of critical system offerings. This developer will work with our Platform team to design and develop tools and platform capabilities. Requirements: ~...Senior
- ...Senior Software Engineer At Commure, we're building the AI Operating System for healthcare, the... ...delivered, documented, and financed. Our platform spans the full care journey: Ambient... ...and platform automation. Improve observability through monitoring, logging, tracing,...SeniorWork at officeLocal areaImmediate start
$120.5k - $243k
...Senior Platform Software Engineer This role has been designed as 'Hybrid' with an expectation that you will work on average 2 days per week from an HPE office. Who We Are: Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people...SeniorWork experience placementWork at officeLocal areaImmediate start2 days per week$168.4k - $193.5k
...operations. We are seeking experience as a mentor, tech lead, or engineering team lead. We require 7+ years of non-internship... ...of SoC subsystems that integrate into our full-system virtual platform for firmware, driver, runtime, and application software teams....SeniorFull timeInternship$168.1k - $261.5k
...operations. We have experience as a mentor, technical lead, or engineering team lead. We have 7 years of non-internship professional... ...of SoC subsystems that integrate into our full-system virtual platform for firmware, driver, runtime, and application software teams....SeniorFull timeInternship$172.5k - $306.63k
...content effortlessly. The AI for Engineering team builds a scalable, production-grade AI platform that powers creativity across... ...around latency, reliability, observability, and cost efficiency.... ...adaptive AI systems. Mentor senior engineers in modern AI system...SeniorTemporary workLocal areaWorldwide$235.03k - $352.29k
...why we're building a universal autonomy platform: self-driving for all roads and all... ...team closely collaborates with autonomy engineers to ensure our labeled data is high-quality... ...distributed systems Experience with observability, monitoring and incident management...SeniorImmediate startFlexible hours$200k - $220k
...This Role: Join Crusoe Energy as a Senior Data Engineer, an early and pivotal hire on our... ...architect and build the foundational data platform infrastructure that powers Crusoe's... ...Implement data validation, monitoring, and observability systems that ensure data integrity...SeniorFull timeTemporary workWork at officeRemote work$176k - $276k
...integrate, and operate the Kubernetes-based platform and shared services used to provision,... ...and upgrades, GitOps delivery, observability, capacity, and service enablement. We... ...environments.We are looking for a hands-on senior engineer to own the lifecycle and automation of...SeniorFull timeRemote workWeekend work$200k - $322k
...impact on the world.Ready to build the platforms that make AI at scale possible? NVIDIA is seeking a Senior Staff Platform Engineer to architect, build, and scale foundational... ...Python, Go, Kubernetes, and Terraform. Observability, capacity analytics, incident learnings,...SeniorFull time$184k - $287.5k
...world.NVIDIA's Silicon Co-Design Group (SCG) is seeking Senior AI Platform Engineers. They will set the technical direction and lead the end-... ...orchestration patterns, authentication and authorization, observability, and SLA enforcement. Driving platform-wide decisions...SeniorFull time$140k - $230k
...technologies; Arene, our software development platform for software-defined vehicles; Woven... ...team is looking for a skilled Software Engineer or Machine Learning Engineer to join... .... ~ Experience with Terraform, AWS, Observability, and Kubernetes in production. ~ Experience...SeniorTemporary workWork at officeFlexible hours3 days per week- ...cybersecurity, physics, mathematics, medicine, engineering, and other specialties. The company... ...The Opportunity We are seeking a Senior Platform Engineer to develop the infrastructure... ...boring, reversible, and auditable. Observability & Infrastructure Quality: Own the...SeniorFull timeSeasonal workRemote workFlexible hours
- ...Platform Project Lead Leads platform projects crossing multiple... ...rollout practices. Drives observability standards, capacity planning... ...provides guidance and coaching to engineers to drive improvements.... ...ensuring review by manager and/or senior technical leaders upon...Temporary workShift work
$245k - $295k
...at Crusoe. About the Role We are seeking a Senior Manager, Infrastructure Platform Engineering to lead a team building core systems that turn large... ...at scale Hands-on background with telemetry and observability platforms at scale (Prometheus, OpenTelemetry,...SeniorTemporary workImmediate start$145k - $182k
...feature flags and cloud costs. The Harness Software Delivery Platform includes modules for CI, CD, Cloud Cost Management, Feature Flags... ...Reliability Management, Security Testing Orchestration, Chaos Engineering, Software Engineering Insights and continues to expand at an...SeniorFull timeLocal areaImmediate startFlexible hours- ...patients worldwide.We’re a team of engineers, clinicians, and innovators... ...building a new AI-powered platform to support how our teams... ...automate. We're looking for a senior full-stack engineer to help... ..., CI/CD, deployment, and observability — asthe platform maturesIterate...SeniorLocal areaWorldwideFlexible hours
$183.6k - $297k
...Cortex is a world-leading cybersecurity platform that provides comprehensive security solutions... ..., including the data pipeline, analytics engine, and user interface. We are a fast-paced... ...in the world. Job Summary As a Senior Principal Backend Engineer in our Cortex...SeniorFull timeWork at office$170k - $277k
...Career We are looking for a visionary Senior Principal Engineer/Architect to serve as the technical authority for our global SRE and Platform Engineering initiatives across the US... .../config deployment frameworks, and observability pipelines scale seamlessly across...SeniorFull timeWork at officeVisa sponsorshipWork visaFlexible hours$104.5k - $234.6k
...smart, motivated, and diverse engineers, empowered with autonomy and... ...matter. As a Senior Principal Software Engineer... ...operating services on public cloud platforms OCI, AWS, Azure, etc. Knowledge... ...breaks. Drive shared observability baselines (SLOs, error budgets...Temporary workFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Platform Engineer, Observability and AIOps. Be the first to apply!
- platform developer Sunnyvale, CA
- platform engineer Sunnyvale, CA
- senior platform engineer Sunnyvale, CA
- senior developer Sunnyvale, CA
- senior aws cloud engineer Sunnyvale, CA
- remote senior salesforce administrator Sunnyvale, CA
- senior marketing operations manager Sunnyvale, CA
- senior manager tax Sunnyvale, CA
- senior tax Sunnyvale, CA
- senior helpdesk technician Sunnyvale, CA




