Observability & SRE Lead — Multi-Cloud Network Reliability
$114k - $253kLam Research Corporation
LAM RESEARCH Corporation is looking for an Observability Lead to join their GIS Infrastructure Platform Engineering team in Fremont, California. The ideal candidate will lead engineers in implementing observability frameworks and managing end-to-end networks across Azure, AWS, and GCP, ensuring reliability and resilience. The role requires a strong background in Site Reliability Engineering and multi-cloud solutions, with at least 12 years of experience in related fields. Salary range for this position is $114,000 - $253,000 annually, depending on experience and location. #J-18808-Ljbffr LAM RESEARCH Corporation
$114k - $253k
...their business objectives. The impact you’ll makeOur team at Lam is seeking a hands-on Observability Lead with a strong Site Reliability Engineering (SRE) and multi-cloud networking foundation to join our GIS Infrastructure Platform Engineering team. You will lead engineers...CloudNetworkLocal areaRemote workFlexible hours2 days per week3 days per week1 day per week- Lam Research is seeking a hands-on Observability Lead to lead a global SRE and Platform Engineering team. You will own observability, reliability, and multi-cloud network initiatives across Azure, AWS, and GCP, driving incident response and automation. You will define SLAs...CloudNetwork
$114k - $253k
...big semiconductor breakthrough. We lead the way in one of the most... ...world together, anything is possible. Observability Lead - Cloud SRE & Network Reliability Date: Jul 21, 2026 Location: Fremont... ...Reliability Engineering (SRE) and multi-cloud networking foundation to join...CloudNetworkLocal areaRemote workFlexible hours2 days per week3 days per week1 day per week$167.7k - $245.2k
...as intended, improving reliability and reducing risks.... ...applications with enhanced observability and control.As a... ...Reliability Engineer (SRE), you will build, operate... ...backbone supporting both cloud and air-gapped... ...spanning Kubernetes, networking, storage, and application...CloudNetworkFull timeTemporary workLocal areaFlexible hours2 days per week$128.6k - $184.9k
...soil.Meet the TeamThe SRE Fleet team is responsible... ...powers our global cloud platform. As a team of... ...focus on automation, reliability, and operational excellence... ...of monitoring, observability, and reliability engineering... ...to that our worldwide network of doers and experts,...CloudNetworkPermanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours$82.3k
...experience. Five9 is a leading provider of cloud contact center... ...experienced Senior Site Reliability Engineer – Compute Platforms... ...with platform and SRE teams to maintain secure, performant, and multi-tenant-isolated... ...Kubernetes, hypervisors, networking, and Linux systems Partner...CloudNetworkTemporary workWork at officeRemote workWorldwide3 days per week$141k - $307k
...capabilities are applied consistently.Lead solution and design reviews,... ...learning, artificial neural networks, etc.) and their real-world... ...models for NLP, CV, or multi‑modal tasks.Hands‑on experience... ...qualificationsExperience developing AI solutions on cloud platforms such as Azure and/or...CloudNetworkWork at officeLocal areaRemote workFlexible hours2 days per week3 days per week1 day per week- ...itD is seeking a Site Reliability Engineer to develop and enhance automation... ...efficiency of large-scale cloud infrastructure. The ideal... .... Attend internal itD networking events (in person and virtual... ...environments. Knowledge of monitoring, observability, and site reliability...CloudNetworkWork experience placementRemote work
$195k - $250k
...Generalist who is able to lead development of a... ...based compute/networking/deployment... ...cycles focusing on reliability enhancements, performance... ...to cloud environments, ensuring... ...Champion Infrastructure Observability (SRE Mindset):... ...management, and debugging multi-threaded...CloudNetworkPermanent employmentRelocationVisa sponsorshipFlexible hoursAfternoon shift3 days per week$130k - $180k
...Why work at Nebius Nebius is leading a new era in cloud computing to serve the global AI economy. We... ...The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team... ..., JBODs, JBOGs, power shelves, network devices, etc. Asset tracking. Hardware...CloudNetworkTemporary workWork at officeImmediate startRemote workFlexible hours$167.7k - $245.2k
...operating market-leading platforms like... ...'s most vital networks. You will join... ...we are building cloud-native, AI-integrated... ...innovation, reliability, and continuous... ...resilient, and observable services in... ...management, security, SRE, DevOps, and... ..., RBAC, and multi-tenant SaaS security...CloudNetworkFull timeTemporary workFlexible hours$209.7k - $272.6k
...transformational Director of Network Engineering to lead the strategy,... ...centers, and cloud environments.... ..., scalability, observability, and operational... ...deliver secure, reliable, and high-... ...engineering, and SRE-aligned operating... ...Establish and execute multi-year network and...CloudNetworkMinimum wage- ...deep expertise in observability, AI, distributed systems, cloud, cybersecurity, and networking. Our team is... ...to provide an AI SRE Teammate that empowers Site Reliability Engineers (SREs)... ...prospective customers, leading technical... ...enterprise environments (multi-cloud,...CloudNetwork
$115.2k - $172.8k
...with an expanded portfolio of leading-edge technologies that... ...device performance, quality, and reliability issues. Onto Innovation strives... ...cross functional AI champion network.Develop reference examples, patterns... ...working across both cloud and on-prem systems. Why Join...CloudNetworkPermanent employmentFull time$65 - $80 per hour
Sr. SRE / DevOps Engineer (Local to Bay Area only) This range... ...Compunnel Inc. Position Senior Site Reliability Engineer / DevOps Engineer —... ..., and automation in AWS cloud environments . Hands-on... ...ECS, EKS Strong knowledge of networking, security groups, firewalls,...CloudNetworkFull timeLocal area$120k - $140k
...Link Systems Inc. is a global provider of reliable networking devices and smart home devices... ...with Managed Service Providers (MSPs), Multi-Dwelling Units (MDUs), Hospitality Service... ...end-to-end portfolio of wired, wireless, cloud, surveillance, and network management/security...CloudNetworkLocal areaRemote workWorldwide- We are a leading global software company dedicated... ...data center, cloud, hybrid environments... ...focus on automation, observability, scalability, and service reliability. The ideal... ...Knowledge of server and network technology is highly... ...communicating with multi‑disciplinary teams...CloudNetworkWork at officeRemote work
$75k - $95k
...Reliability Engineer – Remote Bright Vision... ...company delivering cloud, AI, data, and... ...production. As an SRE you will live at... ...scale, including networking, performance tuning... ...knowledge of observability tooling such as Prometheus... ...experience leading incident response...CloudNetworkFull timeH1bLocal areaImmediate startRemote workVisa sponsorship$120k - $150k
...resources and technology solutions.Senior Platform Engineer (Multi-Cloud & AI Adoption)As a Senior Platform Engineer, you are the... ...-controlled, tested, and deployed with zero downtime.Observability & Reliability: Implement "Self-Healing" infrastructure and advanced monitoring...CloudFull time- Bright Vision Technologies is seeking a Remote Multi-Cloud Solutions Architect to design strategies spanning AWS, Azure, GCP, and OCI... ...portability, abstraction layers, federated identity, cross-cloud networking, and platform offerings that reduce cloud-specific complexity...CloudNetworkRemote job
- ...best practices, including RBAC, network policies, pod security... ...delivery. Build and maintain observability platforms using Prometheus, Grafana, Splunk, and cloud-native monitoring services for... ...measures to improve platform reliability and availability. Define and...CloudNetwork
$186.9k - $267.7k
...driving innovation in networking technologies. Our focus... ...including those in AI, cloud computing, and enterprise... ...the world-class, multi-layered Nexus switches.... ...working relationships.• May lead projects with limited complexity... ...code enabling scale, reliability, and velocity in...CloudNetworkFull timeTemporary workLocal areaFlexible hours$210.6k - $305.1k
...Intersight is the cloud-native operations... ...Joining us means leading a critically important... ...delivers secure, reliable infrastructure... ...remain reliable, observable, and aligned with... ...security, compliance and SRE leadership to... ...that our worldwide network of doers and experts...CloudNetworkFull timeTemporary workLocal areaFlexible hours- ...infrastructure connecting Dexmate’s cloud platform with robots... ...update a growing robot fleet reliably at scale. Responsibilities... ...recovery. -Build fleet telemetry, observability, and remote diagnostics for... ...unreliable or intermittent networks. Requirements -5+ years...CloudNetworkFull timeRemote work
$141k - $307k
...foundations that power secure, reliable, and production-ready... ...a hands-on technical lead, you will define and... ...services, including networking, identity, access,... ...appropriate.Build and enhance observability across Azure AI... ...building, and supporting cloud platforms in Microsoft...CloudNetworkLocal areaRemote workFlexible hours2 days per week3 days per week1 day per week- ...be comfortable working with SRE, observability, cloud, infrastructure,... ...clarity. • Coordinate complex, multi-stakeholder technical deliverables... ...Looking For • Experience leading complex enterprise customer... ...cloud, cybersecurity, and networking. Our mission is to provide...CloudNetworkWork at office
$140k - $170k
...Multi-Cloud Solutions Architect – Remote Bright Vision Technologies... ...identity, cross-cloud networking, and platform offerings that... ...providers. Track record of leading multi-cloud architecture initiatives... ...with cloud-agnostic observability platforms. Experience driving...CloudNetworkFull timeH1bLocal areaImmediate startRemote workVisa sponsorship- ...prem infrastructure that behaves like a cloud environment for internal teams,... ...infrastructure operations: compute, storage, networking, and the automation layered on top Build... ...via Keycloak Build and maintain observability with Prometheus, Grafana, and Alertmanager...CloudNetworkFlexible hours
$145k - $182k
...accounting agents. You'll lead the development of... ...- Ensure reliability: Monitor pipeline... ...Create scalable observability systems-tracking conversation... ...governance for multi-agent runtimes.... ...operations in cloud. Work closely with... ...Understanding of network configurations and...CloudNetworkTemporary workWork at officeShift work3 days per week- ...our campus in Fremont, CA and our public cloud providers. As a Senior Systems... ...and virtual), focusing on scalability, reliability, and security. Perform systems maintenance... ...VMware, KVM). ~ Solid understanding of networking concepts (TCP/IP, DNS, DHCP, firewalls)....CloudNetworkFull timeLocal areaImmediate startRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Observability & SRE Lead — Multi-Cloud Network Reliability. Be the first to apply!
- oracle cloud technical Fremont, CA
- aws cloud Fremont, CA
- cloud Fremont, CA
- senior cloud service delivery manager Fremont, CA
- cloud security Fremont, CA
- network contractor Fremont, CA
- network intern Fremont, CA
- network operations center manager Fremont, CA
- data network cabling Fremont, CA
- IT network Fremont, CA


