Staff Site Reliability Engineer - Observability GCP (New York)
$194k - $267kOkta
Secure Every Identity, from AI to HumanIdentity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.We are seeking a highly technical ObservabilitySite Reliability Engineer with a specialty in Google Cloud, to own and expand our Observability ecosystem into GCP. In this role, you will move beyond simple monitoring to delivering a world class, comprehensive, scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure as code—utilizing Terraform and strong coding proficiency in Go, Python, or Ruby—to automate the deployment of agents and collectors across complex distributed systems.Key ResponsibilitiesAutomated Infrastructure: Design, build, and maintain scalable observability infrastructure using tools like Terraform.GCP Observabilty Engineering: Optimize the collection, processing, and storage of Observabilty data to ensure high reliability and low latency of our Splunk and Grafana servicesIncident Response: Participate in on-call rotations and lead post-incident reviews to drive systemic improvements and observability-driven development.Automation: Eliminate toil by automating the deployment and scaling of observability agents and collectors.Required Skills & Experience (The Essentials)GKE: Minimum 5+ Experience scaling and managing observability in a Google Cloud platform. Visualization: Expertise in creating intuitive, actionable Splunk or Grafana dashboards that correlate data across multiple sources.SRE Mindset: Minimum 3+ years of experience in an SRE, DevOps, or Systems Engineering role with a focus on high-availability systems.Programming Proficiency: Strong coding skills in Python, Go for building internal tools and automating workflows.Distributed Systems: Deep understanding of Linux internals, networking (TCP/IP, DNS, Load Balancing), and container orchestration (Kubernetes/GKE).Problem Solving: A data-driven approach to debugging complex, cross-service performance bottlenecks.Bonus Skills (The Nice-to-Haves)Telemetry Standards: Hands-on experience with OpenTelemetry (OTel), Vector, or similar frameworks for instrumenting applications.Grafana Loki: Experience in migrating Splunk to Grafana LokiOther Cloud Platforms: Experience managing observability native tools within AWS.Additional requirements:This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.#LI-MM#LI-HybridP24517_3387022The annual base salary range for this position for candidates located in the San Francisco Bay area is between: $194,000—$267,000 USDBelow is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: . The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between:$174,000—$239,000 USDThe Okta ExperienceSupporting Your Well-BeingDriving Social Impact Developing Talent and Fostering Connection + CommunityWe are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.
- ...team of researchers, engineers, designers, and... ...offices in London, New York City, Montreal,... ...performance, scalable and reliable machine learning... ...are looking for a Site Reliability... ...Automate environment observability and resilience. Enable... ...with GCP, Azure, AWS, OCI,...SuggestedFull timePart timeWork experience placementWork at officeLocal areaRemote workHome office
$200k - $250k
...is seeking a Senior Site Reliability Engineer to join our growing Enterprise... ..., Kubernetes, and observability; proficiency in... ...platforms (e.g. AWS, Azure, GCP)Prior experience... ...HRT veterans and new hires alike. At HRT we... ...proud of our diverse staff; we have offices all...SuggestedPart timeWork at officeLocal areaImmediate start$296k - $370k
...Datadog’s Cloud Observability group is one of the core data retrieval... ...major hyperscalers (AWS, Azure, GCP, OCI), as well as both... .... As Director, you will own engineering for all clouds, generating more... ...Atlantic org, and are based in New York, Boston, or Paris.Datadog...SuggestedPart timeWork at office$165k - $241.4k
...Collaboration, and Observability portfolios Your... ...looking for talented engineers with a software or... ...to ensure the reliability, performance and security... ...salary range for new hires in this... ...the Cisco careers site to discover more benefits... ...listed below:New York City Metro Area:$1...SuggestedFull timeTemporary workPart timeWork at officeLocal areaFlexible hours1 day per week$194k - $267k
...organizations to safely embrace this new era. This work requires a... ....Position Overview:The Site Reliability Engineer (SRE) will play a key role... ..., security, and observability within the Kubernetes clusters... ...), Colorado, Illinois, New York and Washington. Your actual...SuggestedPermanent employmentPart timeWork at officeLocal areaWorldwideFlexible hours$195.5k - $264.5k
...at the heart of Pave Engineering. Our customers are Pave... ...grade software fast and reliably.This is a high-impact,... ...end, from CI/CD and observability standards to local development... ...cloud architecture (GCP or similar), with a... ...with regional hubs in New York City's Flatiron...Part timeLocal area3 days per week$192k - $240k
...need to grow your career.Engineering at BrexEngineering at... ...power Brex’s release, observability, and incident... ...releases are safe, fast, and reliable, and that our infrastructure... ...will be based in our New York office. We are a... ...platforms (e.g., AWS, GCP, Azure)Deep familiarity...Part timeWork at officeRemote workWork from home$172.5k - $260.1k
...Salesforce.Lead Software Engineer – Platform... ...customers in entirely new ways — while... ...design, testing, observability, scalability, and... ...building scalable, reliable, and secure customer... ...AWS (preferred), GCP, or other public cloud... ...Francisco and New York City metropolitan...Full timePart time$231.4k - $331.8k
...Team The Platform Engineering organization is... ..., automation, observability, security, and... ...services reliably at scale. We are... ...across AWS, Azure, GCP, or hybrid cloud... ...range for new hires in this position... ...Cisco careers site to discover... ...listed below:New York City Metro Area...Permanent employmentFull timeTemporary workPart timeLocal areaFlexible hours$194k - $267k
...organizations to safely embrace this new era. This work requires a... ...too, let's talk.The TeamThe Site Reliability team is dedicated to... ...maximize platform reliability and engineering velocity.The ideal candidate... ...Area), Colorado, Illinois, New York and Washington. Your actual base...Part timeLocal areaWorldwideFlexible hours$244k - $305k
...Summary:Datadog is seeking a Staff Software Engineer to help shape the future of... ...Logs offering by unifying observability pipelines with log... ...observability, diagnostics, reliability, and secure operation across... ...platforms (e.g., AWS, Azure, or GCP) and are comfortable troubleshooting...Part timeWork at office$146.8k - $272.6k
...within our organization.Staff Software Engineer (Full Stack) –... ...stakeholders to bring new interaction models to... ...testing, performance, and observability.Build quick,... ...environments (AWS, Azure, or GCP) and using Docker/... ...of the following: New York City, San Francisco,...Part timeWork at officeLocal areaFlexible hours2 days per week3 days per week$197.3k - $313.7k
...As the Principal Engineer focused on architecture... ..., performance, observability, and user... ...and role based sites such as the developer... ...operate accurately and reliably.Critically evaluate... ...AWS (preferred), GCP, or other public cloud... ...San Francisco and New York City metropolitan...Full timePart time$195k - $275k
...Management, and the Chief Operating Office.The Reliability Operations (RO) within WMT is... ...implementation of operational controls, monitoring, observability, and capacity-management practices.... ..., and production transition for new technology capabilities.Establish and manage...Temporary workPart timeWork at officeWorldwideNight shift$240k - $325k
...experience to a new industry, join our... ...brighter way forward. Staff Engineer (P4) – Research... ...: New York City (On-site)About JLL and JLL... ...deliver high-quality, reliable, and scalable outcomes... ...-as-code, and observability toolingCollaborate... ...(AWS, Azure, or GCP)Experience with relational...Part time$300k - $350k
...Principal AI Platform Engineer to design and... ...applicationsEnable reliable, scalable execution... ...capabilitiesBuild observability and monitoring:... ...cloud platforms (AWS, GCP, Azure) and... ...financial advisor, new parent leave, reproductive... ...: New York, NYType: Full time...Full timeTemporary workPart timeWork experience placementFlexible hours$157.9k - $282.1k
...Principal Full Stack Engineer, AI Platform & AgentsBuild... ...you ship (latency, reliability, hallucination reduction... ...(primary), Azure, GCP; Docker, Terraform, GitHub... ...tooling, CI/CD, and observability for safe, fast iteration... ...Riverwoods, IL; USA - New York City, NY; CAN-Ontario-...Full timePart timeWork at officeRemote work2 days per week- ...Software Engineer (Dashboard) Dashboard is the front door to Browserbase. This team owns... ...and billing to data-rich observability and complex UI. It’s true full stack engineering... ...environment. Are excited to work in New York . Why join us Shape the user-facing...Full timeImmediate start
$144.25k - $256.25k
...2026-06-03Location: New York, NY, United States; Phoenix... ...Function: Engineering & ArchitectureSchedule... ...security, explainability, reliability, and compliance required... ...gatewaysEvaluation, observability, and safety tooling... ...infrastructure: AWS and/or GCP,...Part timeWork at officeVisa sponsorship3 days per week$205k - $235k
...1408890033Location: New York, NY, US, 10001-8604Salary... ...growth. The Software Engineering Director for AI... ...pipelines) optimized for reliability, security, and... ...quality, testing, CI/CD, observability, and performance.Collaborate... ...such as AWS or GCP, using containerization...Part timeSummer holidayWork at officeFlexible hours- ...RoleThe Director of Platform & Reliability Engineering will lead a critical... ...Engineering, Cloud Operations, and Site Reliability Engineering,... ..., reliability engineering, observability, and developer enablement.... ...residents of New York, NY the annual salary range...Part timeWork at officeLocal area2 days per week3 days per week
$245k - $272k
...San Francisco, and New York, we support more... ...prompt management, observability, and safety guardrails... ...the platform is reliable, performant, and... ...easy to adopt. As a Staff MLE, you'll also... ...decisions, and mentor engineers across the... ...cloud platforms (GCP, AWS, or Azure).Strong...Full timePart timeWork at officeLocal areaRemote work2 days per week3 days per week$176.72k - $265.08k
...Posted: 2026-07-13Location: New York, New York, United StatesSalary... ...specializing in AI / AML Engineering to lead the development of enterprise... ...for AI engineeringEnsure observability, monitoring, and performance... ...with cloud platforms (GCP, AWS, Azure)EducationBachelor...Full timePart time$115.2k - $214k
...Drug Discovery, the Software Engineering team builds and operates... ...developing technology and deploying reliable, intuitive solutions that... ...management, evaluation, and observability.Build software that enables... ...For the primary location of New York for the Software Engineer is...Full timePart timeLocal areaWorldwideRelocation package3 days per week- ...service, cluster scheduler, and reliability tooling that powers the... ...the RoleAs a Senior Software Engineer on Spark Platform, you will... ...Francisco, Sunnyvale, Seattle, or New York City for this hybrid... ...cluster lifecycle, shuffle, observability, multi-tenant scheduling) rather...Hourly payWork at officeLocal areaRemote workRelocationFlexible hours
- ...culture thrives on finding new and better ways to... ...Orchestration), HPE OpsRamp (Observability), and HPE Zerto (Data... ...partnering deeply with Engineering, Sales, Alliances, and... ...Hyperscaler (AWS/Azure/GCP) specifically focused... ...00 in California & New York // 170,000 - 412,500 in...Full timePart timeWork experience placementWork at office2 days per week
$175k - $350k
...Platform Infrastructure Engineer to join one of our... ...Engineering teams in Miami or New York. Our team is... ...deployment, configuration, and reliability—for our high frequency... ...is leveraged daily by Site Reliability Engineers... ...and process state observability systems) Collaborate with...Part timeWorldwide- ...Location: London, New York, ChicagoDepartment: TechnologyExperience Level: Experience ProfessionalsContact... ...REQ8338We are seeking a Trading Platform Engineer to join our Systematic Technology team, focused on the reliability, observability, and day-to-day support of a high-...Part time
$220k - $270k
...intelligence, we're setting new industry standards.... ...experienced Software Engineer to help architect and... ..., LLM-driven systems reliable in production.Key ResponsibilitiesOwn... ..., evaluation, observability, and the guardrails... ...Francisco, CA; New York, NYEmployment TypeFull...Part timeLocal areaWork from homeFlexible hoursNight shiftWeekend work- ...service, cluster scheduler, and reliability tooling that powers the... ...About the RoleAs a Software Engineer on Spark Platform, you will... ...lifecycle automation, and the observability and incident automation that... ...Francisco, Sunnyvale, Seattle, or New York City for this hybrid...Hourly payWork at officeLocal areaRemote workRelocationFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer - Observability GCP (New York). Be the first to apply!
- project engineer assistant project manager New York, NY
- assistant chief engineer New York, NY
- staff data engineer New York, NY
- senior staff engineer New York, NY
- engineering aide New York, NY
- software engineer staff New York, NY
- assistant engineer New York, NY
- assistant engineering manager New York, NY
- technology administrator New York, NY
- senior staff systems engineer New York, NY

























