Senior/Staff Platform Engineer
Jobgether
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior/Staff Platform Engineer based in United States .
The Senior/Staff Platform Engineer will build, operate, and evolve large-scale production infrastructure with a strong focus on Kubernetes and reliability.
This hands-on role spans cloud infrastructure, Linux, networking, observability, CI/CD, automation, and production operations.
You will diagnose complex infrastructure challenges across multiple technical layers and build tooling that improves reliability, scalability, and developer experience.
The position offers significant autonomy, with ownership of ambiguous initiatives from initial design through implementation and production operation.
You will work directly with technical stakeholders, contribute to architecture and reliability decisions, and take a leading role in critical incident resolution.
Your impact will extend beyond day-to-day operations as you establish stronger engineering practices, automation, observability, and operational resilience.
This is an ideal opportunity for a highly experienced engineer who combines deep systems expertise with strong technical judgment and communication skills.
- Design, build, operate, and continuously improve production Kubernetes platforms, owning cluster architecture, networking, workload isolation, resource management, security, upgrades, scaling, and reliability.
- Troubleshoot Kubernetes and infrastructure issues beyond the application layer, including CNI and networking problems, scheduling, node behavior, resource constraints, controllers, and cluster-level failures.
- Operate highly available infrastructure across cloud, hybrid, virtualized, and/or bare-metal environments, diagnosing complex issues across Kubernetes, containers, Linux, networking, and underlying infrastructure.
- Develop and maintain production tooling and automation using Go, Python, or Java to improve platform operations, deployment, troubleshooting, reliability, and developer experience.
- Build internal services, APIs, integrations, and operational tools where needed, while applying sound software engineering practices such as testing, code review, maintainability, and documentation.
- Own the reliability and operational health of critical production infrastructure, leading or significantly contributing to incident response, root-cause analysis, durable remediation, and disaster recovery initiatives.
- Define and improve SLOs, SLIs, alerting, and operational processes, using logs, metrics, traces, profiling tools, and system-level diagnostics to improve availability, performance, capacity, resilience, and recovery times.
- Build and maintain infrastructure as code using Terraform or equivalent technologies, creating reusable patterns and improving automation as the platform evolves.
- Develop and improve CI/CD and deployment workflows while balancing delivery speed with reliability, security, scalability, and operational requirements.
- Participate in cloud and infrastructure migrations, including dependency analysis, networking, cutover planning, rollback strategies, and production validation.
- Build and enhance monitoring, dashboards, logging, distributed tracing, and alerting to improve incident detection, diagnosis, and recovery.
- Work directly with customers, engineers, and technical stakeholders to understand requirements, investigate issues, communicate trade-offs, and drive effective technical solutions.
- Own complex infrastructure initiatives from problem definition through design, implementation, and production operation, contributing to RFCs, architecture discussions, design reviews, and broader technical direction.
- Mentor engineers and help strengthen engineering and operational practices while operating independently in ambiguous situations and taking ownership when immediate direction is unavailable.
Requirements
- 10+ years of professional experience in Platform Engineering, Site Reliability Engineering, Infrastructure Engineering, DevOps, or related fields is preferred for Staff-level candidates, with significant hands-on experience operating complex production infrastructure and distributed systems.
- Demonstrated experience building and operating production Kubernetes platforms is required, rather than experience limited to deploying applications onto existing clusters.
- Production programming experience in Go, Python, or Java is required, with the ability to read, debug, maintain, and contribute to production codebases and automation.
- Strong experience with production reliability, incident response, troubleshooting, root-cause analysis, and operational ownership is essential.
- Demonstrated ability to independently own complex technical initiatives from ambiguous starting points through design, implementation, and production operation.
- A degree in Computer Science, Engineering, or a related field is preferred, although equivalent practical experience may be considered.
- Deep understanding of Kubernetes infrastructure, including cluster architecture, networking and CNI, NetworkPolicy, scheduling, resource management, nodes, security, RBAC, and cluster behavior.
- Strong Linux fundamentals and hands-on experience troubleshooting production systems, combined with a solid understanding of DNS, routing, load balancing, connectivity, and cloud/Kubernetes networking.
- Production experience with at least one major cloud platform, such as AWS, GCP, or Alicloud, along with infrastructure-as-code experience using Terraform or equivalent tooling.
- Experience with configuration management and automation technologies such as Ansible, Puppet, or similar platforms.
- Strong observability expertise using metrics, logs, traces, dashboards, and alerting platforms such as Prometheus, Grafana, OpenTelemetry, Datadog, or equivalent technologies.
- Experience with CI/CD infrastructure, Docker, container tooling, modern software delivery practices, high availability, capacity planning, disaster recovery, and production resilience.
- Experience planning and executing production cloud or infrastructure migrations, operating large-scale or multi-cluster Kubernetes environments, or working with hybrid, on-premises, virtualized, or bare-metal infrastructure is preferred.
- Additional experience with Kubernetes controllers or operators, advanced networking, service mesh, mTLS, workload identity, multi-cloud infrastructure, security hardening, IAM, compliance, capacity planning, performance engineering, or developer platform tooling is highly valued.
- Previous technical leadership, mentoring, Staff/Principal-level responsibilities, or experience working directly with external customers in consulting or service-delivery environments is preferred.
- Exceptional written and verbal technical communication skills are required, with the ability to explain architecture, risks, trade-offs, and technical decisions clearly to customers, engineers, and technical leadership.
- Strong analytical and problem-solving abilities, high ownership, sound technical judgment, and the ability to operate independently in ambiguous environments are essential.
- The ideal candidate can influence others without formal management authority and is comfortable taking the lead on complex technical problems without continuous oversight.
- This is a hands-on platform and reliability engineering position. Candidates should have personally built, operated, troubleshot, and improved production infrastructure rather than primarily consuming managed services or deploying applications onto Kubernetes.
- The role requires availability during core hours aligned with Pacific Time, from 8:00 AM to 5:00 PM PST, as well as participation in an on-call rotation approximately every four to five weeks.
Benefits
- Fully remote work arrangement within Canada.
- Opportunity to work on large-scale Kubernetes platforms and complex production infrastructure with significant technical ownership.
- High degree of autonomy to shape infrastructure architecture, reliability initiatives, automation, and operational practices.
- Exposure to cloud, hybrid, networking, observability, CI/CD, security, and distributed systems at scale.
- Opportunity to collaborate directly with experienced engineers, technical stakeholders, and customers on challenging infrastructure initiatives.
- Ability to influence technical direction, mentor other engineers, and contribute to architecture and engineering best practices.
- Remote-first environment designed to support distributed collaboration across North America and Latin America.
- Opportunity to work on technically challenging projects where reliability, scalability, automation, and operational excellence are core priorities.
How Jobgether works:
We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Our system identifies the top-fitting candidates, and this shortlist is then shared directly with the hiring company. The final decision and next steps (interviews, assessments) are managed by their internal team.
We appreciate your interest and wish you the best!
Why Apply Through Jobgether?
Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.
#LI-CL1
$189k - $236k
...requirements vary by role and will be assessed during the interview process.About the Role:We’re hiring seasoned engineers to join our teams that work on core platform capabilities, improving our existing systems for extensibility and scalability, and building the future of...SeniorFull timeWork at officeLocal areaRemote work2 days per week3 days per week$201.3k - $352.3k
Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a... ...business reinvention. Our ServiceNow AI platform brings together any AI, any data, and any... .... We are looking for a seasoned IC5 Senior Staff Engineer with deep expertise in distributed...SeniorPermanent employmentWork experience placementWork at officeImmediate startRemote workFlexible hours2 days per week- ...Senior/ Staff Platform Engineer Type: Remote Coverage: Pacific Hours (8:00 AM – 5:00 PM PST and On-call every 4-5 weeks) Job Description: We are looking for a Senior/Staff Platform Engineer to build, operate, and evolve large-scale production infrastructure....SeniorFull timeTemporary workImmediate startRemote work
$199.75k - $270k
.... Our flagship product, the NodeZeroTM platform, delivers production-safe autonomous pentests... ...Operations cyber operators, startup engineers, and formerly frustrated cybersecurity... ...the Role We are looking for a Senior OR Staff Platform Engineer to join Horizon3’s Internal...SeniorFull timeWork at officeRemote workFlexible hours$230k - $334k
...the 1Password application. It includes Compute Services, Network and Edge Services, and Stateful Services. As a Senior Staff Cloud Platform Engineer, you set the multi-year technical strategy for Infrastructure Services. You decide where the organization invests, define...SeniorFull timeCurrently hiringLocal areaImmediate startRemote workWork from home- ...government agencies worldwide. As we continue to evolve our platform, we are investing in the foundational systems that... ...platform. About the Role We’re looking for a Senior, Staff or Principal Platform Engineer with deep expertise in Python, Linux systems and...SeniorFor contractorsRemote workWorldwide
$201.3k - $352.3k
...Senior Staff Data Platform Engineer - Data Access Team Full-time Employee Type: Regular Region: AMS - North America and Canada Work Persona: Flexible or Remote It all started when engineer Fred Luddy wrote code that automated a tedious task for his coworker...SeniorFull timeWork experience placementWork at officeImmediate startRemote workFlexible hours2 days per week$240k - $360k
About the RoleWe're looking for a Senior Staff Data Engineer to be the technical backbone of our Data & ML Platform team — the foundation powering analytics, product experiences, and machine learning across Hinge Health. This is a high-ownership IC role for someone who...SeniorWork at officeLocal areaImmediate startRemote workWorldwide3 days per week$286.2k - $326.7k
Senior Staff AI Engineer - Agentic AI Platform (Remote Eligible) At Capital One, we are creating responsible and reliable AI systems, changing banking for good. For years, Capital One has been an industry leader in using machine learning to create real-time, personalized...SeniorFull timePart timeLocal areaRemote work- ...Capital One is seeking a Senior Staff AI Engineer to help build scalable AI foundations and production-ready ML systems. You will design and implement AI software components, set architecture direction, and mentor engineers working on AI across Capital One’s products....SeniorRemote job
- ...powers the 1Password application. It includes Compute Services, Network and Edge Services, and Stateful Services. As a Senior Staff Cloud Platform Engineer, you set the multi-year technical strategy for Infrastructure Services. You decide where the organization invests,...SeniorFull timeImmediate startRemote work
$286.2k - $392k
...Computer Science, AI, Electrical Engineering, Computer Engineering, or a... ...Experience architecting AI platforms and making trade-offs involving... ...AI architecture. Mentor senior technical leaders across research... ...: We are hiring a Senior Staff AI Engineer for our Agentic...SeniorFull timeRemote work- ...redditinc. com . Team Overview The Ads ML Platform team builds infrastructure that... ...where appropriate. Our systems help ML engineers move faster across the full development... ...systems reliably. We are looking for a Senior Staff Machine Learning Systems Engineer to lead...SeniorFull timeWork at officeRemote workFlexible hours
$193.93k - $291.15k
...world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides.Founded in 2016, Nuro... ...robotics team is growing and we are looking for an ML Software Engineer to join our Online Mapping team. We are searching for an...SeniorImmediate startFlexible hours$168k - $264.5k
...work. We are looking for an outstanding Senior Art Director to learn, understand, and... ...on the world!NVIDIA is hiring a Senior Staff Engineer - Enterprise Messaging and Exchange to... ...employee communications and collaboration platforms. The core of this role is deep...SeniorFull time$127.63k - $191.2k
...beyond fleeting trends, Marvell is a place to thrive, learn, and lead. Your Team, Your ImpactAs a key CAD member of Marvell Central Engineering, you will play a leading role in developing next-generation Multi-die 3DIC Methodology, plus CAD flow automation and tool support...SeniorPermanent employmentInternshipWork from home$218.4k - $365.2k
...operate at the intersection of product innovation, mobile platform excellence, and large-scale engineering. The features and systems we build support millions... ..., and user experience. The RoleWe’re looking for a Senior Staff Engineer to help shape the future of Slack’s mobile...SeniorFull timeImmediate startRemote workWorldwide$168k - $247k
...C++ rewrites, no TensorRT export. A new policy goes from training to on-vehicle deployment in minutes.About the RoleAs a Senior/Staff Deep RL Engineer, you will design, train, and deploy deep reinforcement learning policies that make real-time driving decisions for our...SeniorHourly payWork at officeLocal areaRemote workFlexible hours- ...shared success where your work truly matters.Job SummaryAs a Senior Low-Level Software Engineer at Cortex, you will be a technical authority and an... ...compatibility issues, or malware-driven crashes.Adjacent Platforms - Exposure to low-level development on macOS (Endpoint...SeniorFull timeLocal areaRemote workVisa sponsorshipWork visa
- ...As a Senior Staff Engineer on the Guest & Host team, this full-time remote position will set the technical direction for backend platforms that manage listing data, ensuring quality and visibility while collaborating with cross-functional teams to design APIs and data...SeniorFull timeRemote work
$152k - $277k
...to build scalable offline and online ML platforms to support ranking and recommendation, and... ...objects that can be easily consumed by engineers and scientists. You will also be... ...your total compensation.The base pay for Staff position ranges from $152K/year to $277K...SeniorTemporary work$111.1k - $207.8k
...environments and is accountable for the successful delivery of platform and infrastructure capabilities at scale. The Assistant... ...delivery teams, and influencing stakeholders across product, engineering, and platform organizations. In addition to hands on technical...SeniorSummer holidayLocal areaFlexible hours$225k - $275k
...that momentum. We're building our next-generation storage platform - designed from the ground up to solve the problems our existing... ...'s architecture. You'll be defining it.On paper this is a "Staff Software Engineer" role. In practice, you're the architect. You'll own the...SeniorRemote work$162.6k - $244k
...Company:Qualcomm Technologies, Inc.Job Area:Engineering Group, Engineering Group Software EngineeringGeneral... ...that redefine mobility.About the role You'll be the senior technical authority for the embedded Linux platform software powering automotive telematics systems —...SeniorWork experience placementWork from home$130k - $260k
## Senior Staff EngineerApply: Hybrid: Remote (United States): Palo Alto, CA: Chicago, IL:... ...We are seeking a Senior Staff Software Engineer to provide technical leadership for critical... ....* Integrate across vendor and bespoke platforms to unify developer experiences.**...SeniorHourly payFull timeWork experience placementLocal areaRemote work- ...Job Title: Senior Staff Engineer Duration: 1 Year, possible extension Location: Remote The Difference You Will Make: As a Senior... ...for the lifecycle management of our Kubernetes compute platform. You will enable the Cloud Infrastructure organization to work...SeniorWork experience placementRemote work
$70k - $110k
...Senior Staff Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud,... ...Deep expertise in at least one of distributed systems, cloud platforms, or large‑scale data systems. Strong programming skills...SeniorFull timeH1bLocal areaRemote workVisa sponsorship- ...SummaryThis is an exciting opportunity in Celestica’s Hardware Platform Solutions (HPS) to make a positive impact and be part of a rapid... ...expediently drive issues/escalations resolutions. The Network Engineer engages with customers to understand their topology and provides...SeniorWork at officeRemote work
$150k - $200k
...Senior Staff Engineer We are hiring a Senior Staff Engineer for a fast-growing Canadian tech startup building AI-powered consumer products... .... This role is heavily focused on data infrastructure and platform engineering. This is a hands-on senior IC role. You will...SeniorRemote workFlexible hours$150k - $190k
...About TripleLift We're TripleLift, an advertising platform on a mission to elevate digital advertising through beautiful creative... ...ecosystem at triplelift.com. Overview The Senior Staff Engineer plays a critical role in driving the performance, scalability...SeniorFull timeFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior/Staff Platform Engineer. Be the first to apply!




