Get new jobs by email
- ...management, governance, stakeholder coordination, resource planning, risk and issue management, reporting, migration planning, and infrastructure modernization support for the Lab Consolidation initiative. Project Manager shall establish and manage structured project...SuggestedRemote workRelocation
- ...Infrastructure Program Director Vancouver, WA Onsite Local person [ EUC& Cloud ] Only Citizen / 100% Onsite / Infrastructure Operation Delivery Director Experience: 15+ Years Job Description The Infrastructure Program Director is responsible for leading large-scale...SuggestedLocal area
$77.8k - $165.3k
...powerhouse of diverse teams and take your career wherever you want it to go. Join EY and help to build a better working world. The Infrastructure Projects Supervising Associate – 1Badge Operations is responsible for leading the firm's 1Badge Physical Security program and...SuggestedFull timeFor contractorsSummer holidayWork at officeRelocationFlexible hours- ...Staff, Hebron, Hebron, KY, United States ZF Active Safety US Inc. About the Team About the Team: The Analyst II Regional Infrastructure Administrator is responsible for supporting, administering, and implementing, server, and network infrastructure as they relate...SuggestedH1bWorldwideFlexible hoursShift work
- ...We\'re building the company which will de-risk the largest infrastructure build-out in history. When people finance GPU clusters, the datacenters housing them, and the infrastructure powering them, they need "offtake" - meaning someone has signed a contract to lease...SuggestedLong term contractContract workFixed term contractWork at officeLocal areaVisa sponsorshipShift work
- ...ecosystem. Oversee the design, build, and lifecycle management of Linux servers, including storage, virtualization, and associated infrastructure. Manage high availability (HA) configurations, clustering, and load-balanced environments to ensure minimal downtime. Drive...SuggestedContract work3 days per week
- ...produced on time according to SLA, proactively managing disruptions and incidents by reasoning across applications, data flows, infrastructure, and dependencies (batch, APIs, integrations). Manage interfaces with upstream data pipelines, connecting signals across...Suggested
- ...for Site Reliability Engineering (SRE), monitoring and observability, domain portfolio management, cloud platform operations, infrastructure automation, web application security, and application availability. The role will work closely with engineering teams in Dallas...Suggested
- ...for Site Reliability Engineering (SRE), monitoring and observability, domain portfolio management, cloud platform operations, infrastructure automation, web application security, and application availability. The role will work closely with engineering teams in Dallas...SuggestedContract work
- ...Job Description Data Infrastructure Site Reliability Engineer (SRE) AWS & Big Data Platforms Dallas, Texas (Remote) Skills : (AWS, Hadoop, on-premises, AI) Skills - Site Reliability Engineering, Big Data, AWS, AWS EMR, AWS EKS, AWS MSK, AWS Athena, AWS Glue, Spark...SuggestedRemote work
$70 per hour
...incident command and blameless postmortem practice. Production Kubernetes and container experience (Docker), with cloud-native infrastructure patterns. Hands-on production ownership on at least one major cloud (AWS, Google Cloud Platform, or Azure). Terraform...SuggestedHourly payRemote work- ...for microservices on Kubernetes/EKS Implement GitOps workflows; enforce trunk-based development and deployment gates Automate infrastructure provisioning using CloudFormation, Terraform, and AWS CDK Manage Akamai CDN configuration: edge rules, caching policies, TLS, WAF...SuggestedFull time
- ...Requirements: UC Citizenship We are looking for a Senior SRE (Linux) to build and operate the systems that run on every host across our infrastructure. Our team owns OS configuration, automation, AMIs, and container base images across a large-scale hybrid environment (AWS + on-...SuggestedRemote work
- ...operational documentation Required Qualifications ~ Experience of 7+ years in Site Reliability Engineering, DevOps, or infrastructure engineering ~3+ years in SRE leadership/Architect roles ~3+ years hands-on experience with Datadog, Splunk, Grafana,...SuggestedRemote work
- ...CD Model Deployment Google Cloud Platform Monitoring & Observability Agentic AI Architecture Secondary Skills Kubernetes Infrastructure Automation Security Controls Site Reliability Engineering (SRE) DevOps Key Responsibilities Design and implement enterprise...Suggested
- ...Travel: Occasional Job Description: Seeking an experienced SRE Solutions Architect with strong expertise in AI/GPU infrastructure, HPC, cloud operations, distributed systems, and production reliability. This role focuses on Day-2 operations, troubleshooting...Contract workRemote work
- ...Location: Remote Duration- 3+ Month Contract ~6+ years of experience as a DevOps Engineer, Site Reliability Engineer, or Infrastructure Operations Engineer with a strong focus on compute ~ Strong hands-on experience operating bare metal compute environments at...Contract workRemote work
- ...and deploy Model Context Protocol (MCP) clients and servers that securely connect enterprise LLMs with engineering tools, cloud infrastructure, and operational platforms. Build custom MCP services using Python, TypeScript, JavaScript, or Node.js to expose...Contract work
- ...responsible for the architecture, design, and standardization of enterprise compute and hypervisor environments spanning bare metal infrastructure, operating systems, hypervisors, private cloud orchestration, and Kubernetes using Infrastructure-as-Code and GitOps practices....Contract workRemote work
$65.12 per hour
...have strong experience in Oracle Enterprise Linux administration, SRE practices, automation with Ansible, and high availability infrastructure and a proven ability to increase platform reliability, scalability, and security through automation and operational excellence....Contract work3 days per week- ...Please find the JD Automation: Automate infrastructure provisioning, deployment and scaling processes using IaC (Infrastructure as Code) methodologies. Observability: Develop and maintain observability systems to monitor the health and performance of applications...
- ...a global environment. The ideal candidate will have strong hands-on experience with SRE, DevOps, production operations, cloud infrastructure, Kubernetes, Terraform, observability, incident management, and automation. This role will focus on improving service stability...Long term contract3 days per week
- ...closely with software engineering teams to: Review performance-critical code paths Propose improvements at code, configuration, or infrastructure level Improve system observability (metrics, logs, traces) Communicate complex technical findings clearly to both engineers and...Long term contractContract work
- ...Service Level Objectives (SLOs), and Error Budgets. Automate operational tasks and repetitive processes using scripting and Infrastructure as Code (IaC). Lead incident response activities, troubleshooting, root cause analysis (RCA), and post-incident reviews....
- ...an immediate need for an experienced Google Cloud Platform Engineer with strong expertise in Google Cloud Platform, Kubernetes, infrastructure automation, observability, and Site Reliability Engineering to design, automate, optimize, and support scalable cloud...Immediate startRemote work
- ...application paths. PostgreSQL query optimization, indexing, connection pooling, and database performance. Cloud Azure Azure infrastructure, networking, and application operations. Capacity Planning & Scalability workload modeling, benchmarking, forecasting, and...
- ...enterprise development and production environments. The ideal candidate combines a strong software engineering mindset with deep infrastructure and reliability engineering expertise . This role will work closely with software development, networking, infrastructure, and...Local area
- ...in large-scale enterprises. - Oversee design, build, and lifecycle management of Linux servers, storage, and virtualization infrastructure. - Manage high availability (HA), clustering, and load-balancing for minimal downtime. - Lead capacity planning and performance...
$55 - $60 per hour
...cloud experience SRE, Reliability Engineering, and DevOps practices OpenTelemetry and modern observability tools Infrastructure-as-Code and automation frameworks Qualified candidates should APPLY NOW for immediate consideration! This position is only...Hourly payFull timeContract workTemporary workWork experience placementImmediate startWorldwideFlexible hours- ..., prefer PST hours Visa Requirements : UC Citizenship About the Team The SRE Platform Engineering team builds and operates the infrastructure that powers our cloud. We focus on delivering reliable, scalable, and simple platforms that enable product teams to move quickly...Remote work