Get new jobs by email
- ...~3+ years serving as a Technical Lead, Senior Engineer, Architect, or equivalent. ~ Strong experience with AWS, Terraform, Kubernetes/EKS, DNS, CDN, AWS WAF, SSL/TLS, observability platforms, CI/CD, and SRE practices. ~ Experience troubleshooting large-scale...Suggested
- ...using AI-powered engineering tools to improve development and operational productivity. Preferred Qualifications : Kubernetes Docker / containerization Terraform / Infrastructure-as-Code AWS, Azure, or Google Cloud Platform Splunk...SuggestedLocal area
- ...including KVM on Ubuntu, OpenStack operations and troubleshooting, Harvester HCI ~ Experience deploying and operating production Kubernetes environments ~ Expertise with enterprise compute hardware, including Cisco UCS, Dell PowerEdge, Supermicro systems and HPE...SuggestedContract workRemote work
- ...Skills MLOps CI/CD Model Deployment Google Cloud Platform Monitoring & Observability Agentic AI Architecture Secondary Skills Kubernetes Infrastructure Automation Security Controls Site Reliability Engineering (SRE) DevOps Key Responsibilities Design...Suggested
- ...working across cloud + on-prem environments ~ U.S. Person required (FedRAMP; U.S.-based work) Nice to Have Docker / Kubernetes AMIs or container image building Go, C, or other systems-level languages Experience with compliance environments (...SuggestedRemote work
- ...Job Title: Site Reliability Engineer (SRE) Kubernetes Platform (FedRAMP High / IL5)-- G8-L1 Client: OpenKyber Location: 100% Remote Visa: Interview Process: Online Work Schedule: 100% Remote Job Description: Level: L1 Location : Remote, prefer PST hours Visa Requirements...SuggestedRemote work
- ...LangGraph. ~ DevSecOps/CI/CD: Azure DevOps, GitHub Actions, Jenkins. ~ IAM: Entra ID/Azure AD, OAuth 2.0, OIDC, SAML, SSO. ~ Kubernetes, containers, networking, observability, HA/DR, and performance optimization. ~ Strong architecture governance, stakeholder...SuggestedContract workRemote work2 days per week
- ...or Google Cloud Platform) and cloud-native services ~8 Required Experience with containerization and orchestration (Docker, Kubernetes) ~8 Required Strong understanding of monitoring, alerting, and logging concepts ~8 Required Experience defining and managing...SuggestedContract workLocal areaRemote work
- ...multi-region AWS deployments for zero-downtime operations Own and evolve Jenkins-based CI/CD pipelines for microservices on Kubernetes/EKS Implement GitOps workflows; enforce trunk-based development and deployment gates Automate infrastructure provisioning...SuggestedFull time
- ...GPU/AI infrastructure preferred ~ GPU Infrastructure, AI Platforms, HPC, Distributed Systems ~ Linux, Python, Bash ~ Kubernetes, Slurm, Containers InfiniBand, NCCL, UFM, High-Speed Ethernet Prometheus, Grafana, OpenTelemetry Terraform, Ansible, Argo...SuggestedContract workRemote work
- ...collaboration, and communication skills across application, platform, and engineering teams. Preferred Qualifications Experience supporting Kubernetes and GKE workloads in production. Experience with GitHub Actions. Experience with Google Cloud Platform Cloud Monitoring,...SuggestedLong term contractLocal areaImmediate startRelocation
- ...R, scikit-learn, TensorFlow, PyTorch, Pandas, NumPy, Power BI/Tableau, SQL, Linux/Unix, AWS/Azure/Google Cloud Platform, Docker, Kubernetes, Prometheus, Grafana, ELK, Splunk, Dynatrace, Datadog, Jenkins, GitHub Actions, GitLab CI, Azure DevOps, Terraform, Ansible, and...SuggestedImmediate start
- ..., OpenTofu, Pulumi, Terragrunt, CloudFormation, or similar technologies. Deploy and manage containerized applications using Kubernetes and Docker across AWS, Azure, or Google Cloud. Build AI-powered observability solutions leveraging Datadog, Prometheus, Grafana...SuggestedContract work
- ...application runtime configurations Middleware and messaging systems Database performance (queries, indexing, pooling) Kubernetes clusters (pods, resources, scaling behavior) Linux OS tuning (CPU, memory, I/O, ulimits, networking) Identify root causes...Suggested
- ...Site Reliability Engineer (FedRAMP | AWS | Kubernetes | Terraform) Level: L2 Contract W2 (No C2C) Location : Remote, prefer PST hours FEDRAMP, AWS, Kubernetes, Terraform, CI/CD systems Role Summary Looking for a Senior Site Reliability Engineer to help build and operate...SuggestedContract workRemote work
- ...service, and perform root cause analysis. Participate in on-call support and production change activities. Platform Support (Kubernetes & Kafka) Support Kubernetes environments, including deployments, scaling, configuration, and troubleshooting. Monitor...Remote work
$20 per hour
...systems on Earth, in orbit, and everything in between. We are looking for an experienced engineer with deep working knowledge of Kubernetes and containerized technologies. You are a hands-on operator and builder who applies first-principles thinking to both software delivery...Permanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work- ...(500 Hrs) We are seeking a highly experienced Sr Site Reliability Engineer Compute Platforms to design, implement, and support Kubernetes on baremetal and hypervisor platforms in a private cloud environment. This role is responsible for the architecture, design, and...Contract workRemote work
- ...with cloud platforms (AWS, Azure, or Google Cloud Platform). Experience with containerization technologies such as Docker and Kubernetes. Knowledge of monitoring and observability tools such as Prometheus, Grafana, ELK, Splunk, Dynatrace, or Datadog....
- ...maintain observability systems to monitor the health and performance of applications and proactively identify and resolve issues. Kubernetes Expertise: Leverage your deep understanding of Kubernetes architecture to design and optimize deployment and orchestration of AI...
- ...Senior Site Reliability Engineer (SRE) Kubernetes Platform (FedRAMP High / IL5) Level: L2 Contract W2 (No C2C) Location : Remote, prefer PST hours About the Team The SRE Platform Engineering team builds and operates the infrastructure that powers our cloud. We focus on...Contract workRemote work
- ...automate cloud platforms across Azure and Google Cloud Platform. The ideal candidate will have hands-on expertise in Terraform, Kubernetes (AKS/GKE), CI/CD pipelines, security automation, vulnerability remediation, and observability. Key Responsibilities Develop...Local area
$70 per hour
...Production on-call experience in a real rotation, with incident command and blameless postmortem practice. Production Kubernetes and container experience (Docker), with cloud-native infrastructure patterns. Hands-on production ownership on at least one...Hourly payRemote work$124k - $271.2k
...What You Can Expect You will be the leader of our Platforms DevOps team. This team is responsible for Zoom's Kubernetes clusters in both datacenters and clouds, as well as for our cloudops infrastructure. The scope and mandate of the team are broad, and you will have...Work at officeRemote workShift workWeekend work- ...Monitoring: Splunk, Grafana, Dynatrace/AppDynamics, ELK. Scripting: Python, Shell, Ansible. Platforms: Linux/Unix, Kubernetes/OpenShift (preferred). Strong experience in Production Support, Incident Management, and Problem Management. Preferred...Long term contract
- ...Performance Analytics Programming: Python, Java, .NET, REST APIs, Automation & Scripting Cloud: AWS, Google Cloud Platform, Kubernetes, Microservices, Container Platforms Analytics: Power BI, ETL/Data Integration, Dashboard Development, KPI & Scorecards...Remote work
- ...Experience: 8+ Years Industry: Telecommunications / Networking / Cloud Infrastructure Primary Skills: SRE | Java | Python | AWS | Kubernetes | Docker | Terraform | Jenkins | CI/CD | Linux | Git | Ansible | Prometheus | Grafana | Splunk | ELK | Microservices | DevOps |...Long term contract
- ...and optimizing high-performance OpenSearch clusters and platforms from the ground up in production environments. Expert with Kubernetes, including troubleshooting, operations, management, and configuration of complex Kubernetes services. Proven hands-on expertise...
- ...and production issues. Distributed Systems designing and optimizing highly available, scalable, fault-tolerant systems. Kubernetes pods, resources, scaling, container behavior, computestorage. Linux Administration & Performance Tuning CPU, memory, IO,...
- ...infrastructure environments. - Knowledge of monitoring, observability, and reliability engineering practices and tooling. - Familiarity with Kubernetes concepts and containerized application platforms. - Experience leveraging AI-assisted development tools to improve software...Worldwide