Get new jobs by email
- ...Infrastructure Program Director Vancouver, WA Onsite Local person [ EUC& Cloud ] Only Citizen / 100% Onsite / Infrastructure Operation Delivery Director Experience: 15+ Years Job Description The Infrastructure Program Director is responsible for leading large-scale...Suggested
- ...management, governance, stakeholder coordination, resource planning, risk and issue management, reporting, migration planning, and infrastructure modernization support for the Lab Consolidation initiative. Project Manager shall establish and manage structured project...SuggestedRemote workRelocation
$90k - $120k
..., NC, NJ, NY, OH, PA, TX, WA Department: Instructor Development & Quality Assurance (IDQA) Position: Director, Critical Infrastructure Instructor Development Reports to: Sr. Director, Instructor Academy Location: National, with occasional travel to on-...Suggested- ...fellow advisors, foster a high-performance culture, and help shape the future growth of the firm. You’ll have the support, infrastructure, and brand strength of a firm with over 175 years of history—while maintaining the freedom to grow your practice and develop a...SuggestedVisa sponsorship
- ..., prefer PST hours Visa Requirements : UC Citizenship About the Team The SRE Platform Engineering team builds and operates the infrastructure that powers our cloud. We focus on delivering reliable, scalable, and simple platforms that enable product teams to move quickly...Suggested
$65.12 per hour
...have strong experience in Oracle Enterprise Linux administration, SRE practices, automation with Ansible, and high availability infrastructure and a proven ability to increase platform reliability, scalability, and security through automation and operational excellence....Suggested- ...a global environment. The ideal candidate will have strong hands-on experience with SRE, DevOps, production operations, cloud infrastructure, Kubernetes, Terraform, observability, incident management, and automation. This role will focus on improving service stability...Suggested
- ...ecosystem. Oversee the design, build, and lifecycle management of Linux servers, including storage, virtualization, and associated infrastructure. Manage high availability (HA) configurations, clustering, and load-balanced environments to ensure minimal downtime. Drive...SuggestedContract work3 days per week
- ...Location: Remote Duration- 3+ Month Contract ~6+ years of experience as a DevOps Engineer, Site Reliability Engineer, or Infrastructure Operations Engineer with a strong focus on compute ~ Strong hands-on experience operating bare metal compute environments at...SuggestedContract workRemote work
- ...disaster recovery strategies. Establish architecture governance and design review processes. Partner with business, infrastructure, cloud, and security teams. Required Skills ~10+ years in enterprise/solution/application architecture. ~5+ years...SuggestedContract workRemote work2 days per week
- ...dynatrace and other tools to reduce recurrence and improve long term stability of GIS solutions. Optimize application and infrastructure configurations with a focus on resource utilization latency and throughput to create efficient systems that support sustainable...Suggested
- ...Job Description Data Infrastructure Site Reliability Engineer (SRE) AWS & Big Data Platforms Dallas, Texas (Remote): Skills : (AWS, Hadoop, on-premises, AI) Skills - Site Reliability Engineering, Big Data, AWS, AWS EMR, AWS EKS, AWS MSK, AWS Athena, AWS Glue, Spark...Suggested
- ...the reliability, availability, performance, and scalability of production systems by applying software engineering practices to infrastructure and operations. Partners with development teams to build resilient, observable, and automated platforms that meet defined...Suggested
- ...microservices on Kubernetes/EKS Implement GitOps workflows; enforce trunk-based development and deployment gates Automate infrastructure provisioning using CloudFormation, Terraform, and AWS CDK Manage Akamai CDN configuration: edge rules, caching policies,...Suggested
- ...application paths. PostgreSQL query optimization, indexing, connection pooling, and database performance. Cloud Azure Azure infrastructure, networking, and application operations. Capacity Planning & Scalability workload modeling, benchmarking, forecasting, and...Suggested
- ...Service Level Objectives (SLOs), and Error Budgets. Automate operational tasks and repetitive processes using scripting and Infrastructure as Code (IaC). Lead incident response activities, troubleshooting, root cause analysis (RCA), and post-incident reviews....
- ...an immediate need for an experienced Google Cloud Platform Engineer with strong expertise in Google Cloud Platform, Kubernetes, infrastructure automation, observability, and Site Reliability Engineering to design, automate, optimize, and support scalable cloud...
$70 per hour
...incident command and blameless postmortem practice. Production Kubernetes and container experience (Docker), with cloud-native infrastructure patterns. Hands-on production ownership on at least one major cloud (AWS, Google Cloud Platform, or Azure). Terraform...- ...platform experience (must have). ~ Deep knowledge of Airflow DAG development, scheduling, and optimisation. ~ CI/CD and Infrastructure as Code (IaC) experience. ~ AWS Cloud proficiency. ~ Monitoring and observability tooling experience. ~ Git for...
- ...reliability, scalability, and deployment efficiency. The ideal candidate has a strong software engineering mindset combined with deep infrastructure expertise. They should understand how applications are built and deployed, be passionate about automation, and have experience...
- ...software engineering teams to: Review performance-critical code paths Propose improvements at code, configuration, or infrastructure level Improve system observability (metrics, logs, traces) Communicate complex technical findings clearly to both...
- ..., GA / New Jersey (Hybrid) Duration: Long-Term Contract Experience: 8+ Years Industry: Telecommunications / Networking / Cloud Infrastructure Primary Skills: SRE | Java | Python | AWS | Kubernetes | Docker | Terraform | Jenkins | CI/CD | Linux | Git | Ansible | Prometheus...Long term contract
- ...cloud resource standards, FinOps reporting, and compliance-oriented automation. You should be able to apply strong practical cloud infrastructure judgment, partner with other teams, and mentor engineers who need help navigating secure and compliant cloud adoption....
- ...and deploy Model Context Protocol (MCP) clients and servers that securely connect enterprise LLMs with engineering tools, cloud infrastructure, and operational platforms. Build custom MCP services using Python, TypeScript, JavaScript, or Node.js to expose...
$55 - $60 per hour
...cloud experience SRE, Reliability Engineering, and DevOps practices OpenTelemetry and modern observability tools Infrastructure-as-Code and automation frameworks Qualified candidates should APPLY NOW for immediate consideration! This position is only...- ...operational documentation Required Qualifications ~ Experience of 7+ years in Site Reliability Engineering, DevOps, or infrastructure engineering ~3+ years in SRE leadership/Architect roles ~3+ years hands-on experience with Datadog, Splunk, Grafana,...
- ...vulnerability remediation, and observability. Key Responsibilities Develop and maintain reusable Terraform modules and Infrastructure-as-Code standards. Integrate security scanning tools into CI/CD pipelines for code, containers, dependencies, and IaC....Local area
- ...replication, and storage utilization Analyze and resolve operational issues, platform instability, and production incidents across infrastructure, platform, and application layers Conduct incident response, root cause analysis, and post-incident remediation to drive...
- ...Requirements: UC Citizenship We are looking for a Senior SRE (Linux) to build and operate the systems that run on every host across our infrastructure. Our team owns OS configuration, automation, AMIs, and container base images across a large-scale hybrid environment (AWS + on-...Remote work
- ...Model Deployment, Google Cloud Platform, Monitoring & Observability, Agentic AI, Architecture Secondary Skills : Kubernetes, Infrastructure Automation, Security Controls, Site Reliability Engineering (SRE), DevOps Key Responsibilities Design and implement...