DevOps / Site Reliability Engineer
AgileEngine
Job Description
AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards. WHY JOIN US If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you! ABOUT THE ROLE We are looking for a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. This role blends platform engineering with incident command, using Terraform, CI/CD pipelines, and CSPM tools like Wiz. You will lead major-incident calls, own remediation follow-through, and build the playbooks that guide response. WHAT YOU WILL DO - Scale and maintain the ability to drive operational stability across multi-cloud environments (Azure, AWS, GCP). - Engineer unified security policies and configuration baselines using IaC (Terraform) to prevent misconfigurations. - Design, maintain, and optimize enterprise CI/CD pipelines to support continuous ASPM ingestion and deployment. - Act on continuous monitoring alerts, utilizing Cloud Security Posture Management (CSPM) tools like Wiz to secure workloads. - Serve as Incident Commander on major and critical incidents - running the bridge, directing technical workstreams, making time-critical decisions, and coordinating cross-functional responders under pressure. - Own the post-incident loop - track remediation items to closure, hold owning teams accountable to timelines, and drive systemic fixes and preventative actions across groups. - Draft and send clear, accurate, audience-appropriate incident notifications and status updates to technical teams, management, and stakeholders throughout the incident lifecycle. - Develop, maintain, and socialize divisional / group-level incident-management playbooks, runbooks, and escalation procedures that standardize response and reduce time-to-resolution. MUST HAVES - You must be authorized to work for ANY employer in the US (e.g., Green card holders, TN visa holders, GC EAD, H4 EAD, U4U with EAD), as we are unable to sponsor or take over employment visa sponsorship at this time; - 5+ years of experience . - In-depth architectural expertise in multi-cloud defense, federated IAM, and zero-trust principles . - Strong practical experience with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting . - Senior-level, hands-on incident-command experience driving major/critical incident calls to resolution in a 24x7 production environment . - Proven track record of remediation follow-up - coordinating with teams and holding owners accountable until issues are fully closed. - Demonstrated skill drafting and issuing incident notification communications to both technical and executive audiences. - Direct experience authoring divisional/group incident-management playbooks and escalation procedures. - Fully autonomous. - Drives the architecture of complex automated runbooks and mentors Middle-level SREs. - Extensive experience deploying and tuning APIs from modern CNAPP/CSPM platforms, ideally Wiz . - Prior experience building platforms subject to strict financial compliance standards ( PCI-DSS, SOC2). - Upper-intermediate English level. NICE TO HAVES - PagerDuty - hands-on experience with on-call scheduling, alert routing, and incident orchestration. - ServiceNow - familiarity with incident, problem, and change management workflows and reporting. PERKS AND BENEFITS - Growth without limits : build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget - Competitive compensation : get recognition that reflects your skills and impact, with regular performance and compensation reviews - Flexibility : work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm - Meaningful, modern projects : build impactful products using modern technologies alongside global teams and leading brands - Collaborative culture : join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized - Well-being & support : access local well-being programs and people-focused support tailored to your location
AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards. WHY JOIN US If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you! ABOUT THE ROLE We are looking for a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. This role blends platform engineering with incident command, using Terraform, CI/CD pipelines, and CSPM tools like Wiz. You will lead major-incident calls, own remediation follow-through, and build the playbooks that guide response. WHAT YOU WILL DO - Scale and maintain the ability to drive operational stability across multi-cloud environments (Azure, AWS, GCP). - Engineer unified security policies and configuration baselines using IaC (Terraform) to prevent misconfigurations. - Design, maintain, and optimize enterprise CI/CD pipelines to support continuous ASPM ingestion and deployment. - Act on continuous monitoring alerts, utilizing Cloud Security Posture Management (CSPM) tools like Wiz to secure workloads. - Serve as Incident Commander on major and critical incidents - running the bridge, directing technical workstreams, making time-critical decisions, and coordinating cross-functional responders under pressure. - Own the post-incident loop - track remediation items to closure, hold owning teams accountable to timelines, and drive systemic fixes and preventative actions across groups. - Draft and send clear, accurate, audience-appropriate incident notifications and status updates to technical teams, management, and stakeholders throughout the incident lifecycle. - Develop, maintain, and socialize divisional / group-level incident-management playbooks, runbooks, and escalation procedures that standardize response and reduce time-to-resolution. MUST HAVES - You must be authorized to work for ANY employer in the US (e.g., Green card holders, TN visa holders, GC EAD, H4 EAD, U4U with EAD), as we are unable to sponsor or take over employment visa sponsorship at this time; - 5+ years of experience . - In-depth architectural expertise in multi-cloud defense, federated IAM, and zero-trust principles . - Strong practical experience with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting . - Senior-level, hands-on incident-command experience driving major/critical incident calls to resolution in a 24x7 production environment . - Proven track record of remediation follow-up - coordinating with teams and holding owners accountable until issues are fully closed. - Demonstrated skill drafting and issuing incident notification communications to both technical and executive audiences. - Direct experience authoring divisional/group incident-management playbooks and escalation procedures. - Fully autonomous. - Drives the architecture of complex automated runbooks and mentors Middle-level SREs. - Extensive experience deploying and tuning APIs from modern CNAPP/CSPM platforms, ideally Wiz . - Prior experience building platforms subject to strict financial compliance standards ( PCI-DSS, SOC2). - Upper-intermediate English level. NICE TO HAVES - PagerDuty - hands-on experience with on-call scheduling, alert routing, and incident orchestration. - ServiceNow - familiarity with incident, problem, and change management workflows and reporting. PERKS AND BENEFITS - Growth without limits : build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget - Competitive compensation : get recognition that reflects your skills and impact, with regular performance and compensation reviews - Flexibility : work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm - Meaningful, modern projects : build impactful products using modern technologies alongside global teams and leading brands - Collaborative culture : join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized - Well-being & support : access local well-being programs and people-focused support tailored to your location
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the DevOps / Site Reliability Engineer in Irving, TX vacancy
- ...cloud-native platforms to advanced release engineering practices, our teams are redefining how... ...accelerate development and improve reliability. Your work will directly influence how... ...scripting (PowerShell, Bash). Expertise in DevOps, pipeline automation, and cloud...DevopsFull timeH1bWork at officeRemote workVisa sponsorshipFlexible hours2 days per week3 days per week
- ...Site Reliability Engineer We are looking for a Site Reliability Engineer for our client location in Dallas TX with the following skills: Java... ...Chain, Warehouse Management, Transportation. 5+ years SRE or DevOps role & Agile Scrum Methodologies. Expertise analyzing...DevopsWork at office
- ...Job Position:- Site Reliability Engineer Duration:- Long Term Client:- UPS This is a Hybrid Work Model (3x a week Onsite... ...Certifications: Google Cloud Professional DevOps Engineer Google Cloud Professional Cloud Architect Red...Devops
- ...Role: Site Reliability Engineer 6+ months Contract role Remote About the Role We are looking for a dynamic and accomplished Site Reliability... ...AI frameworks, self-healing systems, intelligent agents DevOps Tools: Jenkins, Git, Azure DevOps, CI/CD pipelines...DevopsContract workRemote work
$111.61k - $131.3k
...coordinatingcross-functionalresponseteamsanddrivingrestorationactivities.Provideleadership,coaching,mentoring,andworkloadmanagementforSRE,DevOps,andproductionsupportengineers.UtilizeoperationalmetricsincludingMTTR,MTTD,SLAcompliance,backloghealth,incidentvolume,...DevopsFull timeWork experience placementLocal area3 days per week- ...Sr. Dev Ops EngineerWe are looking for an experienced Sr. DevOps Engineer to join our team for a contract role in Irving, TX. Candidates must be available for a face-to-face interview at the client location. Strong experience in DevOps practices, CI/CD pipeline development...DevopsContract work
- Software EngineerLocation: RemotePosition Overview:The application engineer will work with application teams to assist in the modernization... ..., particularly with.NET, Java, Spring boot.• Knowledge of DevOps practices is a bonus.• Familiarity with Azure cloud architecture...Devops
- ...SRE/Devops Engineer Locations can be any of Tampa, Jersey City or Dallas. Below is Detailed JD for reference:. Design... ...planning and performance analysis to ensure Risk platforms scale reliably under high load. Metrics & Continuous Improvement:...Devops
- ...developing, and deploying java based applications. Architect and implement solutions using spring boot framework to ensure efficient and reliable performance. Utilize amazon web services to develop and deploy scalable and secure cloud based solutions. Implement...Devops
- ...Kubernetes; exposure to React.js is a plus.Proficient in CI/CD and DevOps tools such as GitHub, Jenkins, Urban Code Deploy / Harness.... ...best practices (especially in modern architecture and engineering trends).Ability to lead engineering efforts independently, driving...DevopsWork at office3 days per week
- ...Lead Software Engineer Seeking a Senior Application Developer with strong hands-on experience in developing server-side components... ...state-of-the-art solutions using new stack development using Agile/DevOps high standard/Micro services/Docker for application hosting....Devops
- ...performance tuning, troubleshooting, and optimization. Create and maintain technical design documentation using UML. Follow DevOps processes for application build, deployment, and release activities. Analyze and troubleshoot technical issues and provide effective...Devops
- ...input into and lead the technical infrastructure designDevelopment of POCs and development templatesStrong understanding of CI/CD and devops principlesTechnical / Functional Proficiency:The ideal candidate will have a total of 9+ years of experience in software...Devops
- ...description Java Developer Java Backend Engineer Experience Mid to Senior About the... ...will work closely with architects leads DevOps QA and business teams to deliver modern... ...with strong focus on performance reliability and clean architecture Develop and optimize...Devops
- ...Python Developer & API Engineer (GCP Cloud)Client: Brillio || Verizon Location: Irving, TX or Tampa, FL Experience Level: 5+ years Visa... ...on expertise in API development, cloud integrations, and modern DevOps tools, along with strong proficiency in Python and FastAPI.Must...DevopsH1b
- ...stakeholders and testing team. Required Skills Tools/ Framework: Flutter , NodeJS , Appium ,Dart . Integration patterns: REST APIs (OpenAPI 3.0), Graph QL. CICD: Azure DevOps Services (Repos, pipelines), SonarSource (security), Jfrog (artifact management)...Devops
$157k
...Software EngineerConsultant (Software Engineer) needed for Kairos Technologies Inc. located in Irving, TX. Will engage in software engineering... ...services. Will perform CI / CD and version control using Azure DevOps and Github. Will deliver and track user stories using Azure...DevopsRelocation- ...decisions in developing standards and best practices for engineering and large-scale technology solutionsDesign, optimize,... ...certification on GCP and/or AzureExperience with Agile, CI/CD, DevOps concepts, and Site Reliability Engineer (SRE) principlesProficient on container-based...DevopsWork experience placement
- ...enterprises. Job Summary We are seeking an experienced Senior Software Engineer with a strong background in Java-based enterprise application... ...) Experience with Kubernetes Strong Understanding Of CI/CD And DevOps Tools GitHub Jenkins Urban Code Deploy / Harness Expertise in...DevopsLocal area
- ...Infinity (latest versions preferred)Exposure to CI/CD pipelines and DevOps practices for PegaKnowledge of cloudExperience working in... ...Skills8+ years of IT experience5+ years of hands on Pega engineering experienceComplex case management (Jira/Confluence a plus)System...DevopsContract work
- As a DevOps Engineer, you will work collaboratively with software engineering team to deploy and operate the systems and also help in automating and streamlining the operations and processes. You will be required to build and maintain tools for deployment, monitoring and...Devops
- ...Description: We are seeking a seasoned Senior Lead Data Engineer with a dual-threat background in robust backend development and... ...code reviews, and ensure high code quality. • Work with DevOps teams to deploy and monitor applications in production environments...DevopsFull time
- ...DevOps EngineerLocation: Irving, TX - Hybrid Duration - 10 Months An analytics platform engineer supporting Azure Synapse and Azure Data Factory from a DevOps perspective requires a hybrid skill set spanning primarily cloud infrastructure management, automation and CI...Devops
- ...breakers, rate limiters, and fallback mechanisms to enhance system reliability.Validate the effectiveness of these patterns through testing... ...processes.Stay updated with the latest performance engineering tools, techniques, and industry trends.Required SkillsExperience...Devops
- ...Title: AWS DevOps Engineer Location: US Remote Primary Skills (must have): AWS DevOps, Terraform , IAC, GitLab CI/CD, Docker, Kubernetes Secondary Skills : AWS cloudwatch, AWS Service Mesh, API Gateway , Kong Job disruption : • Fully automate...DevopsRemote work
- ...will have strong experience with MongoDB, OpenShift, AWS, CI/CD, DevOps automation, TDD, and cloud deployments. The candidate... ...Generative AI solutions using LLMs , including prompt engineering, RAG, and AI-driven automation use cases. Roles & Responsibilities...DevopsContract work
$60k - $135k
...DevOps Engineer L4Wipro Limited is a leading technology services and consulting company focused on building innovative solutions that address clients' most complex digital transformation needs. Leveraging our holistic portfolio of capabilities in consulting, design, engineering...DevopsMinimum wageLocal area- ...Cloud Engineer Location: Chicago, IL or Minneapolis, MN (other locations - Portland,... ...CI/CD pipelines using GitLab and other DevOps tools. Develop Infrastructure as Code... ...teams to improve deployment frequency, reliability, and rollback strategies. Configured...DevopsContract work
- ...DevOps EngineerThis exciting growth opportunity shall act as a Senior DevOps Engineer working with multiple development teams focused on Cloud, Edge, and On-Prem deployments... ...improvement of SRE practices and system reliability.Implementing and maintaining infrastructure...DevopsWork experience placement
- DevOps EngineerLocation: USARecruiting DevOps Engineers (Mexico Only)- 4+ years of experience with application development using Java- 3.5+ years of experience with DevOps tools like Maven, Jenkins, Puppet, Chef, UrbanCode, etc.- Solid understanding of Continuous Integration...DevopsRelocation
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to DevOps / Site Reliability Engineer. Be the first to apply!
Related searches
- devops aws developer (remote) Irving, TX
- devops engineer Irving, TX
- senior devops cloud engineer Irving, TX
- senior devops engineer remote Irving, TX
- big data devops engineer Irving, TX
- devops visa sponsorship available Irving, TX
- junior devops remote Irving, TX
- azure devops Irving, TX
- senior devops Irving, TX
- linux devops Irving, TX


