Site Reliability Engineer
Infosys Technologies
OverviewThe Infosys Cloud unit is dedicated to empowering enterprises with innovative cloud solutions that drive digital transformation and operational excellence. We specialize in leveraging advanced cloud technologies and AI-driven insights to create scalable, secure, and resilient infrastructures. Our solutions enable organizations to achieve agility, efficiency, and sustainable growth in a hyperconnected world. Join us to be part of a pioneering team at the forefront of cloud and AI innovation. You'll have the opportunity to work with cutting-edge technologies, collaborate with industry experts, and contribute to transformative projects that shape the future of business. We are committed to fostering a culture of continuous learning and growth, ensuring that our team members thrive in a dynamic and supportive environment. If you're passionate about cloud and AI, and eager to make a significant impact, the Infosys Cloud unit is the perfect place for you to grow and excel.Job DescriptionIn the assigned Job Role of Infrastructure Consultant 2, your Area Of Responsibility will be as below: Collaborate with internal and client teams to resolve complex incidents, conduct root cause analyses, and document findings with preventive recommendations Participate in evaluation of client IT infrastructure, prepare actionable assessment reports, and support due diligence to document infrastructure maturity and improvement opportunities Contribute to the design of scalable, cost-effective IT infrastructure solutions, review reusable components, and develop technical documentation for deployed systems Align release schedules and environment readiness, execute deployments as per protocols, perform post-deployment testing, and manage version control to track changes Co-ordinate maintenance schedules, emergency fixes, and technology upgrades while ensuring uninterrupted integration into existing systems and processes Facilitate performance data analysis across systems, coordinate insights on system behavior, and support capacity planning to optimize performance Conduct security checks, recovery drills, and compliance audits, implement security measures, and coordinate continuity plans to maintain adherence to standards Gather feedback to identify automation opportunities, analyze existing infrastructure processes, and propose enhancements for efficiency gains Act as liaison with onsite, offshore, and vendor teams to document project requirements, ensuring effective collaboration Develop a centralized repository of technical and procedural knowledge, leveraging insights from other projects to drive efficiency and retain organizational expertiseYour contribution to the team: A collaborative spirit and excellent communication skills. Ability to handle complex incidents and implement resolutions A knack for conducting IT infrastructure assessment and identifying key optimization opportunities Focused approach towards deployment management, system optimization, and process automation initiatives including sector specific focus The ability to work with cross-functional teamsRequired Skill and ExperienceReliability Engineering· Support SLIs, SLOs, error budgets, and reliability KPIs.· Drive service availability, resiliency, scalability, and performance improvements.· Establish proactive operational models and reliability governance.· Engage in reliability reviews and continuous improvement programs.Terraform & Infrastructure Automation· Drive Infrastructure as Code adoption using Terraform.· Develop reusable modules and platform standards.· Implement policy-as-code and automation frameworks.· Reduce manual infrastructure management through automation.Observability & Datadog Strategy· Define observability standards and monitoring frameworks.· Manage Datadog implementation including dashboards, alerts, APM, logs, and tracing.· Establish observability-as-code practices.· Improve alert quality and operational visibility.Disaster Recovery & Resiliency· Define DR strategy and recovery objectives.· Engage in DR testing, failover planning, and resiliency reviews.· Ensure business continuity readiness and compliance.· Manage recovery runbooks and operational procedures.· Engage in periodic DR exercises, failover validations, recovery evidence collection, post-DR action tracking, and continuous improvement of recovery runbooks.Vulnerability Management & Governance· Support vulnerability management and remediation governance.· Drive onboarding, patching, and vulnerability tracking programs.· Collaborate with security teams to reduce risk exposure.· Establish vulnerability KPIs and executive reporting.· Own vulnerability remediation governance, aging backlog reduction, SLA-based closure tracking, exception governance, patch compliance reporting, and coordination with security and platform teams.CI/CD & Release Reliability· Support release reliability using Harness and Helm Charts.· Improve deployment stability and automation.· Implement quality gates and rollback standards.· Partner with DevOps teams on release governance.Preferred Skill and ExperienceHarness · Helm Charts · Datadog · Incident ManagementPreferred Certifications· AWS DevOps Engineer Professional· AWS Solutions Architect Professional· Terraform Associate· FinOps Practitioner· ITIL Foundation· SRE Practitioner· Datadog CertificationAdditional Required Qualifications• Bachelor’s degree or foreign equivalent required from an accredited institution. Will also consider three years of progressive experience in the specialty in lieu of every year of education. • This position may require relocation and/or travel to work/project location. • Candidates authorized to work for any employer in the United States without employer-based visa sponsorship are welcome to apply. Infosys is unable to provide immigration sponsorship for this role now or in the future.EEO/About UsBenefitsAlong with competitive pay, as a full-time Infosys employee you are also eligible for the following benefits:Medical/Dental/Vision/Life InsuranceLong-term/Short-term DisabilityHealth and Dependent Care Reimbursement AccountsInsurance (Accident, Critical Illness , Hospital Indemnity, Legal)401(k) plan and contributions dependent on salary levelPaid holidays plus Paid Time OffAbout Us Infosys is a global leader in next-generation digital services and consulting. We enable clients in more than 50 countries to navigate their digital transformation. With over four decades of experience in managing the systems and workings of global enterprises, we expertly steer our clients through their digital journey. We do it by enabling the enterprise with an AI-powered core that helps prioritize the execution of change. We also empower the business with agile digital at scale to deliver unprecedented levels of performance and customer delight. Our always-on learning agenda drives their continuous improvement through building and transferring digital skills, expertise, and ideas from our innovation ecosystem.EEO Infosys provides equal employment opportunities to applicants and employees without regard to race; color; sex; gender identity; sexual orientation; religious practices and observances; national origin; pregnancy, childbirth, or related medical conditions; status as a protected veteran or spouse/family member of a protected veteran; or disability.Technical/Domain Skill 1Technology|Cloud Platform|Terraform Technical/Domain Skill 2Technology|DevOps|Site Reliability Engineering(SRE) Technical/Domain Skill 3Technology|Cloud Platform|Terraform Technical/Domain Skill 4Technology|Cloud Security|AWS - Vulnerability Management Technical/Domain Skill 5Technology|Cloud Platform|FinOps Work LocationSan Antonio, TX CountryUSAState / Region / ProvinceTexasCompanyITL USA Interest GroupInfosys Limited Job RoleInfrastructure Consultant 2Career RoleConsultant - Infrastructure Management - USAuto req ID: 150426BR
$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...SuggestedPermanent employmentFull timeContract workRemote workFlexible hours- ...teammates passionate about what we do!What We Need:iHeartMedia Entertainment, Inc. seeks candidates for the position of Senior Site Reliability Engineer (SRE), responsible for leading a talented team of SREs/DevOps Engineers across a wide variety of Cloud Services to ensure...SuggestedFull timeFlexible hours
- ...document and/or identify application optimizations Qualifications: Bachelor’s degree or equivalent experience in an software engineering discipline Proficiency in at least one software language (e.g. Java, Python, etc.) Proficiency in SQL – preferably with...Suggested
- ...Partner with software developers, platform engineers, and IT staff to improve system design,... ...requirements, service quality, reliability, security, and compliance needs. Drive continuous... ...Required: 8+ years of experience in Site Reliability Engineering, DevOps, Platform...SuggestedWork at officeRemote work
- ...Skills Kubernetes and Docker knowledge will be considered high Bachelor’s degree or equivalent experience in a software engineering discipline Mastery in at least two or more software languages (e.g. Python, Java, Go, etc.) with respect to designing, coding,...Suggested
$131.5k
...functionality of Tenable cloud products and ensuring they’re reliable and highly available in cloud environments Responsible for responding... ...with peers on complex projects Collaboration with cloud engineers in understanding new cloud technologies, assessing impact to security...Work experience placementH1bLocal areaRemote workFlexible hours- ...Job Description Job Description Job Description Job Title: Site Reliability Engineer Location: San Antonio, TX (Lackland AFB / Onsite 5x per week) Security Clearance: TS/SCI (or SCI Eligibility) We question. We listen. We adapt. Be honest. Be pragmatic...
- ...to create effective SLOs and engage with exception processes when technical limitations exist. Work with dependent process (Site Reliability Management, SRE, Event Management and Incident Management, CMDB) to create enhancement stories that will improve SLO efficacy...
$114.08k - $218.03k
...recommendations to business leaders on best solutions.Independently experiments with new patterns and technologies.Helps establish and improve engineering best practices, concepts, and patterns with peers and the business.Understands the customer and proactively identifies innovative...Full timeH1bWork at officeRemote workHome officeRelocation packageFlexible hours$91.27k - $121.69k
...lasting impact. We’re looking for top-tier talent ready to take on the challenge. Join us in building the future.The RoleThe IT Systems Engineer II provides advanced Tier II support by troubleshooting and repairing network devices, tools, and services for a nationwide fiber...Temporary workShift workNight shift- ...Reliability Engineer The Reliability Engineer is responsible for overall evaluation and improvement of product reliability, through collecting and analyzing failure data, recommending design and or process improvement, support MTBF modeling activities, interfacing...
- ...innovative ways to help people? Do you like having the autonomy to build new solutions from the ground up? If so, being a Software Engineer III at Frost could be the job for you.At Frost, it’s about more than a job. It’s about having a flourishing career where you can...Full time
$100.17k - $135k
...Today, is a leader in big data solution development and deployment, with expertise in cloud-based services, software and systems engineering, cyber capabilities, and data science. Enlighten provides continued innovation and proactivity in meeting our customers’ greatest...Work experience placementWork at officeWork from home2 days per week- ...collaborative team environment.Position SummaryThe Senior Platforms Engineer for the Refining IT Edge team is responsible for designing,... ...firmware, RF, hardware, cloud, and operations teams to deliver reliable, secure, and observable systems that support real-world...Full timeLocal areaRemote work
- ...As a Hosting and Infrastructure Platform Engineer working at our San Antonio headquarters... ...server hardwareAble and willing to work on-site, in-person at the Valero San Antonio... ...engineering teams to ensure infrastructure reliability, performance, and scalability for AI/ML...Work at officeLocal areaImmediate start
- ...escalate risks, threshold concerns, and recurring issues to senior engineers or the Platform Manager.Participate in incident response as a... ..., and support knowledge articles.Schedule & Presence: This on-site role supports 24/7 operations through real-time collaboration,...Monday to FridayShift work
$98.59k - $173.58k
OverviewGIS Solution Engineers on our commercial team are highly technical, trusted advisors to customers across many different commercial markets working with some of the largest and most complex companies around the world. They inspire customers supporting mission-critical...$149k - $350k
...developing agentic product functionality. We’re looking for engineers with a background in platform engineering or machine learning... ...functionality, while improving the technical quality, performance, and reliability of our AI features. You’ll work alongside developers across...Full timeRemote workWork from home- ...our team in San Antonio, Texas and create practical, scalable tools that turn complex operational data into reliable insights. This role is ideal for an engineer who enjoys working across data systems, application layers, and user-facing utilities to solve meaningful business...
$177k - $185k
OverviewThis is a unique opportunity to grow your technical skills while contributing to projects that matter. You’ll work alongside engineers and mentors in a collaborative, cross-functional environment that values learning, innovation and purpose. From building intuitive...- ...a highly skilled, experienced, and strategic Senior Software Engineer to be a leader within our Software Development Team in the Midstream... ...applications, including maintainability, performance, reliability, security, scalability, and technical debt.Partners with Product...Full timeLocal area
$119.57k - $170k
...Today, is a leader in big data solution development and deployment, with expertise in cloud-based services, software and systems engineering, cyber capabilities, and data science. Enlighten provides continued innovation and proactivity in meeting our customers’ greatest...Work experience placementWork at officeWork from home2 days per week$149k - $350k
...creation. Our work powers products such as AI Assistant, Make, Sites, and other Figma surfaces. This team’s and our work will be... ...ergonomic, scalable, and extendable for the company. We’re hiring engineers to join Code Platform to work on our core code translation...Full timeRemote workWork from home- ...innovative ways to help people? Do you like having the autonomy to build new solutions from the ground up? If so, being a Software Engineer II at Frost could be the job for you.At Frost, it’s about more than a job. It’s about having a flourishing career where you can...Full time
$131.3k - $237.35k
...exciting career opportunity for a Platform Engineering Team Lead to support the USAF Defensive... ...industry-leading business solutions.On-Site Requirement:The position requires on-... ...improvement efforts focused on automation, reliability, scalability, and developer productivity...Full timeWork at office$86.8k - $198k
...with precision and innovation.You’ll take ownership of complex engineering challenges by architecting new mission‑focused capabilities that... ...total benefits by visiting the Resource page on our Careers site and reviewing Our Employee Benefits page.Salary at Booz Allen is...Full timeContract workPart timeWork at officeLocal areaRemote work$99k - $225k
AI Engineer The Opportunity: As an AI Engineer, you will integrate AI-enabled capabilities... ..., ensuring these capabilities operate reliably, securely, and efficiently within distributed... ...the Resource page on our Careers site and reviewing Our Employee Benefits page....Full timeContract workPart timeWork at officeLocal areaRemote work$124k - $280k
...ApplicableSpecialismData, Analytics & AIManagement LevelSenior ManagerJob Description & SummaryAt PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust data solutions for clients. They play a...Full timeH1b$228k - $350k
...organization owns Figma’s build and CI infrastructure, enabling engineers to ship changes to production quickly and safely. We build and... ...initiatives reducing build/test times and improving CI reliability, all while balancing technical excellence with pragmatic delivery...Full timeRemote workWork from home$149k - $350k
...to expand and diversify, and new technological advancements in AI and generative design emerge, the opportunity for AI experience engineering is greater than ever. You’ll collaborate closely with a cutting edge AI Research team and your work will help raise the ceiling...Remote jobFull timeTemporary workWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site services specialist San Antonio, TX
- construction site safety San Antonio, TX
- site leader San Antonio, TX
- official site San Antonio, TX
- website content developer San Antonio, TX
- on site coordinator San Antonio, TX
- IT site lead San Antonio, TX
- site safety San Antonio, TX
- junior website developer San Antonio, TX
- on-site clinical research associate (traveling/remote) San Antonio, TX


