SW Engineer- Developer Systems Reliability Engineering
$88k - $136.9kVisa
About UsVisa is a world leader in payments technology, facilitating transactions between consumers, merchants, financial institutions and government entities across more than 200 countries and territories, dedicated to uplifting everyone, everywhere by being the best way to pay and be paid.At Visa, you'll have the opportunity to create impact at scale — tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.Join Visa and do work that matters – to you, to your community, and to the world. Progress starts with you.Job DescriptionThe Site Reliability Engineering — SRE — role is a critical part of Visa’s Cloud Platform strategy. In this role, you will help ensure Visa’s development platform and tooling enable key stakeholders, including development engineers across the technology organization, to focus more on innovation and less on infrastructure management. You will drive the adoption of observability best practices, automation, and AI-AIOps capabilities to improve reliability, reduce toil, and resolve recurring operational issues.You must be comfortable partnering with software engineering teams and supporting their needs to ensure the security, availability, reliability, and performance of the platform. These software engineering teams are both peer engineering teams supporting the platform and customers within Visa Engineering consuming the platform. This engineer will be expected to triage complex issues, collaborate with peer infrastructure and operations teams, enhance monitoring and alerting coverage, and operationalize automation and AI-driven solutions that improve incident response, service health, and platform efficiency.This is a hands-on engineering role with a strong focus on building and advancing reliability engineering practices for the Visa Cloud Platform.Essential Functions:Help maintain the platform’s defined SLAs and SLOs by driving operational excellence, delivering value-added process and procedure improvements, and partnering with engineering and operations teams to eliminate manual touchpoints through automation and standardization.Own and operationalize end-to-end observability, alerting, and monitoring for the Visa Cloud Platform across IaaS, PaaS, and Container-as-a-Service environments, ensuring telemetry, dashboards, alerts, SLIs, and operational workflows are meaningful, actionable, and effective in supporting production reliability.Own and deliver automation and AI-AIOps initiatives that reduce toil, improve reliability, accelerate incident response, and enhance operational intelligence across the SRE organization.Partner with development teams during release and service transition reviews to define and validate operational requirements, including SLIs-SLOs, monitoring, alerting, dashboards, runbooks, incident response procedures, capacity expectations, and production readiness criteria.Partner with Operations & Infrastructure peers to support ongoing platform maintenance, enhancement, and reliability.Support multiple internal stakeholders across a variety of technical challenges by analyzing recurring issues, identifying patterns, and proposing effective solutions.Support the Visa Cloud SRE team’s 24-7-365 operating model, including shift-based and on-call coverage, with weekend support as required.Visa requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager. QualificationsBasic Qualifications:Bachelor's degree, OR 3+ years of relevant work experiencePreferred Qualifications:Bachelor's degree, OR 3+ years of relevant work experience2 or more years working in a Platform, SRE or Production Engineering group for high availability-critical platforms-applicationsBachelor's degree in IT, CS or related field and-or 3+ Years Working Experience in IT Operations and Delivery.2 years experience with CI-CD tooling such as Jenkins, Github, Bitbucket, ArgoCD, Artifactory, Bitbucket, Azure DevOps in a large-scale environment2 years experience with observability tooling such as Grafana, Prometheus, Splunk, Datadog, New Relic, DynaTrace, Sentry, etc. in a large-scale environment2 years experience supporting relational and non-relational databases [MySQL, MongoDB, PostgreSQL, etc.), including creating and running queries, managing performance and scalingBasic understanding of YAML, JSON, HTML, XML.Hands on experience in Linux and -or Windows systems and good understanding of distributed computing environments.Beginner level programming and-or scripting in 3 or more of the following: Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, Cloud Formation.Experience with AI-enabled solutions and tools such as Claude, ChatGPT, GitHub Copilot, or similar technologies, with practical application in automation, observability, incident response, or operational workflows.Experience managing container infrastructure and supporting development transformation to a container first modelExposure to Virtualization (Hyper-V, VMware, Openshift Virtualization etc.)Experience managing a distributed container platform including but not limited to deployment-release management, provisioning, capacity management, workload managementU.S. Applicants OnlyThe estimated salary range for this position is $88,000.00 to $ 136,900.00 USD per year, which may include potential sales incentive payments (if applicable). Salary may vary depending on job-related factors which may include knowledge, skills, experience, and location. In addition, this position may be eligible for bonus and equity.Visa has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program.Work HoursVaries upon the needs of the department.Travel RequirementsThis position requires travel 5-10% of the time.Mental/Physical RequirementsThis position will be performed in an office setting. The position will require the incumbent to sit and stand at a desk, communicate in person and by telephone, frequently operate standard office equipment, such as telephones and computers.Visa is an EEO EmployerQualified applicants will receive consideration for employment without regard to race, color religion, sex, national origin, sexual orientation, gender identity, disability or protect veteran status. Visa will also consider for employment qualified applicants with criminal histories in a manner consistent with the EEOC guidelines and applicable local law.SummaryLocation: US - Austin, TXType: Full time
- Role Summary: The NVM Reliability team is responsible for characterizing, modeling, and enabling next-generation Non-Volatile Memory technologies... ...:Bachelor's, Masters, or Doctoral degree in Electrical Engineering, Applied Physics, or equivalent, with 5-10 years of industry...SuggestedFull timeWork at officeLocal area
- ...Senior Reliability EngineerThe Senior Reliability Engineer works in cross functional teams with designers, customers... ...engineering team and seek opportunities to develop new/improve existing durability... ...and complex manufacturing systems.Familiar with Reliability Block Diagram...SuggestedWork at officeLocal area
- ...Senior Reliability EngineerThe Annapurna AI Manufacturing, Quality and... ...provider. As a Senior Reliability Engineer you will engage with an... ...solid understanding of computer systems to influence design for... ...development and products in the field.Develop and execute qualification...SuggestedFlexible hours
$111.72k
...Semiconductor Engineering Design PositionDuties: Perform semiconductor engineering design... ...communications industry. Responsible for developing methods for reporting statuses, results... ...and ensure all quality and reliability procedures and requirements are followed...SuggestedInternshipLocal area- ...set become the standard.Role ScopeOwn reliability for named customer workloads: their clusters... ...spin.Turn recurring customer pain into engineering fixes with the production teams.What We... ...technical level.You debug distributed systems methodically across layers you don't...Suggested
- ...Annapurna Labs in Austin, TX is seeking a Senior Manager of Quality & Reliability to lead the QnR function for Trainium AI server products,... ...strategy for liquid and air-cooled platforms and build systems across multiple suppliers. You will lead a global team, drive...
- ...Reliability EngineerLocation: Austin, TexasOnsite/Remote/Hybrid: OnsiteFull time/ ContractVisa: USC and GC and H1b on C2CJob Summary:The engineer will be responsible for leading and executing reliability tests on new client’s and supplier technologies, development of new...H1bRemote work
- ...Reliability EngineerLocation: Austin, Texas Salary Range: *** to 115,000/AnnualJob Summary: The engineer will be responsible for leading and executing reliability tests on new client's and supplier technologies, development of new test procedures to quantify the reliability...
$159k - $230k
Lead system design evaluations and implement reliability plans to assess and mitigate failure risks early in New Product... ...internal groups, while developing in-house test and qualification... ...qualifications:Bachelor's degree in Hardware Engineering, or equivalent practical...Contract workWorldwide$183k - $247.6k
...AI Manufacturing, Quality and Reliability (MQR) Team is part of AWS... ...provider. As a Senior Reliability Engineer you will engage with an... ...solid understanding of computer systems to influence design for reliability... ...products in the field. * Develop and execute qualification...Full timeWork experience placementLocal areaFlexible hours- ...We're looking for a Staff Reliability Engineer for ICON. This role will serve as technical resource... ...our fleet of Vulcan and Titan print systems. This is a high-impact individual... ...program for Vulcan and Titan systems, and develop recommendations for schedules, task...For contractors
- Neuralink is seeking a Mechanical Engineer in Austin to own the simulation-to-test loop for implant mechanical reliability. You will build explicit dynamics FEA models, design and run tests, and calibrate material models against data, ensuring implants meet impact, fatigue...
- ...017, the company has spent the last eight years developing and deploying autonomous logistics systems for real-world operations. Our V2 aircraft... ...The Opportunity Skyways is hiring a Senior Reliability Engineer to lead technical investigations and drive closure...Permanent employmentContract workLocal area
- ...About this role The Astronomer Customer Reliability Engineering (CRE) team is responsible for the success of our customers' usage of our managed... ...either raised by a customer, or from our monitoring system and then taking further steps to ensure problems are permanently...Full timeRemote workWeekend work
- Veridic Solutions is seeking an experienced Database Reliability Engineer to design, deploy, and manage CockroachDB clusters across cloud and on‑prem environments in Austin. You will monitor health, implement DR/BC, and optimize queries while driving SRE practices and...
- ...Overview ACS (Allen Control Systems) is a defense technology... ...former U.S. Navy electrical engineers with deep experience in robotics... ...Bullfrog deployments, and developing the next generation of... ...are looking for a Systems Reliability Engineer to join our team,...Full timeLocal area
$116k - $159.5k
...global leader in materials engineering solutions used to produce virtually... ...encourages you to learn, develop, and grow your career as you... ...on Quality Engineering and Reliability Engineering. We are working... ...in mechanical/Electrical/systems engineering, mechanisms and...Full timeWorldwide- ...moves the world forward.PRODUCT QUALITY & RELIABILITY ENGINEERTHE ROLE:This dynamic, highly... ..., platform, operations, and customer system bring-up and manufacturing to meet product... ....Close coordination with product/test engineering team for test gap debug and coverage...
- ...revolution and the AI era. As a Developer, you'll work on the systems powering the future-from... ...or Hardware Development Engineer, you'll contribute to the... ...and improve system reliability. Automate test suites and... ...calendar year. Job Title SW Developer Intern: Systems...Full timeContract workPart timeFixed term contractInternshipShift work
$167.7k - $245.2k
...very effective.We’re looking for talented engineers with a software or operations background... ...-scale highly available distributed systems in the cloud. You must be willing to work... ...application development teams to ensure the reliability, performance and security of our...Full timeTemporary workWork at officeLocal areaFlexible hours1 day per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible... ...observability and alerting systems.The Fleet Management team... ...that empowers our developers to build and ship products to... ...components that ensure cluster reliability and security (e.g., CoreDNS,...Work at officeLocal areaRemote workWorldwideFlexible hours- ...our team and make an impact?As a Senior Site Reliability Engineer at TeamViewer, you’ll be a key player in ensuring... ....Deploy feature and maintenance releases.Develop and enhance monitoring, alerting, and automation systems to guarantee 24/7 service reliability.Continuously...Temporary workCasual workWorldwide
- ...location(s).We are seeking a Kafka Site Reliability Engineer to help build, operate, and... ...improve platform performance, and enhance system reliability through data-driven decision... ...future of event streaming at Schwab while developing expertise in emerging AIOps, cloud,...Full timeWork at office
- ...the rapidly evolving state of the art to engineer scalable, innovative, and research... ...are focused on leading, designing and developing the platform and services that underpin... ...tooling ecosystemsOwn the operational reliability of developer tooling ecosystems, including...Full timeLocal area
- ...the specified location(s).As a Senior Reliability Engineer, you will help shape the reliability,... ...This role blends software engineering, systems engineering, and operational expertise... ...infrastructure environments.6+ years of experience developing automation solutions, operational...Full timeWork at office
$109.65k - $182.76k
...)Position SummaryWe are seeking a Site Reliability Engineer to ensure the high level of service and... ...as Terraform, Ansible, and Kubernetes. Develop automated CI/CD pipelines via GitLab to... ...Performance & Capacity Planning: Conduct system performance analysis, identify...Full timeLocal area3 days per week- ...guidance.We are seeking a Senior Site Reliability Engineer to join our newly formed Operations Excellence... .... You will work on critical platform systems including EKS infrastructure, Skyway (... ...(SOC 2, PCI, GDPR)Contribute to Developer Experience metrics and platform...Work at officeLocal area
$98.58k - $138.02k
...Region / Denver, COProduct Engineering - DevOps /Full Time /HybridRestaurant... ...CA; or Akron, OH. The Site Reliability Engineer II will be... ...skills in incident response, system monitoring, automation, and... ...cause analysis. Partner with developers, architects, vendors, and IT...Full timeWork at office- ...Oracle products and services. Design and develop designs, architectures, standards, and methods for large-scale distributed systems. Facilitate service capacity planning and... ...accept a Master’s degree in Computer Science, Engineering, or related technical field and 6 years...Temporary workRemote workFlexible hours
- ...the AI era with our groundbreaking P3 system that captures solar energy, stores it as... ...our planet. We’re hiring a Quality & Reliability Engineer to own both the quality of our supply... ...submissions, and process capability reviews Develop inspection plans, control plans, and...RelocationRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SW Engineer- Developer Systems Reliability Engineering. Be the first to apply!
- senior staff systems engineer Austin, TX
- application system engineer Austin, TX
- system safety engineer Austin, TX
- system engineer contract Austin, TX
- operations support system engineer Austin, TX
- sr systems engineer Austin, TX
- visual systems engineer Austin, TX
- active directory systems engineer Austin, TX
- system performance engineer Austin, TX
- systems engineering technician Austin, TX



