Staff Site Reliability Engineer
Medical City Arlington
Staff Site Reliability Engineer
Last year our HCA Healthcare colleagues invested over 156,000 hours volunteering in our communities. As a Staff Site Reliability Engineer with HCA Healthcare you can be a part of an organization that is devoted to giving back!
Job Summary And Qualifications
Position Summary
What makes HCA Healthcare Information Technology Group (ITG) unique as a technology company is that our solutions ultimately impact the care of patients. Although our skills are needed in many industries, we in ITG apply them specifically to the noble cause of healthcare. We are "Healthcare Inspired." This guiding vision pervades and positively influences every level of our organization. It shapes our mission, defines our values, and brings our leaders and employees together in a shared enthusiasm for their work, setting ITG apart as a uniquely purpose-driven company in the IT industry. As a part of that, we exist to raise the bar, unlock possibilities, and care like family.
As a Staff Site Reliability Engineer (SRE), you will provide SRE best practices for mission-critical applications across the enterprise. When these applications fail, you'll have the skills and decision-making capabilities to quickly restore services, investigate the root cause, and develop a plan that mitigates future failures. You will spend time analyzing system performance and identifying ways to enhance the reliability of our environments, from developing dashboards, performing configuration changes, building robust monitoring systems, and learning how to leverage automation to drive efficiencies. You will help drive uptime and reliability across the enterprise.
We are on a mission to change the face of the healthcare industry through value-driven products. These products will create innovation for all healthcare users across HCA's nationwide ecosystem. To do this, we are building both curious and quick teams to adapt to new technologies.
Major Responsibilities:
- Practices and adheres to the "Code of Conduct" philosophy and "Mission and Value Statement."
- Promote a collaborative team environment and work closely with colleagues to achieve business objectives.
- Collaborate with stakeholders (e.g., business stakeholders, product owners, project managers, and end users) to understand functional and non-functional requirements.
- Lead Investigations and solution proposals to development and design problems.
- Participate with team members in scope of work estimation and forecasting.
- Improve performance of existing software by diagnosing and resolving critical issues.
- Prepare technical documentation, including software & architectural design evaluation plans, data flow diagrams, test results, and technical manuals.
- Adhere to and influence established development practices and processes.
- Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault finding.
- Ongoing review of technology, infrastructure, and code to enhance and build resiliency into the applications.
- Create sustainable systems and services through automation and uplifts.
- Balance feature development & deployments with speed, reliability, and well-defined service-level objectives.
- Partner with development teams and vendors of 3rd party applications to improve services through rigorous testing and release procedures.
- Build/Develop automations to "self-heal" applications and reduce the toil of manual operational tasks. Pursuit of operational excellence, uptime, and reliability of our applications
- Participate, lead, and drive in creating postmortem analysis of why services broke or degraded, including recommendations for long-term fixes. It may require going across multiple teams and organizations within the enterprise. Determine root-cause for all production-level incidents and write corresponding high-quality RCA reports.
- Collaborating and building relationships across business and technology organizations, providing sound analysis, and thought leadership.
- Support system upgrades, architecture design, implementations, and deployments.
- Ability to work in a complex organization, navigate multiple verticals of expertise and negotiate, guide direct and influence your peers to provide real solutions.
- Maintain industry knowledge in software development, architecture, and development products, such as databases, security, and observability (Dynatrace), automation products.
Education & Experience:
- Bachelor's degree Computer Science or related field preferred
- 8+ years of experience a Software development or engineering roles required
- Or equivalent combination of education and/or experience
- Microsoft Certified: Azure Solutions Architect Expert preferred
- Microsoft Certified: DevOps Engineer Expert preferred
Knowledge, Skills, Abilities, Behaviors:
- Knowledge of infrastructure, frameworks, and software/cloud design patterns for implementing applications in the cloud preferred
- Experience in the use and implementation of relevant tools and platforms (e.g., cloud platforms (IaaS and PaaS), web technologies, client-server technologies, continuous integration, and deployment) required
- Experience with version control (Git) and open-source practices preferred
- Experience in one or more coding languages. (JavaScript/Typescript, C#, Python, Java, Swift or Kotlin) preferred
- Experience with automation of CI/CD pipelines preferred
- Experience with IaC such as Terraform preferred
- A proactive approach to spotting problems, areas for improvement, and performance bottlenecks required
- Be a creative thinker, not bound by "the way things have always been done". What you know is less important than how well you learn and innovate. We don't need engineers who know all the answers; we need engineers who can invent the answers no one has thought of yet, to the questions yet to be asked required
- Experienced in helping define SLIs, SLOs & SLOs, and the experience to build observability to report on operating against those objectives required
- Strong ability to communicate complex technical information in a condensed manner to various stakeholders verbally and in writing required
- Ability to build and maintain strong cross-functional partnerships at all levels of the organization required
- Strong: Learning and teaching other team members and others external to the team preferred
- Ability to work, make aligned decisions, plan, and accomplish goals without explicit direction/guidance from leadership required
- Experience with system architectures, how software systems interact, and integrate required
- Ability to evaluate new technologies to assist senior leadership align it to the HCA Healthcare strategic roadmap required
- Strong understanding of SRE practices and implementations required
- Expertise in knowledge of Linux and Windows Systems Administration and how to manage through code required
- Ability to determine best practices and articulate authoritative direction required
- Ability to help establish and grow the SRE principles with the team required
- Growth mindset and a willingness to learn new skills, technologies, and frameworks required
Benefits
HCA Healthcare, offers a total rewards package that supports the health, life, career and retirement of our colleagues. The available plans and programs include:
- Comprehensive benefits for medical, prescription drug, dental, vision, behavioral health and telemedicine services
- Wellbeing support, including free counseling and referral services
- Time away from work programs for paid time off, paid family leave, long- and short-term disability coverage and leaves of absence
- Savings and retirement resources, including a 401(k) Plan with a 100% match on 3% to 9% of pay (based on years of service), Employee Stock Purchase Plan, flexible spending accounts, preferred banking partnerships, retirement readiness tools, rollover support and financial wellbeing counseling
- Education support through tuition assistance, student loan assistance, certification support, dependent scholarships and a partnership with Galen College of Nursing
- Additional benefits for fertility and family building, adoption assistance, life insurance, supplemental health protection plans, auto and home insurance, legal counseling, identity theft protection and consumer discounts
Learn more about Employee Benefits
Note: Eligibility for benefits may vary by location.
HCA Healthcare has been recognized as one of the World's Most Ethical Companies® by the Ethisphere Institute more than ten times. In recent years, HCA Healthcare spent an estimated $3.7 billion in cost for the delivery of charitable care, uninsured discounts, and other uncompensated expenses.
"There is so much good to do in the world and so many different ways to do it." - Dr. Thomas Frist, Sr. HCA Healthcare Co-FounderBe a part of an organization that invests in you! We are reviewing applications for our Staff Site Reliability Engineer opening. Qualified candidates will be contacted for interviews. Submit your application and help us raise the bar in patient care!
We are an equal
$81.1k - $187k
...architect infrastructure and service to ensure reliability and functionality. Forecasts demands and... ...impact and develops knowledge of site reliability trends.Only Oracle brings together... ...guidance and mentorship to junior engineers. Communicate status, risks, blockers, and...SuggestedTemporary workFlexible hours$81.1k - $187k
...architect infrastructure and service to ensure reliability and functionality. Forecasts demands and... ...impact and develops knowledge of site reliability trends.Only Oracle brings together... ...Science, Information Technology, Engineering, or a related field, or equivalent practical...SuggestedTemporary workFlexible hours$55k - $151.47k
...ApplicableSpecialismIFS - Internal Firm Services - OtherManagement LevelSenior AssociateJob Description & SummaryThe OpportunityAs a Site Reliability Engineer - Senior Associate, you will play a pivotal role in enhancing the reliability, scalability, and performance of our...SuggestedFull timeH1b$121.5k - $264.1k
...infrastructure and service and shares guidance on practices for reliability and functionality. Provides direction to ensure accurate... ...experimenting with new technology, executing improvements, building site reliability knowledge, and providing clear data.Only Oracle brings...SuggestedTemporary workImmediate startFlexible hours$84.9k - $209.5k
This role combines strategic architecture with practical systems engineering, deployment, automation, patching, troubleshooting, incident response, and compliance support. The Principal Site Reliability Engineer will work across Windows, Linux, Oracle Cloud Infrastructure...SuggestedTemporary workFlexible hours$81.1k - $187k
...As a Senior Site Reliability Engineer, you will help design, deploy, maintain, and improve reliable, secure, and scalable infrastructure and services. You'll proactively identify operational risks and potential failure points, troubleshoot system and application issues...Temporary workFlexible hours- ...As a Senior Site Reliability Engineer, you will help design, deploy, maintain, and improve reliable, secure, and scalable infrastructure and services. You’ll proactively identify operational risks and potential failure points, troubleshoot system and application issues...
- ...Site Reliability Engineer - Compute Focus Duration: 6 Months to Hire Location: On-Site 2-3 days/week in Nashville, TN Job Description: We are seeking a skilled Site Reliability Engineer with a focus on compute infrastructure to join our dynamic team. The ideal...2 days per week3 days per week
- ...on one unified cloud. One cloud for compute, inference, and agents. Role Overview We are seeking a skilled Site Reliability Engineer to join the GMI Global Infrastructure team. This role is hands-on and critical to ensuring the stability, efficiency, and...
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...Role: Site Reliability Engineer (SRE) Location: Brentwood, TN (Onsite) Contract Experience: 6-8+ years Role Description: Combines software engineering and IT operations to ensure the reliability, scalability, and performance of systems, with...Contract work
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$96.3k - $264.1k
...infrastructure and service, ensuring alignment with reliability and functionality standards. Takes full... ...tools and provides expertise in site reliability trends.Only Oracle brings... ...LeadershipDefine and drive the site reliability engineering strategy for large-scale, distributed,...Temporary workFlexible hours$140k - $210k
...Our Mission As the world’s number 1 job site*, our mission is to help people get jobs. We strive to cultivate an... ...Comscore, Total Visits, March 2026) Day to Day As an Engineering Manager in Site Reliability Engineering at Indeed, you will manage and grow a team that...Work experience placementLocal area- ...Last year our HCA Healthcare colleagues invested over 156,000 hours volunteering in our communities. As a Staff Site Reliability Engineer with HCA Healthcare you can be a part of an organization that is devoted to giving back! Job Summary and Qualifications Position...Temporary workFlexible hours
$169.3k - $304.7k
...in building and maintaining fast, efficient, scalable, and reliable routing software and infrastructure that is responsible... ...growth and stability of our global platform. As a Principal Site Reliability Engineer - Network, you will be responsible for: Architecting...Work experience placementWork at office- Position Summary Applied AI Site Reliability Engineer II Role Overview: As an Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity of high...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week
$96.8k - $306.4k
Seeking a senior staff-level engineer with expertise in datacenter platform firmware. This role requires working across product, infrastructure, security and platform teams both within and external to the company in a highly matrixed environment. You will be deeply hands...Temporary workFlexible hours$60 - $73 per hour
...partner for every meaningful moment of health. Within our Retail locations, we bring this promise to life with heart every day. Our Staff Pharmacists play a critical role in cultivating a culture of excellence in their pharmacy by acting as a role model for all, demonstrating...Hourly payFull timeTemporary workWork experience placementLocal areaFlexible hoursShift workNight shift$60 - $73 per hour
...partner for every meaningful moment of health. Within our Retail locations, we bring this promise to life with heart every day. Our Staff Pharmacists play a critical role in cultivating a culture of excellence in their pharmacy by acting as a role model for all, demonstrating...Hourly payFull timeTemporary workWork experience placementLocal areaFlexible hoursShift workNight shift$110.1k - $234.6k
This role combines ethical hacking, vulnerability research, and clean-room reverse-engineering practices to understand how systems work, identify security weaknesses, and help engineering teams remediate them responsibly along with future proofing. This role works closely...Temporary workFlexible hours$15 - $20 per hour
...Outside Golf Staff - Bounty Club Nashville, TN Position: Outside Golf Staff Department: Bounty Club Employment Status:... ...outdoors in various weather conditions Strong work ethic and reliability Positive attitude and professional appearance Weekend and...Hourly payFull timePart timeWork at officeWeekend work$146.3k - $306.4k
...conformance, and rapid incident mitigation. Influences silicon/board/firmware roadmaps to optimize for reliability, performance, and cost at hyperscale. Champions engineering excellence: coding standards, threat modeling, resource management, concurrency, and fault...Temporary workFlexible hoursShift work- ...performance, resolve production issues, and optimize automation reliability and efficiency. Develop technical documentation and... ...communications. Implement retrieval-augmented generation (RAG), prompt engineering, and AI orchestration techniques within UiPath ecosystems....Local areaRemote work
- ...of the wine country, made with local ingredients, and brought to your table fresh from our open-scratch kitchen. Our knowledgeable staff can help you pair each dish with the perfect glass. Because food tastes better with wine. BENEFITS: ~ FLEXIBLE SCHEDULES ~...Seasonal workLocal areaImmediate startFlexible hoursShift work
- ...full-time offer at LBMC upon the conclusion of their internship program. This role will give candidates the opportunity to work on-site with a TAS team that has 80+ years of combined Big 4 experience and $65B+ in completed work. Shadow due diligence engagements related...Full timeInternshipLocal area
- Staff Industrial Hygienist - Nashville, TennesseeIntertek, a leading... ...multiple architecture, engineering and construction disciplines,... ...partner you need to ensure the reliability, safety and performance of your... ..., UST/LUST assessments, site remediation and more. This person...For subcontractorWork at officeWorldwideMonday to FridayShift work
$92.5k - $209.5k
...solutions, edge computing, and more.As a Senior Platform Software Engineer, you will build the platform capabilities that make OCI... ...integration frameworks, or developer tooling, helping teams adopt reliable interfaces and evolve them without disrupting downstream consumers...Temporary workWork at officeWorldwideRelocationRelocation packageFlexible hours$92.5k - $209.5k
...next phase of growth.As a Senior Software Development Engineer, you will design and deliver reliable, secure, and delightful deployment capabilities that improve... ...services, control planes, or workflow orchestration.Site reliability engineering, production incident response,...Temporary workRelocationFlexible hours$92.5k - $209.5k
...execution at scale, with a strong focus on customer experience, reliability, and operational excellence.Only Oracle brings together the... ...Streaming, and observability services.Use modern AI-assisted engineering tools, including Codex and related developer tools, to...Temporary workRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- assistant engineer Nashville, TN
- senior staff systems engineer Nashville, TN
- technology administrator Nashville, TN
- engineering aide Nashville, TN
- staff engineer Nashville, TN
- site reliability engineer sre Nashville, TN
- site reliability engineer Nashville, TN
- official site Nashville, TN
- site services specialist Nashville, TN
- construction site safety Nashville, TN




