Senior Site Reliability Engineer
Wesco
As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production-ready platforms through standardized provisioning, automation, testing, and infrastructure validation activities. Working as part of a holistic team strategy, you will support large, complex customer deployments and ensure infrastructure environments are ready for operational handoff and long-term success.
Responsibilities:
Provide technical expertise and engagement to support infrastructure readiness, platform engineering, and deployment activities across customer environments.
Deploy, configure, and validate AI, GPU, and High Performance Computing (HPC) infrastructure solutions.
Prepare and administer Kubernetes platforms, container runtimes, storage integrations, networking components, and cluster infrastructure.
Install, configure, and validate NVIDIA technologies including GPU drivers, CUDA, GPU Operators, AI Enterprise prerequisites, and telemetry solutions.
Validate accelerated networking technologies including InfiniBand, RoCE, RDMA, and GPU-to-GPU communications.
Perform infrastructure readiness assessments, burn-in testing, operational acceptance testing, and performance validation activities.
Configure and support server infrastructure including iDRAC, iLO, BMC, firmware, storage, and networking components.
Deploy and administer Windows, Linux, VMware ESXi, Hyper-V, and KVM-based environments.
Apply security hardening standards, compliance requirements, and operational best practices throughout deployment and validation activities.
Develop and maintain automation workflows utilizing PowerShell, Python, Bash, and Infrastructure-as-Code methodologies.
Create customer-facing deployment documentation, technical reports, readiness assessments, and operational validation deliverables.
Troubleshoot complex hardware, operating system, virtualization, containerization, networking, and AI platform issues.
Participate in advanced technical training and continued education to maintain expertise in cloud, infrastructure, AI, and platform technologies.
Support technical engagements across customer environments and collaborate with internal engineering, architecture, and service delivery teams.
Qualifications:
Associate degree (U.S.)/College Diploma (Canada) or equivalent combination of education and technical experience required.
Bachelor's degree in Computer Science, Information Technology, Engineering, or related technical discipline preferred.
5+ years of experience in Infrastructure Engineering, Platform Engineering, Site Reliability Engineering (SRE), Systems Administration, or related technical roles.
Experience deploying, supporting, or validating AI, GPU, HPC, or large-scale enterprise infrastructure environments.
Experience with Kubernetes, container platforms, and enterprise Linux administration.
Strong knowledge of server provisioning, virtualization, storage, networking, and infrastructure operations.
Experience with VMware ESXi, Hyper-V, KVM, or related virtualization technologies.
Experience developing automation and scripting solutions using PowerShell, Python, Bash, or similar tools.
Knowledge of Infrastructure-as-Code and automated deployment methodologies.
Experience with NVIDIA GPU technologies, CUDA, AI Enterprise, or related AI infrastructure platforms preferred.
Knowledge of InfiniBand, RDMA, RoCE, or high-performance networking technologies preferred.
Demonstrated troubleshooting, root-cause analysis, and problem-solving skills.
Possess a customer-centric mindset and strong written and verbal communication skills.
Possess intermediate computer skills, including proficiency with Microsoft Office applications.
Ability to travel up to 25%.
Preferred Certifications
Certified Kubernetes Administrator (CKA)
Red Hat Certified System Administrator (RHCSA) or equivalent Linux certification
NVIDIA certifications related to AI, GPU, or DGX platforms
VMware Certified Professional (VCP) or equivalent
#LI-VR1 #Hybrid
At Wesco, we build, connect, power and protect the world. As a leading provider of business-to-business distribution, logistics services and supply chain solutions, we create a world that you can depend on.
Our Company’s greatest asset is our people. Wesco is committed to fostering a workplace where every individual is respected, valued, and empowered to succeed. We promote a culture that is grounded in teamwork and respect. With a workforce of over 20,000 people worldwide, we embrace the unique perspectives each person brings. Through comprehensive benefits ( and active community engagement, we create an environment where every team member has the opportunity to thrive.
Learn more about Working at Wesco here ( and apply online today!
Founded in 1922 and headquartered in Pittsburgh, Wesco is a publicly traded (NYSE: WCC) FORTUNE 500® company.
Wesco International, Inc., including its subsidiaries and affiliates (“Wesco”) provides equal employment opportunities to all employees and applicants for employment. Employment decisions are made without regard to race, religion, color, national or ethnic origin, sex, sexual orientation, gender identity or expression, age, disability, or other characteristics protected by law. US applicants only, we are an Equal Opportunity Employer.
Los Angeles Unincorporated County Candidates Only: Qualified applicants with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance and the California Fair Chance Act.
This posting is for a current, active vacancy intended for immediate hire.
$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior$262k - $364k
...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely with senior technical leads in the development teams.... ...:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software...Senior$104.9k - $174.7k
...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory...SeniorFull timeWork at officeLocal areaRemote workWork from home$136.2k - $214.01k
...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to...SeniorFull timeFlexible hours- ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines... ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the...Senior
- ...Senior Site Reliability Engineer Come join a growing bank at the heart of the innovation, technology, green tech and life sciences space. We continue to expand our global footprint and our banking technology is at the core of everything we do. As a Senior Site Reliability...Senior
$125.7k - $203.1k
...collaborative team of systems and cloud engineers who thrive in a fast-paced environment built... ...maintain efficient platform uptime and reliability. • Lead end-to-end incident response and... ...steps.• Collaborate closely with Site Reliability Engineering (SRE) and Product...Permanent employmentFull timeTemporary workApprenticeshipWork experience placementLocal areaWorldwideFlexible hoursNight shift- ...Senior Site Reliability Engineer (Permanent Role) Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities . Monitoring distribution systems and notifying them of any potential issues. . Assisting with troubleshooting on call....Permanent employmentTemporary workLocal areaFlexible hoursShift workWeekend work
- ...ensure applications are highly available, reliable, and performant at a global scale.... ...Bachelor of Computer Science or related Engineering field required. Master's Degree preferred... ...Minimum of 1 year of lead experience of site reliability engineering team required....Contract workWork at office
- ...Job Title: Site Reliability Engineer Location: Dallas TX (HYBRID) Duration :Full Time Job Description: Skill: Site Reliability Engineer • Ensures supported applications are functioning and available by minimizing downtime and maximizing performance...Full timeWork at office
- ...Site Reliability Engineer Location- Wilmington De, Washington DC, Dallas, TX (Onsite Position) Full time position Minimum Qualifications Bachelor’s degree in computer science, Engineering, or a related technical field. Minimum of 5 years of experience...Full time
- ...Site Reliability Engineer We are looking for a Site Reliability Engineer for our client location in Dallas TX with the following skills: Java Spring boot, Kubernetes, and eCommerce experience required. Key responsibilities include working with the applications, engineering...Work at office
- ...improving platform infrastructure and applications with high reliability, resiliency, performance & quality, and faster time-to-market... ...documentation, including runbooks/playbooks; and, Using Chaos Engineering to test the robustness of the systems and applications....
- ...Role: Site Reliability Engineer 6+ months Contract role Remote About the Role We are looking for a dynamic and accomplished Site Reliability Engineer (SRE) who excels at solving complex reliability challenges and thrives in high-impact environments....Contract workRemote work
- ...Job Position:- Site Reliability Engineer Duration:- Long Term Client:- UPS This is a Hybrid Work Model (3x a week Onsite) and Location is Parsippany, NJ. Job Description: We are looking for a talented Site Reliability Engineer...
- ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through... ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform... ...vendor resources. Willingness to work on-site at stated location in the job opening....For contractorsWork experience placement
- Mandatory Skills: AWS/Azure/GCP (GCP is not used very much ). Kubernetes /Helm,Docker,Gitlab,Grafana,Cyberark/Hashicorp Vault, Terraform etc. Experience utilizing Java, Perl, Python, Go and scripting experience in Shell and Perl to automate reports and monitor enterprise...
- ...Senior Site Reliability Engineer Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities: Monitoring distribution systems and notifying them of any potential issues. Assisting with troubleshooting on call. Managing and tracking...Flexible hoursShift workWeekend work
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...Work at office
- ...will play a strategic role in shaping GM Financials' release engineering and software delivery practices. You'll collaborate with engineering... ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages...Full time
$136.88k - $200.75k
...Make good. Please note that we do not offer visa sponsorship for this position. Role Summary The Senior Cloud Platform & Site Reliability Engineering Lead partners with business and technical stakeholders to lead cloud platform design, engineering, and...SeniorHourly payFull timeWork experience placementWork at officeFlexible hours- ...AI stacks, including frameworks for building autonomous agents, planning, memory, and tool use. Hands-on experience with prompt engineering, model fine-tuning, and deployment of Generative AI applications. Hands on experience with GitHub, Confluence, JIRA, Jenkins, CI/...Senior
- ...NTT DATA Services is seeking a Senior Principal Cloud Architect - AWS & Service Activation to lead design, build, and operation of... ...across cloud environments while ensuring cost efficiency and reliability. The role is hybrid, based in Irving, TX or Charlotte, NC,...Senior
- ...NTT DATA Services seeks a Senior Principal Cloud Architect to lead design, build, and operation of a secure, automated, multi-cloud... ...drive governance, cost optimization, and security while mentoring engineers, with a hybrid work model in Irving, TX or Charlotte, NC and...Senior
$45 - $50 per hour
DescriptionKforce has a client in Dallas, TX that is seeking a Senior Software Engineer.Operational Support & Environment Management:* Support Development, UAT, and Production environments* Monitor application health, system performance, and operational dashboards* Troubleshoot...Senior- ...Seeking a Senior Software Engineer with 5+ years of experience supporting enterprise applications and operational environments, with strong expertise in .NET, AWS, SQL, production support, deployments, and troubleshooting . Roles and Responsibilities Support Development...SeniorContract work
- ...apply now.We are currently seeking a Senior Cloud Platform Engineer (VMware Cloud Foundation) - Hybrid in... ...integration of core IP services and reliable guest OS performance.High Availability... ...locally to NTT DATA offices or client sites. This ensures we can provide timely...SeniorFull timeTemporary workWork at officeRemote workFlexible hours
- Roles and Responsibilities Design, develop, and maintain enterprise applications using Python. Build and deploy document management and document capture applications incorporating OCR and Deep Learning capabilities. Develop and maintain Machine Learning models...SeniorContract work
- Lead by example, driving engineering excellence across the team while actively mentoring and developing technical talent.Partner closely with Engineering Managers, Product Managers, and Technical Program Managers (T/PgMs) to define, refine, and execute the team’s goal,...Senior
$174k - $252k
...business projects.1 year of experience in a technical leadership role.Experience developing accessible technologiesGoogle's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Dallas, TX
- site reliability engineer sre Dallas, TX
- senior living director Dallas, TX
- senior manager customer operations Dallas, TX
- senior support engineer Dallas, TX
- senior java developer Dallas, TX
- senior software engineer ruby on rails Dallas, TX
- sr finance manager Dallas, TX
- sr marketing manager Dallas, TX
- senior customer service Dallas, TX




