Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

Wesco

As the Senior Site Reliability Engineer, you will serve as a trusted technical resource responsible for deploying, validating, and operationalizing AI, HPC, Kubernetes, and enterprise infrastructure environments. This role transforms newly installed hardware into production-ready platforms through standardized provisioning, automation, testing, and infrastructure validation activities. Working as part of a holistic team strategy, you will support large, complex customer deployments and ensure infrastructure environments are ready for operational handoff and long-term success.

Responsibilities:

  • Provide technical expertise and engagement to support infrastructure readiness, platform engineering, and deployment activities across customer environments.

  • Deploy, configure, and validate AI, GPU, and High Performance Computing (HPC) infrastructure solutions.

  • Prepare and administer Kubernetes platforms, container runtimes, storage integrations, networking components, and cluster infrastructure.

  • Install, configure, and validate NVIDIA technologies including GPU drivers, CUDA, GPU Operators, AI Enterprise prerequisites, and telemetry solutions.

  • Validate accelerated networking technologies including InfiniBand, RoCE, RDMA, and GPU-to-GPU communications.

  • Perform infrastructure readiness assessments, burn-in testing, operational acceptance testing, and performance validation activities.

  • Configure and support server infrastructure including iDRAC, iLO, BMC, firmware, storage, and networking components.

  • Deploy and administer Windows, Linux, VMware ESXi, Hyper-V, and KVM-based environments.

  • Apply security hardening standards, compliance requirements, and operational best practices throughout deployment and validation activities.

  • Develop and maintain automation workflows utilizing PowerShell, Python, Bash, and Infrastructure-as-Code methodologies.

  • Create customer-facing deployment documentation, technical reports, readiness assessments, and operational validation deliverables.

  • Troubleshoot complex hardware, operating system, virtualization, containerization, networking, and AI platform issues.

  • Participate in advanced technical training and continued education to maintain expertise in cloud, infrastructure, AI, and platform technologies.

  • Support technical engagements across customer environments and collaborate with internal engineering, architecture, and service delivery teams.

Qualifications:

  • Associate degree (U.S.)/College Diploma (Canada) or equivalent combination of education and technical experience required.

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or related technical discipline preferred.

  • 5+ years of experience in Infrastructure Engineering, Platform Engineering, Site Reliability Engineering (SRE), Systems Administration, or related technical roles.

  • Experience deploying, supporting, or validating AI, GPU, HPC, or large-scale enterprise infrastructure environments.

  • Experience with Kubernetes, container platforms, and enterprise Linux administration.

  • Strong knowledge of server provisioning, virtualization, storage, networking, and infrastructure operations.

  • Experience with VMware ESXi, Hyper-V, KVM, or related virtualization technologies.

  • Experience developing automation and scripting solutions using PowerShell, Python, Bash, or similar tools.

  • Knowledge of Infrastructure-as-Code and automated deployment methodologies.

  • Experience with NVIDIA GPU technologies, CUDA, AI Enterprise, or related AI infrastructure platforms preferred.

  • Knowledge of InfiniBand, RDMA, RoCE, or high-performance networking technologies preferred.

  • Demonstrated troubleshooting, root-cause analysis, and problem-solving skills.

  • Possess a customer-centric mindset and strong written and verbal communication skills.

  • Possess intermediate computer skills, including proficiency with Microsoft Office applications.

  • Ability to travel up to 25%.

Preferred Certifications

  • Certified Kubernetes Administrator (CKA)

  • Red Hat Certified System Administrator (RHCSA) or equivalent Linux certification

  • NVIDIA certifications related to AI, GPU, or DGX platforms

  • VMware Certified Professional (VCP) or equivalent

#LI-VR1 #Hybrid

At Wesco, we build, connect, power and protect the world. As a leading provider of business-to-business distribution, logistics services and supply chain solutions, we create a world that you can depend on. ​

Our Company’s greatest asset is our people. Wesco is committed to fostering a workplace where every individual is respected, valued, and empowered to succeed. We promote a culture that is grounded in teamwork and respect. With a workforce of over 20,000 people worldwide, we embrace the unique perspectives each person brings. Through comprehensive benefits ( and active community engagement, we create an environment where every team member has the opportunity to thrive. ​

Learn more about Working at Wesco here ( and apply online today!​

Founded in 1922 and headquartered in Pittsburgh, Wesco is a publicly traded (NYSE: WCC) FORTUNE 500® company.​

Wesco International, Inc., including its subsidiaries and affiliates (“Wesco”) provides equal employment opportunities to all employees and applicants for employment. Employment decisions are made without regard to race, religion, color, national or ethnic origin, sex, sexual orientation, gender identity or expression, age, disability, or other characteristics protected by law. US applicants only, we are an Equal Opportunity Employer.​

Los Angeles Unincorporated County Candidates Only: Qualified applicants with arrest or conviction records will be considered for employment in accordance with the Los Angeles County Fair Chance Ordinance and the California Fair Chance Act.

This posting is for a current, active vacancy intended for immediate hire.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Dallas, TX vacancy
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Senior

    Google

    Sunnyvale, TX
    15 hours ago
  • $262k - $364k

     ...infrastructure from SRE side, ensuring it is reliable, scalable, cost effective and performant, while working closely with senior technical leads in the development teams....  ...:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software... 
    Senior

    Google

    Sunnyvale, TX
    1 day ago
  • $104.9k - $174.7k

     ...Data Management. You can learn more about LexisNexis Risk at the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Senior
    Full time
    Work at office
    Local area
    Remote work
    Work from home

    RELX Group

    Dallas, TX
    3 days ago
  • $136.2k - $214.01k

     ...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to... 
    Senior
    Full time
    Flexible hours

    Proofpoint

    Dallas, TX
    4 days ago
  •  ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines...  ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the... 
    Senior

    Outshift by Cisco

    Richardson, TX
    4 days ago
  •  ...Senior Site Reliability Engineer Come join a growing bank at the heart of the innovation, technology, green tech and life sciences space. We continue to expand our global footprint and our banking technology is at the core of everything we do. As a Senior Site Reliability... 
    Senior

    Professional Recruiters

    Dallas, TX
    3 days ago
  • $125.7k - $203.1k

     ...collaborative team of systems and cloud engineers who thrive in a fast-paced environment built...  ...maintain efficient platform uptime and reliability. • Lead end-to-end incident response and...  ...steps.• Collaborate closely with Site Reliability Engineering (SRE) and Product... 
    Permanent employment
    Full time
    Temporary work
    Apprenticeship
    Work experience placement
    Local area
    Worldwide
    Flexible hours
    Night shift

    CISCO Systems

    Richardson, TX
    2 days ago
  •  ...Senior Site Reliability Engineer (Permanent Role) Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities . Monitoring distribution systems and notifying them of any potential issues. . Assisting with troubleshooting on call.... 
    Permanent employment
    Temporary work
    Local area
    Flexible hours
    Shift work
    Weekend work

    System One

    Dallas, TX
    a month ago
  •  ...ensure applications are highly available, reliable, and performant at a global scale....  ...Bachelor of Computer Science or related Engineering field required. Master's Degree preferred...  ...Minimum of 1 year of lead experience of site reliability engineering team required.... 
    Contract work
    Work at office

    3B Staffing LLC

    Irving, TX
    2 days ago
  •  ...Job Title: Site Reliability Engineer Location: Dallas TX (HYBRID) Duration :Full Time Job Description: Skill: Site Reliability Engineer • Ensures supported applications are functioning and available by minimizing downtime and maximizing performance... 
    Full time
    Work at office

    Syntricate Technologies

    Dallas, TX
    2 days ago
  •  ...Site Reliability Engineer Location- Wilmington De, Washington DC, Dallas, TX (Onsite Position) Full time position Minimum Qualifications Bachelor’s degree in computer science, Engineering, or a related technical field. Minimum of 5 years of experience... 
    Full time

    Yochana

    Dallas, TX
    20 hours ago
  •  ...Site Reliability Engineer We are looking for a Site Reliability Engineer for our client location in Dallas TX with the following skills: Java Spring boot, Kubernetes, and eCommerce experience required. Key responsibilities include working with the applications, engineering... 
    Work at office

    STIAOS Technologies

    Dallas, TX
    4 days ago
  •  ...improving platform infrastructure and applications with high reliability, resiliency, performance & quality, and faster time-to-market...  ...documentation, including runbooks/playbooks; and, Using Chaos Engineering to test the robustness of the systems and applications.... 

    Software Technology Inc

    Dallas, TX
    3 days ago
  •  ...Role: Site Reliability Engineer 6+ months Contract role Remote About the Role We are looking for a dynamic and accomplished Site Reliability Engineer (SRE) who excels at solving complex reliability challenges and thrives in high-impact environments.... 
    Contract work
    Remote work

    ECHO IT SOLUTIONS INC .

    Farmers Branch, TX
    3 days ago
  •  ...Job Position:- Site Reliability Engineer Duration:- Long Term Client:- UPS This is a Hybrid Work Model (3x a week Onsite) and Location is Parsippany, NJ. Job Description: We are looking for a talented Site Reliability Engineer... 

    Sparktek

    Farmers Branch, TX
    3 days ago
  •  ...Qualifications: 8+ years of software engineering experience, or equivalent demonstrated through...  ...implement and maintain scalable and reliable infrastructure on Google Cloud Platform...  ...vendor resources. Willingness to work on-site at stated location in the job opening.... 
    For contractors
    Work experience placement

    Cedent Life Talent

    Dallas, TX
    3 days ago
  • Mandatory Skills: AWS/Azure/GCP (GCP is not used very much ). Kubernetes /Helm,Docker,Gitlab,Grafana,Cyberark/Hashicorp Vault, Terraform etc. Experience utilizing Java, Perl, Python, Go and scripting experience in Shell and Perl to automate reports and monitor enterprise...

    Omni Inclusive

    Dallas, TX
    1 day ago
  •  ...Senior Site Reliability Engineer Cleveland, OH, Pittsburgh, PA, or Dallas, TX Your future duties and responsibilities: Monitoring distribution systems and notifying them of any potential issues. Assisting with troubleshooting on call. Managing and tracking... 
    Flexible hours
    Shift work
    Weekend work

    System One Holdings, LLC

    Dallas, TX
    20 hours ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex... 
    Work at office

    J.P. Morgan

    Dallas, TX
    1 day ago
  •  ...will play a strategic role in shaping GM Financials' release engineering and software delivery practices. You'll collaborate with engineering...  ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages... 
    Full time

    GMAC Financial Services

    Irving, TX
    4 days ago
  • $136.88k - $200.75k

     ...Make good. Please note that we do not offer visa sponsorship for this position. Role Summary The Senior Cloud Platform & Site Reliability Engineering Lead partners with business and technical stakeholders to lead cloud platform design, engineering, and... 
    Senior
    Hourly pay
    Full time
    Work experience placement
    Work at office
    Flexible hours

    National Life Insurance Company

    Addison, TX
    2 days ago
  •  ...AI stacks, including frameworks for building autonomous agents, planning, memory, and tool use. Hands-on experience with prompt engineering, model fine-tuning, and deployment of Generative AI applications. Hands on experience with GitHub, Confluence, JIRA, Jenkins, CI/... 
    Senior

    Virtusa

    Irving, TX
    4 days ago
  •  ...NTT DATA Services is seeking a Senior Principal Cloud Architect - AWS & Service Activation to lead design, build, and operation of...  ...across cloud environments while ensuring cost efficiency and reliability. The role is hybrid, based in Irving, TX or Charlotte, NC,... 
    Senior

    Jobleads-US

    Irving, TX
    2 days ago
  •  ...NTT DATA Services seeks a Senior Principal Cloud Architect to lead design, build, and operation of a secure, automated, multi-cloud...  ...drive governance, cost optimization, and security while mentoring engineers, with a hybrid work model in Irving, TX or Charlotte, NC and... 
    Senior

    Jobleads-US

    Irving, TX
    20 hours ago
  • $45 - $50 per hour

    DescriptionKforce has a client in Dallas, TX that is seeking a Senior Software Engineer.Operational Support & Environment Management:* Support Development, UAT, and Production environments* Monitor application health, system performance, and operational dashboards* Troubleshoot... 
    Senior

    KForce

    Dallas, TX
    15 hours ago
  •  ...Seeking a Senior Software Engineer with 5+ years of experience supporting enterprise applications and operational environments, with strong expertise in .NET, AWS, SQL, production support, deployments, and troubleshooting . Roles and Responsibilities Support Development... 
    Senior
    Contract work

    Neshent Technologies

    Dallas, TX
    10 days ago
  •  ...apply now.We are currently seeking a Senior Cloud Platform Engineer (VMware Cloud Foundation) - Hybrid in...  ...integration of core IP services and reliable guest OS performance.High Availability...  ...locally to NTT DATA offices or client sites. This ensures we can provide timely... 
    Senior
    Full time
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT DATA

    Irving, TX
    3 days ago
  • Roles and Responsibilities Design, develop, and maintain enterprise applications using Python. Build and deploy document management and document capture applications incorporating OCR and Deep Learning capabilities. Develop and maintain Machine Learning models...
    Senior
    Contract work

    2T Consulting

    Addison, TX
    24 days ago
  • Lead by example, driving engineering excellence across the team while actively mentoring and developing technical talent.Partner closely with Engineering Managers, Product Managers, and Technical Program Managers (T/PgMs) to define, refine, and execute the team’s goal,... 
    Senior

    Google

    Sunnyvale, TX
    1 day ago
  • $174k - $252k

     ...business projects.1 year of experience in a technical leadership role.Experience developing accessible technologiesGoogle's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one... 
    Senior

    Google

    Sunnyvale, TX
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!