Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Insight Global

As Fiber continues to scale across systems, Vendor partners, and our Fiber partner ecosystem, we need dedicated day-to-day technical operational support to help keep the platform stable and responsive to business needs.
This contractor will augment the Fiber Platform team by providing SRE-style production support coverage: monitoring platform health, triaging issues, documenting incidents, supporting root cause analysis, coordinating follow-up across IT, Vendor partners, Fiber partners, and internal teams, and helping resolve issues that impact customers, Sales, Care, and field teams during business hours.
What You'll Do
Platform Monitoring & Triage
• Monitor Fiber platform health, system availability, alerts, logs, and dashboards to identify issues quickly and support timely resolution.
• Provide day-to-day production support for Fiber platform issues, including initial triage, impact assessment, issue routing, and partner follow-up.
• Use logs, system data, dashboards, and operational signals to help identify root cause, quantify customer or order impact, and separate platform issues from partner or downstream system issues.
• Support real-time issue intake and feedback for Sales, field, Care, and Product teams when customer-facing or order-impacting problems arise during business hours.
Incident Management & Operational Support
• Document incidents, timelines, symptoms, owners, decisions, resolution steps, and follow-up actions in a clear and reusable format.
• Coordinate with IT, Vendor Partners, Fiber partners, QA, Product, and operations teams to drive issues toward resolution and ensure handoffs are clear.
• Maintain issue trackers, daily/weekly status updates, and operational reporting so the team has a reliable view of open risks, recurring issues, and resolution progress.
• Support post-incident reviews by identifying patterns, recurring points of failure, and opportunities to improve monitoring, support processes, or platform behavior.
Platform Operations Improvement
• Help maintain and improve operational documentation, support playbooks, escalation paths, and standard operating procedures for Fiber platform support.
• Identify gaps in telemetry, alerting, reporting, or runbooks and partner with the Fiber Platform SRE lead and technical teams to improve coverage.
• Assist with functional validation and production readiness activities for releases, break fixes, partner integrations, and new platform capabilities as needed.
• Help reduce single-thread risk by building enough platform knowledge to provide backup coverage and continuity when primary internal SRE support is unavailable.

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to View email address on click.appcast.io learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy:


Required Skills & Experience
• 5+ years of experience in application support, production support, SRE, DevOps, systems analysis, or technical operations roles.
• Experience monitoring and troubleshooting complex production systems using logs, dashboards, alerts, queries, and issue-management tools.
• Experience with Splunk and Jira.
• Strong incident triage and root-cause analysis skills, with the ability to quickly assess impact, identify likely failure points, and coordinate the right teams.
• Experience working across business, Product, IT, vendor, and operations teams to resolve production issues where ownership or root cause may not be immediately clear.
• Strong written communication skills, including the ability to document incidents, summarize technical findings, and provide concise status updates to non-technical stakeholders.
• Comfort operating in a fast-moving environment where processes and documentation are still maturing, and the ability to bring structure without creating unnecessary overhead.


Nice to Have Skills & Experience
• Experience in broadband, telecom, digital commerce, customer care, order management, billing, provisioning, or partner-integrated platforms.
• Experience with ServiceNow, SQL, API troubleshooting, data validation, or similar operational support tools.
• Experience supporting vendor-managed platforms or coordinating issue resolution with external technology partners.
• Experience working with sales, field, or care support teams on live customer-impacting issues.


Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Herndon, VA vacancy
  • $81.1k - $187k

     ...infrastructure and/or service according to terms for reliability and functionality.- Assists team members...  ...deployments.- Gains basic knowledge of site reliability trends and shares relevant...  ...are seeking a skilled Site Reliability Engineer to design, build, operate, and automate... 
    Suggested
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle Corporation

    Reston, VA
    3 days ago
  • $135.8k - $183.8k

     ...dynamic and flexible work environment with competitive benefits and the ability to grow your career.We are looking for a Site Reliability Engineer to support our team responsible for building, managing, maintaining, deploying, and securing mission-critical services to... 
    Suggested
    Work at office
    Flexible hours

    VeriSign

    Reston, VA
    4 days ago
  • $119.8k - $234.7k

     ...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual...  ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign... 
    Suggested
    Ongoing contract
    Local area
    3 days per week

    Microsoft

    Reston, VA
    1 day ago
  •  ...Site Reliability Engineer Location: Occasional onsite visits to Reston VA (Zip code 20190). Duration-1 year plus Interview process: The final interview is a mandatory, face-to-face interview in Reston VA. Zip code: 20190 Strong... 
    Suggested
    Long term contract
    Temporary work
    H1b
    Immediate start
    Relocation

    3B Staffing LLC

    Reston, VA
    4 days ago
  • $91.4k - $187k

    Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible... 
    Suggested
    Temporary work
    Flexible hours
    Shift work
    Weekend work

    Oracle Corporation

    Reston, VA
    9 hours ago
  • $146k - $194k

     ...focused on positioning Anduril as a lead provider of specialized engineering and products for Intelligence Community (IC) customers. We...  ...pressing national security requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure the health,... 
    Full time
    Work experience placement
    Immediate start
    Remote work

    Anduril Industries

    Reston, VA
    3 days ago
  • $109.18k - $163.77k

     ...channels. As a global company, we have offices in nine countries and can insert advertisements around the world.Job SummaryThe Site Reliability Engineering team is responsible for managing the critical infrastructure that powers FreeWheel's Streaming Hub platform. Streaming... 
    Full time

    Comcast

    Reston, VA
    9 hours ago
  • $150k - $180k

     ...redefining what's possible in remote sensing, you belong here at Umbra. About the Job We are seeking an experienced Senior Site Reliability Engineer to help design, build, operate, and scale the mission- and business-critical infrastructure that powers Umbra's systems.... 
    Permanent employment
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    Umbra

    Reston, VA
    5 days ago
  • $121.5k - $264.1k

     ...sharing guidance on practices and terms for reliability and functionality.- Supervises team...  ...developing and maintaining knowledge of site reliability trends and sharing valuable...  ...Experience:9 years of experience in software engineering, infrastructure management, or related... 
    Temporary work
    Immediate start
    Flexible hours

    Oracle Corporation

    Reston, VA
    3 days ago
  • $102k - $234.6k

     ...members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help...  ...to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data.Only Oracle brings... 
    Temporary work
    Immediate start
    Flexible hours

    Oracle Corporation

    Reston, VA
    1 day ago
  • $84.9k - $209.5k

     .... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers...  ...posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and... 
    Temporary work
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle Corporation

    Reston, VA
    1 day ago
  •  ...Overall requirement is a need for resources with "reliability engineering" experience for multi-region AWS workloads. The resource should have experience in optimizing AWS stack failover. Skills: # AWS experience: # Understanding of Infrastructure... 

    Infinite Computer Solutions

    Reston, VA
    3 days ago
  • $128.83k - $193.25k

     ...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage...  ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile... 
    Full time

    Comcast

    Reston, VA
    9 hours ago
  • $112.5k - $187.5k

     ...We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work... 
    Full time
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    TransUnion

    Reston, VA
    2 days ago
  • $80k - $133k

     ...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and... 
    Permanent employment
    Full time
    Contract work
    Remote work
    Flexible hours

    Guidehouse

    McLean, VA
    2 days ago
  • $62k - $141k

    Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you have... 
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Chantilly, Loudoun County, VA
    4 days ago
  • $87.1k - $157.45k

     ...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across...  ...to come in and help us build systems that stay reliable when things get complicated.We need a Site Reliability Engineer who has experience building, deploying... 
    Full time
    Work from home
    Flexible hours

    Leidos

    Chantilly, Loudoun County, VA
    9 hours ago
  • $133k - $190k

    Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is... 
    Full time

    SOC Support Services

    McLean, VA
    3 days ago
  •  ...Site Reliability Engineer Mc Lean, VA Long Term Client's Enterprise Data Machine Learning (EDML) employs innovative minds like yourself to design and develop software-systems that can meet the demand of our ever-growing customer base. Like a startup inside an enterprise... 
    Immediate start

    Maintec Technologies

    McLean, VA
    2 days ago
  •  ...Detail Description: The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You prepare and take ownership of "day two" operations... 

    InstantServe LLC

    Vienna, VA
    21 hours ago
  • $81.1k - $187k

     ...Infrastructure Engineer Takes proactive steps to design and architect infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity...  ...potential impact and develops knowledge of site reliability trends. Key... 
    Temporary work
    Flexible hours

    Oracle

    Vienna, VA
    1 day ago
  • $118k - $177k

     ...Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation Continuous Diagnostics and Mitigation (CDM) Cyber data solution... 
    Remote work

    ECS

    Fairfax, VA
    3 days ago
  •  ...1 IV with tech SME and PM Period of performance: Up to 2 years in duration MUST HAVES: •Minimum of 8 years of experience as a Site Reliability Engineerwith a strong understanding of SRE principles for highly scalable and reliable systems •Possess a bachelor's degree •Experience... 
    Local area
    Relocation package
    3 days per week

    Beyond SOF

    Vienna, VA
    2 days ago
  • $103.5k - $150k

     ...exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform... 
    Temporary work
    Work experience placement
    Local area
    3 days per week

    Medallia

    McLean, VA
    20 hours ago
  • $117.2k - $176.7k

     ...specific level of U.S. government background investigation and clearance required for this role.Overview of the Role:Join our Site Reliability Engineering (SRE) team, where you'll work alongside Infrastructure and Research & Development (R&D) partners to keep Salesforce... 
    Full time
    Work experience placement

    Salesforce

    Herndon, VA
    9 hours ago
  • $84.24k - $142.48k

    OverviewJoin us to work collaboratively with our talented team of dynamic and passionate engineers to deliver capabilities that enable our customers to make a difference. You'll deploy and operate ArcGIS Velocity and ArcGIS Workflow Manager SaaS solutions. You will also... 
    Worldwide
    Flexible hours

    ESRI

    Vienna, VA
    9 hours ago
  • We're seeking a skilled and proactive Site Reliability Engineer to join our team, ensuring the stability, security, and efficiency of our technological resources as we deliver cutting-edge AI solutions to the government. This is a fully remote position for candidates in... 
    Remote work

    Knexus

    Vienna, VA
    9 hours ago
  • $180k - $230k

    Navstar IT Services Position Would you like to perform rewarding work while contributing to the success of an established, growing company? Navstar is an award-winning organization that has a proven track record of successfully providing IT services and solutions both...
    For subcontractor
    Local area
    Flexible hours

    Navstar

    Chantilly, Loudoun County, VA
    22 hours ago
  • $130k - $200k

    Summary Position Title: Site Reliability Engineer Position ID: TA247 Location(s): On-site; Aurora, CO; Herndon, VA Application Deadline: August 31, 2026 Security Clearance Requirement: TS/SCI Security Clearance with Polygraph Job Description Trusted Space... 
    Full time
    Temporary work
    Local area

    Trusted Space, LLC

    Herndon, VA
    9 hours ago
  •  ...regulated cloud platform audit-ready. This is a hands-on senior engineering role on a FedRAMP-authorised platform: real architecture work...  ...else's, this is it. WHAT YOU'LL OWN ·        Reliability and design decisions.  Across multi-cluster Amazon EKS and AWS... 
    Hourly pay
    Contract work
    Shift work

    C-Serv

    Reston, VA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!