Site Reliability Engineer
Insight Global
As Fiber continues to scale across systems, Vendor partners, and our Fiber partner ecosystem, we need dedicated day-to-day technical operational support to help keep the platform stable and responsive to business needs.
This contractor will augment the Fiber Platform team by providing SRE-style production support coverage: monitoring platform health, triaging issues, documenting incidents, supporting root cause analysis, coordinating follow-up across IT, Vendor partners, Fiber partners, and internal teams, and helping resolve issues that impact customers, Sales, Care, and field teams during business hours.
What You'll Do
Platform Monitoring & Triage
• Monitor Fiber platform health, system availability, alerts, logs, and dashboards to identify issues quickly and support timely resolution.
• Provide day-to-day production support for Fiber platform issues, including initial triage, impact assessment, issue routing, and partner follow-up.
• Use logs, system data, dashboards, and operational signals to help identify root cause, quantify customer or order impact, and separate platform issues from partner or downstream system issues.
• Support real-time issue intake and feedback for Sales, field, Care, and Product teams when customer-facing or order-impacting problems arise during business hours.
Incident Management & Operational Support
• Document incidents, timelines, symptoms, owners, decisions, resolution steps, and follow-up actions in a clear and reusable format.
• Coordinate with IT, Vendor Partners, Fiber partners, QA, Product, and operations teams to drive issues toward resolution and ensure handoffs are clear.
• Maintain issue trackers, daily/weekly status updates, and operational reporting so the team has a reliable view of open risks, recurring issues, and resolution progress.
• Support post-incident reviews by identifying patterns, recurring points of failure, and opportunities to improve monitoring, support processes, or platform behavior.
Platform Operations Improvement
• Help maintain and improve operational documentation, support playbooks, escalation paths, and standard operating procedures for Fiber platform support.
• Identify gaps in telemetry, alerting, reporting, or runbooks and partner with the Fiber Platform SRE lead and technical teams to improve coverage.
• Assist with functional validation and production readiness activities for releases, break fixes, partner integrations, and new platform capabilities as needed.
• Help reduce single-thread risk by building enough platform knowledge to provide backup coverage and continuity when primary internal SRE support is unavailable.
Required Skills & Experience
• 5+ years of experience in application support, production support, SRE, DevOps, systems analysis, or technical operations roles.
• Experience monitoring and troubleshooting complex production systems using logs, dashboards, alerts, queries, and issue-management tools.
• Experience with Splunk and Jira.
• Strong incident triage and root-cause analysis skills, with the ability to quickly assess impact, identify likely failure points, and coordinate the right teams.
• Experience working across business, Product, IT, vendor, and operations teams to resolve production issues where ownership or root cause may not be immediately clear.
• Strong written communication skills, including the ability to document incidents, summarize technical findings, and provide concise status updates to non-technical stakeholders.
• Comfort operating in a fast-moving environment where processes and documentation are still maturing, and the ability to bring structure without creating unnecessary overhead.
Nice to Have Skills & Experience
• Experience in broadband, telecom, digital commerce, customer care, order management, billing, provisioning, or partner-integrated platforms.
• Experience with ServiceNow, SQL, API troubleshooting, data validation, or similar operational support tools.
• Experience supporting vendor-managed platforms or coordinating issue resolution with external technology partners.
• Experience working with sales, field, or care support teams on live customer-impacting issues.
Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.
$81.1k - $187k
...infrastructure and/or service according to terms for reliability and functionality.- Assists team members... ...deployments.- Gains basic knowledge of site reliability trends and shares relevant... ...are seeking a skilled Site Reliability Engineer to design, build, operate, and automate...SuggestedTemporary workImmediate startFlexible hoursShift work$135.8k - $183.8k
...dynamic and flexible work environment with competitive benefits and the ability to grow your career.We are looking for a Site Reliability Engineer to support our team responsible for building, managing, maintaining, deploying, and securing mission-critical services to...SuggestedWork at officeFlexible hours$119.8k - $234.7k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany:... ...opportunity for a Senior Site Reliability Engineer (SRE) to join the Azure Silver and Sovereign...SuggestedOngoing contractLocal area3 days per week- ...Site Reliability Engineer Location: Occasional onsite visits to Reston VA (Zip code 20190). Duration-1 year plus Interview process: The final interview is a mandatory, face-to-face interview in Reston VA. Zip code: 20190 Strong...SuggestedLong term contractTemporary workH1bImmediate startRelocation
$91.4k - $187k
Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible...SuggestedTemporary workFlexible hoursShift workWeekend work$146k - $194k
...focused on positioning Anduril as a lead provider of specialized engineering and products for Intelligence Community (IC) customers. We... ...pressing national security requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure the health,...Full timeWork experience placementImmediate startRemote work$109.18k - $163.77k
...channels. As a global company, we have offices in nine countries and can insert advertisements around the world.Job SummaryThe Site Reliability Engineering team is responsible for managing the critical infrastructure that powers FreeWheel's Streaming Hub platform. Streaming...Full time$150k - $180k
...redefining what's possible in remote sensing, you belong here at Umbra. About the Job We are seeking an experienced Senior Site Reliability Engineer to help design, build, operate, and scale the mission- and business-critical infrastructure that powers Umbra's systems....Permanent employmentWork at officeLocal areaRemote workWorldwideFlexible hours$121.5k - $264.1k
...sharing guidance on practices and terms for reliability and functionality.- Supervises team... ...developing and maintaining knowledge of site reliability trends and sharing valuable... ...Experience:9 years of experience in software engineering, infrastructure management, or related...Temporary workImmediate startFlexible hours$102k - $234.6k
...members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help... ...to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data.Only Oracle brings...Temporary workImmediate startFlexible hours$84.9k - $209.5k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and...Temporary workMonday to FridayFlexible hoursShift workNight shift- ...Overall requirement is a need for resources with "reliability engineering" experience for multi-region AWS workloads. The resource should have experience in optimizing AWS stack failover. Skills: # AWS experience: # Understanding of Infrastructure...
$128.83k - $193.25k
...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage... ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile...Full time$112.5k - $187.5k
...We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work...Full timeWork experience placementWork at officeFlexible hours2 days per week$80k - $133k
...degree, Four (4) years additional experience will be needed.Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud infrastructure and enterprise systems.One(1)+ years of experience deploying and...Permanent employmentFull timeContract workRemote workFlexible hours$62k - $141k
Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you have...Full timeContract workPart timeWork at officeLocal areaRemote work$87.1k - $157.45k
...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across... ...to come in and help us build systems that stay reliable when things get complicated.We need a Site Reliability Engineer who has experience building, deploying...Full timeWork from homeFlexible hours$133k - $190k
Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. Direct Hire Role **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is...Full time- ...Site Reliability Engineer Mc Lean, VA Long Term Client's Enterprise Data Machine Learning (EDML) employs innovative minds like yourself to design and develop software-systems that can meet the demand of our ever-growing customer base. Like a startup inside an enterprise...Immediate start
- ...Detail Description: The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You prepare and take ownership of "day two" operations...
$81.1k - $187k
...Infrastructure Engineer Takes proactive steps to design and architect infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity... ...potential impact and develops knowledge of site reliability trends. Key...Temporary workFlexible hours$118k - $177k
...Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation Continuous Diagnostics and Mitigation (CDM) Cyber data solution...Remote work- ...1 IV with tech SME and PM Period of performance: Up to 2 years in duration MUST HAVES: •Minimum of 8 years of experience as a Site Reliability Engineerwith a strong understanding of SRE principles for highly scalable and reliable systems •Possess a bachelor's degree •Experience...Local areaRelocation package3 days per week
$103.5k - $150k
...exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform...Temporary workWork experience placementLocal area3 days per week$117.2k - $176.7k
...specific level of U.S. government background investigation and clearance required for this role.Overview of the Role:Join our Site Reliability Engineering (SRE) team, where you'll work alongside Infrastructure and Research & Development (R&D) partners to keep Salesforce...Full timeWork experience placement$84.24k - $142.48k
OverviewJoin us to work collaboratively with our talented team of dynamic and passionate engineers to deliver capabilities that enable our customers to make a difference. You'll deploy and operate ArcGIS Velocity and ArcGIS Workflow Manager SaaS solutions. You will also...WorldwideFlexible hours- We're seeking a skilled and proactive Site Reliability Engineer to join our team, ensuring the stability, security, and efficiency of our technological resources as we deliver cutting-edge AI solutions to the government. This is a fully remote position for candidates in...Remote work
$180k - $230k
Navstar IT Services Position Would you like to perform rewarding work while contributing to the success of an established, growing company? Navstar is an award-winning organization that has a proven track record of successfully providing IT services and solutions both...For subcontractorLocal areaFlexible hours$130k - $200k
Summary Position Title: Site Reliability Engineer Position ID: TA247 Location(s): On-site; Aurora, CO; Herndon, VA Application Deadline: August 31, 2026 Security Clearance Requirement: TS/SCI Security Clearance with Polygraph Job Description Trusted Space...Full timeTemporary workLocal area- ...regulated cloud platform audit-ready. This is a hands-on senior engineering role on a FedRAMP-authorised platform: real architecture work... ...else's, this is it. WHAT YOU'LL OWN · Reliability and design decisions. Across multi-cluster Amazon EKS and AWS...Hourly payContract workShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- junior website developer Herndon, VA
- website content developer Herndon, VA
- site leader Herndon, VA
- site recruiter Herndon, VA
- on-site clinical research associate (traveling/remote) Herndon, VA
- official site Herndon, VA
- site safety Herndon, VA
- construction site safety Herndon, VA
- IT site lead Herndon, VA
- junior site reliability engineer


