Data Site Reliability Engineer (SRE)
$111.16k - $150.39kGdit
Type of Requisition: Regular
Clearance Level Must Currently Possess:
NoneClearance Level Must Be Able to Obtain:
NonePublic Trust/Other Required:
BI Full 6C (T4)Job Family:
IT Infrastructure and OperationsJob Qualifications:
Skills:
CI/CD, Containerization, Structured Query Language (SQL) DevelopmentCertifications:
NoneExperience:
5 + years of related experienceUS Citizenship Required:
NoJob Description:
Seize your opportunity to make a personal impact supporting the Case Management Modernization (CMM) Program. The CMM program is an initiative to support the Administrative Office of the US Courts (AO) in developing a modern cloud-based solution to support all 204+ federal courts across the United States.
GDIT is your place to make meaningful contributions to challenging projects and grow a rewarding career. The Data Site Reliability Engineer (SRE) will work as part of the CMM Data Modernization and Governance team responsible for delivering integrated data governance, engineering, data platform, reporting, analytics, and Artificial Intelligence (AI)/Machine Learning (ML) capabilities that support operational decision-making and fulfill the AO's data and analytics objectives in support of the CMM program.
The successful candidate will be responsible for providing technical leadership for the day-to-day operational support, reliability, performance, and continuous improvement of the CMM data platforms, pipelines, applications, and analytics services. This role ensures that data services remain secure, available, reliable, and aligned with established service levels, data governance standards, architecture principles, and operational procedures.
THE DATA SITE RELIABILITY ENGINEER (SRE) WILL EXECUTE THE FOLLOWING RESPONSIBILITIES
Provide comprehensive real-time monitoring, incident and event management, capacity planning, and operational reporting to support application deployments, maintain system health, predict demand, and align cloud operations with evolving business and security objectives.
Maintain and audit user roles and responsibilities in cloud environments.
Integrate Single Sign On (SSO), Multi-Factor Authentication (MFA) and group identity management managed through the Judiciary Enterprise Network Information Exchange (JENIE) for enforcing least privilege access.
Adhere to guidelines prescribed by the Government and continuously assess and improve credential management processes for all user credentials.
Provide Disaster Recovery (DR) and Continuity of Operations (COOP) options. This must include high-availability options, including fault-tolerant and automated failover designs.
Integrate DevSecOps tools and processes seamlessly with enterprise systems (Integrated Development Environments (IDEs), ticketing, monitoring, etc.) to avoid fragmentation and ensure unified security posture.
Provide and manage a centralized secrets management system with automated rotation, access logging, and policy enforcement to securely store, manage, and control access to sensitive information and to prevent unauthorized access and data breaches for any administrative user account.
Integrate security tools (example: SAST, DAST, SCA, CSPM) into pipelines for continuous assessment and remediation.
Implement unified, automated, continuous monitoring (24/7/365) systems and tools for security, performance, and compliance across all environments, leveraging dashboards and alerting for real-time visibility. Provide supplemental monitoring of event response activities beyond normal business hours (7a.m – 6p.m Eastern Time). Systems and tools shall capture data without including a required response to alerts.
Ensure automated generation and management of Software Bill of Materials (SBOM) for all deployed artifacts, supporting transparency and compliance.
Provide diagnostics, metrics’ gathering, and performance tuning services.
Provide canary release function for end-user testing to support beta testing.
Configure an alert mechanism so that the support teams can react in an instance of unusual behavior.
Implement and operate a comprehensive incident and event management process, including integration with enterprise SIEM solutions, automated alerting, escalation workflows, and root cause analysis for all critical incidents.
Provide engineering support to ensure prompt detection, logging, diagnosis, escalation, and resolution of incidents to restore normal service operations as quickly as possible and minimize impact.
Perform systems support in identifying, analyzing, and eliminating the root causes of recurring incidents to minimize continued adverse impacts and potential degradation of services.
Make recommendations for the improvement of Incident and Problem management consistent with industry’s best practices for the cloud.
Maintain knowledge base of known issues, resolutions, and best practices for operational continuity.
Perform automated health checks across the full stack (Operating System, Application, Database and PaaS services) at agreed levels on an agreed frequency.
Provide a monthly issues management report. The report shall include cloud-related incidents, any stability and performance issues, configurations issues, quantity of tickets received, and time duration to resolve tickets.
Develop and implement thresholds, rules, and response procedures based on product team’s recommendation.
Monitor resource utilization (e.g., CPU, Memory, Disk Space) for the cloud hosted Virtual Machines (VMs) and other cloud services.
Manage the resolution procedures for any threshold breaches for cloud resources.
Improves system reliability, observability, automation, scalability, and operational resilience through engineering practices.
Monitors, maintains, and optimizes cloud infrastructure, databases, and platform services for reliability and performance.
Act as FinOps Analyst and perform cost optimization.
QUALIFICATIONS
Education: Bachelor's degree in Computer Science, Software Engineering, or related field. (Or equivalent experience.)
Experience: 5+ years’ experience in IT systems engineering, systems development, systems coding, and programming.
Deep expertise with AWS services, including monitoring, logging, compute, storage, and networking.
Proficiency in Infrastructure as Code (IaC) tools like Terraform, AWS CloudFormation, or Azure Bicep.
Hands-on experience with monitoring and APM tools such as CloudWatch, Azure Monitor, Datadog, Prometheus, Grafana, New Relic, etc.
Solid understanding of incident response, change management, and ITIL-based operational support.
Familiarity with CI/CD toolchains and automation platforms (Jenkins, GitHub Actions, GitLab, ArgoCD).
Strong scripting skills (Python, PowerShell, Bash) for automation and orchestration.
Advanced experience in providing DevSecOps implementation using GitOps, or similar tools.
Experienced in developing, testing, and maintaining containerized applications.
Expert knowledge of source version control, build/release tools and methodologies, CI/CD pipelines and the Software Build process.
Experience in building and maintaining CI/CD pipelines for large enterprises that consist of a large number of complex applications.
Ability to be flexible and work on several different products while supporting multiple teams
Experience with FinOps practices, cost modeling, forecasting, and optimization tools within cloud platforms.
Understanding of federal compliance and security frameworks (e.g., FedRAMP, NIST, JISF Rev 5).
Ability to analyze logs and metrics and conduct performance tuning for cloud-based services and applications.
Experience working across multiple product teams to get a grasp of a product and/or programs overall state of health.
ITIL, AWS SysOps, or Google Professional Cloud DevOps Engineer certifications are a plus.
COMMUNICATION & ORGANIZATIONAL SKILLS
Excellent presentation and communication skills.
Consultant mindset with the ability to work with high level customer stakeholders and build excellent customer relationships.
Experience identifying and applying industry tools, solutions, methods best practices, and emerging technologies.
Strong analytical skills and problem-solving skills with the ability to formulate and communicate recommendations for improvement.
Experience with process design and documentation methodologies, and design and production of quality deliverables, process and use case modeling, business case development.
Demonstrated ability to work effectively, independently, and as part of a team.
Security Clearance Level: Must be able to pass a background check to obtain a position of Public Trust.
Must be a US Person (Green Card Holder, US Permanent Resident Alien, Refugee, Asylee, or US Citizen).
Location: Remote.
GDIT IS YOUR PLACE
At GDIT, the mission is our purpose, and our people are at the center of everything we do.
● Growth: AI-powered career tool that identifies career steps and learning opportunities
● Support: An internal mobility team focused on helping you achieve your career goals
● Rewards: Comprehensive benefits and wellness packages, 401K with company match, and competitive pay and paid time off
● Flexibility: Full-flex work week to own your priorities at work and at home
● Community: Award-winning culture of innovation and a military-friendly workplace
Explore an enterprise IT career at GDIT and you’ll find endless opportunities to grow alongside colleagues who share your desire to drive operations forward.
#GDITLA
The likely salary range for this position is $111,155 - $150,385. This is not, however, a guarantee of compensation or salary. Rather, salary will be set based on experience, geographic location and possibly contractual requirements and could fall outside of this range.Scheduled Weekly Hours:
40Travel Required:
Less than 10%Telecommuting Options:
RemoteWork Location:
Any Location / RemoteAdditional Work Locations:
Total Rewards at GDIT:
Our benefits package for all US-based employees includes a variety of medical plan options, some with Health Savings Accounts, dental plan options, a vision plan, and a 401(k) plan offering the ability to contribute both pre and post-tax dollars up to the IRS annual limits and receive a company match. To encourage work/life balance, GDIT offers employees full flex work weeks where possible and a variety of paid time off plans, including vacation, sick and personal time, holidays, paid parental, military, bereavement and jury duty leave. GDIT typically provides new employees with 15 days of paid leave per calendar year to be used for vacations, personal business, and illness and an additional 10 paid holidays per year. Paid leave and paid holidays are prorated based on the employee’s date of hire. The GDIT Paid Family Leave program provides a total of up to 160 hours of paid leave in a rolling 12 month period for eligible employees. To ensure our employees are able to protect their income, other offerings such as short and long-term disability benefits, life, accidental death and dismemberment, personal accident, critical illness and business travel and accident insurance are provided or available. We regularly review our Total Rewards package to ensure our offerings are competitive and reflect what our employees have told us they value most.Our Identity Verification Process:
As part of the hiring process, we will ask you to complete an identity verification process that leverages advanced biometrics and artificial intelligence to ensure authenticity and protect against identity fraud. You are expected to be on camera during virtual interviews. We reserve the right to take your picture to verify your identity and prevent fraud. By proceeding, you authorize the collection, processing, and use of your biometric data for identity verification and security purposes.About Our Work:
We are GDIT. A global technology and professional services company that delivers technology solutions and mission services to every major agency across the U.S. government, defense and intelligence community. Our 26,000 experts extract the power of technology to create immediate value and deliver solutions at the edge of innovation. We operate across 50+ countries worldwide, offering leading mission-ready capabilities in AI, cloud, cyber and software development.Join our Talent Community to stay up to date on our career opportunities and events atgdit.com/tc.
Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans$74.1k - $148.3k
...performance analysis, and system tuning. As a Site Reliability Engineer, you will solve interesting technical... ...has formed a new organization - Health Data Intelligence Platform. This team will... ...Ownership –You will be part of the SRE team, whose mission is the shared full...DataTemporary workImmediate startFlexible hours$95k - $171k
...infrastructure? Do you want to build your SRE career on one of the most exciting... ..., Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building...SuggestedPermanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$106.9k - $200.6k
...opportunity We are seeking AI Systems Engineers to own the security and trust fabric of... ...with Enterprise Security / Cloud Platform / SRE to ensure independent review, alignment... ...and consumption contracts with platform, data, and runtime teams. Ideally, you’...DataFull timeWork experience placementSummer holidayRemote workFlexible hours$125.5k - $261.6k
...We are seeking AI Systems Engineers to build and operate the foundational... ...at scale, who treats reliability and portability as non-negotiable... ..., so downstream runtime, data, and execution services can run... ...Background in platform or SRE roles where success is measured...DataFull timeContract workSummer holidayFlexible hours- ...using automation and Infrastructure Code. Builds reliability into the ecosystem by applying best practices in resiliency engineering and observability by developing resiliency... ...and software engineering techniques with site reliability engineering practices to create reliable...SuggestedFull time
$106.9k - $200.6k
...opportunity We are seeking an AI Systems Engineer to own the delivery, model-serving,... ...is a distinct discipline from platform, data, and trust engineering. Where Platform Engineering... ..., so every signal is captured and routed reliably. Automate GitOps-based delivery and...DataFull timeSummer holidayFlexible hours- ...Senior Operations Analyst (SRE)CGI's Advantage Cloud Operations is an SRE-driven... ...Senior Operations Analyst is a senior Site Reliability Engineering (SRE) practitioner responsible for driving... ...team in adopting proactive, data-driven operational practices while improving...DataWork at office
- ...wireless, and VPN connectivity for both site-to-site and remote access- Handle endpoint... ...what is happening in the business- Keep data accurate and trustworthy across all master... ...of real, hands-on experience in systems engineering or infrastructure administration - not help...DataRemote work
- ...Overview/ Job Responsibilities Entarian is seeking a Senior Software Engineer for the Commander, Naval Meteorology and Oceanography Command (CNMOC) Data Dissemination SBIR Information Technology Solutions Project. This role will bring strong systems, software, cloud, and...DataInterim role
- ...Energizing Your Tomorrow.The Power Generation Reliability Program Manager leads and advances the... ..., and reducing forced outages through data-driven decision making, predictive... ...depth knowledge of related processes and engineering tools to solve complex problems and drive...DataFull time
- ...implementation of a comprehensive Reliability Centered Maintenance strategy... ...improvements across site.Guide efforts to ensure reliability... ..., develop, and implement engineered improvements and projects for... ...projects by supplying historical data, cost analysis, or other...DataPermanent employmentFull timeContract workWork experience placementFlexible hours
$78k - $163.8k
...0%Type of Travel: Continental US* * *The Opportunity:As a CACI Data Scientist Operations Integrator working at MARFORSOUTH, you will... ....• Bachelor’s degree in Data Science, Computer Science, Engineering, or a related field with at least 2 years of experience in data...DataContract workFor contractorsWork experience placementFlexible hours$103.71k - $138.28k
...AI-driven world. By connecting people, data, and applications quickly, securely, and... ...and experience in system architecture and engineering disciplines. Specific technical knowledge... ...due diligence activities including site surveys, design, design review, bill of materials...DataTemporary workRemote work$103.6k - $155.4k
...they're making history.Northrop Grumman is seeking a Software Engineer to work on-site at Schriever SFB, near Colorado Springs, CO for our Front... ...microservices that interact with a variety of vendor interfaces for data processing. • Deploy new containerized applications in...DataFull timeWork at officeRelocationShift work$293.9k - $406.8k
...outcomes, as a Distinguished Engineer. The team delivers secure, scalable... ..., with a strong emphasis on reliability, interoperability, and long-... ..., policy, privacy, and data protection principles within... ...Please see the Cisco careers site to discover more benefits and...DataFull timeTemporary workLocal areaRemote workFlexible hours$95k - $105.89k
...curiosity.Job DescriptionDevelop real time data pipelines and solve big data problems... ...Scala, SQL, and AWS cloud environment.Write reliable, maintainable, well-documented code that... ...environments to cloud-based data warehouses.Engineer robust streaming architecture using SNS,...DataWork from home- ...alignment with business goals.Confers with data processing or project managers to obtain... ...:Bachelor’s degree in Computer Science, Engineering, Information Technology, Information... ...integration tests, ensuring code quality, reliability, and maintainability.DE collaborating...DataFull time
$112k - $187k
...make it easier for coaches and athletes at any level to capture video, analyze data, share highlights and more. Ready to join us? Your Role We are looking for a Senior Software Engineer with a passion for cloud technology and building highly scalable software...DataFull timeWork at officeLocal areaRemote workFlexible hours- ...are in search of a IT Systems Engineer for our Harahan, LA location.... ...SAP on premise • Review all site Firewalls. • Review all site... ...skills and initiative. ~ Reliable and responsible, able to respond... ...Abilityto define problems, collect data, establish facts, and draw...DataWork experience placementRemote workWeekend work
$197.2k - $255.2k
...shifted. AI is accelerating. The key to AI success is intelligent data and the infrastructure to manage it. Customers are urgently... ...JOB SUMMARY Behind every great sales rep is a Solutions Engineer who makes the magic happen. As an Enterprise Solutions Engineer,...DataFull timeLocal areaImmediate startRemote workShift work$139.3k - $203.6k
...USA.Meet the TeamOur software engineering team develops software using... ..., scalability, security, and reliability of our software.ResponsibilitiesImprove... ...AWS, and Terraform.Work with data technologies such as... ...Please see the Cisco careers site to discover more benefits and...DataFull timeTemporary workLocal areaRemote workFlexible hours$102.3k - $209.5k
...layers, working in DevOps and incident response paradigms to maintain reliability and availability. Work closely with platform engineering, operations, firmware development, silicon/board vendors, and data center teams to support and troubleshoot hardware bring-up,...DataTemporary workImmediate startFlexible hoursShift work$86.4k - $138.6k
...an integral member of an agile software engineer team responsible for building complex scalable... ...in design and analysis of algorithms, data structures, and design patterns in the... ...regularly from the office to various work sites or from site-to-site Occasionally Works...DataFull timeFor contractorsWork at officeLocal area- ...maximize reservoir value. Partner with the best As a Reliability Engineer, you will be responsible for: Demonstrate stand-alone root... ...Maps for specified technologies. Analyze appropriate data to monitor fleet performance Create single technology business...DataWorldwideMonday to FridayFlexible hours
$104k - $125k
...Position at Hood Lumber Hood Industries, Inc is seeking a Reliability Engineer to play a key role in our current operations at our Bogalusa,... ...predictive technologies, continuous improvements methodologies, data and trend development and analysis (FMEA, RCA, RCM, etc.). •...DataFull timeWork at officeFlexible hours$146.22k - $243.69k
...The Director of Infrastructure Engineering & Cloud Operations at Banner... ...services including cloud and data centers, compute, storage,... ...modernization, operational excellence, reliability engineering, and workforce... ..., cloud platform engineers, SRE/platform SRE, cloud security...DataFull timeWork experience placementRemote workShift work$107.48k - $143.31k
...Doing Lead and apply regional reliability engineering strategies to improve equipment performance... ...predictive tools), analyze reliability data and KPIs, and translate insights into... ...with the ability to support multiple sites remotely. Willingness to travel to...DataTemporary workRemote workFlexible hours$250.6k - $362.6k
...security outcomes, as a Principal Engineer. The team delivers secure,... ..., with a strong emphasis on reliability, interoperability, and long-... ...with security, privacy, and data protection principles as applied... ...Please see the Cisco careers site to discover more benefits and...DataFull timeTemporary workLocal areaRemote workFlexible hours- Job Description:Principal Software Engineer Note: Fidelity is not providing immigration sponsorship for this positionThe RoleWe are seeking... ...-on experience is building complex SQL queries required for data validationsIdentify opportunities to improve maintainability of...DataFull time
- ...Provides suggestions and recommendations for new technology solutions. Work with senior level engineers to implement applications and technology solutions. Monitor daily use data and application server backups, upgrades, and patches. Responds to and resolves any critical...DataWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Site Reliability Engineer (SRE). Be the first to apply!
- data center engineer Louisiana
- remote data engineer Louisiana
- etl data engineer Louisiana
- data engineer analytics Louisiana
- data engineer machine learning Louisiana
- finance data engineer Louisiana
- data engineer Louisiana
- junior data engineer remote Louisiana
- data developer Louisiana
- senior data center engineer Louisiana

