Site Reliability Engineer
3B Staffing LLC
Site Reliability Engineer
Location: Occasional onsite visits to Reston VA (Zip code 20190).
Duration-1 year plus
Interview process: The final interview is a mandatory, face-to-face interview in Reston VA. Zip code: 20190
Strong AWS SRE/Platform Engineering with Python/Java, Terraform/Ansible (IaC), Kubernetes/EKS, Docker, CI/CD tools, observability (Datadog/New Relic), Linux, and cloud-native architecture with automation and reliability focus.
Visa Status: H1B, GC, USC, no OPT candidates.
Candidates must already reside in the Washington DC area. The client is not accepting candidates willing to relocate. Final interview is a MANDATORY face to face interview in Reston VA. Zip code 20190
My direct client, a leading healthcare insurance company, is seeking a Site Reliability Engineer in the Washington DC, Maryland and/or Virginia area for a long term contract position with our customer in Reston VA (Hybrid position)
Strong skills are desired in each of the following areas:
Development: Experience programming with one or more languages: Python, Java, Groovy, Go, etc.
IAC Tools for Platform Automation: Strong skills and experience in at least one: Ansible, Terraform, AWS CloudFormation, CDK.
Containers: Docker or other OCI-certified containers- is a Plus
Container Orchestration Platform: Experience with Kubernetes, AWS EKS, AWS ECS is a plus.
CNI Plugins: Calico, Flannel, Weave Net etc.
Service Mesh: Istio, AWS App Mesh, OpenShift Service Mesh etc.
Container Security Tools: Twistlock, Sysdig, Aqua etc. is a plus,
Platform Monitoring, Observability, & Performance Tools: Nginx, New Relic, AppDynamics, Data Dog, Thanos, Jaeger, LogDNA, etc.
DevOps Tools: Git/Repo, Crucible, Bitbucket, Jira, Ansible, Puppet, Jenkins, Circleci, Bamboo, Maven, Artifactory, Nexus etc.
Other Required Skills:
Understanding of Cloud Native Architecture
Linux, Shell scripting, and general admin skills
Network, Security, Plugins, & Storage Skills
AWS skills: EC2, S3, EBS, EFS, IAM, VPC, Lambda, EKS, RedShift etc.
Cloud and DevOps certifications, e.g., AWS
Strong skills are desired in each of the following areas:
Development: Experience programming with one or more languages: Python, Java, Groovy, Go, etc.
IAC Tools for Platform Automation: Strong skills and experience in at least one: Ansible, and Terraform, AWS Cloud formation, CDK.
Containers: Docker or other OCI-certified containers- is a Plus
Container Orchestration Platform: Experience with Kubernetes, AWS EKS, AWS ECS is a plus.
CNI Plugins: Calico, Flannel, Weave Net etc.
Service Mesh: Istio, AWS App Mesh, OpenShift Service Mesh etc.
Container Security Tools: Twistlock, Sysdig, Aqua etc. is a plus,
Platform Monitoring, Observability, & Performance Tools: Nginx, New Relic, AppDynamics, Data Dog, Thanos, Jaeger, LogDNA, etc.
DevOps Tools: Git/Repo, Crucible, Bitbucket, Jira, Ansible, Puppet, Jenkins, Circleci, Bamboo, Maven, Artifactory, Nexus etc.
Other Required Skills:
Understanding of Cloud Native Architecture
Linux, Shell scripting, and general admin skills
Network, Security, Plugins, & Storage Skills
AWS skills: EC2, S3, EBS, EFS, IAM, VPC, Lambda, EKS, RedShift etc.
Cloud and DevOps certifications, e.g., AWS
Communicates Architectural decisions, plans, goals, and strategies, while highlighting short-term trade-offs vs. long-term commitments and costs
Engage in and improve the end-to-end Lifecycle of services, starting from Inception & design, deployment, and operations.
Establish automation capabilities leveraging Cloud native solutions, to improve the Developer experience.
Support activities, including System design consulting, developing software platforms and frameworks, capacity planning, and launch reviews.
Willingness to roll up the sleeves and troubleshoot difficult issues and engage the Customer.
Willingness to learn new AWS Services and other technologies as required.
Systems Scalability and sustainability leveraging automation and strive to improve our systems with changes that improve reliability and velocity.
Experience with Enterprise Cloud transformation and migration efforts.
Actively participate and help guide customers on using Cloud-native design and architecture patterns.
Provide Consultation on Technology infrastructure planning and engineering for assigned systems; Assesses the implications of technology strategies on infrastructure capabilities.
Establish strategies to migrate Legacy applications by conversion to multiple Microservices and hosting on AWS Cloud platform.
Leverage Cloud-native architecture components including Containers, immutable infrastructure, Microservices, Service Mesh etc., to build highly available and Fault tolerant applications.
Conduct research on the global technology trends and their applicability to FEPOC products in support of our internal development teams and business initiatives.
Promotes and ensures Modern application design, applies engineering best practices in the development and operations life cycle and mitigates vulnerabilities.
Monitors and manages the Stability, Availability, and Performance of enterprise systems and platforms across IT domains.
(e.g., Systems, Network, Storage, Security) by analyzing systems to identify problems, trends, and opportunities for improvement.
Automate end to end process to maintain (patches and upgrades) of our AWS Cloud ecosystem.
Makes data-driven recommendations and decisions and continuously improves the overall efficacy and efficiency of our software delivery capabilities.
Mentoring peers as well as engaging with others across teams and socializing solutions.
Additional Required experience
Minimum of One AWS certification is required.
Minimum of 10years of IT experience of which at least 5 years must be in AWS Cloud
Platform engineering and Administration.
Strong Leadership experience with driving Transformation initiatives
3-5 years of experience in a Site Reliability Engineering role
Experience with SRE principles and transformation
3+ years of experience with Containerization (Kubernetes), Cloud technologies (AWS, Azure etc.), DevOps tool chain (Ansible, Jenkins, Artifactory, bitbucket, etc.), and technical patterns (IaC, Automated Provisioning/Release, CI/CD, etc.)
Solid understanding of Software coding techniques and experience with full spectrum of Software engineering (Build, Integration, Test, Releasing and Deployment) leveraging Python.
Experience in Developing and/or challenging engineering solutions/practices and collaborating with peers within and outside of immediate team, including customers (Dev, Architects, Engineers)
Platform Engineering Lead with Hands -on Experience: Building robust Middleware Environments, previous Linux System administration is required.
Must have strong hands-on knowledge of AWS platform and services but not limited to VPC, Networking, Direct Connect, Subnets, NACLs, Security Groups, EC2, S3, IAM, ELBs, Lambda, CloudWatch, CloudTrail, EKS etc.
Must Have Hands on current Implementation and Production level experience in AWS Cloud.
Hands on experience with Automation and Infrastructure Provisioning is a must
Our goal is to only provision infrastructure with Code, and Policy As Code.
Must be familiar with Terraform automation, Ansible playbooks, and Python code.
Experience with AWS Cloud Formation and CDK is required.
Must have hands on experience in writing Lambda functions preferably in Python (Boto3).
Must be well versed in writing Linux Bash scripts.
Hands-on experience with Containerization and Amazon EKS is a big plus.
A great understanding of various DevOps toolchains, including Git/repo, Crucible, Jenkins etc.Solid understanding and experience with a CI/CD tool chain.
Location: Occasional onsite visits to Reston VA (Zip code 20190).
Duration-1 year plus
Interview process: The final interview is a mandatory, face-to-face interview in Reston VA. Zip code: 20190
Strong AWS SRE/Platform Engineering with Python/Java, Terraform/Ansible (IaC), Kubernetes/EKS, Docker, CI/CD tools, observability (Datadog/New Relic), Linux, and cloud-native architecture with automation and reliability focus.
Visa Status: H1B, GC, USC, no OPT candidates.
Candidates must already reside in the Washington DC area. The client is not accepting candidates willing to relocate. Final interview is a MANDATORY face to face interview in Reston VA. Zip code 20190
My direct client, a leading healthcare insurance company, is seeking a Site Reliability Engineer in the Washington DC, Maryland and/or Virginia area for a long term contract position with our customer in Reston VA (Hybrid position)
Strong skills are desired in each of the following areas:
Development: Experience programming with one or more languages: Python, Java, Groovy, Go, etc.
IAC Tools for Platform Automation: Strong skills and experience in at least one: Ansible, Terraform, AWS CloudFormation, CDK.
Containers: Docker or other OCI-certified containers- is a Plus
Container Orchestration Platform: Experience with Kubernetes, AWS EKS, AWS ECS is a plus.
CNI Plugins: Calico, Flannel, Weave Net etc.
Service Mesh: Istio, AWS App Mesh, OpenShift Service Mesh etc.
Container Security Tools: Twistlock, Sysdig, Aqua etc. is a plus,
Platform Monitoring, Observability, & Performance Tools: Nginx, New Relic, AppDynamics, Data Dog, Thanos, Jaeger, LogDNA, etc.
DevOps Tools: Git/Repo, Crucible, Bitbucket, Jira, Ansible, Puppet, Jenkins, Circleci, Bamboo, Maven, Artifactory, Nexus etc.
Other Required Skills:
Understanding of Cloud Native Architecture
Linux, Shell scripting, and general admin skills
Network, Security, Plugins, & Storage Skills
AWS skills: EC2, S3, EBS, EFS, IAM, VPC, Lambda, EKS, RedShift etc.
Cloud and DevOps certifications, e.g., AWS
Strong skills are desired in each of the following areas:
Development: Experience programming with one or more languages: Python, Java, Groovy, Go, etc.
IAC Tools for Platform Automation: Strong skills and experience in at least one: Ansible, and Terraform, AWS Cloud formation, CDK.
Containers: Docker or other OCI-certified containers- is a Plus
Container Orchestration Platform: Experience with Kubernetes, AWS EKS, AWS ECS is a plus.
CNI Plugins: Calico, Flannel, Weave Net etc.
Service Mesh: Istio, AWS App Mesh, OpenShift Service Mesh etc.
Container Security Tools: Twistlock, Sysdig, Aqua etc. is a plus,
Platform Monitoring, Observability, & Performance Tools: Nginx, New Relic, AppDynamics, Data Dog, Thanos, Jaeger, LogDNA, etc.
DevOps Tools: Git/Repo, Crucible, Bitbucket, Jira, Ansible, Puppet, Jenkins, Circleci, Bamboo, Maven, Artifactory, Nexus etc.
Other Required Skills:
Understanding of Cloud Native Architecture
Linux, Shell scripting, and general admin skills
Network, Security, Plugins, & Storage Skills
AWS skills: EC2, S3, EBS, EFS, IAM, VPC, Lambda, EKS, RedShift etc.
Cloud and DevOps certifications, e.g., AWS
Communicates Architectural decisions, plans, goals, and strategies, while highlighting short-term trade-offs vs. long-term commitments and costs
Engage in and improve the end-to-end Lifecycle of services, starting from Inception & design, deployment, and operations.
Establish automation capabilities leveraging Cloud native solutions, to improve the Developer experience.
Support activities, including System design consulting, developing software platforms and frameworks, capacity planning, and launch reviews.
Willingness to roll up the sleeves and troubleshoot difficult issues and engage the Customer.
Willingness to learn new AWS Services and other technologies as required.
Systems Scalability and sustainability leveraging automation and strive to improve our systems with changes that improve reliability and velocity.
Experience with Enterprise Cloud transformation and migration efforts.
Actively participate and help guide customers on using Cloud-native design and architecture patterns.
Provide Consultation on Technology infrastructure planning and engineering for assigned systems; Assesses the implications of technology strategies on infrastructure capabilities.
Establish strategies to migrate Legacy applications by conversion to multiple Microservices and hosting on AWS Cloud platform.
Leverage Cloud-native architecture components including Containers, immutable infrastructure, Microservices, Service Mesh etc., to build highly available and Fault tolerant applications.
Conduct research on the global technology trends and their applicability to FEPOC products in support of our internal development teams and business initiatives.
Promotes and ensures Modern application design, applies engineering best practices in the development and operations life cycle and mitigates vulnerabilities.
Monitors and manages the Stability, Availability, and Performance of enterprise systems and platforms across IT domains.
(e.g., Systems, Network, Storage, Security) by analyzing systems to identify problems, trends, and opportunities for improvement.
Automate end to end process to maintain (patches and upgrades) of our AWS Cloud ecosystem.
Makes data-driven recommendations and decisions and continuously improves the overall efficacy and efficiency of our software delivery capabilities.
Mentoring peers as well as engaging with others across teams and socializing solutions.
Additional Required experience
Minimum of One AWS certification is required.
Minimum of 10years of IT experience of which at least 5 years must be in AWS Cloud
Platform engineering and Administration.
Strong Leadership experience with driving Transformation initiatives
3-5 years of experience in a Site Reliability Engineering role
Experience with SRE principles and transformation
3+ years of experience with Containerization (Kubernetes), Cloud technologies (AWS, Azure etc.), DevOps tool chain (Ansible, Jenkins, Artifactory, bitbucket, etc.), and technical patterns (IaC, Automated Provisioning/Release, CI/CD, etc.)
Solid understanding of Software coding techniques and experience with full spectrum of Software engineering (Build, Integration, Test, Releasing and Deployment) leveraging Python.
Experience in Developing and/or challenging engineering solutions/practices and collaborating with peers within and outside of immediate team, including customers (Dev, Architects, Engineers)
Platform Engineering Lead with Hands -on Experience: Building robust Middleware Environments, previous Linux System administration is required.
Must have strong hands-on knowledge of AWS platform and services but not limited to VPC, Networking, Direct Connect, Subnets, NACLs, Security Groups, EC2, S3, IAM, ELBs, Lambda, CloudWatch, CloudTrail, EKS etc.
Must Have Hands on current Implementation and Production level experience in AWS Cloud.
Hands on experience with Automation and Infrastructure Provisioning is a must
Our goal is to only provision infrastructure with Code, and Policy As Code.
Must be familiar with Terraform automation, Ansible playbooks, and Python code.
Experience with AWS Cloud Formation and CDK is required.
Must have hands on experience in writing Lambda functions preferably in Python (Boto3).
Must be well versed in writing Linux Bash scripts.
Hands-on experience with Containerization and Amazon EKS is a big plus.
A great understanding of various DevOps toolchains, including Git/repo, Crucible, Jenkins etc.Solid understanding and experience with a CI/CD tool chain.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Reston, VA vacancy
- ...prioritize a diverse F5 community where each individual can thrive.Role SummaryWe are seeking a proactive and detail-oriented Site Reliability Engineer II (SRE II) to join our 24/7 Operations team in a hybrid capacity. In this role, you will provide round-the-clock, eyes-on...SuggestedFull timeLocal areaImmediate startShift workNight shiftAfternoon shiftWeekday work
$81.1k - $187k
.... You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers... ...posted.Career Level - IC3Escalation points for junior Site Reliability Engineers during complex or high-impact incidents.Manage and...SuggestedTemporary workMonday to FridayFlexible hoursShift workNight shift$102.1k - $202.2k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewDo... ...? We’re looking for a Site Reliability Engineer II with the right mix of software...SuggestedOngoing contractWork at officeLocal areaShift work3 days per week- ...thrive.Role Overview:Join a growing team securing both leading-edge protection solutions and enterprise infrastructure. As a Site Reliability Engineer II, a part of the Operational Support Systems (OSS) organization under the Security & Distributed Cloud organization, you...SuggestedFull timeLocal area
$92.52k - $138.79k
...insights you need to drive results. FreeWheel’s platform makes TV and video advertising work.Job DescriptionWe're looking for a Site Reliability Engineer to own cloud infrastructure, system reliability, and observability for the Freewheel BuyerCloud and Revenue Science teams....SuggestedFull timeWorldwide$85.4k - $168.1k
...per yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole type: Individual... ...: Software EngineeringDiscipline: Site Reliability EngineeringCompany: MicrosoftOverviewDo... ...? We’re looking for a Site Reliability Engineer with the right mix of software...Ongoing contractWork at officeLocal areaShift work3 days per week- ...diverse F5 community where each individual can thrive.Role SummaryWe are seeking an experienced, security-focused Senior Site Reliability Engineer (Senior SRE) to drive the reliability, architectural design, and continuous compliance of our FedRAMP-authorized cloud platform...Full timeLocal area
$87.1k - $157.45k
...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across... ...to come in and help us build systems that stay reliable when things get complicated. We need a Site Reliability Engineer who has experience building, deploying...Local areaImmediate startWork from homeFlexible hours- ...We are looking for a talented Senior Site Reliability Engineer to join our team to deliver world class search technologies to mobile devices. You will be working with a smart team of Engineers to lead and drive the stability, reliability, and observability of all Seekr...Permanent employmentFlexible hours
- ...Site Reliability Engineer needed for a full time opportunity with SOC's direct client based in Herndon, VA. **Due to federal requirements, candidates must hold and possess an Active DOW TS/SCI security clearance to be considered for this role.** SOC is seeking...Full time
$120k - $140k
...company, was founded in 1997 and started by ensuring the reliability and performance of mission-critical databases. We quickly... .... Why you: Pythian is building a next-generation Site Reliability Engineering team, and we're looking for talented, motivated engineers...Work from home- ...Site Reliability Engineer (SRE) Reston, VA Site Reliability Engineer (SRE) Position: Site Reliability Engineer (SRE) Work Authorization: All Work Authorizations Location: Reston, VA Contract: 24 months Description: Site Reliability Engineer (SRE) roles...Contract work
- ...Job Description We are seeking a Senior Site Reliability Engineer (SRE) to support AWS cloud environments, including AWS GovCloud and Classified Cloud. This role focuses on cloud infrastructure, Linux administration, containerization, CI/CD, automation, security, and...
$136.2k - $214.01k
...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to...Full timeFlexible hours$91.4k - $187k
Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services. Responsible...Temporary workFlexible hoursShift workWeekend work$77.5k - $179k
Site Reliability Engineer I - Sales OperationsThis role has been designed as 'Hybrid' with a requirement that you will work on average 2 days per week from an HPE office.Who We Are:Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live...Full timeWork experience placementInternshipWork at officeLocal areaImmediate start2 days per week$130k - $200k
Summary Position Title: Site Reliability Engineer Position ID: TA247 Location(s): On-site; Aurora, CO; Herndon, VA Application Deadline: October 31, 2026 Security Clearance Requirement: An active TS/SCI Security Clearance with the ability to take and pass a...Full timeTemporary workLocal area$112k - $179k
...Job Locations US Responsibilities Peraton is seeking a Lead Site Reliability Engineer to join our team of qualified, diverse individuals. The ideal candidate will play a critical role in ensuring the reliability, resilience, and recoverability of mission...Contract workShift work- ...dynamic team with a broad knowledge of how Oracle's cloud platform works. You'll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers. Note - this role is not a Monday to Friday core hours role -...Monday to FridayShift workNight shift
$102k - $234.6k
...members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help... ...to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data.Only Oracle brings...Temporary workImmediate startFlexible hours$109.18k - $163.77k
...you will be responsible for ensuring the reliability, scalability, and performance of our data systems. Working closely with data engineers and other operation sub-teams, you will manage... ...and benefits summary on our careers site for more details.EducationBachelor's DegreeWhile...Full time- ...industry leaders and technology innovators. We are looking for an experienced and highly motivated Systems Administrator / Site Reliability Engineer to join our team. Our Systems Administrator and Site Reliability Engineer will ensure that systems stay up, optimally...Full timeWork at officeShift work
$62k - $141k
Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if you have...Full timeContract workPart timeLocal areaRemote work$114.4k - $125.4k
...excellence of the Salesforce GovCloud! Are you passionate about ensuring the reliability and performance of mission-critical cloud services? Salesforce is seeking a talented Site Reliability Engineer to join our dynamic team, supporting our GovCloud environment. As a key...Full timeLocal areaShift workNight shift- ...Catalyst, and In-Q-Tel. Mission | On Site | Full Time | Active TS/SCI with Full... ...government customer site, ensuring the reliability and performance of Twenty's mission-critical... ...technical ownership and customer-facing engineering: you’ll define how we measure...Full timeContract workRemote workFlexible hours
- ...tech SME and PM Period of performance: Up to 2 years in duration MUST HAVES: Minimum of 8 years of experience as a Site Reliability Engineer with a strong understanding of SRE principles for highly scalable and reliable systems Possess a bachelor's degree...Local areaRelocation package3 days per week
$103.5k - $150k
...exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform...Temporary workWork experience placementLocal area3 days per week- ...Site Reliability Engineer Jersey City, New Jersey, United States; McLean, Virginia, United States; Richmond, Virginia, United States Who We Are: Exiger transforms supply chains into a strategic advantage, advancing our mission to make the world a safer and more...
$100k - $160k
...all bring to the team; and empower our employees to create innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and...Temporary work- ...Detail Description: The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You prepare and take ownership of "day two" operations...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
Related searches



