Site Reliability Engineer
3B Staffing LLC
Job Description:
Qualifications:
This position is 60 % SRE and 40% SDE. Also open for candidates to join MTH along with ATL GO team. Must agree to work onsite at either of these two locations based on the team's hybrid schedule. Required Skillset
• Manage and optimize data streaming and API components in OpenShift Onpremise and AWS.
• Proactively review the application's APIs and processes to identify opportunities to optimize the response times for various application components.
• Automate various types of testing including data quality checks, automate delivery to system integration, production and automate deployment for production
• Develop integrations between the application in Onpremise and AWS and our third-party tools (ServiceNow, VersionOne, Sumo)
• Work with teams to create SLI/SLO's .
• Actively monitor and lead troubleshooting of degraded performance and hard to define issues for the platform applications, develop the solution and document artifacts in the back log from root cause analysis.
• Evolve the cloud infrastructure ecosystem for our application suite by experimenting with emerging technologies and completing prototypes to understand benefit
• Design and develop CI/CD pipeline in AWS and Openshift to deploy various application artifacts, including APIs and Data Process Jobs.
• Analyze, design and develop the artifacts to configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and resolve the issues.
• Maintain data integrity and access control by using AWS security tools and services such as HSM, IAM, etc.
• Understand and develop tools to monitor AWS billing for the services, generate cost related reports and help develop and implement cost optimization strategies.
• Work with enterprise security architects to design and implement data security tools, measures, data encryption, key management; design and develop solutions to address the security vulnerabilities discovered by internal security audit team, as well as by the vendors, security community, etc.; design and develop solutions for support team to regularly scan and review to fix security issues
• Regularly and proactively monitor and analyze the capacity and performance of the platform, work with architecture team to design and implement elastic infrastructure to accommodate the irregular burst of user traffic/requests.
• Work with architecture team to develop backup strategy and implement the backup solution for critical data and application components for service restoration and disaster recovery purpose.
• Work with architecture, infrastructure, and application teams to provide input on continuous improvement on the design, performance and security enhancements. Desired Skillset:
• Deep understanding of the operations of AWS cloud platforms.
• Must be well versed in the automation, scripting, monitoring, including use of tools from the major cloud platforms, including but not limited to OpenShift Cloud Formation, Terraform, Ansible, Shell, Python
• Preferable for candidates with significant technical knowledge with infrastructure layers, including but not limited to: Linux OS, major virtualization platforms, Traditional and software defined network, Load Balancers, firewall, API tools, element/performance/intelligent monitoring tools, storage, backup strategy, etc.
• Significant knowledge and experience in end-to-end operations for enterprise systems and applications, including driving issue resolution for mission critical systems.
• Must have experience working to automate, operationalize and improve the Development/QA using CI/CD tools (Gitlab, Github, Jenkins, Maven, Gradle, Nexus)
• Working experience with Software Release Management. Desired Qualification
• BS degree in Computer Science or a related technical field or equivalent practical experience. Minimum Experience
• 3+ years of related DevOps, SysOps engineering experience with focus on major cloud platforms (AWS preferred).
• 2+ years of application development experience including data streaming, deploying/monitoring high availability critical application components.
• 1+ Years in Site Reliability Engineering organization preferred
• Overall 4-6years of experience
Responsibilities:
As a engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to a broad range of industry leading technologies that will help to drive acceleration. The ideal candidate will have expert design and development capabilities and be positioned to contribute to a growing set of services and features for the ecosystem. This role will be supporting highly available, business critical applications. This role will serve as the escalation point for complex and hard to define issues in both on premise and AWS environments. We are seeking talented engineers, well versed in DevOps technologies, automation, infrastructure orchestration, configuration management, continuous integration, troubleshooting of complex issues, who are not constrained by how "things are usually done".
Qualifications:
This position is 60 % SRE and 40% SDE. Also open for candidates to join MTH along with ATL GO team. Must agree to work onsite at either of these two locations based on the team's hybrid schedule. Required Skillset
• Manage and optimize data streaming and API components in OpenShift Onpremise and AWS.
• Proactively review the application's APIs and processes to identify opportunities to optimize the response times for various application components.
• Automate various types of testing including data quality checks, automate delivery to system integration, production and automate deployment for production
• Develop integrations between the application in Onpremise and AWS and our third-party tools (ServiceNow, VersionOne, Sumo)
• Work with teams to create SLI/SLO's .
• Actively monitor and lead troubleshooting of degraded performance and hard to define issues for the platform applications, develop the solution and document artifacts in the back log from root cause analysis.
• Evolve the cloud infrastructure ecosystem for our application suite by experimenting with emerging technologies and completing prototypes to understand benefit
• Design and develop CI/CD pipeline in AWS and Openshift to deploy various application artifacts, including APIs and Data Process Jobs.
• Analyze, design and develop the artifacts to configure the monitoring and alerting metrics so the support engineers can proactively and timely validate, troubleshoot and resolve the issues.
• Maintain data integrity and access control by using AWS security tools and services such as HSM, IAM, etc.
• Understand and develop tools to monitor AWS billing for the services, generate cost related reports and help develop and implement cost optimization strategies.
• Work with enterprise security architects to design and implement data security tools, measures, data encryption, key management; design and develop solutions to address the security vulnerabilities discovered by internal security audit team, as well as by the vendors, security community, etc.; design and develop solutions for support team to regularly scan and review to fix security issues
• Regularly and proactively monitor and analyze the capacity and performance of the platform, work with architecture team to design and implement elastic infrastructure to accommodate the irregular burst of user traffic/requests.
• Work with architecture team to develop backup strategy and implement the backup solution for critical data and application components for service restoration and disaster recovery purpose.
• Work with architecture, infrastructure, and application teams to provide input on continuous improvement on the design, performance and security enhancements. Desired Skillset:
• Deep understanding of the operations of AWS cloud platforms.
• Must be well versed in the automation, scripting, monitoring, including use of tools from the major cloud platforms, including but not limited to OpenShift Cloud Formation, Terraform, Ansible, Shell, Python
• Preferable for candidates with significant technical knowledge with infrastructure layers, including but not limited to: Linux OS, major virtualization platforms, Traditional and software defined network, Load Balancers, firewall, API tools, element/performance/intelligent monitoring tools, storage, backup strategy, etc.
• Significant knowledge and experience in end-to-end operations for enterprise systems and applications, including driving issue resolution for mission critical systems.
• Must have experience working to automate, operationalize and improve the Development/QA using CI/CD tools (Gitlab, Github, Jenkins, Maven, Gradle, Nexus)
• Working experience with Software Release Management. Desired Qualification
• BS degree in Computer Science or a related technical field or equivalent practical experience. Minimum Experience
• 3+ years of related DevOps, SysOps engineering experience with focus on major cloud platforms (AWS preferred).
• 2+ years of application development experience including data streaming, deploying/monitoring high availability critical application components.
• 1+ Years in Site Reliability Engineering organization preferred
• Overall 4-6years of experience
Responsibilities:
As a engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to a broad range of industry leading technologies that will help to drive acceleration. The ideal candidate will have expert design and development capabilities and be positioned to contribute to a growing set of services and features for the ecosystem. This role will be supporting highly available, business critical applications. This role will serve as the escalation point for complex and hard to define issues in both on premise and AWS environments. We are seeking talented engineers, well versed in DevOps technologies, automation, infrastructure orchestration, configuration management, continuous integration, troubleshooting of complex issues, who are not constrained by how "things are usually done".
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Atlanta, GA vacancy
- ...Technical Support Specialist In Site Reliability Engineering (Sre) Mandatory skills: Scripting and programming languages like Python, Java, Ruby. Cloud and infrastructure management – AWS, Google cloud and Azure is a plus- CI/CD Automation, Database Management. The...Suggested
$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...SuggestedWork experience placementWork at office- ...advances cures by helping the world's most important research sites do their best work. Our solutions are now used by over 30,00... ...What You'll Bring to the Team: We are seeking a Site Reliability Engineer (SRE) to join one of our Scrum teams and help ensure the...SuggestedWork at office
$60 - $68 per hour
...Site Reliability Engineer Immediate need for a talented Site Reliability Engineer. This is a 12+ months contract opportunity with long-term potential and is located in Atlanta, GA (Onsite). Please review the job description below and contact me ASAP if you are interested...SuggestedContract workLocal areaImmediate start$152.13k - $162.13k
...challenge the status-quo. Unum is changing, and we’re excited about what’s next. Join us. General Summary: Unum Group seeks Site Reliability Engineers in Atlanta, GA. Applicants who are interested in this position may apply at (Ref #66753) for consideration. Design,...SuggestedTemporary workWork at officeRemote work- ...Lead Engineer, Site Reliability Engineering Team As a lead engineer with Retail, Site Reliability Engineering team, you will be at the forefront of Cloud and Big Data technology. In this role you will establish yourself as a technical leader by exposing yourself to...
$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...Temporary workImmediate startFlexible hoursShift work- ...availability. • Automation Experience with Build/deployment, Software Configuration/Continuous Integration/Continuous Delivery/Release Engineering related tasks in JavaEE/C++ Environments. • Experience in automating manual processes using Python, Ruby, Unix Shell (bash,...Immediate start
$109.5k
...and YouTube. ( Job Description AbbVie Information Security is looking for a highly motivated, diligent, and skillful Site Reliability Engineer to join the Cyber Security Engineering (CSE) Team. The CSE Team, working within the Cyber Security Operations (CSO) function...Temporary workLocal areaRemote work$178.13k - $205.4k
...Bachelor's degree or foreign degree equivalent in Computer Engineering, Computer Science, Engineering, or related field plus five (5)... ...websites that are not Workday Careers. Please be aware of sites that may ask for you to input your data in connection with a job...Work at officeRemote workFlexible hours$165k - $241.4k
...Cisco Meraki, we are responsible for building and growing the cloud that supports these customers and their networks. As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze...Permanent employmentFull timeTemporary workPart timeLocal areaFlexible hours$126k - $248k
...As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and...Local areaRemote workWorldwideFlexible hours- ...in the right place. Job Details: Job Title: SRE Engineer Location: Atlanta GA (Hybrid Duration: 1 year Contract... ...critical application components. 1+ Years in Site Reliability Engineering organization preferred. Overall 4-6 years of...Contract workWork experience placementWork at officeRemote work
- ...enterprise initiatives such as public cloud, data science, AI, engineering innovation, and IoT. Our customers include the world's... ...is founder-led, profitable, and growing. We are hiring a Site Reliability Engineer Our goal is to perfect enterprise infrastructure DevOps...Work at officeLocal areaRemote workWork from homeWorldwide
- ...can create the conditions for educators to teach, students to thrive, and districts to shape the future of education. Site Reliability Engineer (SRE) Overview: We are looking for a Site Reliability Engineer (SRE) to join our Engineering team. This is a build-it...Full timeLive inWork at office
- ...evolving the foundational systems and practices that ensure the reliability, scalability, performance, and efficiency of our critical... ...highly resilient systems. Leverage deep expertise in software engineering, distributed systems, cloud infrastructure, and SRE principles...Early shift
$105k - $130k
...provide the high-speed capabilities our nation and its allies need to maintain a durable, asymmetric advantage. The Mission Systems Engineering (MSE) Team develops the Mission Management System (MMS)—a software platform that integrates mission subsystems, autonomy services...Weekly payPermanent employmentFull timeWork at office- ...We have an immediate need for a Senior Release Train Engineer for a contract assignment located in Carmel, Indiana . The Release Train Engineer (RTE) has a primary purpose of supporting an Agile Release Train (ART) by steering it to success and navigating the complexity...Contract workWork at officeImmediate start
- ...Job Description - Agile -Release Train Engineer (RTE) Duration: FULL TIME Location: Atlanta, GA Role and Responsibilities: Serve as the key facilitator for the Agile Release Train (ART) in Mobile/Web/Services IT projects. Collaborate...Full time
$101.5k - $169.1k
...Company Cox Automotive - USA Job Family Group Engineering / Product Development Job Profile Sr Release Train Engineer... ...organizational AI policies and standards. Monitor AI tool reliability across teams. Create backup plans for system failures....Work at officeRemote workVisa sponsorshipFlexible hoursShift work$140k
...their lives. About the Role: As a Stable Kernel Senior Software Engineer, you play an essential role in setting our portfolio of world-... ...as you learn new technologies. Your knowledgeable practice, reliability, and consultative nature make you an engineer that...Full timeContract workTemporary workVisa sponsorshipWork visaFlexible hours$117.8k - $212.5k
...year-round money coaches. That’s how we’re UNSTOPPABLE for our employees! Are you ready to join the Un-carrier movement? The Engineer, Forward Deployment is a hands-on technical practitioner who embeds within a T-Mobile business unit to design, build, and iterate...Full timeTemporary workPart timeWork experience placementLocal areaFlexible hours$124.6k - $148.2k
...serve them, and shape the consumer experience. Our product and engineering organizations bring together small, empowered teams that move... ...on connected service tools that improve resolution speed and reliability, data and API platforms that unlock growth and decisioning,...Full timeTemporary workLocal areaRelocation- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...
$142k - $214k
...applies practical experience with modern software technologies to redefine robotics using autonomy. We are looking for Full Stack Engineers to join our team. As a Full Stack Engineer at Anduril you will be architecting and building out user interface applications from...Full timeWork experience placementLocal areaRemote workRelocation package$143k - $191k
...computer vision, sensor fusion, and networking technology to the military in months, not years. ABOUT THE TEAM The Reliability Engineering team partners across Anduril's engineering, manufacturing, and operations organizations to ensure our autonomous systems survive...Full timeWork experience placementImmediate start- ...Saviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features. As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform....
$61 - $80 per hour
...Description Anduril is seeking a Reliability Engineer to drive product reliability across the full lifecycle of advanced autonomous defense systems-from early concept development through qualification, production, and field deployment. This individual will partner...Contract workTemporary work- ...Job Description Position Summary The Construction Project Engineer supports the Project Manager and project team in the planning,... ...monthly pay application reviews and progress verification Field & Site Support Participate in site visits, inspections, and progress...Contract workFor contractorsFor subcontractor
$106.1k - $176.8k
...this position We are seeking a DevOps Engineer (P3) to support and enhance cloud infrastructure... ...Operations teams to improve platform reliability, scalability, and automation. Support... ...candidate is expected to work on-site at our Las Colinas office a minimum of two...Full timeH1bWork at officeRemote work2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
Related searches
- site reliability engineer Atlanta, GA
- site reliability engineer sre Atlanta, GA
- construction site safety Atlanta, GA
- site recruiter Atlanta, GA
- on site coordinator Atlanta, GA
- website content developer Atlanta, GA
- website coordinator Atlanta, GA
- on-site clinical research associate (traveling/remote) Atlanta, GA
- site safety Atlanta, GA
- historic site Atlanta, GA




