SRE
Yochana
Required Qualifications:
• 8+ years of Software Engineering experience
• 4+ years of experience in Site Reliability Engineering teams with continued focus on improving Platform health
• Familiar with Agile or other rapid application development practices
• Hands-on expertise in building dashboards using APM tools.
• Experience with distributed (multi-tiered) systems, algorithms, relational databases, and NoSQL databases.
• Knowledge & Exposure caching tools (Redis, memcache) or messaging tools such as MQ, Kafka.
• Must have working knowledge of APM tools such as splunk, GCL, ELK, Grafana, Prometheus etc.
• Able to create Dashboards using GCL/Splunk/ELK and setup alerts.
• Working knowledge of CICD is a plus - Source control like Git, Continuous Integration - Jenkins / UCD Release etc. .
• Ability to work with Engineering teams across the ecosystem such as Security, Networking & Infrastructure challenges which can impact platform health & resiliency.
• Shell Scripting / DevOps tools like Ansible with good knowledge of yaml file to write playbooks .
• Experience with distributed storage technologies like NFS as well as dynamic resource management frameworks PCF, Kubernetes / OpenShift, AWS or Azure.
• Tech Stack: Java/J2EE (Spring, Spring Boot, Python, Shell Scripting, Kafka, Oracle, MongoDB etc.).
• A proactive approach to spotting problems, areas for improvement, and performance bottlenecks.
• 8+ years of Software Engineering experience
• 4+ years of experience in Site Reliability Engineering teams with continued focus on improving Platform health
• Familiar with Agile or other rapid application development practices
• Hands-on expertise in building dashboards using APM tools.
• Experience with distributed (multi-tiered) systems, algorithms, relational databases, and NoSQL databases.
• Knowledge & Exposure caching tools (Redis, memcache) or messaging tools such as MQ, Kafka.
• Must have working knowledge of APM tools such as splunk, GCL, ELK, Grafana, Prometheus etc.
• Able to create Dashboards using GCL/Splunk/ELK and setup alerts.
• Working knowledge of CICD is a plus - Source control like Git, Continuous Integration - Jenkins / UCD Release etc. .
• Ability to work with Engineering teams across the ecosystem such as Security, Networking & Infrastructure challenges which can impact platform health & resiliency.
• Shell Scripting / DevOps tools like Ansible with good knowledge of yaml file to write playbooks .
• Experience with distributed storage technologies like NFS as well as dynamic resource management frameworks PCF, Kubernetes / OpenShift, AWS or Azure.
• Tech Stack: Java/J2EE (Spring, Spring Boot, Python, Shell Scripting, Kafka, Oracle, MongoDB etc.).
• A proactive approach to spotting problems, areas for improvement, and performance bottlenecks.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the SRE in Columbus, OH vacancy
- ...Software Reliability Engineer Location: Columbus, OH Role: SRE Rate: 70 Required Qualifications: ~8+ years of software engineering experience ~4+ years of experience in site reliability engineering teams with continued focus on improving platform health...Suggested
- Primary Skill PCF (Pivotal Cloud Foundry) and Mongo DB Exposure to at least 1 Observability Tool such as AppDynamics, Splunk, Grafana Change Mgmt using CI/CD pipeline. Harness or equivalent tools Secondary Skill SSL Certificate management ...Suggested
- ...qualifications, capabilities, and skills Formal training or certification on software development concepts and 5+ years applied experience, SRE/DevOps, platform engineer, or similar Proficiency in operating and managing cloud-based services using IaC or EaC (infrastructure...Suggested
$91.2k - $136.8k
...Computer Science, Engineering, or a related field. 3+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps. Hands-on experience with observability tools: Splunk, Dynatrace, CloudWatch. Deep knowledge of Infrastructure as Code (...SuggestedFull timeTemporary workWork at office3 days per week$121.4k - $218.6k
...deliver customized solutions for Public Sector customers. This role partners closely with Operations and Site Reliability Engineering (SRE) teams in a highly collaborative environment to deploy, operate, and continuously improve secure, resilient, and high-performance...SuggestedWork experience placementWork at office- ...other regulated environments Experience with tools such as Datadog, Splunk, Dynatrace, AppDynamics, or similar platforms Knowledge of SRE and observability practices Familiarity with automation and integration approaches, including APIs and infrastructure‑as‑code...
$180k - $250k
...Collaboration with multiple layers of contacts within client organizations, including but not limited to CIO, CTO, Platform Engineering, SRE, Developer teams, and procurement to strengthen our overall customer relationship and better understand the goals and objectives they...Work at officeRemote workWorldwideFlexible hours- ...building AI/ML driven applications to drive business goals. Track record of Continuous learning mindset Proficiency in Agile and SRE culture About Us Chase is a leading financial services firm, helping nearly half of America’s households and small businesses...
- ...business objectives. The DevSecOps team is a highly engaged team focused on DevSecOps, Automated Testing, Site Reliability Engineering (SRE), building Self Service Web Portal for Solution Teams and passionate about enabling our mission. We expect all our engineers to be...Remote jobContract workWork experience placementFlexible hours
- ...management dependencies, backup/restore, DR, network paths, application I/O profiles ) and advises on mitigation actions Supports SRE teams as needed through technical consultation and escalation support for complex incidents and problem management, driving long-term...
- ...qualifications, capabilities, and skills A proactive approach to spotting problems, areas for improvement, and performance bottlenecks SRE mindset Culture/Approaches: To run better production systems by creating engineering solutions to operational problems. About...
$120k - $135k
...manages all engineering projects, concerning all the machines, utilities facilities and buildings. Is the Site Responsible Engineer (SRE) and has overall responsibility for site maintenance and engineering activity, to include all safety, quality, continuous process...Permanent employmentContract workWork experience placementFor subcontractorLocal area- ...positive impact on people's lives. Position Summary Reporting to the Director of Engineering, the Lead Site Reliability Engineer (SRE) is a senior technical contributor responsible for building reliable, scalable software systems and the DevOps practices that support...Full timeRemote workMonday to FridayShift workNight shiftWeekend work
$75.7k - $136.3k
...team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and services. Our SRE teams solve reliability, security, and usability at scale for our global fleet while maintaining Akamai's mission at the forefront of...Work experience placementWork at office- ...Site Reliability Engineer (SRE) Level II As a Site Reliability Engineer (SRE) Level II, you will play a key role in maintaining the availability, scalability, and performance of critical infrastructure and services. You will be responsible for building and automating...Work experience placementWork at office
$124k - $186k
...it needs verification particularly for infrastructure code and production-impacting changes. An interest in applying AI tooling to SRE problems: incident triage, runbook generation, log analysis, Terraform authoring, and similar. Ways of Working You...Work at officeLocal areaImmediate start2 days per week$51.9 per hour
...automation, continuous improvement, collaboration, and patient safety. Develops core metrics for monitoring and maintaining system health for SRE practitioners (e.g., latency, traffic, errors, and saturation) leveraging industry practices, manufacturer guidance, and other...For contractorsLocal area$95k - $171k
...Are you passionate about cutting-edge AI infrastructure? Do you want to build your SRE career on one of the most exciting platforms in cloud computing? Join the Akamai Inference Cloud Team The Akamai Inference Cloud team is part of Akamai's Cloud Technology Group...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$150k - $165k
...workflows, and improving the developer experience through internal tooling and infrastructure standards. As Immuta continues to grow, the SRE team plays a key role in ensuring platform stability, performance, and compliance for enterprise-scale environments. YOUR ROLE...Full timeWork at officeLocal areaWorldwide3 days per week$84.9k - $209.5k
...in an open, diverse, and productive environment Responsibilities What You'll Do Service Ownership –You will be part of the SRE team, whose mission is the shared full stack ownership of a collection of services and/or technology areas, with our Development partners...Temporary workImmediate startFlexible hours- ...certification on site reliability engineering concepts and 5+ years applied experience Proficient in site reliability engineering (SRE) culture and principles, with experience implementing SRE practices within applications and platforms; strong observability...
- ...Lead SRE Job Code: SRES04 • Grade: 603 Assume a critical role in defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer...Local area
$124k - $155k
...open to applicants authorized to work for any employer within the United States. What you will be doing: Lead a 6 member SRE team supporting production infrastructure and services Manage backlog, sprint planning, and team velocity Own reliability,...Seasonal workRemote work$159.6k - $287.4k
...balanced against factors including staffing, reliability risks, cost, and schedule. Partnering with Architecture, Engineering, and SRE leaders to define the strategy, vision, and development roadmap for the product line. Documenting platform service usage while...Work experience placementWork at office- ...qualifications, capabilities, and skills Formal training or certification on software development concepts and 5+ years applied experience, SRE/DevOps, platform engineer, or similar Proficiency in operating and managing cloud-based services using IaC or EaC (infrastructure...
- ...triage, perform RCA, and deliver preventative engineering and resilience improvements. Partner with infrastructure, application, and SRE teams to align platform capabilities to SLIs/SLOs, operational readiness, and continuous improvement goals. Contribute to a...
$94.9k - $135.6k
...across teams, platforms, environments, and business priorities. Coordinate with Solution Owners, Scrum Masters, Engineering, Testing, SRE, and Operations to align scope, sequencing, dependencies, and deployment timelines. Partner closely with the Cutover Lead to align...Full timeTemporary workLocal areaImmediate startFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!

