SRE
TriOptus LLC
Site Reliability Engineer (SRE)
As a part of the FRDC Site Reliability Engineer (SRE) team, you will help identify resilience challenges, build reusable, foundational software and infrastructure components to improve, influence, and validate the resilience and reliability for technologies that move trillions of dollars per day. Responsibilities include, but are not limited to:
- Participate in the design of build, refactor major software components that improve the availability, resilience, performance of our system
- Design, code, test, and deliver software to automate manual operational work
- Support incident responses, blameless postmortem, design and implement the product improvement to prevent incident reoccurring
- Implement application patterns in support of better service level objectives
- Implement self-healing, resiliency patterns
- Exercise failure cases regularly to validate resilience assumptions
- Engage with development teams throughout the life cycle of incident, ensure lessons learned are translated into automated or process adjust responses to help develop software for reliability and scale, ensuring minimal refactoring or changes
- Code, test and deliver software to automate manual operational work
- Troubleshoot incidents, participate in blameless post-incident evaluations and ensure permanent closure of incidents
- Identify application patterns and analytics in support of better service level objectives
- Analyze self-healing and resiliency patterns and contribute to software which can use these outcomes
- Implement best in class monitoring frameworks to accomplish end to end flow monitoring and noiseless alerting
Requirements & Qualifications:
- Bachelor’s degree or equivalent experience in a software engineering discipline
- 2+ years of hands-on software engineer experience
- Curious about solving resilience problems in run time at scale
- Expertise in at least one technology stack designing, coding, testing, and delivering software
- Knowledge in a few of infrastructure components (e.g. routers, load balancers, cloud products, container systems, compute, storage, and networks)
- Experience in cloud native, distributed application design and implementation
- Demonstrated communication and ownership skills
- Debugging and trouble shooting skills
- Collaboration with a diversified high-performing multi-location team
- Excellent analytical, interpersonal and communication skills
- Understanding of SRE methodologies/practices
Required Skills: Technical expertise of 4+ years the below areas, overall IT experience of 6+ years:
- Proficiency in Java / JVM based system design & implementation
- Infrastructure knowledge required including Unix, Windows, networking, and scripting (e.g. Perl / Python)
- Experience with orchestration tools like Jenkins CI/CD, or Jules
- Experience following source control best practices: Git/bitbucket
- Experience with database development (MySQL / Oracle)
- Understanding of architecture and design across distributed systems
Prefer Skills:
- Knowledge of SpringBoot / Microservices architecture
- Experience using Pivotal Cloud Foundry
- Experience with Public Cloud: AWS
- Enterprise platforms using Big Data tools and technologies (e.g. Hadoop, Spark, Hive, Impala, Dremio, Nifi, Ignite)
- Experience setting up & building solutions for Containers e.
Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the SRE in San Antonio, TX vacancy
- ...SLOs and engage with exception processes when technical limitations exist. Work with dependent process (Site Reliability Management, SRE, Event Management and Incident Management, CMDB) to create enhancement stories that will improve SLO efficacy and value when passing...Suggested
- Job Title Responsibilities: Develop, test, and debug automated tasks (Apps, Systems, Infrastructure) Troubleshoot minor incidents and contribute to resolution through post-mortems Participate in the application or service development lifecycle through code ...Suggested
- ...product and technology strategy into successful outcomes, as well as a strong working knowledge of public cloud platforms, CI/CD, and SRE practices, is essential. Experience in healthcare technology and EHR integrations, particularly with HL7, FHIR, and APIs, is a...Suggested
- ...sports web based software engineering gaming sports betting and analytics analytics engineering igaming casino sre and devops Business Classifications B2C Mobile About the Role The Company is seeking a Vice President for iGaming...Suggested
- ...experience; Terraform experience including writing custom modules and collaboration at scale; Ansible experience. Proficient in Linux; SRE experience for a mid-to-large enterprise system. Designing, implementing, and maintaining AWS cloud architecture; creating/...SuggestedFull timeTemporary workWork at officeImmediate startFlexible hours
$115k - $135k
...degree in Computer Science or related field or equivalent practical experience ~4+ years of experience in DevOps, Cloud Engineering, or SRE roles ~ Strong expertise in: ~ CI/CD tools (Azure DevOps, GitHub Actions, Jenkins) ~ Cloud platforms (Azure strongly preferred;...Full timeTemporary workWork at officeRemote work$77.12k - $147.39k
...technology domains. ~ Experience in an Enterprise Availability Command Center (ACC), Production Operations, Site Reliability Engineering (SRE), or comparable service restoration environment, with direct involvement in high-severity incident response and recovery activities....H1bWork at officeRemote workRelocation packageFlexible hours- ...enabling rapid delivery without compromising architecture integrity Engineering Collaboration & Enablement Partner with Engineering, SRE, Ops, and Data teams on roadmap execution and delivery practices Provide hands-on architectural guidance and mentorship to...Work at officeLocal areaRemote workWorldwide
- ...infrastructure using AWS / Azure services. Responsibilities Work across multiple teams to define and implement end-end DevOps /SRE strategy Expert in automation by writing Shell, Perl & Python scripts to monitor Production Applications Expertise in...Work experience placement
- ...what we do! What We Need: iHeartMedia Entertainment, Inc. seeks candidates for the position of Senior Site Reliability Engineer (SRE), responsible for leading a talented team of SREs/DevOps Engineers across a wide variety of Cloud Services to ensure the reliability,...Full timeWork at officeFlexible hours
$116.35k - $210.33k
...and continuous improvement, including complex troubleshooting, root-cause analysis, and implementation of corrective actions.Integrate SRE practices with DevSecOps and cybersecurity, embedding reliability, security, automated testing, and operational readiness into CI/CD...Full timeContract work- ...AWS/Azure/GCP) bility to interpret architecture diagrams, data flows, and system design Knowledge of DevOps, CI/CD, monitoring, SRE fundamentals Familiarity with data engineering concepts (ETL, data models, analytics). Product visioning, roadmap planning, and...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SRE. Be the first to apply!

