Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

SRE

Keylent Inc

Site Reliability Engineer

As a Site Reliability Engineer manages our production environment, providing a highly available and scalable platform for Ekata to serve our customers. The infrastructure team provides a resource for Engineering to help diagnose production issues and provide guidance on improving the availability and performance of our applications. This position also develops systems, automation, and tools to help make it easier for Engineering teams to deploy services in a fast, automated, and reliable fashion.

In this role you will:

  • Build, scale and support high-availability Ubuntu Linux production and development systems in a public cloud environment.
  • Work with tools such as Jenkins, Ansible, Argo CD, Terraform, CloudFormation, Resource Manager and many more to ensure that our stack is well represented as Infrastructure as Code.
  • Manage and improve security and availability monitoring for all services, ensure defined security policies are consistently implemented across all environments.
  • Deploy workloads to multiple cloud environments, proven experience with all of the core services within AWS, Azure or GCP, including instance management, IAM configuration, Database, Caching and general support/troubleshooting.
  • Have a developed understanding of the core components required to run Kubernetes and be able to build a cluster from scratch if needed.
  • Have perfected the fundamentals of load balancing, service mesh and always looking for ways to improve availability and uptime.
  • Maintain quality documentation for systems owned by the Infrastructure team.
  • Use monitoring tools to identify and resolve issues before they happen. Have familiarity with Prometheus.
  • Help other teams troubleshoot and solve failures and performance problems, participate in on-call rotations.
  • Have a passion for working with Go, Python, Rust or even Bash to build custom tools and improve system integration. Take code ownership to the next level and act as an advocate for writing code that aligns with industry best practice.
  • Have a solid grasp on networking fundamentals and can easily explain how DNS, DHCP and routing work in most environments.

All About You:

  • Excellent spoken and written English skills. Is a team player and values collaboration.
  • BS degree in Computer Science or equivalent experience.
  • Proven skills with Linux or UNIX systems and related protocols/software with 3+ years' experience.
  • A command of Linux systems including troubleshooting, memory management, tuning, I/O subsystem, RAID, and security.
  • Experience with provisioning tools such as Ansible/Chef/Terraform.
  • Experience with Jenkins or other CI/CD tools.
  • Programming aptitude in Go, Python, and Bash.
  • Working knowledge of database systems such as MySQL or PostgreSQL.
  • Experience building and deploying Containers, including orchestration tools such as Kubernetes, Mesos, or Docker Swarm.
  • Experience with cloud providers (AWS, Azure, GCP)
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the SRE in Saint Louis, MO vacancy
  • London Stock Exchange Group (LSEG) invites applications for a Manager-level Platform Engineering role focused on the LOCP platform. You will lead the design, development, and evolution of enterprise-grade platform services on a cutting-edge container platform, with emphasis...
    Suggested

    Jobleads-US

    Saint Louis, MO
    2 days ago
  •  ...United States is seeking a Senior Platform Engineer to design, build and operate LOCP platform capabilities at scale. You will apply SRE practices, IaC, and observability to deliver reliable, high-performance infrastructure for real-time systems and cloud-native deployments... 
    Suggested

    Jobleads-US

    Saint Louis, MO
    2 days ago
  •  ...infrastructure and platform capabilities that operate at extraordinary scale. Working alongside engineering, product, architecture, and SRE specialists, you will contribute to the engineering excellence that underpins some of the world's most demanding technology... 
    Suggested
    Worldwide

    Jobleads-US

    Saint Louis, MO
    2 days ago
  •  ...ceremonies, including standups, retrospectives, and planning sessions, to drive team efficiency. Collaborate closely with Operations, SRE, and Platform teams to diagnose and resolve production issues swiftly, ensuring minimal disruption to users and maintaining system... 
    Suggested
    Temporary work
    Part time
    Work at office
    Flexible hours
    3 days per week

    Jobleads-US

    Saint Louis, MO
    1 day ago
  •  ...with strategic objectives and engineering standards. Reliability & Operational Excellence Champion Site Reliability Engineering (SRE) principles across the platform, driving improvements in reliability, availability, resilience, and scalability. Implement and improve... 
    Suggested
    Worldwide

    Jobleads-US

    Saint Louis, MO
    2 days ago
  •  ...model after launch. Partner with Engineering to strengthen the connection between delivery governance and CI/CD, automated testing, SRE, observability, quality, and operating support. Ensure maintenance, defects, technical debt, platform work, and reliability needs... 

    Jobleads-US

    Saint Louis, MO
    3 days ago
  •  ...and secure releases. Champion Infrastructure-as-Code (Terraform, Ansible, etc.) practices. Reliability & Observability: Establish SRE best practices to drive uptime, system resilience, and performance. Create and maintain metrics and reporting performance dashboards,... 
    Work at office

    Alberici Group, LLC

    Saint Louis, MO
    27 days ago
  • $70 - $80 per hour

     ...scanning, GitOps, Helm, or Kustomize. Background in software engineering or system design, with progression into cloud, platform, or SRE-focused roles. Prior experience in smaller or midsize organizations, lean cloud teams, central platform teams, or consulting... 
    Contract work
    Immediate start

    Mondo

    Saint Louis, MO
    21 days ago
  •  ...Required: Bachelor’s degree in IT or related field, or equivalent experience.Minimum Required: 5+ years' of experience in IT Operations, SRE, or similar roles.Licenses & CredentialsMinimum Required: None.Systems & TechnologyDemonstrated experience with observability... 

    Stifel Financial

    Saint Louis, MO
    3 days ago
  •  ...with hybrid cloud, multi-cloud, and edge networking architectures. Minimum 2+ years of experience working within DevOps, NetOps, SRE, and platform engineering operating models. Minimum 4+ years of experience integrating network infrastructure with cloud, security... 
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Saint Louis, MO
    5 days ago
  •  ...interaction for both internal and external partners by communicating effectively with all key stakeholders. Ultimately, the role of SRE is to align Product and Customer Focused priorities with Operational needs. We regularly review our run state not only from an... 
    Long term contract
    Shift work
    3 days per week

    Futran Tech Solutions Pvt. Ltd.

    Saint Louis, MO
    4 days ago
  • $64 - $74 per hour

     ...clear "why" and "why now." Facilitate collaboration across teams, coordinating dependencies and rollout plans with functional teams (UX, SRE, Analytics, etc.). Lead delivery and iteration cycles: run weekly demos, track progress toward value, document learnings, and shape... 
    Hourly pay
    Full time
    Contract work
    Temporary work
    Work experience placement
    Remote work

    Jobs via Dice

    Saint Louis, MO
    2 days ago
  •  ...regulatory requirements * Conduct security assessments and vulnerability management Leadership & Collaboration * Mentor junior SRE team members and promote SRE culture across the organization * Partner with software engineering teams to improve system... 
    Full time
    Part time

    Federal Reserve Bank of San Francisco

    Saint Louis, MO
    2 days ago
  •  ...reliability into systems. Build proprietary tools to mitigate weaknesses in incident management or software delivery. Implement SRE best practices to increase system reliability and performance. Automate processes for improved collaborative response and... 
    Local area
    Immediate start

    Momento USA

    Saint Louis, MO
    2 days ago
  •  ...guidance, SLO practices, and platform workflows. What You will Bring: Experience supporting production platforms or services in an SRE, platform engineering, infrastructure, DevOps, or operations engineering role. Experience using observability data such as... 
    Worldwide

    LSEG (London Stock Exchange Group)

    Creve Coeur, MO
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to SRE. Be the first to apply!