Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Full-time

Vynca

Join the dynamic journey at Vynca, where we're passionate about transforming care for individuals with complex needs. We’re more than just a team; we're a close-knit community. Our shared commitment to caring for each other and those we serve is what sets us apart. Guided by our unwavering core values: Excellence, Compassion, Curiosity, and Integrity, we forge paths of success together. Join us in this transformative movement where you can contribute to making a profound difference every day. At Vynca, our mission is to provide comprehensive care for more quality days at home. About the job We're looking for a Site Reliability Engineer (E3) to help build and operate the infrastructure that powers Vynca's healthcare technology platform. In this role, you'll work at the intersection of software engineering, cloud infrastructure, and operations to ensure our systems are reliable, scalable, secure, and performant. As a member of the Technology team, you'll design and manage cloud infrastructure in AWS, operate Kubernetes-based workloads, improve observability across our platform, and automate operational processes that enable engineering teams to move quickly and safely. You'll play a critical role in maintaining the health of our production environment while helping shape the future architecture of our systems. This is a hands-on engineering role with significant ownership and impact. You'll partner closely with Software Engineers, Product teams, and Data teams to build resilient systems that support our mission of delivering comprehensive care for more quality days at home. This position is remote and requires working East Coast business hours (EST). What you'll do Design, provision, and manage AWS infrastructure using Terraform as the source of truth. Operate, maintain, and scale production workloads running on Kubernetes. Package, deploy, and manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity. Develop automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce manual effort and improve system resilience. Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions. Lead incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability. Partner with Product and Engineering teams on capacity planning, performance optimization, and resilient system design. Implement and maintain security best practices to support HIPAA, SOC 2, and other compliance requirements. Participate in an on-call rotation and provide operational support for production systems. Your experience and qualifications Experience: Three to five (3–5) years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, Cloud Infrastructure Engineering, or similar infrastructure-focused roles, preferably within healthcare, SaaS, or high-growth technology environments. Education: Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a related technical field; equivalent professional experience will also be considered. Strong hands-on experience operating production workloads within AWS environments. Proven experience managing infrastructure as code using Terraform, including module development, state management, and deployment automation. Experience operating and supporting production Kubernetes environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault tolerance. Experience establishing and managing observability practices including monitoring, logging, tracing, alerting, and incident response. Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals. Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation. Strong problem-solving skills with the ability to troubleshoot complex infrastructure and application issues. Excellent written and verbal communication skills with the ability to collaborate effectively across technical and non-technical teams. High level of ownership, accountability, and initiative with a proactive approach to reliability and operational excellence. Ability and willingness to participate in an on-call rotation supporting production systems. Preferred Qualifications Strong programming or scripting experience with Python, Go, or similar languages. Experience with observability platforms such as Prometheus, Grafana, Datadog, CloudWatch, SigNoz, or OpenTelemetry. Experience with GitOps tools such as ArgoCD or Flux. Experience managing databases such as PostgreSQL, MySQL, Redshift, or ClickHouse. Experience implementing secrets management solutions such as AWS Secrets Manager or HashiCorp Vault. Experience supporting healthcare technology platforms or other highly regulated environments. Familiarity with data infrastructure technologies including Snowflake, Redshift, and ETL/ELT pipelines. Experience with database performance tuning and optimization. At this time we are only considering applicants in the following states: Arizona, California, Colorado, Florida, Georgia, Illinois, Nevada, North Carolina, Oregon, Texas, Utah and Washington. Additional Information The hiring process for this role may consist of applying, followed by a phone screen, online assessment(s), interview(s), an offer, and background/reference checks. Background Screening: A background check, which may include a drug test or other health screenings depending on the role, will be required prior to employment. Job Description Scope: This job description is not exhaustive and may include additional activities, duties, and responsibilities not listed herein. Vaccination Requirement: Employees in patient, client, or customer-facing roles must be vaccinated against influenza. Requests for religious or medical accommodations will be considered but may not always be approved. Employment Eligibility: Compliance with federal law requires identity and work eligibility verification using E-Verify upon hire. Equal Opportunity Employer: At Vynca Inc., we embrace diversity and are committed to fostering an inclusive workplace. We value all applicants regardless of race, color, religion, age, national origin, ancestry, ethnicity, gender, gender identity, gender expression, sexual orientation, marital status, veteran status, disability, genetic information, citizenship status, or membership in any other protected group under federal, state, or local law.

Vacancy posted 6 hours ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in California vacancy
  • $71.6k - $119.4k

     ...support application teams. Our services provide applications with reliability, security, and better customer experiences. About the Job:...  ...automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You’ll gain exposure to a... 
    Suggested
    Full time
    Temporary work
    Internship
    Local area

    RELX

    California
    1 day ago
  • $90k - $180k

     ...medicines. Our 115,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: About the Role This Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division. We... 
    Suggested
    Full time
    Remote work
    Shift work

    Abbott

    Sunnyvale, CA
    3 days ago
  •  ...our Series B and have grown 800% over the last 12 months. Engineering at Ivo Engineers at Ivo are inventors. Ivo was first-to-...  ...still expect us to hit our SLAs. What? We’re looking for a Site level Reliability Engineer as part of Infrastructure team to: Own uptime,... 
    Suggested
    Full time
    Contract work
    Work at office
    Remote work
    Visa sponsorship
    Relocation package
    Flexible hours

    Ivo Inc.

    San Francisco, CA
    6 hours ago
  •  ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data... 
    Suggested

    Tik Tok

    Mountain View, CA
    20 hours ago
  •  ...company valued at $10 billion. We work in‑person five days a week in our new SanFrancisco headquarters. About the Role As a Site Reliability Engineer (SRE) at Mercor, you’ll own production reliability across our most critical systems, partnering directly with... 
    Suggested

    Mercor Inc

    San Francisco, CA
    1 day ago
  • $150k - $195k

     ...customers worldwide. Our team is growing, and we are looking for engineers with passion for automation. You will help support the...  ...alongside engineering/operations teams to improve the scalability and reliability of internal processes. Participate in an on‑call rotation.... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    13 hours ago
  • $187.04k - $359.72k

     ...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum...  ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas....  ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company... 
    Temporary work
    Local area
    Overseas
    Shift work

    Tik Tok

    San Jose, CA
    20 hours ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 

    Bolt Graphics, Inc.

    Sunnyvale, CA
    19 hours ago
  •  ..., Elise AI, IBM and Accern. Position Summary We are hiring for a hands‑on Head of SRE to establish, lead, and scale our Site Reliability Engineering function. This role combines strategic ownership with deep technical execution. You will be responsible for defining reliability... 
    Shift work

    Wand AI

    Palo Alto, CA
    19 hours ago
  • $100k - $200k

     ...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    2 days ago
  •  ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud... 
    Full time

    Saransh

    Sunnyvale, CA
    20 hours ago
  •  ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable... 

    Epic Games

    Sunnyvale, CA
    20 hours ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 

    TechDigital Group

    Santa Clara, CA
    19 hours ago
  •  ...CoreWeave seeks a Reliability Lead for Common Services in Sunnyvale, CA. You will establish and lead reliability engineering, production operations, and an observability-driven culture across multiple teams. Partner with engineering leaders to define SLOs/SLIs, drive... 

    CoreWeave

    Sunnyvale, CA
    19 hours ago
  • $210k - $240k

     ...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $24... 
    Full time

    Alembic Technologies

    San Francisco, CA
    1 day ago
  •  ...ActiveHours is looking for an experienced DevOps Engineer to enhance platform automation and site reliability in a collaborative environment. The role involves automating key systems, improving visibility through metrics, and troubleshooting critical problems. You'll... 

    ActiveHours

    Palo Alto, CA
    19 hours ago
  •  ...Acryl Data seeks a Site Reliability Engineering (SRE) Tech Lead to enhance the reliability and scalability of its DataHub platform. The role involves leading infrastructure design, optimizing system performance, and driving continuous improvement across cloud deployments... 

    Acryl Data

    Palo Alto, CA
    19 hours ago
  • $145k - $165k

     ...Your Ego : Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to... 
    Work at office
    Immediate start

    Bolt Graphics, Inc.

    Sunnyvale, CA
    2 days ago
  • $135.6k - $180k

     ...operations team. This role involves overseeing 24/7 operational stability, enhancing processes and systems, and mentoring a diverse engineering team. The ideal candidate will have over 8 years of technical operations experience, proficiency in infrastructure automation... 

    Ruckus

    Sunnyvale, CA
    19 hours ago
  •  ...exceptional professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of...  ...and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase... 

    J.P. Morgan

    Palo Alto, CA
    19 hours ago
  • $175k - $250k

     ...00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance of...  ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build... 
    Full time
    Remote work
    Relocation
    Relocation package

    The Recruiting Guy

    San Francisco, CA
    1 day ago
  •  ...JOB DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will help Shipt understand where we can improve stability and reliability. There will be a focus on the intersection of systems... 

    BayOne Solutions

    San Francisco, CA
    2 days ago
  • $260k - $300k

     ...agents. We're the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding...  ...than anyone expects. You will own both the production reliability of our user-facing products and the platform engineering that... 

    Cognition Corp

    San Francisco, CA
    13 hours ago
  • $148.5k - $223.9k

     ...you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1...  ...is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts... 
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    13 hours ago
  • $166k - $220k

     ...technology to the military in months, not years. ABOUT THE TEAM We are seeking a highly skilled and mission-driven Site Reliability Engineer (SRE) to join our Mission Autonomy team. In this critical role, you will be responsible for ensuring the reliability,... 
    Full time
    Work experience placement
    Immediate start
    Remote work

    Anduril Industries

    Costa Mesa, CA
    4 days ago
  •  ...SchoolsFirst FCU seeks an experienced Splunk Administrator/Site Reliability Engineer to deploy, optimize, and monitor our IT enterprise platforms. You will onboard data, craft SPL searches, and build dashboards powering IT applications, security operations, and observability... 

    SchoolsFirst Federal Credit Union

    Tustin, CA
    1 day ago
  •  ...globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You... 
    Temporary work
    Worldwide

    Air Apps

    San Francisco, CA
    2 days ago
  •  ...Arena Intelligence Engineer Arena Intelligence is looking for an engineer to build the core infrastructure that sits beneath our online...  ...foundational infrastructure for our users that scales, is reliable, and makes the complexities of operating this infrastructure at... 
    Permanent employment
    Shift work

    Arena AI

    San Francisco, CA
    4 days ago
  •  ...A tech startup in San Francisco is looking for Site Reliability Engineers to enhance system reliability and performance. Ideal candidates have over 5 years of relevant experience and strong expertise in cloud infrastructure, including AWS and Kubernetes. The role involves... 

    Breakout Tools

    San Francisco, CA
    20 hours ago
  • A leading technology firm is looking for a Manager to expand their Cloud Site Reliability team. The ideal candidate will have extensive Linux administration experience, a passion for automation, and be comfortable in a remote, diverse workplace. This position emphasizes... 
    Remote work

    mbhsobana

    San Francisco, CA
    20 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!