Site Reliability Engineer
Vynca
Join the dynamic journey at Vynca, where we're passionate about transforming care for individuals with complex needs. We’re more than just a team; we're a close-knit community. Our shared commitment to caring for each other and those we serve is what sets us apart. Guided by our unwavering core values: Excellence, Compassion, Curiosity, and Integrity, we forge paths of success together. Join us in this transformative movement where you can contribute to making a profound difference every day. At Vynca, our mission is to provide comprehensive care for more quality days at home. About the job We're looking for a Site Reliability Engineer (E3) to help build and operate the infrastructure that powers Vynca's healthcare technology platform. In this role, you'll work at the intersection of software engineering, cloud infrastructure, and operations to ensure our systems are reliable, scalable, secure, and performant. As a member of the Technology team, you'll design and manage cloud infrastructure in AWS, operate Kubernetes-based workloads, improve observability across our platform, and automate operational processes that enable engineering teams to move quickly and safely. You'll play a critical role in maintaining the health of our production environment while helping shape the future architecture of our systems. This is a hands-on engineering role with significant ownership and impact. You'll partner closely with Software Engineers, Product teams, and Data teams to build resilient systems that support our mission of delivering comprehensive care for more quality days at home. This position is remote and requires working East Coast business hours (EST). What you'll do Design, provision, and manage AWS infrastructure using Terraform as the source of truth. Operate, maintain, and scale production workloads running on Kubernetes. Package, deploy, and manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity. Develop automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce manual effort and improve system resilience. Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions. Lead incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability. Partner with Product and Engineering teams on capacity planning, performance optimization, and resilient system design. Implement and maintain security best practices to support HIPAA, SOC 2, and other compliance requirements. Participate in an on-call rotation and provide operational support for production systems. Your experience and qualifications Experience: Three to five (3–5) years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, Cloud Infrastructure Engineering, or similar infrastructure-focused roles, preferably within healthcare, SaaS, or high-growth technology environments. Education: Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a related technical field; equivalent professional experience will also be considered. Strong hands-on experience operating production workloads within AWS environments. Proven experience managing infrastructure as code using Terraform, including module development, state management, and deployment automation. Experience operating and supporting production Kubernetes environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault tolerance. Experience establishing and managing observability practices including monitoring, logging, tracing, alerting, and incident response. Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals. Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation. Strong problem-solving skills with the ability to troubleshoot complex infrastructure and application issues. Excellent written and verbal communication skills with the ability to collaborate effectively across technical and non-technical teams. High level of ownership, accountability, and initiative with a proactive approach to reliability and operational excellence. Ability and willingness to participate in an on-call rotation supporting production systems. Preferred Qualifications Strong programming or scripting experience with Python, Go, or similar languages. Experience with observability platforms such as Prometheus, Grafana, Datadog, CloudWatch, SigNoz, or OpenTelemetry. Experience with GitOps tools such as ArgoCD or Flux. Experience managing databases such as PostgreSQL, MySQL, Redshift, or ClickHouse. Experience implementing secrets management solutions such as AWS Secrets Manager or HashiCorp Vault. Experience supporting healthcare technology platforms or other highly regulated environments. Familiarity with data infrastructure technologies including Snowflake, Redshift, and ETL/ELT pipelines. Experience with database performance tuning and optimization. At this time we are only considering applicants in the following states: Arizona, California, Colorado, Florida, Georgia, Illinois, Nevada, North Carolina, Oregon, Texas, Utah and Washington. Additional Information The hiring process for this role may consist of applying, followed by a phone screen, online assessment(s), interview(s), an offer, and background/reference checks. Background Screening: A background check, which may include a drug test or other health screenings depending on the role, will be required prior to employment. Job Description Scope: This job description is not exhaustive and may include additional activities, duties, and responsibilities not listed herein. Vaccination Requirement: Employees in patient, client, or customer-facing roles must be vaccinated against influenza. Requests for religious or medical accommodations will be considered but may not always be approved. Employment Eligibility: Compliance with federal law requires identity and work eligibility verification using E-Verify upon hire. Equal Opportunity Employer: At Vynca Inc., we embrace diversity and are committed to fostering an inclusive workplace. We value all applicants regardless of race, color, religion, age, national origin, ancestry, ethnicity, gender, gender identity, gender expression, sexual orientation, marital status, veteran status, disability, genetic information, citizenship status, or membership in any other protected group under federal, state, or local law.
$71.6k - $119.4k
...support application teams. Our services provide applications with reliability, security, and better customer experiences. About the Job:... ...automation, troubleshoot issues, and work closely with senior engineers to learn and apply best practices. You’ll gain exposure to a...SuggestedFull timeTemporary workInternshipLocal area$90k - $180k
...medicines. Our 115,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: About the Role This Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division. We...SuggestedFull timeRemote workShift work- ...our Series B and have grown 800% over the last 12 months. Engineering at Ivo Engineers at Ivo are inventors. Ivo was first-to-... ...still expect us to hit our SLAs. What? We’re looking for a Site level Reliability Engineer as part of Infrastructure team to: Own uptime,...SuggestedFull timeContract workWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours
- ...Site Reliability Engineer, Data Platform - USDS Responsibilities Engage in and improve the whole lifecycle of service, from inception and design, through to deployment, operation and refinement. Ensure reliable, fault-tolerant, efficiently scalable and cost-effective data...Suggested
- ...company valued at $10 billion. We work in‑person five days a week in our new SanFrancisco headquarters. About the Role As a Site Reliability Engineer (SRE) at Mercor, you’ll own production reliability across our most critical systems, partnering directly with...Suggested
$150k - $195k
...customers worldwide. Our team is growing, and we are looking for engineers with passion for automation. You will help support the... ...alongside engineering/operations teams to improve the scalability and reliability of internal processes. Participate in an on‑call rotation....Full timeWorldwide$187.04k - $359.72k
...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum... ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas.... ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company...Temporary workLocal areaOverseasShift work$145k - $165k
...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key...- ..., Elise AI, IBM and Accern. Position Summary We are hiring for a hands‑on Head of SRE to establish, lead, and scale our Site Reliability Engineering function. This role combines strategic ownership with deep technical execution. You will be responsible for defining reliability...Shift work
$100k - $200k
...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about...Full time- ...Overview Title: Site Reliability Engineer SRE – ML platform Location: Austin, TX or Sunnyvale, CA Employment type: Full-time • Seniority: Mid-Senior level • ONLY W2 Responsibilities Continuous Deployment using GitHub Actions, Flux, Kustomize Design and implement cloud...Full time
- ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable...
- ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless...
- ...CoreWeave seeks a Reliability Lead for Common Services in Sunnyvale, CA. You will establish and lead reliability engineering, production operations, and an observability-driven culture across multiple teams. Partner with engineering leaders to define SLOs/SLIs, drive...
$210k - $240k
...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $24...Full time- ...ActiveHours is looking for an experienced DevOps Engineer to enhance platform automation and site reliability in a collaborative environment. The role involves automating key systems, improving visibility through metrics, and troubleshooting critical problems. You'll...
- ...Acryl Data seeks a Site Reliability Engineering (SRE) Tech Lead to enhance the reliability and scalability of its DataHub platform. The role involves leading infrastructure design, optimizing system performance, and driving continuous improvement across cloud deployments...
$145k - $165k
...Your Ego : Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to...Work at officeImmediate start$135.6k - $180k
...operations team. This role involves overseeing 24/7 operational stability, enhancing processes and systems, and mentoring a diverse engineering team. The ideal candidate will have over 8 years of technical operations experience, proficiency in infrastructure automation...- ...exceptional professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...and position yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at JPMorgan Chase...
$175k - $250k
...00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance of... ...while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build...Full timeRemote workRelocationRelocation package- ...JOB DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will help Shipt understand where we can improve stability and reliability. There will be a focus on the intersection of systems...
$260k - $300k
...agents. We're the makers of Devin, the first AI software engineer. Our team is extremely talent-dense. Among our founding... ...than anyone expects. You will own both the production reliability of our user-facing products and the platform engineering that...$148.5k - $223.9k
...you are not duplicating efforts. Job Category Software Engineering Job Details About Salesforce Salesforce is the #1... ...is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts...WorldwideWeekend work$166k - $220k
...technology to the military in months, not years. ABOUT THE TEAM We are seeking a highly skilled and mission-driven Site Reliability Engineer (SRE) to join our Mission Autonomy team. In this critical role, you will be responsible for ensuring the reliability,...Full timeWork experience placementImmediate startRemote work- ...SchoolsFirst FCU seeks an experienced Splunk Administrator/Site Reliability Engineer to deploy, optimize, and monitor our IT enterprise platforms. You will onboard data, craft SPL searches, and build dashboards powering IT applications, security operations, and observability...
- ...globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You...Temporary workWorldwide
- ...Arena Intelligence Engineer Arena Intelligence is looking for an engineer to build the core infrastructure that sits beneath our online... ...foundational infrastructure for our users that scales, is reliable, and makes the complexities of operating this infrastructure at...Permanent employmentShift work
- ...A tech startup in San Francisco is looking for Site Reliability Engineers to enhance system reliability and performance. Ideal candidates have over 5 years of relevant experience and strong expertise in cloud infrastructure, including AWS and Kubernetes. The role involves...
- A leading technology firm is looking for a Manager to expand their Cloud Site Reliability team. The ideal candidate will have extensive Linux administration experience, a passion for automation, and be comfortable in a remote, diverse workplace. This position emphasizes...Remote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- construction site safety California
- website content developer California
- on-site clinical research associate (traveling/remote) California
- site safety California
- historic site California
- IT site lead California
- site leader California
- junior website developer California
- official site California
- site services specialist California

