Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

Full-time

Vynca

Join the dynamic journey at Vynca, where we're passionate about transforming care for individuals with complex needs. We’re more than just a team; we're a close-knit community. Our shared commitment to caring for each other and those we serve is what sets us apart. Guided by our unwavering core values: Excellence, Compassion, Curiosity, and Integrity, we forge paths of success together. Join us in this transformative movement where you can contribute to making a profound difference every day. At Vynca, our mission is to provide comprehensive care for more quality days at home. About the job We're looking for a Site Reliability Engineer (E3) to help build and operate the infrastructure that powers Vynca's healthcare technology platform. In this role, you'll work at the intersection of software engineering, cloud infrastructure, and operations to ensure our systems are reliable, scalable, secure, and performant. As a member of the Technology team, you'll design and manage cloud infrastructure in AWS, operate Kubernetes-based workloads, improve observability across our platform, and automate operational processes that enable engineering teams to move quickly and safely. You'll play a critical role in maintaining the health of our production environment while helping shape the future architecture of our systems. This is a hands-on engineering role with significant ownership and impact. You'll partner closely with Software Engineers, Product teams, and Data teams to build resilient systems that support our mission of delivering comprehensive care for more quality days at home. This position is remote and requires working East Coast business hours (EST). What you'll do Design, provision, and manage AWS infrastructure using Terraform as the source of truth. Operate, maintain, and scale production workloads running on Kubernetes. Package, deploy, and manage applications using Helm and infrastructure automation tools. Build, operate, and improve distributed and event-driven systems, including event sourcing, partitioning, event ordering, replay, and failure recovery mechanisms. Define, monitor, and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability and engineering velocity. Develop automation for deployment, scaling, monitoring, incident response, and operational workflows to reduce manual effort and improve system resilience. Own platform observability by implementing and maintaining metrics, logging, tracing, monitoring, and alerting solutions. Lead incident response efforts, facilitate blameless postmortems, and drive long-term corrective actions that improve system reliability. Partner with Product and Engineering teams on capacity planning, performance optimization, and resilient system design. Implement and maintain security best practices to support HIPAA, SOC 2, and other compliance requirements. Participate in an on-call rotation and provide operational support for production systems. Your experience and qualifications Experience: Three to five (3–5) years of experience in Site Reliability Engineering, DevOps Engineering, Platform Engineering, Cloud Infrastructure Engineering, or similar infrastructure-focused roles, preferably within healthcare, SaaS, or high-growth technology environments. Education: Bachelor's degree in Computer Science, Information Systems, Software Engineering, or a related technical field; equivalent professional experience will also be considered. Strong hands-on experience operating production workloads within AWS environments. Proven experience managing infrastructure as code using Terraform, including module development, state management, and deployment automation. Experience operating and supporting production Kubernetes environments. Hands-on experience deploying and managing applications using Helm. Experience working with distributed systems, event-driven architectures, or event-sourcing platforms, including concepts such as partitioning, event ordering, replay, and fault tolerance. Experience establishing and managing observability practices including monitoring, logging, tracing, alerting, and incident response. Strong understanding of Linux systems administration, networking, cloud architecture, and distributed systems fundamentals. Experience designing, implementing, and maintaining CI/CD pipelines and deployment automation. Strong problem-solving skills with the ability to troubleshoot complex infrastructure and application issues. Excellent written and verbal communication skills with the ability to collaborate effectively across technical and non-technical teams. High level of ownership, accountability, and initiative with a proactive approach to reliability and operational excellence. Ability and willingness to participate in an on-call rotation supporting production systems. Preferred Qualifications Strong programming or scripting experience with Python, Go, or similar languages. Experience with observability platforms such as Prometheus, Grafana, Datadog, CloudWatch, SigNoz, or OpenTelemetry. Experience with GitOps tools such as ArgoCD or Flux. Experience managing databases such as PostgreSQL, MySQL, Redshift, or ClickHouse. Experience implementing secrets management solutions such as AWS Secrets Manager or HashiCorp Vault. Experience supporting healthcare technology platforms or other highly regulated environments. Familiarity with data infrastructure technologies including Snowflake, Redshift, and ETL/ELT pipelines. Experience with database performance tuning and optimization. At this time we are only considering applicants in the following states: Arizona, California, Colorado, Florida, Georgia, Illinois, Nevada, North Carolina, Oregon, Texas, Utah and Washington. Additional Information The hiring process for this role may consist of applying, followed by a phone screen, online assessment(s), interview(s), an offer, and background/reference checks. Background Screening: A background check, which may include a drug test or other health screenings depending on the role, will be required prior to employment. Job Description Scope: This job description is not exhaustive and may include additional activities, duties, and responsibilities not listed herein. Vaccination Requirement: Employees in patient, client, or customer-facing roles must be vaccinated against influenza. Requests for religious or medical accommodations will be considered but may not always be approved. Employment Eligibility: Compliance with federal law requires identity and work eligibility verification using E-Verify upon hire. Equal Opportunity Employer: At Vynca Inc., we embrace diversity and are committed to fostering an inclusive workplace. We value all applicants regardless of race, color, religion, age, national origin, ancestry, ethnicity, gender, gender identity, gender expression, sexual orientation, marital status, veteran status, disability, genetic information, citizenship status, or membership in any other protected group under federal, state, or local law.

Vacancy posted 17 hours ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Texas vacancy
  •  ...dynamic team with a broad knowledge of how Oracle’s cloud platform works. You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers. Note - this role is not a Monday to Friday core hours... 
    Suggested
    Full time
    Monday to Friday
    Flexible hours
    Shift work
    Night shift

    Oracle

    Austin, TX
    3 days ago
  •  ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team, you hold a leadership role in your team,... 
    Suggested
    Full time
    Local area

    JPMorgan Chase & Co.

    Plano, TX
    17 hours ago
  •  ...Administrator / SRE in Dallas to own production Java environments, middleware, and cloud automation. You will optimize performance, drive reliability, and mentor teammates while aligning with enterprise security and AI-enabled integrations. You will work across Java apps, IBM... 
    Suggested

    Motion Recruitment Partners LLC

    Dallas, TX
    1 day ago
  • $127k - $249k

     ...MongoDB, Inc. is seeking an experienced Senior or Staff Engineer for their SRE, InfraSec team, responsible for guiding the security of cloud-based infrastructure. The role involves hands-on technical work and mentorship of a small team while collaborating with engineering... 
    Suggested
    Remote work
    Flexible hours

    RTL2 Fernsehen GmbH & Co. KG

    Austin, TX
    1 day ago
  •  ...Electric Reliability Council of Texas seeks a data reliability and platform stewardship lead to ensure reliability, availability, and efficiency of ERCOT’s data platforms across cloud and on‑prem environments. You will champion incident response, automation, monitoring... 
    Suggested

    Electric Reliability Council of Texas Inc

    Taylor, TX
    1 day ago
  •  ...The Depository Trust & Clearing Corporation (DTCC) is seeking a Senior Application Support Engineer (SRE) to enhance reliability, scalability, and performance of mission-critical applications. You will apply SRE principles across engineering, infrastructure, and operations... 

    The Depository Trust & Clearing Corporation

    Dallas, TX
    1 day ago
  •  ...METRIX IT SOLUTIONS INC is seeking a senior database engineer to design, deploy, and manage multi-region CockroachDB clusters in production. The role focuses on high availability, data consistency, and scalable capacity planning for global deployments. You will monitor... 

    METRIX IT SOLUTIONS INC

    Austin, TX
    1 day ago
  •  ...Caterpillar is seeking a Digital Technical Support Analyst to ensure platform stability and cloud service reliability. You will own incidents end-to-end, coordinate across engineering and product teams, and drive improvements in runbooks, monitoring, and incident response. The... 

    Caterpillar Brazil

    Irving, TX
    1 day ago
  • $174.35k - $210k

     ...Site Reliability Engineer, IBM Corporation, Austin, TX (Up to 80% telecommuting permitted): Analyze business needs, determine problems, and advise on design and solutions. Design, build, test, deploy and maintain well-engineered information systems and ecosystems. Guide... 
    Remote work

    IBM

    Austin, TX
    5 days ago
  •  ...General Motors Financial Company, Inc. has one opening in Fort Worth, TX: Site Reliability Engineer II (Ref#22029.90.2). Req. MS in Engineering Management, Computer Engineering, or a related field & 2 yrs exp. in job offered or in operations engineering/site reliability... 
    Remote work

    General Motors Financial Company, Inc.

    Fort Worth, TX
    4 days ago
  •  ...Site Reliability Engineer- W2 Role* Technical proficiency: Strong Proficiency in Java, Strong understanding of Database concepts (Oracle, SQL, Dynamo DB etc.) Industry standard SRE Tools like Prometheus, Grafana, Data Dog Etc Good to have skills: Cloud Concepts / AWS,... 

    RSA Tech Group

    Dallas, TX
    1 day ago
  •  ...operational performance and availability of critical business platforms and cloud services. With a strong technical background in site reliability engineering, the ideal applicant will have excellent communication skills and a focus on continuous improvement through automation.... 

    Take-Two Interactive

    Austin, TX
    1 day ago
  • $127k - $249k

     ...A leading technology company is seeking an experienced Senior or Staff Engineer for their SRE, InfraSec team in Austin. This role focuses on leading the design and implementation of security solutions for cloud platforms while mentoring a team. Candidates should have... 

    MongoDB

    Austin, TX
    1 day ago
  • $110.7k - $171.8k

     ...components Participation in oncall rotation as a platform reliability escalation point Incident response, postincident reviews...  ..., and internal control requirements. Collaborate with engineering teams across the organization to influence platform adoption,... 
    Work experience placement
    Work at office
    Local area

    Visa

    Austin, TX
    4 days ago
  • $152.5k - $219.2k

     ...global cloud platform. As a team of six engineers distributed across the US, Canada, and the...  ...with a strong focus on automation, reliability, and operational excellence. We are one...  ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure... 
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    CISCO, Inc.

    Shenandoah, TX
    2 days ago
  • $152k - $195k

     ...Senior Site Reliability Engineer Austin, TX (Hybrid) SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million companies continuously rated, operating in 64 countries. Founded in 2013 by security and risk experts Dr. Alex Yampolskiy and... 

    SecurityScorecard

    Austin, TX
    3 days ago
  • $75.7k - $136.3k

     ...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and... 
    Work experience placement
    Work at office

    Akamai

    Austin, TX
    4 days ago
  • $98.58k - $138.02k

     ...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering... 
    Work at office

    Restaurant365

    Austin, TX
    4 days ago
  •  ...This role requires a seasoned engineer who combines deep knowledge of F5/AVI load balancing...  ...network, and platform teams to improve reliability, performance, and scalability....  ...performance, and availability risks. Implement Site Reliability Engineering best practices... 

    3B Staffing LLC

    Roanoke, TX
    2 days ago
  •  ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability... 

    Chase

    Houston, TX
    5 days ago
  •  ...Site Reliability Engineer Location: Fischer, Texas Schedule: Full-Time Pay Range: Competitive pay, based on experience and qualifications. Make a Difference as a Site Reliability Engineer in Fischer! Are you a problem-solving Site Reliability Engineer... 
    Full time
    Flexible hours

    MLee Medical Employment

    Fischer, TX
    3 days ago
  •  ...Site Reliability Engineer (SRE) The successful applicant may be performing work in FedRAMP High or IL-5 environments, and therefore, must be a U.S. Person (i.e. U.S. citizen, U.S. national, lawful permanent resident, asylee, or refugee). This position may also perform... 
    Permanent employment
    Worldwide
    Shift work

    Webex Events (formerly Socio)

    Richardson, TX
    5 days ago
  • $96.8k - $145.2k

     ...want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX), United States (US). Job Responsibilities Include:  Own... 
    Temporary work
    Work at office
    Remote work
    Flexible hours

    NTT America

    Plano, TX
    4 days ago
  •  ...TypeScript , with solid foundations in software engineering principles, debugging, and performance...  ..., manage error budgets, and drive reliability improvements. Hands-on coding...  ...7 years hands-on experience in Site Reliability Engineering, DevOps, or software... 
    Work experience placement

    3B Staffing LLC

    Fort Worth, TX
    1 day ago
  • $106k - $176k

     ...Travel Required: Up to 10% Clearance Required: Ability to Obtain Public Trust What You Will Do Lead implementation of Site Reliability Engineering practices across enterprise environments. Define and manage SLIs, SLOs, error budgets, and reliability metrics. Drive... 
    Temporary work
    Work experience placement
    Flexible hours

    Guidehouse

    San Antonio, TX
    3 days ago
  • $65 - $71 per hour

     ...NTT Data is looking for a Site Reliability Engineer to join their team in Westlake, Texas. Site Reliability Engineer - Westlake, TX - 26-01183 Job Description: We are seeking a highly skilled Site Reliability Engineer (SRE) with strong expertise in Terraform,... 
    Hourly pay
    Temporary work
    Remote work
    Flexible hours

    NTT Data Americas, Inc.

    Roanoke, TX
    6 days ago
  •  ...Title: Site Reliability Engineer (SRE) Location: Austin, TX Description: We're searching for a driven Site Reliability Engineer (SRE) to join our innovative team. As an SRE, you'll be a cornerstone of our production software, ensuring our systems... 
    Work experience placement

    United IT Solutions

    Austin, TX
    5 days ago
  •  ...generative AI and cloud-native platforms to advanced release engineering practices, our teams are redefining how financial technology...  ...AI-driven solutions that accelerate development and improve reliability. Your work will directly influence how GM Financial leverages... 
    Full time
    H1b
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours
    2 days per week
    3 days per week

    GMAC Financial Services

    Irving, TX
    2 days ago
  • $100k - $115k

     ...Internal Developer Platform Engineer Analytic Partners is a global leader in commercial...  ...teams as customers and optimizing for reliability, usability, and delivery velocity. Define...  ...of experience in Platform Engineering, Site Reliability Engineering, DevOps, or... 
    Temporary work

    Analytic Partners

    Dallas, TX
    1 day ago
  • Mandatory Skills: AWS/Azure/GCP (GCP is not used very much ). Kubernetes /Helm,Docker,Gitlab,Grafana,Cyberark/Hashicorp Vault, Terraform etc. Experience utilizing Java, Perl, Python, Go and scripting experience in Shell and Perl to automate reports and monitor enterprise...

    Omni Inclusive

    Dallas, TX
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!