Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer (SRE)

Fluix AI

Site Reliability Engineer (SRE)

FLUIX is building the AI operating system that plans, designs, and optimizes AI infrastructure. We are based in Silicon Valley. We specialize in providing AI-driven solutions for data centers and power providers, leveraging cutting-edge Machine Learning (ML) and Artificial Intelligence (AI) technologies. Our mission is to double America's compute capacity without building new data centers.

We are seeking a skilled Site Reliability Engineer to join our growing team. The ideal candidate will help ensure the reliability, scalability, and performance of our hybrid-based (Cloud & On-Prem) platform while supporting our AI/ML infrastructure. You will work closely with our engineering, AI, and operations teams to build and maintain robust systems that support our cutting-edge solutions. Your expertise in ML/AI and experience with data center sites will be crucial in driving the success of our platform.

Who You'll Work Closely With

Abhi Sastri

Founder & CEO

Chase Overcash

CTO

What You'll Do
  • Design, implement, and maintain scalable systems while optimizing performance, ensuring high availability and disaster recovery, and assisting with codebase refactoring for modular deployment.
  • Develop and maintain automation tools to streamline operations, improve efficiency, and automate repetitive tasks to enhance system reliability.
  • Collaborate with engineering and data science teams to integrate ML and AI models into production environments, while ensuring seamless integration and high performance of cutting-edge models within our technology stack.
  • Identify areas for improvement and drive initiatives to enhance system reliability and performance, while staying updated on industry trends and advancements in SRE practices, ML, and AI technologies.
  • Respond to and resolve incidents to minimize impact and ensure timely resolution, while conducting post-incident reviews and implementing improvements to prevent recurrence.
  • Create and manage multiple cloud instances (dev, staging, test), optimize cloud infrastructure and data center operations, and ensure the security and compliance of both infrastructure and applications.
Your Background
  • Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent experience).
  • Proven experience as a Site Reliability Engineer or similar role in a SaaS environment, with a strong background in managing and optimizing cloud infrastructure (AWS preferred, or GCP, Azure), experience with ML and AI technologies, and familiarity with data center operations integrations.
  • Proficiency in programming and scripting languages (e.g., Python), experience with containerization and orchestration tools (Kubernetes), a strong understanding of networking, security, and performance optimization, and knowledge of CI/CD pipelines and DevOps practices.
  • Excellent problem-solving skills with attention to detail, strong communication and collaboration abilities, and the capacity to thrive in a fast-paced, dynamic startup environment.
Culture Fit
  • We are looking for obsessed individuals who want to give it their all.
  • We are not afraid to get our hands dirty with physical and software systems.
  • We are eager to visit and work with clients and understand the importance and gravitas of their mission-critical work.
  • We are eager to come into the office and on-site, as our work directly affects physical environments.
  • Due to our mission-critical work, we understand and are eager to help our teammates and co-workers during holidays, weekends, and emergencies.
  • We are cordial and over-communicate with teammates, co-workers, and management.
Benefits

Competitive Salary

Attractive compensation package, including equity options.

Benefits

Comprehensive health, dental, and vision insurance, along with other standard benefits.

Work Environment

A dynamic and collaborative San Francisco Bay Area work environment.

Growth Opportunities

Opportunities for professional growth and development, with the chance to shape the future of technology in the industry.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer (SRE) in San Francisco, CA vacancy
  • $170k - $250k

     ...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000-$250,000 + Competitive Equity Company Description... 
    Suggested
    Work at office
    Visa sponsorship
    Flexible hours

    Recruiting from Scratch

    San Francisco, CA
    1 day ago
  • $170k - $230k

     ...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,... 
    Suggested
    Work at office
    Local area
    1 day per week

    Mithril

    San Francisco, CA
    4 days ago
  •  ...globe. Join us on this journey to redefine resource management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the reliability, availability, and scalability of our systems. You... 
    Suggested
    Temporary work
    Worldwide

    Air Apps

    San Francisco, CA
    1 day ago
  • $186.9k - $267.7k

     ...requiring approximately 2 days per week on-site at Cisco offices in either San...  ...AI agents behave as intended, improving reliability and reducing risks. This unified approach...  ...and control.As a Staff Site Reliability Engineer (SRE), you will provide technical leadership... 
    Suggested
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    San Francisco, CA
    2 hours ago
  • $350k

     ...and novel use-cases. We’re hiring to grow the platform alongside the Tinker community. About the Role We're looking for a Site Reliability Engineer to drive the reliability of Tinker end-to-end. You'll work alongside the engineers building the platform and research... 
    Suggested
    Full time
    Visa sponsorship
    Work visa
    Relocation package

    Thinking Machines Lab

    San Francisco, CA
    2 days ago
  • $163.71k - $306k

     ...their own infrastructure, behind their own controls, with the reliability and operational clarity they would expect from any critical system...  ..., Support, and TAMs to trust. Partner with product engineers on infrastructure requirements for new Retool products, especially... 

    Retool

    San Francisco, CA
    5 days ago
  • $300k

     ...experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and...  .... Skills / Must Have: ~7+ years of experience in SRE, DevOps, or Infrastructure Engineering roles supporting... 
    Permanent employment
    San Francisco, CA
    more than 2 months ago
  •  ...OpportunityTo achieve our ambitious goals, we’re looking for an SRE to join our infrastructure team. This role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them, and planning for our future growth.... 
    Worldwide
    Home office
    Flexible hours

    Superhuman

    San Francisco, CA
    5 hours ago
  • $165k - $241.4k

     ...portfolios Your ImpactThe FedRAMP SRE team is focused on our...  ...effective.We’re looking for talented engineers with a software or operations...  ...teams to ensure the reliability, performance and security of...  ...Please see the Cisco careers site to discover more benefits and... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours
    1 day per week

    CISCO Systems

    San Francisco, CA
    4 days ago
  • $117k - $209.33k

     ...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,...  ...cloud services for Autodesk GovCloud products.As part of a new SRE team supporting Autodesk GovCloud, you will have a unique... 
    Full time
    For contractors

    Autodesk

    San Francisco, CA
    1 day ago
  • $113.4k - $162k

     ...conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd,...  ...practices. Contribute to the design and implementation of new SRE best practices.You'll be a great fit if you have:Experienced... 
    Temporary work

    TextNow

    San Francisco, CA
    3 days ago
  • $152.5k - $205k

     ...is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate...  ...public-cloud environments. This role is for an experienced SRE or infrastructure engineer who enjoys solving hard distributed... 
    Flexible hours

    Circle

    San Francisco, CA
    4 days ago
  •  ...what’s next.About the teamThe Engineering team at Airwallex is a diverse...  ...working together to build scalable, reliable, and secure products that...  ...to grow without borders.Our SRE team is breaking new engineering...  ....What you’ll doAs a Senior Site Reliability Engineer, you’ll work... 
    Temporary work
    Local area
    Worldwide

    Airwallex

    San Francisco, CA
    5 days ago
  • $148.5k - $223.9k

     ...future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts...  ...and our customers protected. The ExperienceAs an SRE, you will be a technical leader of the team driving... 
    Full time
    Worldwide
    Weekend work

    Salesforce

    San Francisco, CA
    4 days ago
  • About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the... 

    Alembic

    San Francisco, CA
    5 days ago
  • $165k - $225.6k

     ...we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability...  ...engineering teams to champion DevOps and SRE best practices, deliver excellent internal... 
    Permanent employment
    Local area
    Worldwide
    Flexible hours

    Okta

    San Francisco, CA
    4 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support...  ...alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    San Francisco, CA
    2 days ago
  •  ...builds the platforms and tooling that help engineering teams develop, deploy, and operate...  ...default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll...  ...experience in backend systems, SRE, or platform engineering roles.Proven track... 
    Permanent employment
    Work experience placement
    Work at office
    Local area

    Plaid Financial

    San Francisco, CA
    4 days ago
  • $155k - $222.6k

     ...soil. Meet the Team The SRE Fleet team is responsible for...  ...platform. As a team of six engineers distributed across the US,...  ...strong focus on automation, reliability, and operational excellence....  ...~2+ years of experience in Site Reliability Engineering, DevOps... 
    Permanent employment
    Full time
    Temporary work
    Local area
    Worldwide
    Flexible hours

    Cisco

    Daly City, CA
    1 day ago
  •  ...About the job Senior Site Reliability Engineer About the Company Stellar is a decentralized, public blockchain that gives developers the tools...  ...of working in cloud-based systems operations, as a SRE or DevOps engineer. ~ First-hand experience with configuration... 

    TechChain Talent

    San Francisco, CA
    5 days ago
  • $166.9k - $225.9k

     ...and career news. Job Summary: Drata's SRE team operates as both a central engineering function and an embedded reliability practice. You'll be part of a close-knit SRE...  ...you'll bring: ~6+ years of experience in Site Reliability Engineering, Cloud Engineering,... 
    Work at office
    Immediate start
    Worldwide
    Monday to Friday
    Flexible hours

    Drata Inc

    San Francisco, CA
    1 day ago
  • $260k - $300k

     ...makers of Devin, the first AI software engineer. Our team is extremely talent-dense...  .... You will own both the production reliability of our user-facing products and the...  ...Strong software engineering fundamentals; SRE at Cognition means writing real code, not... 

    Cognition Corp

    San Francisco, CA
    4 days ago
  • $100k - $170k

     ...Site Reliability Engineer Houston; San Francisco; Seattle About Nscale Nscale is the GPU cloud built for AI. We run high-performance, cost...  ...that makes AI work. The Role This is a career-level SRE role for someone who wants to own systems, not just watch them... 
    Flexible hours
    Shift work

    Nscale

    San Francisco, CA
    4 days ago
  •  ...DESCRIPTION Project Outline: We are looking for a Site Reliability Engineer with experience in incident response. In this role, you will...  ...Skill Requirements: - Engineering Background: 4+ years in SRE, DevOps, or Systems Engineering roles managing production... 

    BayOne Solutions

    San Francisco, CA
    1 day ago
  •  ...SRE Location: San Francisco, CA (5 Days In-Office) You are the infrastructure...  ...treatment. What We Look for in a Great Engineer You have the intensity and technical...  ...feature release while maintaining the highest reliability. DevX Support: Support Developer... 
    Work at office

    Latent

    San Francisco, CA
    11 hours ago
  • $150k

     ...Site Reliability Engineer San Francisco, CA About The Role We are seeking an experienced Site Reliability Engineer (SRE) with a strong focus on DevSecOps to join our growing engineering team. In this role, you will oversee and maintain the reliability, security... 

    VantageScore®

    San Francisco, CA
    5 days ago
  •  ...Site Reliability Engineer We are looking for a dynamic engineer to join our rapidly growing SRE team. As an SRE, you will report to our VP of Technical Operations and be responsible for operating an extremely high performance and scalable, low latency platform built... 
    Relocation package

    1872 Consulting

    San Francisco, CA
    4 days ago
  •  ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering Apple services...  ...define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role includes... 

    Socket

    San Francisco, CA
    1 day ago
  • $120k - $168.49k

     ...Site Reliability Engineer, Cloud Infrastructure About Quizlet At Quizlet, our mission is to help every learner achieve their outcomes in the most...  ...About The Role We are looking for a Site Reliability Engineer (SRE) to join our infrastructure team and help build reliable,... 
    Internship
    Work at office
    3 days per week

    Quizlet

    San Francisco, CA
    5 days ago
  •  ...The role We're looking for a world-class Site Reliability Engineer to ensure the reliability, performance, and scalability of our AI infrastructure...  ...your decisions. Required skills ~3+ years in SRE, DevOps, or infrastructure engineering roles ~ Strong... 

    Blaxel, Inc

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer (SRE). Be the first to apply!