Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, ML Serving - Rime Ai

Full-time

Unusual Ventures


Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences at scale. Our models are purpose-built for high-volume conversational deployments, engineered for the accuracy, performance, and deployment flexibility that production environments actually demand.

We started from a different premise than the rest of the field: build voice AI for human connection, not slop. Before we trained a single model, we built our own corpus: full-duplex, studio-quality conversational speech of normal people, recorded and annotated by linguists. It's why our models are unparalleled in naturalism, and it's why enterprises pick Rime when pilots need to make it to production.

  Role Overview

We're hiring a Software Engineer to own the serving infrastructure that connects Rime's inference engines to the world. This role sits at the intersection of ML systems and cloud infrastructure — you'll work directly on model inference and cloud infrastructure to build, harden, and scale the systems that stream voice at real-time latency. As Rime moves toward its next-generation architecture, you'll be a core architect of how our models get served.

 

What You'll Own




  • Architecture and implementation of Rime's TTS serving infrastructure, from GPU-backed inference engines to the API surface.



  • Model optimization from a single-node to disaggregated fleet serving.



  • Compatibility with different NVIDIA hardwares from Hopper to Blackwell and beyond for on-prem and cloud deployments.



  • Continuous integration and deployment workflows for the model serving pipeline.



  • Site reliability: on-call rotation, monitoring, alerting, and observability across the serving stack.



  • Resource provision, cost management across our GPU fleet.


What We're Looking For




  • Hands-on experience with real-time multinode ML serving infrastructure — ML serving framework experience: NVIDIA Dynamo/Triton, vLLM, SGLang, or equivalent.



  • Experience with distributed or disaggregated model serving (Tensor Parallel, Pipeline Parallel, or equivalent).



  • Strong cloud infrastructure fundamentals: Linux internals, networking, containerization (Docker, Kubernetes).



  • IaC experience — Terraform, Packer, or comparable. You should have opinions about how to do this right.



  • On-call is part of the job. You treat production reliability as a shared responsibility.


Nice to Have




  • Experience with multinode training (DDP, FSDP, etc.).



  • Experience with gRPC or other bidirectional binary streaming protocols.



  • Experience with audio streaming and related technologies (WebRTC, WebSockets, etc.).



  • Experience with a multilingual monorepo where you pick the best language out of merit more than personal experience.



  • Experience with multi-cloud infrastructures (AWS, GCP, OCI, etc.).



  • Comfort with configuration management tooling (Ansible, Chef, Puppet, or similar).



  • SRE, DevOps, or platform engineering background at a startup.



  • Experience at an early-stage company.


Why Join Rime




  • Build the serving infrastructure behind a category-defining voice AI company from the ground up.



  • You will bring in experience that no one else currently has at the company: you can help us set the vision.



  • Direct collaboration with the inference, platform, and ML teams — no handoff culture.



  • The systems you build determine what experiences our customers can deploy at scale.



  • Meaningful equity upside at an early stage.



  • High ownership, high standards, low bureaucracy.



  • SF / Bay Area.


At Rime, we...




  • Are outliers



  • Cut through the hype to focus on the craft



  • Move fast with agency and freedom



  • Maintain a growth mindset, finding joy in the struggle



  • Do the right things, knowing that it'll lead to making money


If that sounds like you too, you'll be a great fit for Rime!

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted 18 hours ago
Similar jobs that could be interesting for youBased on the Software Engineer, ML Serving - Rime Ai in San Francisco, CA vacancy
  •  ...Rime is a foundation modeling company that builds voice AI for enterprises running customer...  ...conversational deployments, engineered for the accuracy,...  ...a generalist Software Engineer to work across...  ...model serving, bidirectional streaming...  ...Comfort working within ML systems Why... 
    Suggested
    Full time

    Unusual Ventures

    San Francisco, CA
    18 hours ago
  •  ...Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences...  ...conversational deployments, engineered for the accuracy,...  ...Overview We're hiring a Software Engineer to own Rime's...  ...solutions that serve both the user and the business... 
    Suggested
    Full time
    Casual work

    Unusual Ventures

    San Francisco, CA
    18 hours ago
  • $230k - $385k

     ...the TeamThe Cooperative AI team is scaling OpenAI with...  .... We iterate fast, and engineer for reliable long-term...  ...re looking for a Backend Software Engineer to help...  ...and backend services that serve as the foundation for how...  ...timeCuriosity about AI/ML and excitement to work alongside... 
    Suggested
    Internship
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    4 days ago
  • $130k - $150k

     ...Notion is the collaborative AI workspace where teams and agents...  ...About the Role As an engineer at Notion, you’ll help shape core...  ...infrastructure that supports model serving and experimentation...  ...and you’ve started exploring AI/ML through coursework, projects,... 
    Suggested
    Full time
    Internship
    Local area

    Notion

    San Francisco, CA
    18 hours ago
  • $325k

     ...interpretable, and steerable AI systems. We want AI to...  ...committed researchers, engineers, policy experts, and...  ...our most critical serving paths -- every hop from...  ...for reliability-minded software engineers and SREs....  ...experience with one or more ML hardware accelerators (... 
    Suggested
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    18 hours ago
  • $160k - $240k

     ...businesses function in the age of AI. The simple task of buying software, services, or tools at...  ...Role As a Software Engineer on ZipX, Zip's AI team, you...  ...maintainable code that serves as the foundation for Zip'...  ...interest in working with AI/ML teams to build intelligent... 
    Full time
    Contract work
    Home office
    Flexible hours

    Zip

    San Francisco, CA
    18 hours ago
  •  ...for the world's most dynamic AI companies, like Cursor, Notion...  ...us and help build the platform engineers turn to to ship AI products....  ...Develop world-class model serving stack for state-of-the-art open...  ...oriented tools Interest in ML/AI infrastructure and willingness... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    18 hours ago
  •  ...inference for the world's most dynamic AI companies, like Cursor, Notion,...  ...us and help build the platform engineers turn to to ship AI products....  ...external teams deploy and serve AI models, your focus is inward...  ...k) ~ Exposure to a variety of ML startups, offering unparalleled... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    18 hours ago
  •  ...We’re seeking an Agent Engineer to design and build agentic...  ...Friendli Agent API, which serves as the core developer...  ...agent systems and making AI easy for developers to adopt...  ...+ years of experience in software engineering, preferably in backend, ML systems, or API development... 
    Full time
    Worldwide
    Flexible hours

    FriendliAI

    San Francisco, CA
    18 hours ago
  • $145.4k - $188.1k

     ...personalized guidance to AI tools and a best-in-...  ...detection, and more. The ML Infrastructure team is...  ...experiences. We empower product engineering teams by providing...  ...The challenge As a Software Engineer on the ML...  ...ML model training and serving systems, feature and data... 
    Full time
    Seasonal work
    Local area

    Thumbtack

    San Francisco, CA
    18 hours ago
  •  ...Gateway, open-weights model serving and batch inference,...  ...for Generative AI at DoorDash, with a primary...  ...This role is ideal for an engineer who enjoys building...  ...excellence.Partner closely with ML engineers, product...  ...industry experience in software engineeringStrong... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    18 hours ago
  •  ...inference for the world's most dynamic AI companies, like Cursor, Notion,...  ...us and help build the platform engineers turn to to ship AI products....  ...configured and working. Self-serve infrastructure so teams build...  ...k) - Exposure to a variety of ML startups, offering unparalleled... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    18 hours ago
  •  ...for agentic team chat—a workspace where AI agents and humans collaborate as peers. Our...  ...Opportunity We’re looking for an AI/ML Engineer with a strong foundation in large language...  ...data ingestion and model integration to serving and monitoring—and understand what it takes... 
    Local area
    Immediate start
    Flexible hours

    Glue

    San Francisco, CA
    2 days ago
  • $180k - $201k

     ...Notion is the collaborative AI workspace where teams...  ...a team of talented engineers focused on making speed...  ...You’ve worked on LLM, ML platform, data, or infrastructure...  ...of experience as a Software Engineer ~ Experience...  ...building MLOps and ML serving infrastructure.... 
    Full time
    Local area

    Notion

    San Francisco, CA
    18 hours ago
  •  ...the Role Roger is an AI platform that frees home...  ...busywork that organizations serving these vulnerable...  ...for a Senior Applied AI Engineer to build the intelligence...  ...years of professional software engineering experience,...  ...meaningful depth in AI/ML. ~ Experience training... 
    Full time
    Remote work
    Work from home

    Roger Healthcare

    San Francisco, CA
    18 hours ago
  •  ...Python and Kubernetes Software Engineer - Data, AI/ML & Analytics Join to apply for the Python and Kubernetes Software Engineer - Data, AI/ML &...  ...Kubernetes, on developer desktops, or as web services. We serve the needs of individuals and community members as much as... 
    Full time
    Freelance
    Internship
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    San Francisco, CA
    1 day ago
  •  ...Arcade is building the world's first AI physical product creation platform, where...  ...We're looking for a Staff Applied AI Engineer to architect the ML systems at the core of Arcade's...  ...pipelines, training infrastructure, and serving systems that let diffusion models, LLMs... 
    Full time

    Arcade

    San Francisco, CA
    18 hours ago
  • $198k - $280k

    About the teamThe Applied AI Engineering team is responsible for ensuring the safe and effective...  ...use cases being built on our API, serving as a critical partner in collecting...  ...you:Have 5+ years of experience as a software engineer, ML engineer or equivalent, ideally in a... 
    Work at office
    Local area
    Relocation package
    Flexible hours

    OpenAI

    San Francisco, CA
    2 days ago
  • $150k - $200k

    San Francisco, CaliforniaEngineering - AI Engineering /US Full-time Salaried /On-siteSamba is a media...  ...) serversOpen-source contributions to AI/ML projectsFamiliarity with observability...  ...one another, reflect the communities we serve, and tackle meaningful projects that make... 
    Full time
    Local area

    Samba TV

    San Francisco, CA
    4 days ago
  • $250k - $300k

     ...vertically integrated AI infrastructure company...  ...getting deep into the serving code when the defaults...  ...directly with customer engineering teams to tailor deployments...  ...methods across many kinds of ML models, with an...  ....Build and support the software and product features around... 
    Temporary work

    Crusoe

    San Francisco, CA
    2 days ago
  • $225k - $320k

     ...speed and accuracy. Our AI-enabled platform turns...  ...states and two countries, serving more than 125 million...  .... As the Applied AI Engineering Lead, you will own how...  ...into core workflowsAI/ML FoundationsOwn and guide...  ...quality measurementStrong software engineering fundamentals... 
    Local area

    Peregrine Technologies

    San Francisco, CA
    4 days ago
  • $150k - $170k

     ...motivated, mission-oriented Applied AI Engineer to help develop, integrate, and deploy AI/ML-powered capabilities across our...  ...with other data scientists, software engineers, product teams, and mission...  ...for acquisition, model serving, monitoring, lifecycle management... 
    Work experience placement
    Casual work
    Work at office
    Relocation package
    2 days per week

    CHAOS Industries

    San Francisco, CA
    18 hours ago
  • $264.66k - $369.8k

     ...Washington D.C., London and Amsterdam.AI and intelligent systems are driving the...  ...those systems to life.Work with other AI engineers, software engineers and machine learning...  ...spaceNice-to-HavesExperience training and/or serving ML models in production, or fine-tuning LLMs... 
    Full time
    Work experience placement
    Local area
    Shift work

    Plaid Financial

    San Francisco, CA
    18 hours ago
  • $286.2k - $326.7k

    Sr. Distinguished AI Engineer (Remote Eligible) At Capital One,...  ...time, our applications of AI & ML are bringing humanity and...  ...capabilities to reimagine how we serve our customers and businesses who...  ...test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area
    Remote work

    Capital One Financial Corporation

    San Francisco, CA
    18 hours ago
  • $314.8k - $359.3k

     ...description": "Senior Distinguished AI Engineer At Capital One, we...  ..., our applications of AI & ML are bringing humanity and simplicity...  ...to reimagine how we serve our customers and businesses who...  ...test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area

    Capital One Financial Corporation

    San Francisco, CA
    18 hours ago
  •  ...Forward Deployed AI Engineer The opportunity We are looking for a Forward...  ...Deployed AI Engineer to serve as the critical bridge between...  ...are You have a strong CS or ML educational background. You hold...  ...You have a solid grounding in software engineering principles and... 
    Full time
    Shift work

    Latent Labs

    San Francisco, CA
    18 hours ago
  •  .... Put simply, we build software for the people who enable...  ...Fieldguide is building AI agents for the most...  ...investors. As a Senior AI Engineer, Quality , you will...  ...at the intersection of ML engineering,...  ...evaluation platform that serves as the single source of... 
    Remote job
    Full time
    Work at office
    Work from home
    Flexible hours

    Field Guide Inc

    San Francisco, CA
    18 hours ago
  •  ...data science-first growth engine that gives B2C teams...  ...building alongside leading AI companies. We're...  ...features and pipelines that serve enterprise customers at...  ...on AI systems Backend software engineering experience...  ...production systems beyond the ML/AI layer Exposure to... 
    Full time
    Shift work
    Night shift
    Weekend work

    Hilbert's Ai

    San Francisco, CA
    18 hours ago
  • $7.5k

     ...AI Engineer Location: San Francisco, CA or Phoenix, AZ (In-Office) Partnership: EQL Tech...  ...largest smart warehousing company globally, serving 1M customers/day at age 19. ~ The...  ...experience at one of the world's leading ML research labs — to build and ship the intelligence... 
    Full time
    Work at office
    Relocation
    Visa sponsorship
    Relocation package

    Eql Tech

    San Francisco, CA
    18 hours ago
  • $100k - $250k

     ...Solutions Engineer Specializing In Ai Infrastructure And Platforms We are seeking...  ...platforms both Hardware and Software technologies. You will collaborate...  .... Role Description Serve as the primary AI...  ...infrastructure, preferably with AI/ML/HPC workloads. ~ Proven... 
    Work experience placement
    Work at office
    Flexible hours

    SHI GmbH

    San Francisco, CA
    5 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, ML Serving - Rime Ai. Be the first to apply!