Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, ML Serving - Rime Ai

Full-time

Unusual Ventures

Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences at scale. Our models are purpose-built for high-volume conversational deployments, engineered for the accuracy, performance, and deployment flexibility that production environments actually demand.

We started from a different premise than the rest of the field: build voice AI for human connection, not slop. Before we trained a single model, we built our own corpus: full-duplex, studio-quality conversational speech of normal people, recorded and annotated by linguists. It's why our models are unparalleled in naturalism, and it's why enterprises pick Rime when pilots need to make it to production.

  Role Overview

We're hiring a Software Engineer to own the serving infrastructure that connects Rime's inference engines to the world. This role sits at the intersection of ML systems and cloud infrastructure — you'll work directly on model inference and cloud infrastructure to build, harden, and scale the systems that stream voice at real-time latency. As Rime moves toward its next-generation architecture, you'll be a core architect of how our models get served.

 

What You'll Own



  • Architecture and implementation of Rime's TTS serving infrastructure, from GPU-backed inference engines to the API surface.



  • Model optimization from a single-node to disaggregated fleet serving.



  • Compatibility with different NVIDIA hardwares from Hopper to Blackwell and beyond for on-prem and cloud deployments.



  • Continuous integration and deployment workflows for the model serving pipeline.



  • Site reliability: on-call rotation, monitoring, alerting, and observability across the serving stack.



  • Resource provision, cost management across our GPU fleet.


What We're Looking For



  • Hands-on experience with real-time multinode ML serving infrastructure — ML serving framework experience: NVIDIA Dynamo/Triton, vLLM, SGLang, or equivalent.



  • Experience with distributed or disaggregated model serving (Tensor Parallel, Pipeline Parallel, or equivalent).



  • Strong cloud infrastructure fundamentals: Linux internals, networking, containerization (Docker, Kubernetes).



  • IaC experience — Terraform, Packer, or comparable. You should have opinions about how to do this right.



  • On-call is part of the job. You treat production reliability as a shared responsibility.


Nice to Have



  • Experience with multinode training (DDP, FSDP, etc.).



  • Experience with gRPC or other bidirectional binary streaming protocols.



  • Experience with audio streaming and related technologies (WebRTC, WebSockets, etc.).



  • Experience with a multilingual monorepo where you pick the best language out of merit more than personal experience.



  • Experience with multi-cloud infrastructures (AWS, GCP, OCI, etc.).



  • Comfort with configuration management tooling (Ansible, Chef, Puppet, or similar).



  • SRE, DevOps, or platform engineering background at a startup.



  • Experience at an early-stage company.


Why Join Rime



  • Build the serving infrastructure behind a category-defining voice AI company from the ground up.



  • You will bring in experience that no one else currently has at the company: you can help us set the vision.



  • Direct collaboration with the inference, platform, and ML teams — no handoff culture.



  • The systems you build determine what experiences our customers can deploy at scale.



  • Meaningful equity upside at an early stage.



  • High ownership, high standards, low bureaucracy.



  • SF / Bay Area.


At Rime, we...



  • Are outliers



  • Cut through the hype to focus on the craft



  • Move fast with agency and freedom



  • Maintain a growth mindset, finding joy in the struggle



  • Do the right things, knowing that it'll lead to making money


If that sounds like you too, you'll be a great fit for Rime!

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, ML Serving - Rime Ai in San Francisco, CA vacancy
  •  ...Rime is a foundation modeling company that builds voice AI for enterprises running customer...  ...conversational deployments, engineered for the accuracy,...  ...a generalist Software Engineer to work across...  ...model serving, bidirectional streaming...  ...Comfort working within ML systems Why... 
    Suggested
    Full time

    Unusual Ventures

    San Francisco, CA
    1 day ago
  •  ...Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences...  ...deployments, engineered for the accuracy, performance...  ...We're hiring a Software Engineer to own the serving infrastructure that...  ...the intersection of ML systems and cloud infrastructure... 
    Suggested
    Full time

    Unusual Ventures

    San Francisco, CA
    1 day ago
  •  ...Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences...  ...conversational deployments, engineered for the accuracy,...  ...Overview We're hiring a Software Engineer to own Rime's...  ...solutions that serve both the user and the business... 
    Suggested
    Full time
    Casual work

    Unusual Ventures

    San Francisco, CA
    1 day ago
  • $266k

     ...the TeamThe Cooperative AI team is scaling OpenAI with...  .... We iterate fast, and engineer for reliable long-term...  ...re looking for a Backend Software Engineer to help...  ...and backend services that serve as the foundation for how...  ...timeCuriosity about AI/ML and excitement to work alongside... 
    Suggested
    Internship
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    1 day ago
  • $130k - $150k

     ...Notion is the collaborative AI workspace where teams and agents...  ...About the Role As an engineer at Notion, you’ll help shape core...  ...infrastructure that supports model serving and experimentation...  ...and you’ve started exploring AI/ML through coursework, projects,... 
    Suggested
    Full time
    Internship
    Local area

    Notion

    San Francisco, CA
    1 day ago
  • $325k

     ...interpretable, and steerable AI systems. We want AI to...  ...committed researchers, engineers, policy experts, and...  ...our most critical serving paths -- every hop from...  ...for reliability-minded software engineers and SREs....  ...experience with one or more ML hardware accelerators (... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    1 day ago
  •  ...for the world's most dynamic AI companies, like Cursor, Notion...  ...us and help build the platform engineers turn to to ship AI products....  ...Develop world-class model serving stack for state-of-the-art open...  ...oriented tools Interest in ML/AI infrastructure and willingness... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • $160k - $240k

     ...businesses function in the age of AI. The simple task of buying software, services, or tools at...  ...Role As a Software Engineer on ZipX, Zip's AI team, you...  ...maintainable code that serves as the foundation for Zip'...  ...interest in working with AI/ML teams to build intelligent... 
    Full time
    Contract work
    Home office
    Flexible hours

    Zip

    San Francisco, CA
    1 day ago
  •  ...inference for the world's most dynamic AI companies, like Cursor, Notion,...  ...us and help build the platform engineers turn to to ship AI products....  ...external teams deploy and serve AI models, your focus is inward...  ...k) ~ Exposure to a variety of ML startups, offering unparalleled... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  •  ...We’re seeking an Agent Engineer to design and build agentic...  ...Friendli Agent API, which serves as the core developer...  ...agent systems and making AI easy for developers to adopt...  ...+ years of experience in software engineering, preferably in backend, ML systems, or API development... 
    Full time
    Worldwide
    Flexible hours

    FriendliAI

    San Francisco, CA
    1 day ago
  • $145.4k - $188.1k

     ...personalized guidance to AI tools and a best-in-...  ...detection, and more. The ML Infrastructure team is...  ...experiences. We empower product engineering teams by providing...  ...The challenge As a Software Engineer on the ML...  ...ML model training and serving systems, feature and data... 
    Full time
    Seasonal work
    Local area

    Thumbtack

    San Francisco, CA
    1 day ago
  •  ...Gateway, open-weights model serving and batch inference,...  ...for Generative AI at DoorDash, with a primary...  ...This role is ideal for an engineer who enjoys building...  ...excellence.Partner closely with ML engineers, product...  ...industry experience in software engineeringStrong... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    2 days ago
  •  ...inference for the world's most dynamic AI companies, like Cursor, Notion,...  ...us and help build the platform engineers turn to to ship AI products....  ...configured and working. Self-serve infrastructure so teams build...  ...k) - Exposure to a variety of ML startups, offering unparalleled... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • $156.06k - $211.14k

    Afresh, the AI platform for grocery, began by tackling the most...  ...them is not.As a Senior AI Software Engineer, you build the AI and data platform...  ...it, and the evaluation and serving infrastructure underneath....  ...software, data, or ML systems; an excellent engineer... 
    Full time
    Live in
    Work at office
    Local area
    Remote work
    Work from home
    Home office
    Flexible hours
    3 days per week

    Afresh

    San Francisco, CA
    1 day ago
  •  ...for agentic team chat—a workspace where AI agents and humans collaborate as peers. Our...  ...Opportunity We’re looking for an AI/ML Engineer with a strong foundation in large language...  ...data ingestion and model integration to serving and monitoring—and understand what it takes... 
    Local area
    Immediate start
    Flexible hours

    Glue

    San Francisco, CA
    7 days ago
  • $180k - $201k

     ...Notion is the collaborative AI workspace where teams...  ...a team of talented engineers focused on making speed...  ...You’ve worked on LLM, ML platform, data, or infrastructure...  ...of experience as a Software Engineer ~ Experience...  ...building MLOps and ML serving infrastructure.... 
    Full time
    Local area

    Notion

    San Francisco, CA
    1 day ago
  • $200k - $400k

     ...Decagon is the leading conversational AI platform empowering every brand to deliver...  ...that power Decagon: networking, data, ML serving, developer platform, and real-time...  ...the Role We're looking for a Senior Software Engineer to help build and evolve our internal developer... 
    Full time
    Work at office
    Local area

    Decagon

    San Francisco, CA
    1 day ago
  •  ...the Role Roger is an AI platform that frees home...  ...busywork that organizations serving these vulnerable...  ...for a Senior Applied AI Engineer to build the intelligence...  ...years of professional software engineering experience,...  ...meaningful depth in AI/ML. ~ Experience training... 
    Full time
    Remote work
    Work from home

    Roger Healthcare

    San Francisco, CA
    1 day ago
  •  ...Python and Kubernetes Software Engineer - Data, AI/ML & Analytics Join to apply for the Python and Kubernetes Software Engineer - Data, AI/ML &...  ...Kubernetes, on developer desktops, or as web services. We serve the needs of individuals and community members as much as... 
    Full time
    Freelance
    Internship
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    San Francisco, CA
    3 days ago
  •  ...Arcade is building the world's first AI physical product creation platform, where...  ...We're looking for a Staff Applied AI Engineer to architect the ML systems at the core of Arcade's...  ...pipelines, training infrastructure, and serving systems that let diffusion models, LLMs... 
    Full time

    Arcade

    San Francisco, CA
    1 day ago
  • $197k - $278k

    About the teamThe Applied AI Engineering team is responsible for ensuring the safe and effective...  ...and technical platform customers, serving as their technical thought partner in...  ...consists of former startup founders, software and ML engineers, technical product managers... 
    Work at office
    Local area
    Relocation package
    Flexible hours

    OpenAI

    San Francisco, CA
    4 days ago
  •  ...Possible. At Pinterest, AI isn't just a feature, it's...  ...The Trends & Insights Engineering team builds the products that...  ...for advertisers. As a Staff Software Engineer on this team, you...  ...large-scale data or ML-powered platforms that serve customer-facing products.... 
    Full time
    Work at office
    Remote work
    Relocation
    Relocation package
    Day shift

    Pinterest

    San Francisco, CA
    1 day ago
  • $215k - $260k

     ...vertically integrated AI infrastructure company...  ...getting deep into the serving code when the defaults...  ...directly with customer engineering teams to tailor deployments...  ...methods across many kinds of ML models, with an...  ....Build and support the software and product features around... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  • $150k - $200k

    San Francisco, CaliforniaEngineering - AI Engineering /US Full-time Salaried /On-siteSamba is a media...  ...) serversOpen-source contributions to AI/ML projectsFamiliarity with observability...  ...one another, reflect the communities we serve, and tackle meaningful projects that make... 
    Full time
    Local area

    Samba TV

    San Francisco, CA
    1 day ago
  • $251k - $335k

    About the TeamThe Applied AI Engineering team partners closely with customers to help them move...  ...impact from frontier AI.The Startups segment serves fast-moving, high-growth companies that...  ...across APIs, platform products, AI/ML systems, developer workflows, and production... 
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    2 days ago
  •  ...Spellbrush, the world’s leading generative AI studio behind niji・journey , is looking for an AI Infrastructure Engineer to join us in building out end-to-end ML infrastructure to run our models on...  ...of-the-art image generation models serving over 16 million users You might... 
    Work experience placement
    Work at office
    Visa sponsorship

    Spellbrush

    San Francisco, CA
    29 days ago
  • $225k - $320k

     ...speed and accuracy. Our AI-enabled platform turns...  ...states and two countries, serving more than 125 million...  .... As the Applied AI Engineering Lead, you will own how...  ...into core workflowsAI/ML FoundationsOwn and guide...  ...quality measurementStrong software engineering fundamentals... 
    Local area

    Peregrine Technologies

    San Francisco, CA
    1 day ago
  • $264.66k - $369.8k

     ...Washington D.C., London and Amsterdam.AI and intelligent systems are driving the...  ...those systems to life.Work with other AI engineers, software engineers and machine learning...  ...spaceNice-to-HavesExperience training and/or serving ML models in production, or fine-tuning LLMs... 
    Full time
    Work experience placement
    Local area
    Shift work

    Plaid Financial

    San Francisco, CA
    2 days ago
  •  ...Forward Deployed AI Engineer The opportunity We are looking for a Forward...  ...Deployed AI Engineer to serve as the critical bridge between...  ...are You have a strong CS or ML educational background. You hold...  ...You have a solid grounding in software engineering principles and... 
    Full time
    Shift work

    Latent Labs

    San Francisco, CA
    1 day ago
  • $286.2k - $326.7k

    Sr. Distinguished AI Engineer (Remote Eligible) At Capital One,...  ...time, our applications of AI & ML are bringing humanity and...  ...capabilities to reimagine how we serve our customers and businesses who...  ...test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area
    Remote work

    Capital One Financial Corporation

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, ML Serving - Rime Ai. Be the first to apply!