Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, ML Serving - Rime Ai

Full-time

Unusual Ventures


Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences at scale. Our models are purpose-built for high-volume conversational deployments, engineered for the accuracy, performance, and deployment flexibility that production environments actually demand.

We started from a different premise than the rest of the field: build voice AI for human connection, not slop. Before we trained a single model, we built our own corpus: full-duplex, studio-quality conversational speech of normal people, recorded and annotated by linguists. It's why our models are unparalleled in naturalism, and it's why enterprises pick Rime when pilots need to make it to production.

  Role Overview

We're hiring a Software Engineer to own the serving infrastructure that connects Rime's inference engines to the world. This role sits at the intersection of ML systems and cloud infrastructure — you'll work directly on model inference and cloud infrastructure to build, harden, and scale the systems that stream voice at real-time latency. As Rime moves toward its next-generation architecture, you'll be a core architect of how our models get served.

 

What You'll Own




  • Architecture and implementation of Rime's TTS serving infrastructure, from GPU-backed inference engines to the API surface.



  • Model optimization from a single-node to disaggregated fleet serving.



  • Compatibility with different NVIDIA hardwares from Hopper to Blackwell and beyond for on-prem and cloud deployments.



  • Continuous integration and deployment workflows for the model serving pipeline.



  • Site reliability: on-call rotation, monitoring, alerting, and observability across the serving stack.



  • Resource provision, cost management across our GPU fleet.


What We're Looking For




  • Hands-on experience with real-time multinode ML serving infrastructure — ML serving framework experience: NVIDIA Dynamo/Triton, vLLM, SGLang, or equivalent.



  • Experience with distributed or disaggregated model serving (Tensor Parallel, Pipeline Parallel, or equivalent).



  • Strong cloud infrastructure fundamentals: Linux internals, networking, containerization (Docker, Kubernetes).



  • IaC experience — Terraform, Packer, or comparable. You should have opinions about how to do this right.



  • On-call is part of the job. You treat production reliability as a shared responsibility.


Nice to Have




  • Experience with multinode training (DDP, FSDP, etc.).



  • Experience with gRPC or other bidirectional binary streaming protocols.



  • Experience with audio streaming and related technologies (WebRTC, WebSockets, etc.).



  • Experience with a multilingual monorepo where you pick the best language out of merit more than personal experience.



  • Experience with multi-cloud infrastructures (AWS, GCP, OCI, etc.).



  • Comfort with configuration management tooling (Ansible, Chef, Puppet, or similar).



  • SRE, DevOps, or platform engineering background at a startup.



  • Experience at an early-stage company.


Why Join Rime




  • Build the serving infrastructure behind a category-defining voice AI company from the ground up.



  • You will bring in experience that no one else currently has at the company: you can help us set the vision.



  • Direct collaboration with the inference, platform, and ML teams — no handoff culture.



  • The systems you build determine what experiences our customers can deploy at scale.



  • Meaningful equity upside at an early stage.



  • High ownership, high standards, low bureaucracy.



  • SF / Bay Area.


At Rime, we...




  • Are outliers



  • Cut through the hype to focus on the craft



  • Move fast with agency and freedom



  • Maintain a growth mindset, finding joy in the struggle



  • Do the right things, knowing that it'll lead to making money


If that sounds like you too, you'll be a great fit for Rime!

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Software Engineer, ML Serving - Rime Ai in San Francisco, CA vacancy
  •  ...Job Description Rime is a foundation modeling...  ...that builds voice AI for enterprises...  ...conversational deployments, engineered for the accuracy,...  ...a generalist Software Engineer to work across...  ...model serving, bidirectional streaming...  ...Comfort working within ML systems Why Join... 
    Suggested

    Unusual Ventures

    San Francisco, CA
    22 days ago
  •  ...Job Description Rime is a foundation modeling...  ...company that builds voice AI for enterprises...  ...conversational deployments, engineered for the accuracy,...  ...We're hiring a Software Engineer to own the serving infrastructure that connects...  ...the intersection of ML systems and cloud... 
    Suggested

    Unusual Ventures

    San Francisco, CA
    24 days ago
  •  ...Description Job Description Rime is a foundation...  ...that builds voice AI for enterprises running...  ...conversational deployments, engineered for the accuracy,...  ...Overview We're hiring a Software Engineer to own Rime's...  ...building solutions that serve both the user and the business... 
    Suggested
    Casual work

    Unusual Ventures

    San Francisco, CA
    22 days ago
  • $150k - $170k

     ...motivated, mission-oriented Applied AI Engineer to help develop, integrate, and deploy AI/ML-powered capabilities across our...  ...with other data scientists, software engineers, product teams, and mission...  ...for acquisition, model serving, monitoring, lifecycle management... 
    Suggested
    Work experience placement
    Casual work
    Work at office
    Relocation package
    2 days per week

    CHAOS Industries

    San Francisco, CA
    2 days ago
  •  ...for agentic team chat—a workspace where AI agents and humans collaborate as peers. Our...  ...Opportunity We’re looking for an AI/ML Engineer with a strong foundation in large language...  ...data ingestion and model integration to serving and monitoring—and understand what it takes... 
    Suggested
    Local area
    Immediate start
    Flexible hours

    Glue

    San Francisco, CA
    22 days ago
  •  ...Python and Kubernetes Software Engineer - Data, AI/ML & Analytics Join to apply for the Python and Kubernetes Software Engineer - Data, AI/ML &...  ...Kubernetes, on developer desktops, or as web services. We serve the needs of individuals and community members as much as... 
    Full time
    Freelance
    Internship
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    San Francisco, CA
    3 days ago
  • $220k - $300k

     ...entrepreneurial Founding Applied AI Engineer to join our early engineering...  ...roadmap. Mentor future AI/ML engineers as the team grows....  ...professional experience in software engineering, machine learning...  ...optimization, or model serving.li ~ ~ Experience with LLM... 
    Full time

    Weekday (YC W21)

    San Francisco, CA
    13 hours ago
  • $266k

     ...the TeamThe Cooperative AI team is scaling OpenAI with...  .... We iterate fast, and engineer for reliable long-term...  ...re looking for a Backend Software Engineer to help...  ...and backend services that serve as the foundation for how...  ...timeCuriosity about AI/ML and excitement to work alongside... 
    Internship
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    1 day ago
  • $150k - $200k

     ...'ll Do Build and deploy AI agents using modern agent SDKs...  ...from each Develop context engineering strategies—understanding how to...  ...Open-source contributions to AI/ML projects Familiarity with...  ...another, reflect the communities we serve, and tackle meaningful... 
    Local area

    Samba TV

    San Francisco, CA
    4 days ago
  • $200k - $250k

     ...the Role Roger is an AI platform that frees home...  ...busywork that organizations serving these vulnerable...  ...for a Senior Applied AI Engineer to build the intelligence...  ...years of professional software engineering experience,...  ...meaningful depth in AI/ML. ~ Experience training... 
    Remote work
    Work from home

    Roger Healthcare

    San Francisco, CA
    23 days ago
  • $162.15k - $259.44k

     ...ecosystem of devices and cloud software. Like our products, we work...  ...where you matter.Senior Software Engineer, AI Platform (Corporate AI Team)...  ..., OpenAI, AWS Bedrock, Azure ML) for fit, performance, and...  ...that reflect the communities we serve.Studies have shown that women... 
    Work experience placement
    Work at office

    Axon

    San Francisco, CA
    2 days ago
  •  ...Smart Bricks is a frontier AI lab building autonomous reasoning...  ...About The Role As an AI Engineer at Smart Bricks, you will work...  ...store quality, coverage, and serving latency Building and maintaining...  ...Have strong production ML engineering experience - model... 
    Live in

    Smart Bricks

    San Francisco, CA
    3 days ago
  •  ...sources, resolved into a single record and served to humans and AI agents alike. This role sits on a...  ...billions of them. The Role As an Applied AI Engineer on this team, you build the multi-step...  ...Must Haves 5+ years of applied AI or ML engineering, with work that reached... 
    Remote work
    Shift work

    Firmable

    San Francisco, CA
    4 days ago
  •  ...for the world's most dynamic AI companies, like Cursor, Notion...  ...us and help build the platform engineers turn to to ship AI products....  ...Develop world-class model serving stack for state-of-the-art open...  ...oriented tools Interest in ML/AI infrastructure and willingness... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • $164k

    About the roleAs a Senior Software Engineer on the AI Enablement team, you will build the internal platforms...  ...-on experience building with LLMs or ML services and shipping production-...  ...embedded in every aspect of our business, serving as a north star that guides us as we... 
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    Chime

    San Francisco, CA
    1 day ago
  • $130k - $150k

     ...Notion is the collaborative AI workspace where teams and agents...  ...About the Role As an engineer at Notion, you’ll help shape core...  ...infrastructure that supports model serving and experimentation...  ...and you’ve started exploring AI/ML through coursework, projects,... 
    Full time
    Internship
    Local area

    Notion

    San Francisco, CA
    1 day ago
  •  ...inference for the world's most dynamic AI companies, like Cursor, Notion,...  ...us and help build the platform engineers turn to to ship AI products....  ...external teams deploy and serve AI models, your focus is inward...  ...k) ~ Exposure to a variety of ML startups, offering unparalleled... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  •  ...company at the forefront of AI-native innovation. We...  ...-powered workflows engineered to scale in real-world...  ...operational environments. Serve as the technical bridge...  ...: Minimum 10 years software engineering experience...  ...with modern AI/ML platforms such as: OpenAI... 
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    San Francisco, CA
    1 day ago
  •  ...Opportunity We are looking for a Principal AI Engineer who builds things that matter. You will...  ..., LLM applications, and production ML pipelines. This is a hands‑on engineering...  ...milestones across cross‑functional teams. Serve as the primary technical escalation point... 

    H2O.ai

    San Francisco, CA
    1 day ago
  •  ...Why Blee Blee is an AI-first marketing content...  ...Role This is an AI Engineer role specializing in agent...  ...optimize applied AI/ML pipelines, work across...  ...proficiency with modern software engineering practices (...  ...~ Production ML: model serving, monitoring, A/B testing... 
    Work at office
    Remote work
    Work visa
    Flexible hours
    3 days per week

    Blee

    San Francisco, CA
    2 days ago
  •  ...The Role You'll build the AI systems that power Surf's core product — turning...  ...model performance through prompt engineering, fine-tuning, and rigorous eval-driven...  ...crypto/DeFi protocols Background in ML infrastructure, model serving, or fine-tuning You've built side... 

    Cyber Services Inc

    San Francisco, CA
    4 days ago
  • $229.9k - $262.4k

    Senior Lead AI Engineer,(MLX, Agentic AI, Gen AI platform Services) Overview...  ..., our applications of AI & ML are bringing humanity and...  ...capabilities to reimagine how we serve our customers and businesses who...  ...test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    1 day ago
  •  ...Gateway, open-weights model serving and batch inference,...  ...for Generative AI at DoorDash, with a primary...  ...This role is ideal for an engineer who enjoys building...  ...excellence.Partner closely with ML engineers, product...  ...industry experience in software engineeringStrong... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    2 days ago
  •  ...We’re seeking an Agent Engineer to design and build agentic...  ...Friendli Agent API, which serves as the core developer...  ...agent systems and making AI easy for developers to adopt...  ...+ years of experience in software engineering, preferably in backend, ML systems, or API development... 
    Full time
    Worldwide
    Flexible hours

    FriendliAI

    San Francisco, CA
    1 day ago
  • $145.4k - $188.1k

     ...personalized guidance to AI tools and a best-in-...  ...detection, and more. The ML Infrastructure team is...  ...experiences. We empower product engineering teams by providing...  ...The challenge As a Software Engineer on the ML...  ...ML model training and serving systems, feature and data... 
    Full time
    Seasonal work
    Local area

    Thumbtack

    San Francisco, CA
    1 day ago
  •  ...Spellbrush, the world’s leading generative AI studio behind niji・journey, is looking for an AI Infrastructure Engineer to join us in building out end-to-end ML infrastructure to run our models on...  ...-of-the-art image generation models serving over 16 million users. You... 
    Work at office
    Visa sponsorship

    Visa Hunt

    San Francisco, CA
    2 days ago
  •  ...Forward Deployed Principal Engineer We are seeking a...  ...enterprise-grade Generative AI solutions within a...  ...requirements. This person will serve as a recognized GenAI...  ...Qualifications 7+ years of software engineering experience,...  ...significant depth in AI/ML, data-intensive systems,... 

    Ontrac Solutions Inc

    San Francisco, CA
    4 days ago
  • $229.9k - $262.4k

     ...Overview Senior Lead AI Engineer At Capital One, we are creating...  ...time, our applications of AI & ML are bringing humanity and simplicity...  ...to reimagine how we serve our customers and businesses who...  ...test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    2 days ago
  • $154.39k - $247.02k

     ...ecosystem of devices and cloud software. Like our products, we...  ...Axon’s Corporate AI Team sits within Business...  ...for an AI Infrastructure Engineer to help move internal AI...  ...models or develop novel ML techniques. You should understand...  ...the communities we serve. Studies have shown... 
    Work experience placement

    Axon

    San Francisco, CA
    4 days ago
  •  ...Description Job Description AI Backend Engineer, Work From Home   We are...  ...operate backend systems that serve AI-powered insurance...  ...boundaries between backend systems, ML components, and product APIs....  ...-Source Models, Python, SQL, Software Engineer, Software Developer,... 
    Remote work
    Work from home

    Parallel Partners

    San Francisco, CA
    13 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, ML Serving - Rime Ai. Be the first to apply!