Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer, ML Serving - Rime Ai

Unusual Ventures

Job Description

Job Description

Rime is a foundation modeling company that builds voice AI for enterprises running customer experiences at scale. Our models are purpose-built for high-volume conversational deployments, engineered for the accuracy, performance, and deployment flexibility that production environments actually demand.

We started from a different premise than the rest of the field: build voice AI for human connection, not slop. Before we trained a single model, we built our own corpus: full-duplex, studio-quality conversational speech of normal people, recorded and annotated by linguists. It's why our models are unparalleled in naturalism, and it's why enterprises pick Rime when pilots need to make it to production.

  Role Overview

We're hiring a Software Engineer to own the serving infrastructure that connects Rime's inference engines to the world. This role sits at the intersection of ML systems and cloud infrastructure — you'll work directly on model inference and cloud infrastructure to build, harden, and scale the systems that stream voice at real-time latency. As Rime moves toward its next-generation architecture, you'll be a core architect of how our models get served.

 

What You'll Own

  • Architecture and implementation of Rime's TTS serving infrastructure, from GPU-backed inference engines to the API surface.

  • Model optimization from a single-node to disaggregated fleet serving.

  • Compatibility with different NVIDIA hardwares from Hopper to Blackwell and beyond for on-prem and cloud deployments.

  • Continuous integration and deployment workflows for the model serving pipeline.

  • Site reliability: on-call rotation, monitoring, alerting, and observability across the serving stack.

  • Resource provision, cost management across our GPU fleet.

What We're Looking For

  • Hands-on experience with real-time multinode ML serving infrastructure — ML serving framework experience: NVIDIA Dynamo/Triton, vLLM, SGLang, or equivalent.

  • Experience with distributed or disaggregated model serving (Tensor Parallel, Pipeline Parallel, or equivalent).

  • Strong cloud infrastructure fundamentals: Linux internals, networking, containerization (Docker, Kubernetes).

  • IaC experience — Terraform, Packer, or comparable. You should have opinions about how to do this right.

  • On-call is part of the job. You treat production reliability as a shared responsibility.

Nice to Have

  • Experience with multinode training (DDP, FSDP, etc.).

  • Experience with gRPC or other bidirectional binary streaming protocols.

  • Experience with audio streaming and related technologies (WebRTC, WebSockets, etc.).

  • Experience with a multilingual monorepo where you pick the best language out of merit more than personal experience.

  • Experience with multi-cloud infrastructures (AWS, GCP, OCI, etc.).

  • Comfort with configuration management tooling (Ansible, Chef, Puppet, or similar).

  • SRE, DevOps, or platform engineering background at a startup.

  • Experience at an early-stage company.

Why Join Rime

  • Build the serving infrastructure behind a category-defining voice AI company from the ground up.

  • You will bring in experience that no one else currently has at the company: you can help us set the vision.

  • Direct collaboration with the inference, platform, and ML teams — no handoff culture.

  • The systems you build determine what experiences our customers can deploy at scale.

  • Meaningful equity upside at an early stage.

  • High ownership, high standards, low bureaucracy.

  • SF / Bay Area.

At Rime, we...

  • Are outliers

  • Cut through the hype to focus on the craft

  • Move fast with agency and freedom

  • Maintain a growth mindset, finding joy in the struggle

  • Do the right things, knowing that it'll lead to making money

If that sounds like you too, you'll be a great fit for Rime!

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Software Engineer, ML Serving - Rime Ai in San Francisco, CA vacancy
  •  ...Job Description Rime is a foundation modeling...  ...that builds voice AI for enterprises...  ...conversational deployments, engineered for the accuracy,...  ...a generalist Software Engineer to work across...  ...model serving, bidirectional streaming...  ...Comfort working within ML systems Why Join... 
    Suggested

    Unusual Ventures

    San Francisco, CA
    4 days ago
  •  ...Description Job Description Rime is a foundation...  ...that builds voice AI for enterprises running...  ...conversational deployments, engineered for the accuracy,...  ...Overview We're hiring a Software Engineer to own Rime's...  ...building solutions that serve both the user and the business... 
    Suggested
    Casual work

    Unusual Ventures

    San Francisco, CA
    4 days ago
  • $230k - $385k

     ...the TeamThe Cooperative AI team is scaling OpenAI with...  .... We iterate fast, and engineer for reliable long-term...  ...re looking for a Backend Software Engineer to help...  ...and backend services that serve as the foundation for how...  ...timeCuriosity about AI/ML and excitement to work alongside... 
    Suggested
    Internship
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    1 day ago
  • $162.15k - $259.44k

     ...ecosystem of devices and cloud software. Like our products, we work...  ...where you matter.Senior Software Engineer, AI Platform (Corporate AI Team)...  ..., OpenAI, AWS Bedrock, Azure ML) for fit, performance, and...  ...that reflect the communities we serve.Studies have shown that women... 
    Suggested
    Work experience placement
    Work at office

    Axon

    San Francisco, CA
    22 hours ago
  •  ...ourselves — real-time GPU serving, high-throughput batch...  ...for Generative AI at DoorDash, leading the...  ...serving and inference engines, fine-tuning and training...  ...the technical bar for — ML engineers, product engineers...  ...industry experience in software engineeringDeep backend... 
    Suggested
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    3 days ago
  •  ...Software Engineer, Applied AICHAOS Industries is redefining modern defense with...  ...motivated, mission-oriented Applied AI Engineer to help develop, integrate, and deploy AI/ML-powered capabilities across our...  ...for acquisition, model serving, monitoring, lifecycle management... 
    Work experience placement
    Casual work
    Work at office
    Relocation package
    2 days per week

    Chaos Industries

    San Francisco, CA
    1 day ago
  •  ...for agentic team chat—a workspace where AI agents and humans collaborate as peers. Our...  ...Opportunity We’re looking for an AI/ML Engineer with a strong foundation in large language...  ...data ingestion and model integration to serving and monitoring—and understand what it takes... 
    Local area
    Immediate start
    Flexible hours

    Glue

    San Francisco, CA
    4 days ago
  •  ...AI Software EngineerBishop Fox is the leading authority in offensive...  ...You AreThis isn't just another engineering role. You'll be joining what'...  ...language models and cutting-edge AI/ML techniquesCreate systems that...  ...your creations to serve Fortune 100 clients with enterprise... 
    Work at office
    Local area
    Remote work

    Bishop Fox

    San Francisco, CA
    1 day ago
  • $200k - $400k

     ...DecagonDecagon is the leading conversational AI platform empowering every brand to...  ...that power Decagon: networking, data, ML serving, developer platform, and real-time voice...  ...the RoleWe're looking for a Senior Software Engineer to help build and evolve our internal developer... 
    Full time
    Work at office
    Local area

    decagon

    San Francisco, CA
    2 days ago
  • $180k - $201k

     ...AreNotion is the collaborative AI workspace where teams...  ...a team of talented engineers focused on making speed...  ...You’ve worked on LLM, ML platform, data, or...  ...years of experience as a Software EngineerExperience with...  ...building MLOps and ML serving infrastructure.Notion is... 
    Local area

    Notion Labs

    San Francisco, CA
    1 day ago
  •  ...inference for the world's most dynamic AI companies, like Cursor, Notion,...  ...us and help build the platform engineers turn to to ship AI products....  ...configured and working. Self-serve infrastructure so teams build...  ...k) - Exposure to a variety of ML startups, offering unparalleled... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    17 days ago
  •  ...Python and Kubernetes Software Engineer - Data, AI/ML & Analytics Join to apply for the Python and Kubernetes Software Engineer - Data, AI/ML &...  ...Kubernetes, on developer desktops, or as web services. We serve the needs of individuals and community members as much as... 
    Full time
    Freelance
    Internship
    Local area
    Remote work
    Work from home
    Worldwide

    Canonical

    San Francisco, CA
    2 days ago
  • $198k - $280k

    About the teamThe Applied AI Engineering team is responsible for ensuring the safe and effective...  ...use cases being built on our API, serving as a critical partner in collecting...  ...you:Have 5+ years of experience as a software engineer, ML engineer or equivalent, ideally in a... 
    Work at office
    Local area
    Relocation package
    Flexible hours

    OpenAI

    San Francisco, CA
    4 days ago
  • $250k - $300k

     ...vertically integrated AI infrastructure company...  ...getting deep into the serving code when the defaults...  ...directly with customer engineering teams to tailor deployments...  ...methods across many kinds of ML models, with an...  ....Build and support the software and product features around... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  • $150k - $200k

    San Francisco, CaliforniaEngineering - AI Engineering /US Full-time Salaried /On-siteSamba is a media...  ...) serversOpen-source contributions to AI/ML projectsFamiliarity with observability...  ...one another, reflect the communities we serve, and tackle meaningful projects that make... 
    Full time
    Local area

    Samba TV

    San Francisco, CA
    1 day ago
  • $225k - $320k

     ...speed and accuracy. Our AI-enabled platform turns...  ...states and two countries, serving more than 125 million...  .... As the Applied AI Engineering Lead, you will own how...  ...into core workflowsAI/ML FoundationsOwn and guide...  ...quality measurementStrong software engineering fundamentals... 
    Local area

    Peregrine Technologies

    San Francisco, CA
    1 day ago
  • $220k - $293.33k

     ...We own the data centers, software, and applications that power today’s AI stack using sustainable technology...  ...for a Staff AI Product Engineer to drive technical...  ...Experience building AI/ML product platforms: inference...  ...or compute platforms serving ML workloads; experience... 
    Full time
    Contract work
    Flexible hours

    Nscale

    San Francisco, CA
    3 days ago
  • $264.66k - $369.8k

     ...Washington D.C., London and Amsterdam.AI and intelligent systems are driving the...  ...those systems to life.Work with other AI engineers, software engineers and machine learning...  ...spaceNice-to-HavesExperience training and/or serving ML models in production, or fine-tuning LLMs... 
    Full time
    Work experience placement
    Local area
    Shift work

    Plaid Financial

    San Francisco, CA
    2 days ago
  • $100k - $250k

     ...Solutions Engineer Specializing In Ai Infrastructure And Platforms We are seeking...  ...platforms both Hardware and Software technologies. You will collaborate...  .... Role Description Serve as the primary AI...  ...infrastructure, preferably with AI/ML/HPC workloads. ~ Proven... 
    Work experience placement
    Work at office
    Flexible hours

    SHI GmbH

    San Francisco, CA
    1 day ago
  •  ...The RoleWe are hiring a Software Engineer focused on Applied AI to build intelligent systems that directly power our underwriting and pricing engine....  ...dataBuild production-grade AI infrastructure for model serving, inference, evaluation, and monitoringCreate internal AI... 
    Work experience placement
    Work at office

    Shepherd

    San Francisco, CA
    2 days ago
  •  ...building a human-level, AI-powered tutor in your pocket...  ...app, and we now serve learners across many markets...  ...role As an AI Product Engineer at Speak, you'll play a...  ...teams, Applied ML, Product, Design, and Content...  ...backend, product-focused software engineering ~ Proficiency... 
    Live in
    Work at office
    Worldwide

    Speak LLC

    San Francisco, CA
    1 day ago
  • $110.7k - $372.9k

     ...need. Deloitte has a new AI-first effort, backed by...  ...and for the patients they serve — so that care is faster...  ...a lab. As an Agentic AI Engineer, you will design, build,...  ...AI systems across software, data, models, and cloud...  ...expect strong software/ML fundamentals plus substantial... 
    Local area
    Visa sponsorship

    Deloitte

    San Francisco, CA
    5 days ago
  •  ...a business—until generative AI. Today, AI is the number one...  ...Are:As a Snowflake Advanced AI Engineer, you will design, build, and...  ...in building and deploying AI/ML based software to a cloud environment.Bachelor...  ...approximately 791,000 people serving clients in more than 120... 
    Full time
    Work experience placement
    Live in
    Work at office
    Local area
    Shift work

    Accenture

    San Francisco, CA
    4 days ago
  • $214k - $285k

     ...About the RoleHex is an AI-powered platform for modern Data Science and...  ...Data Analytics workflows. AI Research Engineers at Hex partner with product teams...  ...If You Have:Experience getting AI/ML capabilities into production and serving real users – we're not currently looking... 
    Full time
    Work at office
    Flexible hours

    HEX

    San Francisco, CA
    1 day ago
  • $200k - $250k

     ...the Role Roger is an AI platform that frees home...  ...busywork that organizations serving these vulnerable...  ...for a Senior Applied AI Engineer to build the intelligence...  ...years of professional software engineering experience,...  ...meaningful depth in AI/ML. ~ Experience training... 
    Remote work
    Work from home

    Roger Healthcare

    San Francisco, CA
    4 days ago
  • LanceDB Inc. is seeking a Senior Solutions Engineer to blend AI/ML infrastructure expertise with strong communication skills, serving as a trusted advisor to customers and partners. You will bridge cutting-edge AI engineering with real-world business use cases in production... 

    LanceDB Inc.

    San Francisco, CA
    2 days ago
  • $13 per hour

     ...SalesforceSalesforce is the #1 AI CRM, where humans with...  ...the same founders and engineers who built the original...  ...."As a Senior/Lead AI Software Engineer, for the...  ...for distributed systems serving millions of users with...  ...improve code quality, AI/ML practices, and system designCollaborate... 
    Full time
    Immediate start

    Salesforce

    San Francisco, CA
    1 day ago
  • $99k - $232k

     ...individuals analyse client needs, implement software solutions, and provide training and...  ...part of the Data and Analytics Engineering team, you will serve as both a technical leader and a trusted advisor to clients, combining AI/ML knowledge with business acumen to design... 
    Full time
    H1b

    PwC

    San Francisco, CA
    4 days ago
  • $174.92k - $209.91k

     ...ready to query, with no engineering or maintenance required...  ...the context teams and AI systems rely on. Fivetran...  ...for a Senior R&D Software Engineer to join our fast...  ...holdingComfortable reading AI/ML research and turning...  ...company to better serve our customers, our people... 
    Full time
    Work at office
    Remote work
    Shift work

    Fivetran

    Oakland, CA
    3 days ago
  • $250.8k - $286.2k

    Senior Lead AI Engineer (AI Foundations: LLM Customization, Finetuning,...  ...time, our applications of AI & ML are bringing humanity and simplicity...  ...to reimagine how we serve our customers and businesses who...  ...test, deploy, and support AI software components including foundation... 
    Full time
    Part time
    Local area

    Capital One

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer, ML Serving - Rime Ai. Be the first to apply!