Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Full-Stack ML Systems Engineer

Success Matcher Recruitment

Job Description

Job Description

About Us

We are an elite, research-backed AI infrastructure startup building the defining workflow intelligence and policy management layer for enterprise multi-agent applications. While the market has flooded with tools to build simple AI agents, nobody has cracked the infrastructure gap required to scale them efficiently, safely, and cost-effectively in production. We bridge this exact gap by applying deep systems programming, software-defined networking, and OS-level primitives to AI workflow orchestration.

We recently closed a $5M Seed round and operate as a flat, hyper-technical team of three. Our leadership includes:

  • The CEO: A 3x deep-tech founder (PhD, ex-Bell Labs, ex-Qualcomm) who successfully scaled his last venture from Seed to Series B.
  • The Co-Founder: A Regents Chair Professor at UT Austin and world-renowned ML systems researcher with a pedigree spanning CMU and Stanford.

Our core technology originates from a top-tier academic research group. This is not another API wrapper—this is a foundational infrastructure play designed for the next frontier of enterprise software.

What You'll Be Doing
  • Bridge Research & Production: Work directly alongside our Chief Architect and CEO to translate cutting-edge ML systems research into an enterprise platform that product developers love to use.
  • Build the Orchestration Fabric: Design and implement full-stack ML systems for multi-agent workflow orchestration, merging deep backend systems capability with clean, satisfying developer interfaces.
  • Architect the Scaling Layer: Develop the policy-driven scaling mechanics that govern multi-agent interactions, resolving the infrastructure bottlenecks that limit agentic autonomy at scale.
  • Own the Developer Experience: Take research-grade ideas and make them legible by authoring technical documentation, code examples, sample applications, and intuitive onboarding flows.
  • Systems Integration: Collaborate across distributed systems design, networking configurations, and asynchronous event streams applied directly to runtime AI operations.
What We're Looking For Experience & Seniority
  • 5+ years of production experience engineering ML systems, OR a PhD from a top-tier institution in a relevant ML Systems/Distributed Computing field.
Core Technical Competencies (Critical)
  • Hands-on Agentic Frameworks: Proven experience building and deploying multi-agent applications to live production environments using frameworks like LangGraph, LangChain, CrewAI, AutoGen, Semantic Kernel, ADK, or custom orchestration systems.
  • Production Scaling & Compute: Direct exposure handling scale on either the orchestration side (managing complex state, cyclical routing, and memory loops) or the infrastructure side (scaling AI compute fabrics, asynchronous jobs, queues, and container clusters).
  • Model-Serving Platforms: Practical production experience deploying and optimizing models via frameworks such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo.
  • Full-Stack Implementation: Strong Python background combined with full-stack capabilities, including constructing front-end visual telemetry dashboards (using Grafana, Tableau, Streamlit, or similar).
  • Open Source Focus: Experience utilizing, developing upon, or contributing directly to open-source software repositories.
Mindset & Soft Skills
  • Enterprise Product Instincts: The ability to look at powerful backend capabilities and figure out how to make them readable, structured, and satisfying for enterprise end-users.
  • Thriving in Ambiguity: An early-stage startup mentality—comfortable shipping fast, iterative code within dynamic, shifting, and initially ill-defined parameters.
  • Clear Communication: Strong written skills to help transform complex technical primitives into clear developer documentation and clean product language.
  • A background as a Solutions Architect or Forward Deployed Engineer is a significant plus.

    Why You Should Join Us
    • Solve an Unsolved Problem: Nobody has cracked the gap between building a toy agent app and scaling it across a global enterprise. You will build technology that fundamentally does not exist anywhere else.
    • Unmatched Pedigree: Learn from and build with a world-class founding team combining serial entrepreneurship with elite academic systems research.
    • Founding-Level Impact: As Employee #4, your code dictates the core blueprint of the platform. You get a direct up-to-1% equity block in a venture backed by a highly technical $5M validation check.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Full-Stack ML Systems Engineer in Sunnyvale, CA vacancy
  • $174.9k - $261.3k

     ...the world!The Data Labeling Engineering team designs, builds, and operates...  ..., data engineering, and AI/ML, defining the strategies,...  ...and consumers.We own a modern full‑stack architecture including TypeScript...  ..., and direct impact on systems that unblock the next generation... 
    Fullstack
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $90.1k - $191.8k

     ...the world!The Data Labeling Engineering team designs, builds, and operates...  ..., data engineering, and ML, defining labeling strategies...  ...autonomous features.We own a modern full‑stack architecture including...  ...leadership, and work directly on systems that unblock the next... 
    Fullstack
    Full time
    Work experience placement
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • Google in Sunnyvale seeks software engineers to develop next-generation AI and large-scale systems. You will contribute across the full stack, from model deployment to distributed computing...  ...on design reviews, implement scalable ML solutions, and help build products that... 
    Fullstack

    Google

    Sunnyvale, CA
    12 hours ago
  • $224k - $356.5k

     ...phase, we are building agentic systems that can reason about, build,...  ...creating the meta-layer of modern ML: the agents, tooling,...  ...We are looking for exceptional engineers who are passionate about the idea...  ...protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...recommendation models that shape user experience, content ecosystem, and platform security. You’ll deliver end-to-end ML solutions and own full-stack systems, collaborating with cross-functional teams to grow TikTok in strategic markets. We value bold ideas and rapid... 
    Fullstack

    TikTok

    San Jose, CA
    1 day ago
  • $207k - $300k

     ...generated content through advanced context engineering and agentic feedback loops to identify...  ...improve the performance of the AIGC stack.Resolve key system-level bottlenecks in recommending and...  ...to take on new problems across the full-stack as we continue to push technology... 
    Fullstack

    Google

    Mountain View, CA
    12 hours ago
  • $207k - $300k

     ...years of experience leading ML design and optimizing ML infrastructure...  ...:Master’s degree or PhD in Engineering, Computer Science, or a...  ...computing, large-scale system design, networking and data storage...  ...on new problems across the full-stack as we continue to push technology... 
    Fullstack

    Google

    Mountain View, CA
    1 day ago
  •  ...innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product teams to... 
    Fullstack

    Realmlabs

    Sunnyvale, CA
    1 day ago
  •  ...with a group of highly skilled engineers and scientists to bring new...  ...Designing and implementing AI/ML systems to solve real-world problems...  ...systems into a larger software stack Minimum Qualifications MS/PhD...  ..., etc.) Familiarity with the full robotics stack (controls,... 
    Fullstack

    Apple

    Cupertino, CA
    3 days ago
  • $150.4k - $277.6k

     ...with a group of highly skilled engineers and scientists to bring new...  ...Designing and implementing AI/ML systems to solve real-world problems...  ...systems into a larger software stack Minimum Qualifications MS/PhD...  ..., etc.) Familiarity with the full robotics stack (controls,... 
    Fullstack
    Relocation

    Apple Inc.

    Cupertino, CA
    4 days ago
  • $147.4k - $272.1k

     .... Because when advertising is done right, it benefits everyone. The Ads ML Experimentation team is looking for a full‑stack ML Engineer to help shape the future of how Apple's advertising systems experiment to connect millions of global users to content from publishers... 
    Fullstack
    Relocation

    Apple Inc.

    Cupertino, CA
    12 hours ago
  • NVIDIA’s Cosmos team seeks engineers to build agentic AI-native software and tooling that accelerates...  .../PyTorch codebases, design end-to-end ML pipelines, and create self-improving...  .... We expect deep expertise in ML systems, robust software engineering, and hands-on... 

    Thomas To

    Santa Clara, CA
    2 days ago
  • $204k - $259k

     ...autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal... 

    Waymo

    Mountain View, CA
    12 hours ago
  • Rhoda is building the next generation of generalist robotic systems in Mountain View, CA. We are seeking a senior or staff-level Research Engineer or ML Systems Engineer to make the robot-learning pipeline reliable and measurable from end to end. You will own the supported... 

    RHODA

    Mountain View, CA
    3 days ago
  • NVIDIA is seeking an ML and Agentic Systems Engineer (Finance) to help build the meta-layer of modern ML: agents, tooling, pipelines, and feedback loops that accelerate model development. You will own large Python/PyTorch codebases and design end-to-end agentic workflows... 

    Nvidia Corporation in

    Santa Clara, CA
    12 hours ago
  • Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through...  ..., and real-robot evaluation. You will build validation systems, observability, and robust operating practices to distinguish model... 

    Socket.dev

    Mountain View, CA
    2 days ago
  • General Motors’ Data Labeling Engineering team is building cutting‑edge...  ...that power autonomous vehicle ML models. The role sits at the intersection...  ...focusing on scalable labeling systems and foundations for foundation...  ...engineering across the stack. #J-18808-Ljbffr General... 

    General Motors

    Mountain View, CA
    2 days ago
  • $105k - $115k

     ...delivers world-class end-to-end engineering solutions by leveraging our...  ...cloud platformsAI Engineer/ ML Engineer with python or typescript...  ...delivering NLP or LLM-based systems, with knowledge of multi-...  ...teamsExperience as a backend or full stack dev preferably in Python or... 
    Fullstack
    Temporary work

    Quest Global Services

    Sunnyvale, CA
    4 days ago
  • Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model... 

    Rhoda AI

    Mountain View, CA
    4 days ago
  • $122.6k - $185k

     ...performance and scalability in AI/ML and HPC workloads.You are...  ...like you. The AWS Hardware Engineering team creates server designs...  ...are knowledgeable of the full technical stack - vertically from baremetal...  ...cloud scale and curious how systems and software decisions impact... 
    Fullstack
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $210.16k - $271.98k

    Senior Principal Systems Development EngineerHelp architect and deliver...  ...Principal Systems Development Engineer, you'll own the system-level...  ...software, and diagnostics for full hardware/software...  ...root causes across the full rack stack and its manufacturing and deployment... 
    Fullstack

    Dell Technologies

    Santa Clara, CA
    3 days ago
  • $207k - $300k

     ...infrastructure that powers our AI/ML applications. You will operate...  ...to build high-throughput systems that enable seamless creative...  ...infrastructure, including creative data engines, generation and rendering...  ...on new problems across the full-stack as we continue to push technology... 
    Fullstack

    Google

    Mountain View, CA
    1 day ago
  • $150.4k - $277.6k

     ...Senior System Engineer Apple's Compositing, Color, and Display Software organization provides...  ...you're excited about working across the full spectrum of compositing challenges—from...  ...play a central role within our graphics stack. Your work will span multiple areas of compositing... 
    Fullstack
    Worldwide
    Relocation

    Apple

    Cupertino, CA
    2 days ago
  • $165k - $242k

     ...seeking a highly skilled and motivated Systems Kernel Engineer to join the HAVOCK Team, reporting to the...  ...fixes and features that improve stack performance and reliability. This position...  ...Strong problem‑solving abilities with a full‑stack systems perspective. Preferred Contributions... 
    Fullstack
    Permanent employment
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    12 hours ago
  • $120k - $275k

     ...We are developing vertically integrated full-stack solutions from silicon to systems including hardware and software to train and run the largest ML workloads for AGI. MatX is seeking silicon micro-architects and design engineers to join our team as we create best-in-class... 
    Fullstack
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    12 hours ago
  • $169k - $338k

     ...$169000 - $338000/yearType: Full time / Regular/PermanentCompany...  ......As a Distinguished AI/ML Engineer within Walmart Global Tech's...  ...of next-generation agentic AI systems and intelligent automation solutions...  ...systems built on modern tech stacks with intelligent capacity... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    3 days ago
  • $150k - $230k

     ...advanced AI, recommendation systems, and adtech.Recognized by Fast...  ...a hands-on Machine Learning Engineer to drive the post-training...  ...learning (RL). You will own the full post-training stack — continuous pre-training (...  ...Strong data engineering for ML. You can independently... 
    Fullstack
    Full time
    Local area
    Work from home

    News Break

    Mountain View, CA
    1 day ago
  • $207k - $300k

     ...Influence and coach a distributed team of engineers.Facilitate alignment and clarity across...  ...working with embedded operating systems.3 years of experience with software design...  ...enthusiastic to take on new problems across the full-stack as we continue to push technology... 
    Fullstack
    Worldwide

    Google

    Sunnyvale, CA
    12 hours ago
  • $170.6k - $261.3k

     ...join us. About the team: The AV ML Infra team at GM builds end-...  ...the productivity of ML engineers, and drive the adoption of cutting...  .... Together, these tools and systems empower GM to tackle the complexities...  ...Overview: As a Senior AI/ML Full-Stack Engineer, you will design and... 
    Fullstack
    Full time
    Local area
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    12 hours ago
  • $120k - $275k

     ...We are developing vertically integrated full-stack solutions from silicon to systems including hardware and software to train and run the largest ML workloads for AGI.About the Role:We are looking for a Hardware Systems Engineer to join our Hardware Systems team. In this... 
    Fullstack
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Full-Stack ML Systems Engineer. Be the first to apply!