Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Full-Stack ML Systems Engineer

Success Matcher Recruitment

Job Description

Job Description

About Us

We are an elite, research-backed AI infrastructure startup building the defining workflow intelligence and policy management layer for enterprise multi-agent applications. While the market has flooded with tools to build simple AI agents, nobody has cracked the infrastructure gap required to scale them efficiently, safely, and cost-effectively in production. We bridge this exact gap by applying deep systems programming, software-defined networking, and OS-level primitives to AI workflow orchestration.

We recently closed a $5M Seed round and operate as a flat, hyper-technical team of three. Our leadership includes:

  • The CEO: A 3x deep-tech founder (PhD, ex-Bell Labs, ex-Qualcomm) who successfully scaled his last venture from Seed to Series B.
  • The Co-Founder: A Regents Chair Professor at UT Austin and world-renowned ML systems researcher with a pedigree spanning CMU and Stanford.

Our core technology originates from a top-tier academic research group. This is not another API wrapper—this is a foundational infrastructure play designed for the next frontier of enterprise software.

What You'll Be Doing
  • Bridge Research & Production: Work directly alongside our Chief Architect and CEO to translate cutting-edge ML systems research into an enterprise platform that product developers love to use.
  • Build the Orchestration Fabric: Design and implement full-stack ML systems for multi-agent workflow orchestration, merging deep backend systems capability with clean, satisfying developer interfaces.
  • Architect the Scaling Layer: Develop the policy-driven scaling mechanics that govern multi-agent interactions, resolving the infrastructure bottlenecks that limit agentic autonomy at scale.
  • Own the Developer Experience: Take research-grade ideas and make them legible by authoring technical documentation, code examples, sample applications, and intuitive onboarding flows.
  • Systems Integration: Collaborate across distributed systems design, networking configurations, and asynchronous event streams applied directly to runtime AI operations.
What We're Looking For Experience & Seniority
  • 5+ years of production experience engineering ML systems, OR a PhD from a top-tier institution in a relevant ML Systems/Distributed Computing field.
Core Technical Competencies (Critical)
  • Hands-on Agentic Frameworks: Proven experience building and deploying multi-agent applications to live production environments using frameworks like LangGraph, LangChain, CrewAI, AutoGen, Semantic Kernel, ADK, or custom orchestration systems.
  • Production Scaling & Compute: Direct exposure handling scale on either the orchestration side (managing complex state, cyclical routing, and memory loops) or the infrastructure side (scaling AI compute fabrics, asynchronous jobs, queues, and container clusters).
  • Model-Serving Platforms: Practical production experience deploying and optimizing models via frameworks such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo.
  • Full-Stack Implementation: Strong Python background combined with full-stack capabilities, including constructing front-end visual telemetry dashboards (using Grafana, Tableau, Streamlit, or similar).
  • Open Source Focus: Experience utilizing, developing upon, or contributing directly to open-source software repositories.
Mindset & Soft Skills
  • Enterprise Product Instincts: The ability to look at powerful backend capabilities and figure out how to make them readable, structured, and satisfying for enterprise end-users.
  • Thriving in Ambiguity: An early-stage startup mentality—comfortable shipping fast, iterative code within dynamic, shifting, and initially ill-defined parameters.
  • Clear Communication: Strong written skills to help transform complex technical primitives into clear developer documentation and clean product language.
  • A background as a Solutions Architect or Forward Deployed Engineer is a significant plus.

    Why You Should Join Us
    • Solve an Unsolved Problem: Nobody has cracked the gap between building a toy agent app and scaling it across a global enterprise. You will build technology that fundamentally does not exist anywhere else.
    • Unmatched Pedigree: Learn from and build with a world-class founding team combining serial entrepreneurship with elite academic systems research.
    • Founding-Level Impact: As Employee #4, your code dictates the core blueprint of the platform. You get a direct up-to-1% equity block in a venture backed by a highly technical $5M validation check.
Vacancy posted 24 days ago
Similar jobs that could be interesting for youBased on the Full-Stack ML Systems Engineer in Sunnyvale, CA vacancy
  • $174.9k - $261.3k

     ...the world!The Data Labeling Engineering team designs, builds, and operates...  ..., data engineering, and AI/ML, defining the strategies,...  ...and consumers.We own a modern full‑stack architecture including TypeScript...  ..., and direct impact on systems that unblock the next generation... 
    Fullstack
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • $90.1k - $191.8k

     ...the world!The Data Labeling Engineering team designs, builds, and operates...  ..., data engineering, and ML, defining labeling strategies...  ...autonomous features.We own a modern full‑stack architecture including...  ...leadership, and work directly on systems that unblock the next... 
    Fullstack
    Full time
    Work experience placement
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  •  ...this exact gap by applying deep systems programming, software-defined...  ...UT Austin and world-renowned ML systems researcher with a...  ...Fabric: Design and implement full-stack ML systems for multi-agent workflow...  ...of production experience engineering ML systems, OR a PhD from a... 
    Fullstack
    Shift work

    Success Matcher Recruitment

    Sunnyvale, CA
    2 days ago
  • General Motors is seeking a Staff ML Systems Engineer to help our Data Labeling Engineering team advance self-driving vehicle capabilities...  ...The role emphasizes cloud-scale, distributed systems, modern full-stack development, and collaboration with cross-disciplinary... 
    Fullstack
    Remote job

    General Motors

    Sunnyvale, CA
    3 days ago
  • Google in Sunnyvale seeks software engineers to develop next-generation AI and large-scale systems. You will contribute across the full stack, from model deployment to distributed computing...  ...on design reviews, implement scalable ML solutions, and help build products that... 
    Fullstack

    Google

    Sunnyvale, CA
    10 hours ago
  • $224k - $356.5k

     ...phase, we are building agentic systems that can reason about, build,...  ...creating the meta-layer of modern ML: the agents, tooling,...  ...We are looking for exceptional engineers who are passionate about the idea...  ...protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...recommendation models that shape user experience, content ecosystem, and platform security. You’ll deliver end-to-end ML solutions and own full-stack systems, collaborating with cross-functional teams to grow TikTok in strategic markets. We value bold ideas and rapid... 
    Fullstack

    TikTok

    San Jose, CA
    1 day ago
  • $207k - $300k

     ...generated content through advanced context engineering and agentic feedback loops to identify...  ...improve the performance of the AIGC stack.Resolve key system-level bottlenecks in recommending and...  ...to take on new problems across the full-stack as we continue to push technology... 
    Fullstack

    Google

    Mountain View, CA
    10 hours ago
  • $207k - $300k

     ...years of experience leading ML design and optimizing ML infrastructure...  ...:Master’s degree or PhD in Engineering, Computer Science, or a...  ...computing, large-scale system design, networking and data storage...  ...on new problems across the full-stack as we continue to push technology... 
    Fullstack

    Google

    Mountain View, CA
    1 day ago
  •  ...innovative AI startup is seeking a Founding ML Infrastructure Engineer to take charge of deploying and optimizing production-grade LLM systems. In this core role, you will be responsible for building and managing a full ML serving stack, working closely with product teams to... 
    Fullstack

    Realmlabs

    Sunnyvale, CA
    1 day ago
  • Apple Inc. in Cupertino, California, is seeking a full-stack ML Engineer to enhance its advertising systems. The ideal candidate will design intuitive user interfaces, partner with cross-functional teams, and build production-ready RAG Machine Learning models. Required... 
    Fullstack

    Apple Inc.

    Cupertino, CA
    5 days ago
  •  ...with a group of highly skilled engineers and scientists to bring new...  ...Designing and implementing AI/ML systems to solve real-world problems...  ...systems into a larger software stack Minimum Qualifications MS/PhD...  ..., etc.) Familiarity with the full robotics stack (controls,... 
    Fullstack

    Apple

    Cupertino, CA
    3 days ago
  •  ...position is for a self-motivated robotics engineer at Apple, where you will collaborate...  ...background Strong proficiency with modern ML approaches such as LLMs, VLMs, and...  ...robotic algorithms Familiarity with the full robotics stack including controls, perception, and motion... 
    Fullstack

    Apple Inc.

    Cupertino, CA
    3 days ago
  • $150.4k - $277.6k

     ...with a group of highly skilled engineers and scientists to bring new...  ...Designing and implementing AI/ML systems to solve real-world problems...  ...systems into a larger software stack Minimum Qualifications MS/PhD...  ..., etc.) Familiarity with the full robotics stack (controls,... 
    Fullstack
    Relocation

    Apple Inc.

    Cupertino, CA
    4 days ago
  • $147.4k - $272.1k

     .... Because when advertising is done right, it benefits everyone. The Ads ML Experimentation team is looking for a full‑stack ML Engineer to help shape the future of how Apple's advertising systems experiment to connect millions of global users to content from publishers... 
    Fullstack
    Relocation

    Apple Inc.

    Cupertino, CA
    10 hours ago
  • NVIDIA’s Cosmos team seeks engineers to build agentic AI-native software and tooling that accelerates...  .../PyTorch codebases, design end-to-end ML pipelines, and create self-improving...  .... We expect deep expertise in ML systems, robust software engineering, and hands-on... 

    Thomas To

    Santa Clara, CA
    2 days ago
  • $204k - $259k

     ...autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal... 

    Waymo

    Mountain View, CA
    10 hours ago
  • Rhoda is building the next generation of generalist robotic systems in Mountain View, CA. We are seeking a senior or staff-level Research Engineer or ML Systems Engineer to make the robot-learning pipeline reliable and measurable from end to end. You will own the supported... 

    RHODA

    Mountain View, CA
    3 days ago
  • General Motors’ Data Labeling Engineering team is building cutting‑edge...  ...that power autonomous vehicle ML models. The role sits at the intersection...  ...focusing on scalable labeling systems and foundations for foundation...  ...engineering across the stack. #J-18808-Ljbffr General... 

    General Motors

    Mountain View, CA
    2 days ago
  • NVIDIA is seeking an ML and Agentic Systems Engineer (Finance) to help build the meta-layer of modern ML: agents, tooling, pipelines, and feedback loops that accelerate model development. You will own large Python/PyTorch codebases and design end-to-end agentic workflows... 

    Nvidia Corporation in

    Santa Clara, CA
    10 hours ago
  • Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through...  ..., and real-robot evaluation. You will build validation systems, observability, and robust operating practices to distinguish model... 

    Socket.dev

    Mountain View, CA
    2 days ago
  • NVIDIA Gruppe is seeking a Senior Engineer in Santa Clara, CA, to join the Cosmos team. This role focuses on creating AI-native systems that enhance the efficiency of machine learning workflows. Candidates should have extensive Python and PyTorch experience, along with... 

    NVIDIA Gruppe

    Santa Clara, CA
    5 days ago
  • $105k - $115k

     ...delivers world-class end-to-end engineering solutions by leveraging our...  ...cloud platformsAI Engineer/ ML Engineer with python or typescript...  ...delivering NLP or LLM-based systems, with knowledge of multi-...  ...teamsExperience as a backend or full stack dev preferably in Python or... 
    Fullstack
    Temporary work

    Quest Global Services

    Sunnyvale, CA
    4 days ago
  • Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model... 

    Rhoda AI

    Mountain View, CA
    4 days ago
  • $122.6k - $185k

     ...performance and scalability in AI/ML and HPC workloads.You are...  ...like you. The AWS Hardware Engineering team creates server designs...  ...are knowledgeable of the full technical stack - vertically from baremetal...  ...cloud scale and curious how systems and software decisions impact... 
    Fullstack
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $210.16k - $271.98k

    Senior Principal Systems Development EngineerHelp architect and deliver...  ...Principal Systems Development Engineer, you'll own the system-level...  ...software, and diagnostics for full hardware/software...  ...root causes across the full rack stack and its manufacturing and deployment... 
    Fullstack

    Dell Technologies

    Santa Clara, CA
    3 days ago
  • $207k - $300k

     ...infrastructure that powers our AI/ML applications. You will operate...  ...to build high-throughput systems that enable seamless creative...  ...infrastructure, including creative data engines, generation and rendering...  ...on new problems across the full-stack as we continue to push technology... 
    Fullstack

    Google

    Mountain View, CA
    1 day ago
  • $150.4k - $277.6k

     ...Senior System Engineer Apple's Compositing, Color, and Display Software organization provides...  ...you're excited about working across the full spectrum of compositing challenges—from...  ...play a central role within our graphics stack. Your work will span multiple areas of compositing... 
    Fullstack
    Worldwide
    Relocation

    Apple

    Cupertino, CA
    2 days ago
  • $50 - $70 per hour

    *Description* Hybrid Systems Engineer/DC (AI & Infrastructure Automation) We are seeking a highly motivated Hybrid Systems Engineer...  ...across North America, Europe and Asia. As an industry leader in Full-Stack Technology Services, Talent Services, and real-world... 
    Fullstack
    Contract work
    Temporary work

    TEKsystems

    Santa Clara, CA
    5 days ago
  • $165k - $242k

     ...seeking a highly skilled and motivated Systems Kernel Engineer to join the HAVOCK Team, reporting to the...  ...fixes and features that improve stack performance and reliability. This position...  ...Strong problem‑solving abilities with a full‑stack systems perspective. Preferred Contributions... 
    Fullstack
    Permanent employment
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Full-Stack ML Systems Engineer. Be the first to apply!