Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Systems Engineer

$150k - $350k

ChipAgents

About ChipAgents ChipAgents is redefining the future of chip design and verification with agentic AI workflows. Our platform leverages cutting‑edge generative AI to assist engineers in RTL design, simulation, and verification, dramatically accelerating chip development. Founded by experts in AI and semiconductor engineering, we partner with top semiconductor firms, cloud providers, and innovative startups to build intelligent AI agents. The company is a Series A company backed by tier‑1 VC firms. ChipAgents is deployed in production to companies that have shipped 16B chips. Position Overview We are seeking an ML Systems Engineer to optimize the performance and efficiency of large language model inference powering our agentic AI platform. This is a technical role focused on low‑level systems optimization. You will implement performance optimizations, build evaluation harnesses, and architect multi‑node clusters for training and inference that push the limits of LLM throughput and latency. Your work will directly impact the responsiveness and cost‑efficiency of AI agents used by leading semiconductor companies to design chips. Key Responsibilities Design, deploy, and optimize LLM inference systems across multi‑node clusters, maximizing throughput and minimizing latency for production workloads. Implement and benchmark concrete inference optimizations. Profile and analyze inference bottlenecks at the systems level—from GPU kernel execution to memory bandwidth constraints. Build robust evaluation harnesses and benchmarking frameworks that measure accuracy, throughput, latency, and resource utilization across various parallelism strategies. Collaborate with research scientists to integrate new model architectures and optimizations into production inference infrastructure. Investigate and apply emerging techniques from research papers and open‑source projects to continuously improve inference performance. Qualifications B.S., M.S., or PhD in Computer Science, Electrical Engineering, or related field (or equivalent experience). Experience with large‑scale ML systems, GPU computing, or high‑performance inference optimization. Strong proficiency in Python and C++/CUDA; hands‑on experience with SGLang, vLLM, PyTorch, or similar inference frameworks. Deep understanding of GPU architecture, memory hierarchies, and parallel computing paradigms. Experience deploying and optimizing LLMs in production: model serving, batching strategies, distributed inference, or quantization. Strong systems‑level debugging and profiling skills; comfort working at multiple layers of the stack from CUDA kernels to application logic. Familiarity with distributed computing frameworks (Ray, multi‑node training/inference) is a plus. Self‑directed problem solver who is interested in working on ambitious optimization challenges. Why Join Us Work on cutting‑edge LLM inference optimization problems with real‑world production impact. Access to substantial GPU compute resources for experimentation and benchmarking. Collaborate with a world‑class team spanning AI research, systems engineering, and EDA. Shape the performance characteristics of AI systems used by leading semiconductor companies. What we offer $150K/yr – $350K/yr + Offers Equity. We are open to discuss above‑scale compensation with exceptional candidates on a case‑by‑case basis. Unlimited PTO and full benefits (medical, vision, dental, 401k). Two engineering‑centric offices with free parking, private gym, and free lunch, drinks and snacks. #J-18808-Ljbffr

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the ML Systems Engineer in San Jose, CA vacancy
  • $224k - $356.5k

    ML and Agentic Systems Engineer (Finance) At NVIDIA, we're not just building the future, we're generating it! Our Cosmos team is pushing the boundaries of multimodal AI, simulation, and world models. As we enter the next phase, we are building agentic systems that can reason... 
    Suggested

    Nvidia Corporation in

    Santa Clara, CA
    2 days ago
  • NVIDIA is seeking an ML and Agentic Systems Engineer (Finance) to help build the meta-layer of modern ML: agents, tooling, pipelines, and feedback loops that accelerate model development. You will own large Python/PyTorch codebases and design end-to-end agentic workflows... 
    Suggested

    Nvidia Corporation in

    Santa Clara, CA
    2 days ago
  • $174.9k - $261.3k

     ...and understand the world!The Data Labeling Engineering team designs, builds, and operates hybrid...  ...engineering, data engineering, and AI/ML, defining the strategies, tooling, and quality...  ...leadership, and direct impact on systems that unblock the next generation of AV capabilities... 
    Suggested
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  • $90.1k - $191.8k

     ...and understand the world!The Data Labeling Engineering team designs, builds, and operates high‑...  ...engineering, data engineering, and ML, defining labeling strategies, tooling, and...  ...technical leadership, and work directly on systems that unblock the next generation of AV models... 
    Suggested
    Full time
    Work experience placement
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    4 days ago
  • GlobalFoundries in the United States is seeking a Staff AI/ML Systems Engineer to drive workload‑driven architecture decisions across hardware and software boundaries. You will define how AI/ML workloads are studied, modeled, and optimized, and align HW/SW teams while serving... 
    Suggested

    GlobalFoundries

    Santa Clara, CA
    3 days ago
  •  ...ML Systems Engineer Role Overview: We're seeking an experienced engineer to build our ML data infrastructure platform. You'll create the systems and tools that enable efficient data preparation, feature engineering, and dataset management for machine learning. This... 
    Local area

    My3Tech Inc

    Sunnyvale, CA
    12 hours ago
  • $142.2k - $213.2k

     ...Technologies, Inc.Job AreaEngineering Group, Engineering Group Multimedia SystemsGeneral...  ...The successful candidate will work with systems, software, and integration/test engineers...  ...for single and multi-sensor fusion using ML/AI techniques. Typical development flow uses... 
    Work experience placement
    Work from home

    Qualcomm

    Santa Clara, CA
    4 days ago
  •  ...ML Systems Engineer Role OverviewWe’re seeking an experienced engineer to build our ML data infrastructure platform. You’ll create the systems and tools that enable efficient data preparation, feature engineering, and dataset management for machine learning. This role... 

    My3Tech Inc

    Sunnyvale, CA
    4 days ago
  • $204k - $259k

     ...autonomous driving technology company is looking for an experienced engineer to improve compute performance in machine learning systems. This hybrid role involves collaboration with a world-class ML team and requires strong expertise in ML software or systems. The ideal... 

    Waymo

    Mountain View, CA
    2 days ago
  • Rhoda is building the next generation of generalist robotic systems in Mountain View, CA. We are seeking a senior or staff-level Research Engineer or ML Systems Engineer to make the robot-learning pipeline reliable and measurable from end to end. You will own the supported... 

    RHODA

    Mountain View, CA
    5 days ago
  • General Motors’ Data Labeling Engineering team is building cutting‑edge labeling tools and pipelines that power autonomous vehicle ML models. The role sits at the intersection of software...  ...and ML, focusing on scalable labeling systems and foundations for foundation‑model... 

    General Motors

    Mountain View, CA
    4 days ago
  • $144.7k - $261.3k

     ...challenges for autonomous vehicle development. We engineer high-performance tools that identify top-...  ...models and partner with data-intensive ML teams to drive rapid innovation. Why Join...  ...of next-generation autonomous systems. About the Role As a Senior Engineer in the... 
    Local area
    Remote work
    Work from home
    Flexible hours

    General Motors

    Sunnyvale, CA
    2 days ago
  • General Motors in the United States seeks a Senior Engineer on the Embodied AI Scaling Foundations team to measure and visualize autonomous vehicle model performance across the stack. You will design, implement, and iterate evaluation and introspection tools, collaborate... 
    Remote job
    Relocation package

    General Motors

    Sunnyvale, CA
    2 days ago
  • Rhoda AI is hiring a Senior/Staff-level Research Engineer to ensure our robot-learning pipeline is reliable from data collection through...  ..., and real-robot evaluation. You will build validation systems, observability, and robust operating practices to distinguish model... 

    Socket.dev

    Mountain View, CA
    4 days ago
  • Rhoda AI in Mountain View is seeking a Staff / Principal ML Training Systems Engineer to lead the performance of large-scale multimodal training systems. This role involves improving training efficiency and collaborating closely with research teams to accelerate model... 

    Rhoda AI

    Mountain View, CA
    6 days ago
  • Apple in Cupertino is seeking a robotics engineer to design and implement AI/ML systems for real-world products. You will work with a team of engineers and scientists to bring new experiences to Apple hardware and software. The role requires a strong foundation in robotics... 

    Apple

    Cupertino, CA
    5 days ago
  • $125k - $165k

     ...understand the world.  The  Data Labeling Engineering team designs, builds, and operates hybrid...  ...,  data engineering , and  AI/ML , defining the strategies, tooling, and...  ...machine learning integrations, and quality systems used by labelers, ML engineers, and operations... 
    Full time
    Internship
    Work at office
    Local area
    Work from home
    Relocation package

    General Motors

    Sunnyvale, CA
    1 day ago
  • $169k - $338k

     ...Segment: Home OfficePosition Summary...As a Distinguished AI/ML Engineer within Walmart Global Tech's Site Reliability Engineering organization...  ...lead the technical development of next-generation agentic AI systems and intelligent automation solutions that ensure mission-... 
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    5 days ago
  •  ...We bridge this exact gap by applying deep systems programming, software-defined networking,...  ...at UT Austin and world-renowned ML systems researcher with a pedigree spanning...  ...Seniority ~5+ years of production experience engineering ML systems, OR a PhD from a top-tier... 
    Shift work

    Success Matcher Recruitment

    Sunnyvale, CA
    14 days ago
  • $184k - $287.5k

     ...building intelligent video analytics and perception solutions for smart cities, industrial automation, and autonomous systems at scale. We seek a Senior ML Engineer to compose and deliver next-generation Metropolis solutions powered by Cosmos, NVIDIA's world foundation model... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $165.2k - $223.6k

     ...forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions...  ...of software, hardware, and machine learning systems, you'll bring expertise in low-level... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $165.2k - $223.6k

     ...team is actively seeking skilled compiler engineers to join our efforts in developing a state...  ...forefront of AWS innovation for advanced ML capabilities, powering solutions like...  ...in compiler technology and deep-learning systems software. Additionally, you'll collaborate... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  •  ...Position: Client Engineer Location: Cupertino, CA (Onsite) Duration: C2C Contract Experience: 12+ Years Job Description...  ...this email by mistake and delete this e-mail from your system. You have received this email as we have your email address shared... 
    Contract work
    Immediate start

    Syntricate Technologies

    Cupertino, CA
    4 days ago
  •  ...Machine Learning Engineer The position is for a Machine Learning Engineer, with a high...  ...or platform. Proficiency in Java, Python, ML Platforms, LLMs, and Machine Learning is...  ...practices, including the use of version control systems like Git, code reviews, and testing... 
    Remote work

    Argyle Infotech

    San Jose, CA
    2 days ago
  •  ...ML Engineer Cupertino, California, United States About the Job Our client is a rapidly growing Tier 1 VC backed startup based...  ...long-term growth trajectory in the evolving world of intelligent systems. Location: New York, NY Work type: Full Time... 
    Full time

    Catalyst Labs, LLC

    Cupertino, CA
    5 days ago
  •  ...proficiency in programming languages such as Java, Python, ML platforms, LLMs, Machine Learning, etc. -...  ...large enterprises. - Solid knowledge of software engineering best practices, including version control systems (e.g., Git), code reviews, and testing methodologies... 

    Argyle Infotech

    San Jose, CA
    2 days ago
  •  ...A tech-driven AI company is seeking an ML Engineer to design and deploy production-grade ML systems with full ownership of the model lifecycle. The ideal candidate will possess a degree in Computer Science and 1-6 years of experience in ML engineering, with robust programming... 

    Catalyst Labs, LLC

    Cupertino, CA
    5 days ago
  • $55 - $60 per hour

     ...chunking, embedding generation, and retrieval systems. Design, develop, and deploy Custom AI...  ...memory management, dynamic prompt engineering, and secure data handling. Build multi...  ...secure data handling. Skills ML. Python. TensorFlow. PyTorch.... 

    Cynet Systems

    Santa Clara, CA
    1 day ago
  •  ...An innovative AI startup is seeking an experienced Machine Learning Engineer to design and deploy production-grade ML systems. This role involves using AI to enhance sales data insights through real-time audio understanding. Candidates should have a relevant degree and... 

    Catalyst Labs, LLC

    San Jose, CA
    4 days ago
  • $130.7k - $261.3k

     ...executives, and scientists.THE OPPORTUNITYThis Senior Staff ML Ops Engineer position can work out of our Santa Clara, CA location.Senior...  .... Experience with microservices architecture and distributed systems.Experience in reviewing and selecting Technical and Applications... 
    Shift work

    Abbott

    Santa Clara, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Systems Engineer. Be the first to apply!