Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Systems Engineer

$150k - $350k

ChipAgents

About ChipAgents ChipAgents is redefining the future of chip design and verification with agentic AI workflows. Our platform leverages cutting‑edge generative AI to assist engineers in RTL design, simulation, and verification, dramatically accelerating chip development. Founded by experts in AI and semiconductor engineering, we partner with top semiconductor firms, cloud providers, and innovative startups to build intelligent AI agents. The company is a Series A company backed by tier‑1 VC firms. ChipAgents is deployed in production to companies that have shipped 16B chips. Position Overview We are seeking an ML Systems Engineer to optimize the performance and efficiency of large language model inference powering our agentic AI platform. This is a technical role focused on low‑level systems optimization. You will implement performance optimizations, build evaluation harnesses, and architect multi‑node clusters for training and inference that push the limits of LLM throughput and latency. Your work will directly impact the responsiveness and cost‑efficiency of AI agents used by leading semiconductor companies to design chips. Key Responsibilities Design, deploy, and optimize LLM inference systems across multi‑node clusters, maximizing throughput and minimizing latency for production workloads. Implement and benchmark concrete inference optimizations. Profile and analyze inference bottlenecks at the systems level—from GPU kernel execution to memory bandwidth constraints. Build robust evaluation harnesses and benchmarking frameworks that measure accuracy, throughput, latency, and resource utilization across various parallelism strategies. Collaborate with research scientists to integrate new model architectures and optimizations into production inference infrastructure. Investigate and apply emerging techniques from research papers and open‑source projects to continuously improve inference performance. Qualifications B.S., M.S., or PhD in Computer Science, Electrical Engineering, or related field (or equivalent experience). Experience with large‑scale ML systems, GPU computing, or high‑performance inference optimization. Strong proficiency in Python and C++/CUDA; hands‑on experience with SGLang, vLLM, PyTorch, or similar inference frameworks. Deep understanding of GPU architecture, memory hierarchies, and parallel computing paradigms. Experience deploying and optimizing LLMs in production: model serving, batching strategies, distributed inference, or quantization. Strong systems‑level debugging and profiling skills; comfort working at multiple layers of the stack from CUDA kernels to application logic. Familiarity with distributed computing frameworks (Ray, multi‑node training/inference) is a plus. Self‑directed problem solver who is interested in working on ambitious optimization challenges. Why Join Us Work on cutting‑edge LLM inference optimization problems with real‑world production impact. Access to substantial GPU compute resources for experimentation and benchmarking. Collaborate with a world‑class team spanning AI research, systems engineering, and EDA. Shape the performance characteristics of AI systems used by leading semiconductor companies. What we offer $150K/yr – $350K/yr + Offers Equity. We are open to discuss above‑scale compensation with exceptional candidates on a case‑by‑case basis. Unlimited PTO and full benefits (medical, vision, dental, 401k). Two engineering‑centric offices with free parking, private gym, and free lunch, drinks and snacks. #J-18808-Ljbffr

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the ML Systems Engineer in San Jose, CA vacancy
  • $224k - $356.5k

     ...the next phase, we are building agentic systems that can reason about, build, evaluate, and...  ...about creating the meta-layer of modern ML: the agents, tooling, pipelines, and feedback...  .... We are looking for exceptional engineers who are passionate about the idea of AI-native... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    10 hours ago
  •  ...of what’s next.About the JobAI for Member Systems (AIMS) runs the AI systems behind every...  ...remarkably effective at doing so. But AI/ML is moving fast, and the infrastructure that...  ...end-to-end.Platform Systems is the engineering foundation of AIMS, owning reliability, scalability... 
    Suggested
    Hourly pay
    Full time
    Immediate start
    Flexible hours

    Netflix

    Los Gatos, CA
    2 days ago
  • $152k - $241.5k

     ...transparency, and broad scientific scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered tooling that helps find,...  ...is simple: evidence first. We are looking for an Evaluation/ML-Systems Engineer to own how we measure the program. You will play a... 
    Suggested
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    10 hours ago
  • $164k - $313.3k

    The OpportunityPhotoshop ART is seeking a Senior Machine Learning (ML) Systems & Efficiency Engineer to join our R&D team focused on delivering practical, production-ready improvements in inference performance, latency, and cost efficiency across image editing applications... 
    Suggested
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Jose, CA
    3 days ago
  • $169k - $338k

     ...Segment: Home OfficePosition Summary...As a Distinguished AI/ML Engineer within Walmart Global Tech's Site Reliability Engineering organization...  ...lead the technical development of next-generation agentic AI systems and intelligent automation solutions that ensure mission-... 
    Suggested
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    1 day ago
  • $174k - $253k

     ...agents and LLM-powered journeys that evaluate, gate, and improve engineering artifacts.Engineer and refine skills and context provided to...  ...information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence... 

    Google

    San Jose, CA
    2 days ago
  • $165.2k - $223.6k

     ...forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions...  ...of software, hardware, and machine learning systems, you'll bring expertise in low-level... 
    Internship
    Local area
    Work from home
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $165.2k - $223.6k

     ...team is actively seeking skilled compiler engineers to join our efforts in developing a state...  ...forefront of AWS innovation for advanced ML capabilities, powering solutions like...  ...in compiler technology and deep-learning systems software. Additionally, you'll collaborate... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $108k - $192k

     ...before you apply.Job Description:We need an engineer to own the end-to-end silicon diagnostic...  ...goal is simple but ambitious: Use AI/ML to transform millions of data points into...  ...with industry-standard yield management systems (e.g., PDF Solutions, Synopsys YieldExplorer... 
    Full time
    Local area

    Broadcom

    San Jose, CA
    4 days ago
  •  ...data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and...  ...looking for a Principal Machine Learning Engineer to join our Models and Applications team....  ...stakeholders.PREFERRED EXPERIENCE:Experience with ML/DL frameworks such as PyTorch, JAX, or... 

    AMD

    San Jose, CA
    1 day ago
  •  ...A tech-driven AI company is seeking an ML Engineer to design and deploy production-grade ML systems with full ownership of the model lifecycle. The ideal candidate will possess a degree in Computer Science and 1-6 years of experience in ML engineering, with robust programming... 

    Catalyst Labs, LLC

    Cupertino, CA
    1 day ago
  •  ...Responsibilities Design, build, and deploy production‑grade ML systems with end‑to‑end ownership of the model lifecycle from conception...  ...a related field. 1-6 years of professional experience in ML engineering. Strong programming skills in Python (TypeScript experience... 
    Full time

    Catalyst Labs, LLC

    San Jose, CA
    1 day ago
  • Entefy’s vision is simplifying how people interact digitally. To make it happen, we’re seeking computer vision-aries. Our hyper-talented Product team is seeking a Machine Learning and AI expert with the skills to redefine how computers make sense of the visual world. ...
    Remote work

    GrabJobs

    San Jose, CA
    2 days ago
  •  ...An innovative AI startup is seeking an experienced Machine Learning Engineer to design and deploy production-grade ML systems. This role involves using AI to enhance sales data insights through real-time audio understanding. Candidates should have a relevant degree and... 

    Catalyst Labs, LLC

    San Jose, CA
    5 days ago
  • $136k - $218.5k

     ...essential part of our Power Team, you'll closely collaborate with HW/ML experts and infrastructure teams. You'll work together to...  ...foundational knowledge of data structures, algorithms, and software engineering principles.Familiarity with training and fine-tuning large... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $100k

     ...Tenstorrent is seeking an Physical Design Engineer to lead cross-functional efforts to solve...  ...will architect, integrate, and deploy AI/ML-driven solutions into production physical...  ...and/or indirect access to information, systems, or technologies subject to these laws, the... 
    Permanent employment

    Tenstorrent

    Santa Clara, CA
    1 day ago
  • $185k - $254k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced...  ...and focus during later stages of completion.Ensures that all systems engineering projects and programs assigned are completed in accordance... 
    Full time

    Applied Materials

    Santa Clara, CA
    3 days ago
  • $193.3k - $261.5k

     ...looking for a Senior Software Development Engineer to own the design and implementation of...  ...and optimize compute kernels for a custom ML accelerator architecture, targeting production...  ...proficiency in C/C++- Strong Linux systems knowledge- Experience developing compute... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  •  ...Description: We are seeking a talented and experienced Deep Learning Engineer with a specialization in Large Language Models (LLMs) to join our dynamic team. The ideal candidate will have a deep understanding of state-of-the-art LLM architectures, such as GPT, BERT... 

    Xforia Inc

    San Jose, CA
    2 days ago
  • $128k - $192k

     ...and more efficient. Job Summary & Responsibilities Job Summary & Responsibilities The core responsibilities of the Senior Systems Engineer are to develop new optical metrology technologies, to optimize the performance of existing metrology products, and to communicate... 
    Permanent employment
    Full time
    Local area

    Onto Innovation

    Milpitas, CA
    1 day ago
  • $110k - $140k

     ...connections amongst all team members. We are a team of dedicated Engineers and Designers with an average of 7 years’ tenure at Grantek....  ...design, SCADA design, Network design, Database design, Vision system design, Machine Safety system design, Servo motion control, Information... 
    Full time
    Work at office
    Local area
    Remote work
    Flexible hours

    Grantek Systems Integration

    San Jose, CA
    14 hours ago
  •  ...Job Summary: Job Description: PayPal, Inc. seeks Sr Machine Learning Engineer in San Jose, CA Job Duties: Design and deploy scalable traditional and generative AI solutions and productize ML models that enhance PayPal's ability to provide a seamless customer experience... 
    Full time
    Work at office
    Local area
    Immediate start
    Remote work
    Flexible hours

    PayPal

    San Jose, CA
    2 days ago
  •  ...Employment Type: Full TimeIndustry: Computer SoftwareClient: WiproContact: Meghana GorusuCompany: SRI Tech SolutionsJob Title: Systems Engineer (C, Python)Location: San Jose, CAWill not accept submission without screening questions and skill matrix on mandatory skills:... 
    Full time
    Work at office
    3 days per week

    SRI Tech

    San Jose, CA
    4 days ago
  • $200.4k - $290.1k

     ...Role We are seeking a Machine Learning Engineer to help drive the development, optimization...  ...engineers, and IP developers to design ML models, optimize inference pipelines, and...  ...learning development, model optimization or ML systems engineering. Experience with C++ and... 
    Full time
    Local area
    Shift work

    Altera

    San Jose, CA
    1 day ago
  •  ...network. It's the foundation for autonomous networking, giving engineers and AI agents the ability to know the impact of every change before...  ...things have always been done.Forward Networks is looking for a Systems Sales Engineer ManagerDo you want to create a category and help... 
    Work experience placement

    Forward Networks

    Santa Clara, CA
    2 days ago
  •  ...We are seeking a motivated AI / Machine Learning Engineer with hands-on experience in Intelligent Systems and Generative AI to join our growing technology team...  ...eager to work on real-world AI solutions and scalable ML models. Experience: 6 Months – 2 Years... 
    Full time
    Internship
    Relocation

    Hudson Manpower

    San Jose, CA
    14 hours ago
  • $224k - $308k

     ...Consultant Manager is the evolution of the traditional Channel Sales Engineer Manager role, aligning how we lead teams to best serve our GSI...  ...level required; 5 years preferredExperience as a pre-sales System Engineer ManagerExperience as a Senior System Engineer/... 
    Full time
    Remote work
    Visa sponsorship
    Work visa

    Palo Alto Networks

    Santa Clara, CA
    3 days ago
  • $193.3k - $261.5k

    We are seeking an experienced engineer to work on distributed AI/ML systems. This role involves working on collective operations - the fundamental operations that enable AI to scale across multiple accelerators & servers. Most of our stack is C/C++ and relatively low level... 
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  • $136.5k - $276.5k

    AI/ML Engineer - AgenticThis role has been designed as ‘Hybrid’ with an expectation that you will work on average 2 days per week from an...  ...retrieval and memory services, and high‑performance backend systems that power agent execution. This position owns reliability, observability... 
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...scale? NVIDIA is seeking a Senior MLOps Engineer to join our Autonomous Driving organization...  ...design and operate end‑to‑end data and ML pipelines for NVIDIA’s autonomous driving...  ...focus, and engineering judgment to scale systems and solve problems across teams. What You... 
    Full time

    NVIDIA

    Santa Clara, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Systems Engineer. Be the first to apply!