Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff, ML Performance

Odyssey

Who we are Odyssey is an AI lab pioneering general-purpose world models—a new form of multimodal intelligence unlocking entirely new consumer, enterprise, and intelligence applications. World models are the next major frontier in AI, and Odyssey is leading the way with breakthrough models like Odyssey-2 Pro. What we're looking for We’re seeking those who are obsessed with gaining every last drop of performance from complex systems. We're building inference infrastructure to scale to hundreds of thousands of users within a year, while also working with massive, ever-growing datasets and models in training. Your focus will be ensuring our models deliver exceptional speed, reliability, and scalability in both the training and inference phases, optimizing efficiency to minimize TFLOPS per user and training compute cost. What you'll do Optimize models that will be used in real-time by hundreds of thousands of users. Design and implement distributed training strategies to reduce training time and resource consumption on large GPU clusters. Partner with our elite team of ML researchers and engineers to ensure model architectures are highly performant from conception. Develop sophisticated tools to identify performance bottlenecks and stability issues in both training and serving environments. Pioneer innovative approaches, frameworks, and system designs that enhance performance metrics across our model development and inference infrastructure. Have significant autonomy in technical decisions. Use the latest-generation GPUs. Who you are 8+ years of software engineering experience, with significant work in ML performance. Deep insight into modern machine learning architectures with a natural instinct for performance optimization, particularly distributed training and inference. Track record of owning projects end to end. Problem-solving mindset with the ability to acquire new skills as needed. Proficiency with PyTorch (or TF/JAX) and Triton as well as NVIDIA GPU ecosystems and optimization stacks. Highly metric-based. #J-18808-Ljbffr Odyssey

Vacancy posted 9 hours ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff, ML Performance in Santa Clara, CA vacancy
  •  .... About the Role We are looking for a Member of Technical Staff (MTS - Research) to help us build the...  ...runs (CPT, SFT, RL) to deliver frontier performance on open-source models using curated,...  ...Iteration: Debug and iterate across the full ML stack, from infrastructure to model... 
    Performance

    Collinear AI

    Sunnyvale, CA
    7 days ago
  • $230k

     ...users to effortlessly run large-scale ML applications, without the hassle of managing...  ...Inc. has multiple openings for Sr. Member of Technical Staff. Title: Sr. Member of Technical Staff...  ...low-latency and scalable system performance. Develop Python-based scripts and APIs... 
    Performance
    Remote work

    Cerebras Systems

    Sunnyvale, CA
    20 hours ago
  •  ...partners. We are looking for an exceptional Member of Technical Staff to help design, build, and scale core...  ..., and infrastructure. Emphasis on performance, scalability, and reliability. AI...  ...software, distributed systems, compilers, ML systems, hardware‑software co‑design,... 
    Performance

    DensityAI

    Mountain View, CA
    2 days ago
  • Member of Technical Staff (Software Engineer) Sunnyvale, CA Cerebras Systems builds the world’s largest...  ...learning users to run large‑scale ML applications without managing hundreds...  ...Implement infrastructure to support high‑performance, low‑latency inference service.... 
    Performance
    Full time
    Part time
    Internship

    Cerebras Systems

    Sunnyvale, CA
    20 hours ago
  • $120k - $200k

     ...AI provides the foundation for building high-performance AI systems. About the Role As a Member of Technical Staff, Platform, you'll build full-stack product features...  ...consumer-facing products at scale. ML and Deep Learning knowledge and experience... 
    Performance
    Full time
    Flexible hours

    Abaka AI

    Mountain View, CA
    6 days ago
  • Member of Technical Staff — Kernel / Compiler / Communication RadixArk is seeking a deeply technical engineer who pushes the limits of performance for frontier AI systems. You will work at the lowest layers...  ...compiler and runtime stacks for ML systems Improve communication... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  • About The Role RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train frontier AI models...  ...100k+ of GPUs. This role sits at the intersection of ML, systems, and performance engineering. Your work will directly impact how next-... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  • RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference...  ...frontier models at scale, optimizing performance, latency, throughput, and cost across...  ...intersection of systems engineering, ML infrastructure, and performance... 
    Performance
    Worldwide
    Flexible hours

    Dormont Manufacturing Co

    Palo Alto, CA
    9 hours ago
  • About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and build the...  ...Implement and monitor observability and performance metrics across services and pipelines....  ...data pipelines , job scheduling , or ML experimentation workflows . Excellent... 
    Performance

    NeoCognition

    Palo Alto, CA
    9 hours ago
  • Member of Technical Staff — Kernel / Compiler / Communication About the Role RadixArk is seeking a Member...  .../ Communication to push the limits of performance for frontier AI systems. You will...  ...systems Contributions to kernel/compiler/ML systems open source Experience... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  • About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll be part of the core...  ...Design and execute experiments, benchmark performance, and analyze model behaviors to...  ...in Python and familiarity with modern ML frameworks (e.g., PyTorch, JAX, or TensorFlow... 
    Performance

    NeoCognition Inc.

    Palo Alto, CA
    9 hours ago
  •  ...product definition. The ideal candidate has a proven track of record of pursuing ML systems research, and is very familiar with industry-standard LLM inference systems. This role will be performed on‑site from one of our offices in Santa Clara, CA or Boston, MA. Essential... 
    Performance
    Visa sponsorship
    Relocation package

    Netpreme

    Santa Clara, CA
    3 days ago
  •  ...execute investigations into how AI models perform and fail across real-world scenarios....  ...tooling. Contribute beyond your immediate technical domain — this is a founding role that requires...  ...equivalent demonstrated depth. Strong ML, statistics, and data science... 
    Performance
    Immediate start
    Visa sponsorship
    Shift work

    Jobtailor

    Sunnyvale, CA
    9 hours ago
  • About The Role RadixArk is seeking a Member of Technical Staff: Accelerator Systems to push the limits of performance for frontier AI systems. Most performance engineering assumes...  ...of experience in systems, performance, or ML infrastructure engineering Deep expertise in... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    7 days ago
  • We are looking for a Member of Technical Staff with strong Python skills and a passion for building scalable platforms for AI and ML workloads. As MTS, you'll influence strategic decisions...  ...infrastructure optimized for high-performance AI and ML workloads, ensuring... 
    Performance

    S27a

    Mountain View, CA
    3 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You’ll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect,...  ...Strong skills for design automation, performance modeling, regression infrastructure,... 
    Performance

    Kindredventures

    Palo Alto, CA
    5 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You'll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect,...  ...development on energy-efficient, high-performance HW accelerators on your block-of-expertise... 
    Performance

    Architect

    Palo Alto, CA
    2 days ago
  • RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train frontier AI models. You will work...  ...of GPUs. This role sits at the intersection of ML, systems, and performance engineering. Your work will directly impact how next... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  •  ...training. About The Role We are hiring Members of Technical Staff to build the data and evaluation...  ...applied role focused on real-world system performance , not purely theoretical research. Key...  ...building or iterating on applied ML systems at scale Comfortable operating... 
    Performance
    Shift work

    Orbifold AI

    Palo Alto, CA
    2 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You’ll Do As a Founding Member of the Technical Staff at Architect, you'll be at the forefront...  ..., ensuring that theoretical performance translates into production-ready implementations... 
    Performance

    Architect Labs

    Palo Alto, CA
    2 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You’ll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect,...  ...bandwidth vs. area, coherence overhead vs. performance), and feed area/timing/power... 
    Performance
    Night shift

    Kindredventures

    Palo Alto, CA
    5 days ago
  • $180k - $250k

    Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk is looking for a TPU Systems Engineer to build high-performance inference and training systems using JAX, XLA, and Pallas....  ...experience building production ML systems with JAX, XLA, or TPU... 
    Performance
    Full time
    Flexible hours

    RadixArk

    Palo Alto, CA
    9 hours ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You’ll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect,...  ...microarchitecture and RTL design of high-performance networking subsystems going into... 
    Performance

    Doist

    Palo Alto, CA
    3 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Apple and Intel. What You’ll Do As Member of the Technical Staff - Software at Architect, you’ll build...  ...full stack, developing intuitive, high‑performance systems using TypeScript, Python, and... 
    Performance

    Architect

    Palo Alto, CA
    2 days ago
  • $180k

    Member of Technical Staff - Multimodal Understanding About xAI xAI’s mission is to create AI systems that...  ...paradigms for state‑of‑the‑art performance. Build research tooling, user‑friendly...  ...or optimizing large‑scale distributed ML systems (training/inference optimisation... 
    Performance
    Temporary work

    xAI

    Palo Alto, CA
    1 day ago
  • $148.5k - $223.9k

    Senior Member of Technical Staff - AI ResearchSkip to main content#Senior Member of Technical Staff -...  ...exceptional engineering skills.** *Has deep ML knowledge with meaningful...  ...for correctness, quality, security, and performance** *Strong software engineering fundamentals... 
    Performance
    Work at office

    Salesforce, Inc.

    Palo Alto, CA
    1 day ago
  • Member of Technical Staff — Developer Technology About the Role RadixArk is seeking a Member of Technical...  ...served and trained: SGLang is a high-performance inference engine that serves trillions...  ...; contributions to open‑source AI/ML projects. About RadixArk RadixArk is... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    9 hours ago
  •  ...tackle, formulate them into benchmarkable ML objectives, build reproducible pipelines...  ...coach junior researchers by providing technical guidance, rigorous code reviews, and career...  ...to inspire, guide, and elevate a high-performance research culture. #J-18808-Ljbffr Axiom... 
    Performance

    Axiom Math

    Palo Alto, CA
    3 days ago
  •  ...reinforcement/preference‑based fine‑tuning for interactivity and policy performance. Run a high cadence of ablations across WM and policy, from...  ...training and inference efficiency. Take ownership of the full ML stack, including core frameworks that Odyssey researchers and... 
    Performance
    Remote work
    Flexible hours

    Odyssey

    Palo Alto, CA
    4 days ago
  •  ...looking for We’re looking for a deeply technical and creative researcher who thrives on invention...  ...generative models. Who you are A staff-level or senior researcher with deep...  ...Motivated by understanding as much as by performance: you care about how and why models work,... 
    Performance

    Odyssey

    Santa Clara, CA
    9 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff, ML Performance. Be the first to apply!