Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff, ML Performance

Odyssey

Who we are Odyssey is an AI lab pioneering general-purpose world models—a new form of multimodal intelligence unlocking entirely new consumer, enterprise, and intelligence applications. World models are the next major frontier in AI, and Odyssey is leading the way with breakthrough models like Odyssey-2 Pro. What we're looking for We’re seeking those who are obsessed with gaining every last drop of performance from complex systems. We're building inference infrastructure to scale to hundreds of thousands of users within a year, while also working with massive, ever-growing datasets and models in training. Your focus will be ensuring our models deliver exceptional speed, reliability, and scalability in both the training and inference phases, optimizing efficiency to minimize TFLOPS per user and training compute cost. What you'll do Optimize models that will be used in real-time by hundreds of thousands of users. Design and implement distributed training strategies to reduce training time and resource consumption on large GPU clusters. Partner with our elite team of ML researchers and engineers to ensure model architectures are highly performant from conception. Develop sophisticated tools to identify performance bottlenecks and stability issues in both training and serving environments. Pioneer innovative approaches, frameworks, and system designs that enhance performance metrics across our model development and inference infrastructure. Have significant autonomy in technical decisions. Use the latest-generation GPUs. Who you are 8+ years of software engineering experience, with significant work in ML performance. Deep insight into modern machine learning architectures with a natural instinct for performance optimization, particularly distributed training and inference. Track record of owning projects end to end. Problem-solving mindset with the ability to acquire new skills as needed. Proficiency with PyTorch (or TF/JAX) and Triton as well as NVIDIA GPU ecosystems and optimization stacks. Highly metric-based. #J-18808-Ljbffr Odyssey

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff, ML Performance in Santa Clara, CA vacancy
  •  ...About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and build the...  ...Implement and monitor observability and performance metrics across services and pipelines....  ...data pipelines , job scheduling , or ML experimentation workflows . Excellent... 
    Performance

    NeoCognition

    Palo Alto, CA
    3 days ago
  •  ...About the Role As a Member of Technical Staff [Research] at NeoCognition , you’ll be part of the core...  ...Design and execute experiments, benchmark performance, and analyze model behaviors to...  ...in Python and familiarity with modern ML frameworks (e.g., PyTorch, JAX, or TensorFlow... 
    Performance

    NeoCognition Inc.

    Palo Alto, CA
    3 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ..., Apple and Intel. What You’ll Do As Member of the Technical Staff - Software at Architect, you’ll build...  ...full stack, developing intuitive, high‑performance systems using TypeScript, Python, and... 
    Performance

    Architect Labs

    Palo Alto, CA
    3 days ago
  • $200k - $400k

    Member of Technical Staff — Cluster Infrastructure & Supercomputing About the Role RadixArk is looking...  ...design and operate highly reliable, high-performance GPU/TPU clusters, build next-...  ...Strong Plus: Experience with large-scale ML/AI workloads Familiarity with RDMA, InfiniBand... 
    Performance

    RadixArk

    Palo Alto, CA
    3 days ago
  • RadixArk is seeking a Member of Technical Staff — Inference to push the limits of large-scale AI inference...  ...frontier models at scale, optimizing performance, latency, throughput, and cost across...  ...intersection of systems engineering, ML infrastructure, and performance... 
    Performance
    Worldwide
    Flexible hours

    Dormont Manufacturing Co

    Palo Alto, CA
    6 days ago
  • Member of Technical Staff — Kernel / Compiler / Communication About the Role RadixArk is seeking a Member...  .../ Communication to push the limits of performance for frontier AI systems. You will...  ...systems Contributions to kernel/compiler/ML systems open source Experience... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    2 days ago
  •  ...training. About The Role We are hiring Members of Technical Staff to build the data and evaluation...  ...applied role focused on real-world system performance , not purely theoretical research. Key...  ...building or iterating on applied ML systems at scale Comfortable operating... 
    Performance
    Shift work

    Orbifold AI

    Palo Alto, CA
    4 days ago
  • RadixArk is seeking a Member of Technical Staff — Training to build and scale the systems that train frontier AI models. You will work...  ...of GPUs. This role sits at the intersection of ML, systems, and performance engineering. Your work will directly impact how next... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You’ll Do As a Founding Member of the Technical Staff at Architect, you'll be at the forefront...  ..., ensuring that theoretical performance translates into production-ready implementations... 
    Performance

    Architect Labs

    Palo Alto, CA
    3 days ago
  • About The Role RadixArk is seeking a Member of Technical Staff: Accelerator Systems to push the limits of performance for frontier AI systems. Most performance engineering assumes...  ...of experience in systems, performance, or ML infrastructure engineering Deep expertise in... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  • Member of Technical Staff, LLM Post-Training, Applied Sanas is pioneering the future of human communication...  ...optimization — into models that perform reliably in high‑stakes, real‑world, on...  ...scale Proficiency with the open‑source ML ecosystem (Hugging Face, PyTorch) and... 
    Performance

    Sanas

    Palo Alto, CA
    6 days ago
  • We are looking for a Member of Technical Staff with strong Python skills and a passion for building scalable platforms for AI and ML workloads. As MTS, you'll influence strategic decisions...  ...infrastructure optimized for high-performance AI and ML workloads, ensuring... 
    Performance

    S27a

    Mountain View, CA
    4 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You’ll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect,...  ...Strong skills for design automation, performance modeling, regression infrastructure,... 
    Performance

    Kindredventures

    Palo Alto, CA
    6 days ago
  •  ...design custom ASICs alongside evolving ML workloads, and enable a new era of...  ...Intel. What You'll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect,...  ...development on energy-efficient, high-performance HW accelerators on your block-of-expertise... 
    Performance

    Architect

    Palo Alto, CA
    3 days ago
  • $180k - $250k

    Member of Technical Staff -- TPU Systems (JAX / XLA / PALLAS) About the Role RadixArk is looking for a TPU Systems Engineer to build high-performance inference and training systems using JAX, XLA, and Pallas....  ...experience building production ML systems with JAX, XLA, or TPU... 
    Performance
    Full time
    Flexible hours

    RadixArk

    Palo Alto, CA
    6 days ago
  •  ...partners. We are looking for an exceptional Member of Technical Staff to help design, build, and scale core...  ..., and infrastructure. Emphasis on performance, scalability, and reliability. AI...  ...software, distributed systems, compilers, ML systems, hardware‑software co‑design,... 
    Performance

    DensityAI

    Mountain View, CA
    6 days ago
  • $140k - $200k

     ...Abaka AI provides the foundation for building high-performance AI systems. About the Role As a Member of Technical Staff, Infra, you'll own the scalability and...  ...product teams. Partner closely with Product and AI/ML teams to translate product and model requirements... 
    Performance
    Flexible hours

    Abaka AI

    Mountain View, CA
    5 days ago
  • $180k

    Member of Technical Staff - Multimodal Understanding SpaceXAI’s mission is to create AI systems that can...  ...paradigms for state‑of‑the‑art performance. Build research tooling, user‑friendly...  ...or optimizing large‑scale distributed ML systems (training/inference optimization... 
    Performance
    Temporary work

    SpaceXAI

    Palo Alto, CA
    5 days ago
  •  ...inefficient abstraction. The Role As a Member of Technical Staff, AI Training Platform, you will be a...  ...Architect, scale, and maintain the core AI/ML/RL training platform and...  ...researchers to optimize models for both performance and energy consumption. Minimum Qualifications... 
    Performance
    Work at office
    Flexible hours

    Unconventional, Inc.

    Mountain View, CA
    3 days ago
  • Member of Technical Staff — Developer Technology About the Role RadixArk is seeking a Member of Technical...  ...served and trained: SGLang is a high-performance inference engine that serves trillions...  ...; contributions to open‑source AI/ML projects. About RadixArk RadixArk is... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    6 days ago
  • Member of Technical Staff, ML Inference Engineering Sanas is pioneering the future of human communication. Founded by a team of Stanford researchers...  ...the level of the engineers working alongside them. Performance Optimization Optimize system and GPU performance for high... 
    Performance

    Sanas

    Palo Alto, CA
    6 days ago
  •  ...Engineer to develop and maintain high-performance, low-latency inference infrastructure...  ...system reliability. Author detailed technical documentation for infrastructure...  ...Student/Intern (Software Developer), Member of Technical Staff (Software Engineer), Software Engineer... 
    Performance
    Full time
    Part time
    Internship

    Cerebras Systems, Inc.

    Sunnyvale, CA
    6 days ago
  • $140k - $200k

     ...and evolvability, and lead high-stakes technical decisions and design reviews. Establish...  ...teams. Partner with Product and AI/ML teams to turn product and model requirements...  ...into production systems. Improve the performance, reliability, and observability of... 
    Performance
    Full time
    Flexible hours

    Embedding VC

    Mountain View, CA
    16 days ago
  •  ...and diffusion algorithms interact. Implement state of the art ML algorithms, define metrics, and relentlessly iterate on leaderboards...  ...software engineering experience, with significant work in ML performance. 2+ years of ML experience. Track record of owning projects... 
    Performance

    Odyssey

    Palo Alto, CA
    6 days ago
  •  ...looking for We’re looking for a deeply technical and creative researcher who thrives on invention...  ...generative models. Who you are A staff-level or senior researcher with deep...  ...Motivated by understanding as much as by performance: you care about how and why models work,... 
    Performance

    Odyssey

    Santa Clara, CA
    6 days ago
  •  ...reinforcement/preference‑based fine‑tuning for interactivity and policy performance. Run a high cadence of ablations across WM and policy, from...  ...training and inference efficiency. Take ownership of the full ML stack, including core frameworks that Odyssey researchers and... 
    Performance
    Remote work
    Flexible hours

    Odyssey

    Palo Alto, CA
    5 days ago
  • RadixArk is hiring a Performance Engineer in Palo Alto, CA — someone who can push LLM inference...  ...customers and cloud partners on deep technical evaluations Contribute performance insights...  ...distributed systems, inference serving, ML runtimes, or high-performance computing... 
    Performance
    Flexible hours

    RadixArk

    Palo Alto, CA
    5 days ago
  •  ...and the decisions made now will shape how product, growth, and ML operate for years. As our first Data Engineer, you’ll own the architecture...  ...and optimize SQL queries - profiling, tuning, and setting the performance standard Own data quality from day one: monitoring, alerting,... 
    Performance

    Astrocade

    Palo Alto, CA
    5 days ago
  •  ...closely with a strong engineering team on technically challenging problems. Design, build, and...  ...reliability, observability, and performance Collaborate across engineering to deliver...  ...Alto) ⭐ Nice to Have Experience with AI/ML frameworks (e.g. PyTorch, TensorFlow) Strong... 
    Performance
    Flexible hours
    3 days per week

    DeepRec.ai

    Palo Alto, CA
    4 days ago
  •  ...high‑speed inference. About the Role We are seeking a Sr. Member of Technical Staff to design and develop software features that support system...  ...and collaborate across engineering teams to deliver high‑performance software solutions. Responsibilities Design and develop software... 
    Performance

    Cerebras Systems, Inc.

    Sunnyvale, CA
    6 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff, ML Performance. Be the first to apply!