Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff — Developer TechnologyPalo Alto, CA

RadixArk

About The Role RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to make LLM inference and training dramatically faster, cheaper, and more accessible on modern GPU hardware. Our systems sit at the center of how modern AI is served and trained: SGLang is a high-performance inference engine that serves trillions of tokens daily across leading AI companies and research labs, and Miles is our reinforcement-learning post-training framework for large-scale LLM and MoE models. Your work directly advances our mission to democratize AI: every improvement you ship lowers the cost and raises the ceiling of what developers everywhere can build. About The Role RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to make LLM inference and training dramatically faster, cheaper, and more accessible on modern GPU hardware. Our systems sit at the center of how modern AI is served and trained: SGLang is a high-performance inference engine that serves trillions of tokens daily across leading AI companies and research labs, and Miles is our reinforcement-learning post-training framework for large-scale LLM and MoE models. Your work directly advances our mission to democratize AI: every improvement you ship lowers the cost and raises the ceiling of what developers everywhere can build. As our technical face to a community of expert users and partners, you\'ll push the performance of SGLang and Miles through the lens of real production workloads. You\'ll profile and optimize GPU performance, enable new models and hardware, build kernels, deliver day-0 model support, and push the limits of inference and training. Working in close partnership with leading teams across the ecosystem, you\'ll turn their hardest, most ambiguous problems into concrete wins and clear guidance, and feed those improvements back into our systems and future roadmap. Key Responsibilities Accelerate AI workloads. Profile and optimize GPU performance for real production workloads on current and next-generation hardware, root-causing bottlenecks from kernels to distributed multi-node systems. Go deep in one or two focus areas. The team collectively covers the full stack; each engineer specializes in one or two tracks: Inference performance: engine tuning, benchmarking, long-context and multi-turn optimization, parallelism strategy, production debugging Kernels and model/hardware enablement: custom CUDA/ROCm/Triton kernels, low-precision quantization, day-0 support for new models on new silicon Speculative decoding: draft-model training, acceptance-rate tuning, cross-platform kernel adaptation Training systems: RL post-training with Miles, FP8 training, elasticity, long-rollout and long-context efficiency Partner directly with the ecosystem. Turn ambiguous, high-stakes problems from expert engineers at our key partners into concrete wins, clear technical guidance, and reproducible cookbooks. Enhance SGLang and Miles. Feed user-driven improvements back into our open-source systems and roadmap, so every win compounds across the ecosystem. Qualifications Minimum Requirements 4+ years of experience in GPU systems, LLM infrastructure, or performance engineering. Strong profiling and debugging skills: able to root-cause performance and correctness issues across the stack. Hands-on GPU programming experience in at least one of CUDA, ROCm, or Triton, and willingness to work across platforms. Strong programming skills in Python plus C++ or CUDA. Comfortable making progress on hard, ambiguous problems with little context to start from, and fast to ramp into unfamiliar systems, codebases, and domains. Ability to translate ambiguous asks into clear technical plans, verified cookbooks, and actionable recommendations, and to communicate credibly with expert engineering audiences. Preferred (Bonus) Qualifications Deep familiarity with LLM inference internals: distributed serving, parallelism, routing, KV-cache management, scheduling. Experience with low-precision quantization and inference/training (FP8, INT8/INT4; NVFP4 or MXFP4 a strong plus). Experience writing and optimizing custom GPU kernels. Practical familiarity with speculative decoding methods such as Eagle, DFlash, or DSpark. Working knowledge of large-scale distributed training: pre-training, SFT, RL post-training, elasticity, long-context workloads. Experience optimizing across both NVIDIA and AMD platforms. Hands-on experience with SGLang, Miles, vLLM, TensorRT-LLM, Megatron, or comparable frameworks; contributions to open-source AI/ML projects. About RadixArk RadixArk is an infrastructure-first company built by engineers who\'ve shipped production AI systems, created SGLang (20K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). We\'re on a mission to democratize frontier-level AI infrastructure by building world-class open systems for inference and training. Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs. We\'re backed by well-known infrastructure investors and partner with Nvidia, Google, AWS, and frontier AI labs. Join us in building infrastructure that gives real leverage back to the AI community. Compensation We offer competitive compensation with meaningful equity, comprehensive benefits, and flexible work arrangements. Compensation depends on location, experience, and level. Equal Opportunity RadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. #J-18808-Ljbffr RadixArk

Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff — Developer TechnologyPalo Alto, CA in Palo Alto, CA vacancy
  • About The Role RadixArk is seeking a Developer Advocate to build and engage our technical community around SGLang, Miles, and our open source infrastructure. SGLang already has 20K+ GitHub stars and serves billions of tokens daily across leading AI companies and research... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    13 hours ago
  • About The Role RadixArk is seeking a Member of Technical Staff: Accelerator Systems to push the limits of performance for frontier AI systems....  ...K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). We\'re on a mission... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  • About The Role RadixArk is seeking a Member of Technical Staff — Diffusion Model to advance the frontier of generative modeling. You will work...  ...next‑generation generative AI systems used by researchers, developers, and real‑world applications. This is a high‑impact role... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    1 day ago
  • About The Role RadixArk is hiring a Member of Technical Staff — CI Engineer to own the infrastructure that keeps SGLang moving. Our CI system...  ..., separate PR smoke tests from nightly full runs Improve developer experience — faster feedback, clearer failure messages, workflow... 
    Suggested
    Flexible hours
    Night shift

    RadixArk

    Palo Alto, CA
    1 day ago
  • $200k - $420k

    Member of Technical Staff, Hardware, Kernel Engineer (Custom Silicon) At River, our mission is to create personal AI owned and shaped by each individual...  ...Location: This role is based in Austin, Texas or Palo Alto, California . Compensation: Depending on background,... 
    Suggested
    Local area
    Visa sponsorship
    Work visa
    Relocation package

    River AI Inc.

    Palo Alto, CA
    2 days ago
  • $200k - $420k

    Member of Technical Staff, Hardware, Compiler Engineer At River, our mission is to create personal AI...  ...frameworks. Custom Backend Development: Develop and maintain the backend toolchain for...  ...is based in Austin, Texas or Palo Alto, California . Compensation: Depending... 
    Local area
    Visa sponsorship
    Work visa
    Relocation package
    Flexible hours

    River AI Inc.

    Palo Alto, CA
    2 days ago
  • We are looking for a Member of Technical Staff with strong Python skills and a passion for building scalable platforms for AI and ML workloads....  ...AI infrastructure. What You Will Be Doing Architect and Develop: Design and build scalable software infrastructure optimized... 

    S27a

    Mountain View, CA
    2 days ago
  •  ...Description MosaixSoft, Inc. is recruiting for our Los Altos, CA office: Member of the Technical Staff (job code #37659). Design, architect, and implement...  ...easy to deploy and configure and without conflicts. Develop ad‑hoc tools to aid with testing. Responsible for the... 
    Work at office

    MosaixSoft, Inc.

    Los Altos, CA
    4 days ago
  • Member of Technical Staff Physical AI (Robotics / World Models) Palo Alto, CA About Orbifold AI Orbifold AI advances the frontier of physical AI and world model companies through rigorous evaluation and curated, real-world data. We work directly with leading robotics and... 
    Shift work

    Bonfirevc

    Palo Alto, CA
    1 day ago
  • $300k - $350k

    Member of Technical Staff Level 1 - Engineering Bellevue | Hybrid NTT DATA AIVista, Inc., a wholly owned subsidiary of NTT DATA, is based in Silicon Valley. We develop AI products that operationalize AI across enterprises operating in complex regulatory environments. We... 
    Work experience placement
    Local area
    Flexible hours

    NTT DATA AIVista

    Palo Alto, CA
    3 days ago
  • Member of Technical Staff, Hardcore Engineer Employment Type Full time Location Type On-site Compensation Actual salary will depend on qualifications, including experience, skillset, and location. We are looking for exceptional engineers and researchers who are confident... 
    Full time

    Epochal

    Palo Alto, CA
    13 hours ago
  •  ...AI Startup in Stealth | Mountain View, CA AI Startup in Stealth is a full-stack semiconductor company founded by pioneers...  ...investors and strategic partners. We are looking for an exceptional Member of Technical Staff to help design, build, and scale core components of our next... 

    DensityAI

    Mountain View, CA
    2 days ago
  •  ...projects full lifecycle Design discussions & technical scoping Implementation & testing Post-...  ...experience as a frontend or fullstack developer, ideally deeply embedded with design and...  ...: Hybrid - we’re in the office in Palo Alto, CA near the Caltrain station from Tuesday to... 
    Work at office
    Remote work
    Visa sponsorship
    Monday to Friday

    ProductNow

    Palo Alto, CA
    1 day ago
  • Member of Technical Staff, ML Inference Engineering Sanas is pioneering the future of human communication. Founded by a team of Stanford researchers...  ...entrepreneurs with deep industry experience, Sanas has developed the world's first real-time speech AI platform capable of... 

    Sanas

    Palo Alto, CA
    4 days ago
  • $220k - $405k

    Location San Francisco; New York City; Palo Alto Employment Type Full time Department Product Engineering Compensation $220K - $405...  ..., projects and products end-to-end, from problem definition to technical design, implementation, and launch. Hill climb on hard problems... 
    Full time
    Local area

    Kindredventures

    Palo Alto, CA
    3 days ago
  • $13 per hour

     ...deeply curious Wants to own features from design to development to deployment to maintenance Is willing to put the work in to solve the hardest of problems Location: Palo Alto, CA Base Salary Range: $140,000/yr to $220,000/yr + Equity + Benefits #J-18808-Ljbffr Pylon
    Immediate start

    Pylon

    Palo Alto, CA
    2 days ago
  • $180k

     ...Area (San Francisco and Palo Alto). Candidates are expected to be...  ...Engineer long‑context data recipes. Develop robust and diverse evaluation...  ...interview”) during which a member of our team will ask some...  ...process, which consists of four technical interviews: Coding assessment... 
    Temporary work
    Relocation

    Pantera Capital

    Palo Alto, CA
    2 days ago
  • $180k

     ...government projects. In this role, you will develop and manage training and inference...  ...This is an in‑person role based in Palo Alto, CA or Washington, DC, with up to 50% travel...  ...curiosity, and enthusiasm for tackling complex technical challenges in secure environments.... 
    Temporary work

    Pantera Capital

    Palo Alto, CA
    3 days ago
  • $180k

    Member of Technical Staff, Pre-training Data Infrastructure xAI’s mission is to create AI systems that can accurately understand the universe and...  ...Ray Location Bay Area, including San Francisco and Palo Alto. Candidates are expected to be located nearby or open to relocation... 
    Temporary work
    Relocation

    xAI

    Palo Alto, CA
    1 day ago
  • $140k - $200k

     ...for building high-performance AI systems. About the Role As a Member of Technical Staff, Infra, you'll own the scalability and reliability of the...  ...dental, vision, PTO, flexible work schedule). This position is based onsite in Mountain View, CA. #J-18808-Ljbffr Abaka AI
    Flexible hours

    Abaka AI

    Mountain View, CA
    3 days ago
  • $140k - $220k

     ...deeply curious Wants to own features from design to development to deployment to maintenance Is willing to put the work in to solve the hardest of problems Location: Palo Alto, CA • Base Salary Range: $140,000/yr to $220,000/yr + Equity + Benefits #J-18808-Ljbffr Pylon

    Pylon

    Palo Alto, CA
    2 days ago
  •  ...is hiring a Performance Engineer in Palo Alto, CA — someone who can push LLM inference and...  ...customers and cloud partners on deep technical evaluations Contribute performance insights...  ...AI systems, created SGLang, and developed Miles, our large-scale RL framework. We'... 
    Flexible hours

    RadixArk

    Palo Alto, CA
    3 days ago
  • $300k - $400k

    Member of Technical Staff - Applied AI Engineering Location: San Francisco Bay Area | Hybrid Role Description AIVista turns foundation models into specialized, governed agents that run mission-critical, regulated enterprise operations reliably at scale. But most enterprise... 
    Work experience placement
    Local area
    Flexible hours

    NTT DATA AIVista

    Palo Alto, CA
    4 days ago
  • Epochal is seeking a Member of Technical Staff, Hardcore Engineer for on-site work in Palo Alto. You will tackle ambitious research and engineering challenges at the edge of model intelligence and usefulness, collaborating with world-class researchers and engineers. The... 

    Epochal

    Palo Alto, CA
    13 hours ago
  •  ...parallel computing, or in physics simulation. Responsibilities Develop and optimize algorithms for geometry import, repair,...  ...performance and memory efficiency. Experience working in a relevant technical domain such as computational geometry, graphics, simulation, or... 

    Getvinci

    Palo Alto, CA
    3 days ago
  • Architect Labs, based in Palo Alto, is seeking a Founding Member of the Technical Staff (Applied AI) to revolutionize chip design using AI. The role combines hardware expertise with AI to build agent systems that enhance chip-design tasks. Applicants should have an MS or... 

    Architect Labs

    Palo Alto, CA
    1 day ago
  • Who we are Odyssey is an AI lab pioneering general world models: causal, multimodal systems that learn to predict and interact with the world over long horizons. This foundational technology promises to revolutionize robotics, science, healthcare, education, gaming, defense...

    Odyssey

    Palo Alto, CA
    23 hours ago
  •  ...We are seeking a Software Engineer to develop and maintain high-performance, low-...  ...system reliability. Author detailed technical documentation for infrastructure configurations...  ...Student/Intern (Software Developer), Member of Technical Staff (Software Engineer), Software... 
    Full time
    Part time
    Internship

    Cerebras Systems, Inc.

    Sunnyvale, CA
    4 days ago
  •  ...to build new products to help us improve our MVP. Responsibilities You’ll be a force multiplier for our team and will design and develop the pipelines and tools that make our product a product. This includes accelerating our development by developing the system that enables... 

    Getvinci

    Palo Alto, CA
    3 days ago
  • $220k - $405k

     ...experience for the user. Build the data and evaluation foundations that let these systems learn and improve with usage. Help shape the technical direction of ranking, recommendations, and personalization at Perplexity. What We're Looking For Deep, hands‑on experience... 

    Perplexity

    Palo Alto, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff — Developer TechnologyPalo Alto, CA. Be the first to apply!