Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Member of Technical Staff — Developer Technology

RadixArk

Member of Technical Staff — Developer Technology About the Role RadixArk is seeking a Member of Technical Staff, Developer Technology (DevTech) to make LLM inference and training dramatically faster, cheaper, and more accessible on modern GPU hardware. Our systems sit at the center of how modern AI is served and trained: SGLang is a high-performance inference engine that serves trillions of tokens daily across leading AI companies and research labs, and Miles is our reinforcement-learning post-training framework for large-scale LLM and MoE models. Your work directly advances our mission to democratize AI: every improvement you ship lowers the cost and raises the ceiling of what developers everywhere can build. As our technical face to a community of expert users and partners, you'll push the performance of SGLang and Miles through the lens of real production workloads. You'll profile and optimize GPU performance, enable new models and hardware, build kernels, deliver day-0 model support, and push the limits of inference and training. Working in close partnership with leading teams across the ecosystem, you'll turn their hardest, most ambiguous problems into concrete wins and clear guidance, and feed those improvements back into our systems and future roadmap. Key Responsibilities Accelerate AI workloads . Profile and optimize GPU performance for real production workloads on current and next-generation hardware, root-causing bottlenecks from kernels to distributed multi-node systems. Go deep in one or two focus areas. The team collectively covers the full stack; each engineer specializes in one or two tracks: Inference performance: engine tuning, benchmarking, long-context and multi-turn optimization, parallelism strategy, production debugging Kernels and model/hardware enablement: custom CUDA/ROCm/Triton kernels, low‑precision quantization, day‑0 support for new models on new silicon Training systems: RL post‑training with Miles, FP8 training, elasticity, long‑rollout and long‑context efficiency Partner directly with the ecosystem. Turn ambiguous, high‑stakes problems from expert engineers at our key partners into concrete wins, clear technical guidance, and reproducible cookbooks. Enhance SGLang and Miles . Feed user‑driven improvements back into our open‑source systems and roadmap, so every win compounds across the ecosystem. Qualifications 4+ years of experience in GPU systems, LLM infrastructure, or performance engineering. Strong profiling and debugging skills: able to root‑cause performance and correctness issues across the stack. Hands‑on GPU programming experience in at least one of CUDA, ROCm, or Triton, and willingness to work across platforms. Strong programming skills in Python plus C++ or CUDA. Comfortable making progress on hard, ambiguous problems with little context to start from, and fast to ramp into unfamiliar systems, codebases, and domains. Ability to translate ambiguous asks into clear technical plans, verified cookbooks, and actionable recommendations, and to communicate credibly with expert engineering audiences. Preferred (Bonus) Qualifications Deep familiarity with LLM inference internals: distributed serving, parallelism, routing, KV‑cache management, scheduling. Experience with low‑precision quantization and inference/training (FP8, INT8/INT4; NVFP4 or MXFP4 a strong plus). Experience writing and optimizing custom GPU kernels. Practical familiarity with speculative decoding methods such as Eagle, DFlash, or DSpark. Working knowledge of large‑scale distributed training: pre‑training, SFT, RL post‑training, elasticity, long‑context workloads. Experience optimizing across both NVIDIA and AMD platforms. Hands‑on experience with SGLang, Miles, vLLM, TensorRT‑LLM, Megatron, or comparable frameworks; contributions to open‑source AI/ML projects. About RadixArk RadixArk is an infrastructure‑first company built by engineers who've shipped production AI systems, created SGLang (20K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large‑scale RL framework).We're on a mission to democratize frontier‑level AI infrastructure by building world‑class open systems for inference and training.Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.We're backed by well‑known infrastructure investors and partner with Nvidia, Google, AWS, and frontier AI labs. Join us in building infrastructure that gives real leverage back to the AI community. Compensation We offer competitive compensation with meaningful equity, comprehensive benefits, and flexible work arrangements. Compensation depends on location, experience, and level. RadixArk is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more. #J-18808-Ljbffr RadixArk

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Member of Technical Staff — Developer Technology in Palo Alto, CA vacancy
  • Member of Technical Staff - Developer Experience About the Role RadixArk is seeking a Developer Advocate to build and engage our technical community around SGLang, Miles, and our open source infrastructure. SGLang already has 20K+ GitHub stars and serves billions of tokens... 
    Suggested
    Flexible hours

    RadixArk

    Palo Alto, CA
    4 days ago
  •  ...world models: causal, multimodal systems that learn to predict and interact with the world over long horizons. This foundational technology promises to revolutionize robotics, science, healthcare, education, gaming, defense, and beyond. Odyssey's founders previously pioneered... 
    Suggested

    Doist

    Palo Alto, CA
    2 days ago
  • $200k - $420k

    Member of Technical Staff, Hardware, Compiler Engineer At River, our mission is to create personal AI owned and shaped by each individual. To achieve...  ...dialects and LLVM frameworks. Custom Backend Development: Develop and maintain the backend toolchain for our custom silicon,... 
    Suggested
    Local area
    Visa sponsorship
    Work visa
    Relocation package
    Flexible hours

    River AI Inc.

    Palo Alto, CA
    2 days ago
  •  ...breakthrough lays the foundation for a massive shift in multiple technologies—from 3D sensing and imaging to optical networking, free space...  ..., wake sources) and validate current/latency targets. Develop and optimize drivers and middleware for common interconnects (... 
    Suggested
    Shift work

    Lumotive

    Milpitas, CA
    more than 2 months ago
  • $200k - $420k

    At River AI, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, bespoke training infrastructure, next‑generation UIs, and frontier deep learning...
    Suggested
    Local area
    Visa sponsorship
    Relocation package

    Doist

    Palo Alto, CA
    1 day ago
  • $142.8k - $274.8k

     ...5%Profession: Software EngineeringDiscipline: Software EngineeringCompany: MicrosoftOverviewMicrosoft AI is looking for a Member of Technical Staff - Full Stack - Engineering Manager to help build the next wave of capabilities of our personalized AI assistant, Copilot.... 
    Ongoing contract
    Work at office
    Local area

    Microsoft

    Mountain View, CA
    2 days ago
  •  ...realities of engineering design workflows. Role Summary Hands-on technical role spanning validation, training data, and customers. Vinci's...  ...3 days a week in person. We hire remote for the right person — members of the team already work this way. Travel is occasional.... 
    Full time
    Remote work
    2 days per week
    3 days per week

    Vinci4d

    Palo Alto, CA
    2 days ago
  • $200k - $420k

    At River, our mission is to create personal AI owned and shaped by each individual. To achieve this, we are rewriting the entire stack from scratch: personal hardware for local inference, custom training infrastructure, next-generation UIs, and frontier deep learning research...
    Local area
    Visa sponsorship
    Work visa
    Relocation package

    Doist

    Palo Alto, CA
    1 day ago
  • $140k - $160k

     ...Dimitri Stiliadis, and is backed by leading VC firms such as Dell Technology Capital, Lightspeed, and Sierra Ventures. Sound interesting?...  .... Why Endor Labs We\'re building at the intersection of developer productivity and security — one of the fastest-growing spaces... 
    Shift work

    Endor Labs

    Palo Alto, CA
    9 hours ago
  • Member of Technical Staff - Applied AI Research San Francisco, CA; Sunnyvale, CA DoorDash’s mission...  ...huge opportunity to leverage agentic technology to help businesses grow their sales,...  ...research at DoorDash. In this role you will develop realistic agent environments and... 
    Hourly pay
    Work at office
    Local area
    Immediate start
    Remote work
    Flexible hours

    DoorDash

    Sunnyvale, CA
    2 days ago
  •  ...world models: causal, multimodal systems that learn to predict and interact with the world over long horizons. This foundational technology promises to revolutionize robotics, science, healthcare, education, gaming, defense, and beyond. Odyssey’s founders previously pioneered... 
    Remote work
    Flexible hours

    Odyssey

    Palo Alto, CA
    3 days ago
  • Build technology that truly matters. We’re working with a high-growth, venture-backed company...  ...Generative AI and healthcare. They’re developing advanced infrastructure that enables...  ...closely with a strong engineering team on technically challenging problems. Design, build,... 
    Flexible hours
    3 days per week

    DeepRec.ai

    Palo Alto, CA
    2 days ago
  • $300k - $400k

    Member of Technical Staff - Applied AI Engineering Location: San Francisco Bay Area | Hybrid Role Description AIVista turns foundation models into specialized, governed agents that run mission-critical, regulated enterprise operations reliably at scale. But most enterprise... 
    Work experience placement
    Local area
    Flexible hours

    NTT DATA AIVista

    Palo Alto, CA
    4 days ago
  •  ...parallel computing, or in physics simulation. Responsibilities Develop and optimize algorithms for geometry import, repair,...  ...performance and memory efficiency. Experience working in a relevant technical domain such as computational geometry, graphics, simulation, or... 

    Getvinci

    Palo Alto, CA
    3 days ago
  • $120k - $200k

     ...high-performance AI systems. About the Role As a Member of Technical Staff, Platform, you'll build full-stack product features Expers...  ...customer outcomes as the measure of success, not the technology you shipped. AI-native by default: you already use AI... 
    Full time
    Flexible hours

    Abaka AI

    Mountain View, CA
    1 day ago
  •  ...MosaixSoft, Inc. is recruiting for our Los Altos, CA office: Member of the Technical Staff (job code #37659). Design, architect, and implement...  ...solution is easy to deploy and configure and without conflicts. Develop ad‑hoc tools to aid with testing. Responsible for the... 
    Work at office

    MosaixSoft, Inc.

    Los Altos, CA
    1 day ago
  • $324k - $396k

     ...About the Role Member of Technical Staff (X.AI LLC; Palo Alto, CA): Introduce innovative techniques and analyses to the AI field to facilitate...  ...distributed system engineers and AI researchers in developing technologies in the area of natural language processing, computer... 
    Remote work

    Xai

    Palo Alto, CA
    3 days ago
  •  ...experience for the user. Build the data and evaluation foundations that let these systems learn and improve with usage. Help shape the technical direction of ranking, recommendations, and personalization at Perplexity. What we're looking for Deep, hands-on experience... 

    Pantera Capital

    Palo Alto, CA
    3 days ago
  • Member of Technical Staff (Software Engineer) Sunnyvale, CA Cerebras Systems builds the world’s largest AI chip, 56 times larger than GPUs. Its...  ...multi‑region deployments and disaster recovery strategies. Develop Python‑based scripts and APIs to streamline data... 
    Full time
    Part time
    Internship

    Cerebras Systems

    Sunnyvale, CA
    9 hours ago
  • $180k - $220k

     ...build new products to help us improve our MVP. Your Responsibilities You’ll be a force multiplier for our team and will design and develop the pipelines and tools that make our product a product. This includes accelerating our development by developing the system that... 

    Vinci4D.ai

    Palo Alto, CA
    2 days ago
  •  ...Apple and Intel. What You’ll Do As Member of the Technical Staff - Software at Architect, you’ll build...  ...design. You’ll own the full stack, developing intuitive, high‑performance systems...  .... Demonstrated ability to learn new technologies and frameworks on the fly. What We Offer... 

    Architect Labs

    Palo Alto, CA
    1 day ago
  • $220k - $405k

     ...build, and own product and platform systems for Computer Lead features, projects and products end-to-end, from problem definition to technical design, implementation, and launch. Hill climb on hard problems, continuously iterating to improve for ourselves and customers.... 
    Full time
    Local area

    Kindredventures

    Palo Alto, CA
    3 days ago
  •  ...iterate on cutting-edge AI models powering our core experience. As an expert in machine learning and artificial intelligence, you will develop scalable and impactful solutions for user personalization, query understanding, and content discovery - fulfilling the curiosity of... 
    Full time

    Kindredventures

    Palo Alto, CA
    2 days ago
  •  ...world over long horizons. This foundational technology promises to revolutionize robotics,...  ...TLM role sits at the intersection of deep technical research and scientific leadership: you’...  ...direction of the field. Who you are A staff‑level or senior ML researcher with a... 

    Odyssey

    Palo Alto, CA
    2 days ago
  •  ...fully in-person , between our Palo Alto HQ and San Francisco office. We’re a flat, talent-dense organization dedicated to solving technical and creative problems. We seek like-minded individuals who share our passion for applying science, creativity, and consistency to... 
    Work at office
    Visa sponsorship
    Flexible hours

    Parallel Web Systems

    Palo Alto, CA
    3 days ago
  •  ...founding team from Anthropic, Google DeepMind, Meta SuperIntelligence, xAI, Apple and Intel. What You'll Do As a Founding Member of the Technical Staff on the RTL Design team at Architect, you'll own the AI-driven microarchitecture and RTL design of mission-critical SoC... 

    Architect

    Palo Alto, CA
    1 day ago
  • $200k - $300k

     ...drug discovery and biomedical research. We are looking for a Member of Technical Staff to help design, build, and deploy the systems that power...  ..., including AI agents, scientific data infrastructure, developer tools, and product systems. This is a broad engineering role... 
    Work experience placement
    Work at office
    Visa sponsorship

    GXL

    Palo Alto, CA
    1 day ago
  •  ...engaging, more meaningful, and more fun than passive content — and we think everyone should be able to make them. As a Frontend Member of Technical Staff (MTS), you’ll help make that a reality. You’ll own the creation and play experiences that millions of users touch every... 

    Astrocade

    Palo Alto, CA
    3 days ago
  • About the Role As a Member of Technical Staff [Platform] at NeoCognition , you’ll design and build the internal systems that power everything...  ...deployments. You’ll create the tooling, infrastructure, and developer experience that enable our team to iterate rapidly, deploy... 

    NeoCognition

    Palo Alto, CA
    4 days ago
  •  ...SuperIntelligence, xAI, Apple and Intel. What You’ll Do As a Founding Member of the Technical Staff at Architect, you'll be at the forefront of training AI...  ...-grade chips going into production at leading foundry technologies like TSMC. Responsible for co-designing and implementing... 

    Architect Labs

    Palo Alto, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Member of Technical Staff — Developer Technology. Be the first to apply!