Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff DevTech: Accelerate AI Inference on GPUs

RadixArk

RadixArk is seeking a Member of Technical Staff to accelerate LLM inference and training on modern GPUs. You will profile, optimize, and extend SGLang and Miles across production workloads, collaborating with partner teams to push the performance envelope and broaden model support. You will work on kernel development, day-0 model support, and enablement for new silicon, while contributing to the roadmap and ecosystem integrations that power scalable AI inference and training. #J-18808-Ljbffr RadixArk

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Staff DevTech: Accelerate AI Inference on GPUs in Palo Alto, CA vacancy
  •  ...is seeking a Member of Technical Staff, Developer Technology (DevTech) to accelerate LLM inference and training on modern GPU...  ...researchers and partners to scale AI workloads. You will profile, optimize...  ...extend SGLang and Miles across GPUs, collaborate with ecosystem... 
    Suggested

    RadixArk

    Palo Alto, CA
    2 days ago
  •  ...is seeking a Senior hardware architect to design and implement accelerators for codesigned systems. You will own subsystems from requirements...  ...with ML model design and numerics. Google DeepMind values safety and ethics in AI development. #J-18808-Ljbffr Google DeepMind
    Suggested

    Google DeepMind

    Mountain View, CA
    4 days ago
  •  ...Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras...  ...deliver industry-leading training and inference speeds; over 10 times faster than...  ...roleAs demand for AI continues to accelerate, intelligent capacity management becomes... 
    Suggested
    Remote work

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  • MatX is building custom silicon for LLM inference and training. You will develop the host-side interface library, manage device memory, DMA, streams and events, and extend the executable format to enable safe evolution of compiler-runtime contracts. You will design the... 
    Suggested

    MatX

    Mountain View, CA
    2 days ago
  • $207k - $235k

     ...build and deploy proprietary technology, AI, and analytics to help us scale fast....  ...scalable analytics tools and processes that accelerate decision-making across the Growth team...  ...understanding of experimentation design, causal inference, and marketing measurement best practices... 
    Suggested
    Local area

    Quince

    Palo Alto, CA
    3 days ago
  • $230k - $250k

     ...Cerebras Systems in Sunnyvale, CA, seeks a Sr. Member of Technical Staff to develop resilient software for their AI chip. Responsibilities include designing robust software features, maintaining deployment workflows using AWS, and debugging software issues. Candidates... 
    Remote work

    Cerebras

    Sunnyvale, CA
    4 days ago
  • $197k - $316k

     ...are seeking an experienced, visionary Sr. Staff Enterprise Architect with deep expertise...  ...decision frameworks, driving operational AI enablement, and leading complex enterprise...  ...integrations, and automation agents to accelerate internal operational efficiency and productivity... 
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Mountain View, CA
    10 hours ago
  • $207k - $300k

     ...engineers on a mission to tackle climate change by developing novel AI reasoning capabilities that enable stakeholders to target their...  ...-world enterprise applications. How you will have 10X impact:Accelerate the automated creation of scientific and techno-economic models,... 
    Full time

    X Company

    Mountain View, CA
    4 days ago
  • $188k - $274k

     ...thermals. Support critical user journeys, including on-device AI for audio and ecosystem interoperability across Google’s platforms...  ...domains.Experience with machine learning models and AI accelerators, such as Neural Processing Units (NPUs).Experience in ambient and... 

    Google

    Mountain View, CA
    3 days ago
  • $197k - $285k

     ...design, validation and manufacturing readiness.Mentor engineering staff in system-level thinking, architecture methods, and technical...  ...architecture, and/or network designDeep understanding of AI/ML accelerators and energy efficient compute designBackground in performance... 
    Work at office
    Local area
    3 days per week

    Aurora Innovation

    Mountain View, CA
    2 days ago
  • $148.7k - $201.2k

     ...backbone of Generative AI at AWS? Do you want to...  ...cloud for AI training and inference, delivering continuous...  ...infrastructure for our accelerated (AI/ML) server...  ...accelerator servers with GPUs. Located in Seattle, Cupertino...  ..., supervisors, and staff; adhere to standards of... 
    Internship
    Local area
    Worldwide
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  •  ...Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to...  ...deliver industry-leading training and inference speeds; over 10 times faster than...  ...functionsRole OverviewWe are looking for a Staff Business Systems Analyst who can... 

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  • NVIDIA in Santa Clara, CA is seeking an ATE Test Engineer to design, develop, and support automated test solutions for GPUs and AI server products. You will work with design, DFT and manufacturing teams to improve yields, test coverage, and reduce costs across continents... 

    NVIDIA

    Santa Clara, CA
    10 hours ago
  • $238k - $302k

     ...growth. Timeliness: As Waymo’s deployments accelerate across the globe, time to first detection...  ...development teams including large AI deployments, System engineers and Data scientists...  ...but not limited to: human triage, VLM inference, clustering. Key workstreams that the TLM... 
    Full time
    Remote work

    Waymo

    Mountain View, CA
    2 days ago
  • $272k - $425.5k

     ...platforms bring together the full power of NVIDIA GPUs, NVLink, InfiniBand networking, Grace CPUs, and our optimized AI/HPC software stack. This deep technical...  ...experience working with complex system software for accelerators such as GPUs, DPUs, or FPGAs.* Strong... 

    NVIDIA Corporation

    Santa Clara, CA
    10 hours ago
  • $216k - $414k

     ...transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s an...  ...tapping into the unlimited potential of AI to define the next era of computing....  ...performance networking technologies, DPUs and GPUs to join the NVIDIA Network Software... 

    NVIDIA

    Santa Clara, CA
    1 day ago
  •  ...the compiler→runtime contract. You will design the custom-kernel ABI, and implement Python bindings to move tensors from Python to accelerator hardware. The role involves working with CUDA/ROCm-style accelerators, memory models, and high-performance computing stacks,... 
    Contract work

    MatX Inc.

    Mountain View, CA
    2 days ago
  • $252k - $274k

    Identify the next wave of user needs in generative AI. Move beyond tactical testing to lead foundational research that informs the...  ...advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our... 

    Google

    Mountain View, CA
    10 hours ago
  • $177k - $349k

    About the RoleWe are seeking a Staff Enterprise Architect, Data to lead the strategy, design, and modernization of our enterprise data landscape...  ...at the intersection of data architecture, engineering, and AI enablement, defining solutions to integrate our Data Lake and... 
    Contract work
    Work at office
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Palo Alto, CA
    4 days ago
  •  ...Clara seeks a creative ATE Test Engineer to help transfer GPU and AI server products from design to mass production. You will work...  ...server products. You will define and support ATE test programs for GPUs and AI servers; develop test methods, scripting, and automation;... 
    Overseas

    Thomas To

    Santa Clara, CA
    3 days ago
  • $186k - $233k

     ...team is seeking a highly technicalSenior Staff Solutions Architect, Field Applications to...  ...will design scalable solutions that accelerate delivery, improve platform utilization, reduce...  ...assess use cases and implement compliant AI and automation features across field... 
    Full time
    Local area

    Revolution Medicines

    Redwood City, CA
    10 hours ago
  • $192k - $278k

     ...s mission is to organize the world's information and make it universally accessible and useful. Our team combines the best of Google AI, Software, and Hardware to create radically helpful experiences. We research, design, and develop new technologies and hardware to make... 
    Worldwide

    Google

    Mountain View, CA
    4 days ago
  • NVIDIA is seeking a Test Methodology Engineer to join our Santa Clara team. You will transfer GPU/AI server products from design to mass production, define and support ATE test programs, and collaborate with overseas manufacturing to improve yields and reduce costs. The... 
    Overseas

    Nvidia Corporation in

    Santa Clara, CA
    4 days ago
  • $176.1k - $308.2k

     ...meaningful work. Today, ServiceNow is the AI control tower for business reinvention....  ...About the role We are looking for a Staff FinOps AI Governance Lead to drive financial...  ..., context length management, batch inference, and caching. Familiarity with provider... 
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    2 days ago
  • $189k - $210k

    Lightmatter is leading the revolution in AI data center infrastructure, enabling the...  ...valuation of $4.4 billion. We will continue to accelerate the development of data center photonics...  ...role is open to two different levels. Staff: $189,000 - $210,000 Sr Staff: $217,000 -... 
    Full time
    Contract work
    Temporary work
    Flexible hours
    Shift work

    Socket.dev

    Mountain View, CA
    4 days ago
  • Architect Labs in Palo Alto seeks a founding member of the Technical Staff focused on Formal Methods to own the formal stack end-to-end—spec language, IR, proof obligations, solver integration, and empirical evidence against the spec. This hands-on role collaborates with... 

    Architect

    Palo Alto, CA
    1 day ago
  • Intuit is seeking a Sr Staff Technical Program Manager to drive AI-native planning for the Tech Ecosystem. You will partner with program managers, business operations and Finance to design processes and leadership mechanisms that inform senior leaders and shape C-suite... 

    Intuit

    Mountain View, CA
    2 days ago
  • Intuit is seeking a Staff Product Manager for Benefits & Insurance Services in Mountain View, CA. You will define the vision, strategy, and roadmap for AI-native benefits enrollment, administration, and compliance, ensuring seamless integration with QuickBooks Online services... 

    Intuit

    Mountain View, CA
    1 day ago
  • $250k - $300k

     ...Job Description Crusoe is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the...  ...This Role Crusoe is looking for a Senior Staff Network Architect to help define and evolve... 
    Temporary work

    Crusoe

    Sunnyvale, CA
    17 days ago
  • $272k - $431.25k

     ...Today, we are increasingly known as “the AI computing company.” We're looking to grow...  ...large clusters and data centers deploying GPUs and Grace solution from Nvidia. Work...  ...characteristic protected by law. NVIDIA pioneered accelerated computing. Today, our AI infrastructure... 
    Full time
    Work at office
    Remote work

    NVIDIA

    Santa Clara, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff DevTech: Accelerate AI Inference on GPUs. Be the first to apply!