Get new jobs by email
  • $2,000 per month

     ...the Etched serving front-end. Representative projects Build scheduling logic for handling continuous batching and real time inference Implement inference-time acceleration techniques such as speculative decoding, tree search, KV cache sharing, etc. Implement... 
    Suggested
    Full time
    Work at office
    Relocation package

    Etched

    Cupertino, CA
    7 hours ago
  • $8k

     ...TS/SCI) clearance with polygraph is required.  Visionist has an exciting new, fully FUNDED opportunity for a Software Engineer - Inference on our largest PRIME contract. Our team of Analysts and Engineers is motivated by the direct impact on the mission, crafting... 
    Suggested
    Permanent employment
    Full time
    Contract work
    Temporary work
    Immediate start
    Flexible hours

    Visionist, Inc.

    Remote
    7 hours ago
  • $135k - $160k

     ...the technologies to make this possible, with the ultimate goal of enabling human life on Mars. APPLICATION SOFTWARE ENGINEER, INFERENCE The application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Remote work
    Worldwide
    Weekend work

    Spacex

    Palo Alto, CA
    7 hours ago
  •  ...About the Role We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model... 
    Suggested
    Full time
    Work at office
    3 days per week

    Pika

    Palo Alto, CA
    7 hours ago
  • $188k - $275k

     ...Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at  . What You’ll Do: Inference Platform Team The Inference team builds and operates CoreWeave’s Kubernetes-native inference platform, powering low-latency, high... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    Core Weave

    Sunnyvale, CA
    7 hours ago
  • $190k - $225k

     ...AssemblyAI builds the best-in-class Voice AI models powering the next generation of voice applications. Our models serve 600M+ inference calls monthly, process 1M+ hours of audio daily, and power 2 billion+ end-user experiences. The Voice AI space is at an inflection... 
    Suggested
    Full time

    Assemblyai, Inc.

    United States
    7 hours ago
  •  ...the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is... 
    Suggested
    Full time

    Cerebras Systems

    Sunnyvale, CA
    7 hours ago
  • $133k - $150k

     ...learning. Thanks for your interest in joining our team!   Key statistics: 950+ Consultants 640+ Ph.D.s 90+ Disciplines 30+ Offices globally...  ...projects involving exploratory data analysis, statistical inference, and predictive modelingApplying risk analysis methodologies to... 
    Suggested
    Work at office
    Flexible hours

    Exponent

    Bellevue, WA
    7 hours ago
  • $160k - $230k

     ...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and... 
    Suggested
    Full time

    Together Ai

    San Francisco, CA
    7 hours ago
  •  ...About the Team Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprise and developers alike to use and access our start-of-the-art AI models, allowing them to do things that they’ve never... 
    Suggested
    Full time

    OpenAI

    San Francisco, CA
    7 hours ago
  •  ...the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is... 
    Suggested
    Full time

    Cerebras Systems

    United States
    7 hours ago
  •  ...of learning. Thanks for your interest in joining our team! Key statistics:950+ Consultants640+ Ph.D.s90+ Disciplines30+ Offices globally...  ...projects involving exploratory data analysis, statistical inference, and predictive modelingApplying risk analysis methodologies to... 
    Suggested
    Work experience placement
    Work at office
    Flexible hours

    Exponent

    Phoenix, AZ
    7 hours ago
  • $320k

     ...policy experts, and business leaders working together to build beneficial AI systems. About the Role Our mandate is to make inference deployment boring and unattended. Anthropic serves Claude to millions of users across GPUs, TPUs, and Trainium — and every... 
    Suggested
    Full time
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    7 hours ago
  •  ...About the Team Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks and drive faster, cheaper inference. We combine systems profiling, benchmarking, and analysis to understand where time and cost are spent... 
    Suggested
    Full time

    OpenAI

    San Francisco, CA
    7 hours ago
  •  ...MassachusettsCompany: Planet Pharma GroupContact: ApplicationsEmail: ****@*****.*** Market RateSummary of Position:We seek a statistical geneticist to leverage cutting-edge human genetic datasets and methodologies to identify and validate therapeutic targets and... 
    Suggested

    Planet Pharma

    Cambridge, MA
    3 days ago
  •  ...their organizations. We're building the products and infrastructure that power the next generation of AI. The Foundation Model Inference team is the backbone of Databricks’ generative AI capabilities. We build the infrastructure that enables our customers to serve, scale... 
    Full time

    Databricks

    San Francisco, CA
    7 hours ago
  •  ...About the Team We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hiring... 
    Full time

    OpenAI

    San Francisco, CA
    7 hours ago
  • Job-ID31618032Reference25-04664 Responsibilities: Develop SAS programs and statistical output for the management and reporting of clinical trial data managed by GCD Guarantee quality of statistical output produced by external provider, to program tools to support data review... 

    Katalyst Healthcares & Life Sciences

    Tampa, FL
    7 hours ago
  • $320k

     ...engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role The Cloud Inference team scales and optimizes Claude to serve the massive audiences of developers and enterprise companies across AWS, GCP, Azure, and... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    7 hours ago
  •  ...the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is... 
    Full time

    Cerebras Systems

    Sunnyvale, CA
    7 hours ago
  • Job-ID27995022Reference24-01027Responsibilities and Requirements:Experience working in various disease areas.They also must have specific expertise in SDTM programming, have submissions experience, and MUST possess strong skill in graphing.They need to be comfortable working...

    Katalyst Healthcares & Life Sciences

    Trumbull, CT
    7 hours ago
  • Job-ID31508322Reference25-04459 Requirements: B.Sc., M.Sc., or equivalent education in Biostatistics/Statistics, Bioinformatics, Life Sciences, Mathematics, Physical Sciences, or related fields Advanced SAS macro programming experience is required for the development of... 
    Work experience placement

    Katalyst Healthcares & Life Sciences

    Atlanta, GA
    7 hours ago
  • Job-ID31590458Reference25-04599 Responsibilities: The SAS/Statistical Programmer will also be responsible for ensuring the accuracy and consistency of clinical data. The successful candidate will work closely with the Biostatisticians, Data Managers, and Clinical Research... 

    Katalyst Healthcares & Life Sciences

    North Chicago, IL
    7 hours ago
  • $143.7k - $194.4k

     ...Inferentia and Trainium cloud-scale machinelearning accelerators. This role is for a senior software engineer in the Machine Learning Inference Applications team. This role is responsible for development and performance optimization of core building blocks of LLM Inference... 
    Internship
    Flexible hours

    Amazon

    Seattle, WA
    2 days ago
  •  ...'s Private Cloud AI organization is seeking a Principal Software Engineer to lead the model runtime within HPE AI Essentials, the inference platform used by enterprises to operate large language models on infrastructure they own, including air-gapped and sovereign environments... 
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    Remote work
    2 days per week

    Hewlett Packard Enterprise

    Durham, NC
    7 hours ago
  • $151.8k - $332.2k

    What you can expect We are looking for an AI Inference Engineer with a solid background in speech recognition and model inference. In this role, you will develop state-of-the-art automatic speech recognition system and ship it to various Zoom products. You will work on... 
    Full time
    Work at office
    Remote work

    Zoom

    Seattle, WA
    7 hours ago
  •  ...About the Team OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production,... 
    Full time

    OpenAI

    San Francisco, CA
    7 hours ago
  •  ...289Reference26-00699 Job Summary: The Contract Lead Programmer serves as a senior, hands-on programmer responsible for managing statistical programming activities across multiple clinical trials and therapeutic areas. The role oversees external CRO partners to ensure high... 
    Contract work
    Shift work

    Katalyst Healthcares & Life Sciences

    Boston, MA
    1 day ago
  •  ...motivated, independent, committed to the generation of high-quality data within a fast-paced and innovative research team. The Senior Statistical Programmer will serve as the lead statistical programmer and is responsible for statistical programming activities such as... 
    Shift work

    Katalyst Healthcares & Life Sciences

    Baltimore, MD
    1 day ago
  •  ...development, market access, and commercial use cases for our life sciences partners Independently translate analytic specifications from a statistical analysis plan into R code to create analytic datasets, generating descriptive and inferential statistics, data visualizations,... 

    Katalyst Healthcares & Life Sciences

    Raleigh, NC
    7 hours ago