Get new jobs by email
$2,000 per month
...the Etched serving front-end. Representative projects Build scheduling logic for handling continuous batching and real time inference Implement inference-time acceleration techniques such as speculative decoding, tree search, KV cache sharing, etc. Implement...SuggestedFull timeWork at officeRelocation package$8k
...TS/SCI) clearance with polygraph is required. Visionist has an exciting new, fully FUNDED opportunity for a Software Engineer - Inference on our largest PRIME contract. Our team of Analysts and Engineers is motivated by the direct impact on the mission, crafting...SuggestedPermanent employmentFull timeContract workTemporary workImmediate startFlexible hours$135k - $160k
...the technologies to make this possible, with the ultimate goal of enabling human life on Mars. APPLICATION SOFTWARE ENGINEER, INFERENCE The application software team is the central nervous system of SpaceX – we create mission critical applications that are used throughout...SuggestedPermanent employmentFull timeTemporary workRemote workWorldwideWeekend work- ...About the Role We are seeking Senior/Staff level Inference Engineers to accelerate the performance of Pika's AI-driven products. In this highly technical role, you will operate at the intersection of cutting-edge inference acceleration, GPU parallelism, advanced model...SuggestedFull timeWork at office3 days per week
$188k - $275k
...Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at . What You’ll Do: Inference Platform Team The Inference team builds and operates CoreWeave’s Kubernetes-native inference platform, powering low-latency, high...SuggestedPermanent employmentFull timeTemporary workCasual workWork at officeRemote workFlexible hours$190k - $225k
...AssemblyAI builds the best-in-class Voice AI models powering the next generation of voice applications. Our models serve 600M+ inference calls monthly, process 1M+ hours of audio daily, and power 2 billion+ end-user experiences. The Voice AI space is at an inflection...SuggestedFull time- ...the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is...SuggestedFull time
$133k - $150k
...learning. Thanks for your interest in joining our team! Key statistics: 950+ Consultants 640+ Ph.D.s 90+ Disciplines 30+ Offices globally... ...projects involving exploratory data analysis, statistical inference, and predictive modelingApplying risk analysis methodologies to...SuggestedWork at officeFlexible hours$160k - $230k
...About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and...SuggestedFull time- ...About the Team Our Inference team brings OpenAI’s most capable research and technology to the world through our products. We empower consumers, enterprise and developers alike to use and access our start-of-the-art AI models, allowing them to do things that they’ve never...SuggestedFull time
- ...the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is...SuggestedFull time
- ...of learning. Thanks for your interest in joining our team! Key statistics:950+ Consultants640+ Ph.D.s90+ Disciplines30+ Offices globally... ...projects involving exploratory data analysis, statistical inference, and predictive modelingApplying risk analysis methodologies to...SuggestedWork experience placementWork at officeFlexible hours
$320k
...policy experts, and business leaders working together to build beneficial AI systems. About the Role Our mandate is to make inference deployment boring and unattended. Anthropic serves Claude to millions of users across GPUs, TPUs, and Trainium — and every...SuggestedFull timeWork at officeVisa sponsorshipFlexible hoursShift work- ...About the Team Our team analyzes inference stack performance across the application, model, and fleet layers to identify bottlenecks and drive faster, cheaper inference. We combine systems profiling, benchmarking, and analysis to understand where time and cost are spent...SuggestedFull time
- ...MassachusettsCompany: Planet Pharma GroupContact: ApplicationsEmail: ****@*****.*** Market RateSummary of Position:We seek a statistical geneticist to leverage cutting-edge human genetic datasets and methodologies to identify and validate therapeutic targets and...Suggested
- ...their organizations. We're building the products and infrastructure that power the next generation of AI. The Foundation Model Inference team is the backbone of Databricks’ generative AI capabilities. We build the infrastructure that enables our customers to serve, scale...Full time
- ...About the Team We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hiring...Full time
- Job-ID31618032Reference25-04664 Responsibilities: Develop SAS programs and statistical output for the management and reporting of clinical trial data managed by GCD Guarantee quality of statistical output produced by external provider, to program tools to support data review...
$320k
...engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role The Cloud Inference team scales and optimizes Claude to serve the massive audiences of developers and enterprise companies across AWS, GCP, Azure, and...Full timeWork at officeVisa sponsorshipFlexible hours- ...the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is...Full time
- Job-ID27995022Reference24-01027Responsibilities and Requirements:Experience working in various disease areas.They also must have specific expertise in SDTM programming, have submissions experience, and MUST possess strong skill in graphing.They need to be comfortable working...
- Job-ID31508322Reference25-04459 Requirements: B.Sc., M.Sc., or equivalent education in Biostatistics/Statistics, Bioinformatics, Life Sciences, Mathematics, Physical Sciences, or related fields Advanced SAS macro programming experience is required for the development of...Work experience placement
- Job-ID31590458Reference25-04599 Responsibilities: The SAS/Statistical Programmer will also be responsible for ensuring the accuracy and consistency of clinical data. The successful candidate will work closely with the Biostatisticians, Data Managers, and Clinical Research...
$143.7k - $194.4k
...Inferentia and Trainium cloud-scale machinelearning accelerators. This role is for a senior software engineer in the Machine Learning Inference Applications team. This role is responsible for development and performance optimization of core building blocks of LLM Inference...InternshipFlexible hours- ...'s Private Cloud AI organization is seeking a Principal Software Engineer to lead the model runtime within HPE AI Essentials, the inference platform used by enterprises to operate large language models on infrastructure they own, including air-gapped and sovereign environments...Full timeWork experience placementWork at officeLocal areaImmediate startRemote work2 days per week
$151.8k - $332.2k
What you can expect We are looking for an AI Inference Engineer with a solid background in speech recognition and model inference. In this role, you will develop state-of-the-art automatic speech recognition system and ship it to various Zoom products. You will work on...Full timeWork at officeRemote work- ...About the Team OpenAI’s Inference team powers the deployment of our most advanced models - including our GPT models, 4o Image Generation, and Whisper - across a variety of platforms. Our work ensures these models are available, performant, and scalable in production,...Full time
- ...289Reference26-00699 Job Summary: The Contract Lead Programmer serves as a senior, hands-on programmer responsible for managing statistical programming activities across multiple clinical trials and therapeutic areas. The role oversees external CRO partners to ensure high...Contract workShift work
- ...motivated, independent, committed to the generation of high-quality data within a fast-paced and innovative research team. The Senior Statistical Programmer will serve as the lead statistical programmer and is responsible for statistical programming activities such as...Shift work
- ...development, market access, and commercial use cases for our life sciences partners Independently translate analytic specifications from a statistical analysis plan into R code to create analytic datasets, generating descriptive and inferential statistics, data visualizations,...
