Senior Research Engineer / Scientist - Storage for LLM
$202.16k - $368.22kByteDance
Senior Research Engineer / Scientist - Storage for LLM Location: Seattle Team: Infrastructure Employment Type: Regular Job Code: A152690 About the Team The Infrastructure System Lab is a hybrid research and engineering group focused on building next‑generation AI‑native data infrastructure. Positioned at the intersection of databases, large‑scale systems, and AI, the team leads innovation in areas such as vector and multi‑modal databases, infrastructure optimization through machine learning, and LLM‑based tooling like NL2SQL and NL2Chart. They also develop high‑performance cache systems, including multi‑engine key‑value stores and LLM inference KV caches. The team thrives on collaboration, with researchers and engineers working closely to take ideas from paper to prototype to production. Their work supports key products used by millions and is regularly published and deployed at scale. About the Role We are seeking a systems researcher or engineer with deep expertise in large‑scale distributed storage and caching infrastructure to design and maintain a high-performance KV cache layer for large language model (LLM) inference. This role focuses on improving latency, throughput, and cost‑efficiency in transformer‑based model serving by optimizing the reuse of attention key‑value states and prompt embeddings. You’ll work on cutting‑edge AI systems problems with real‑world impact, alongside a world‑class team. The role offers opportunities to publish, contribute to open‑source, attend top conferences, and enjoy competitive compensation, generous research resources, and an innovation‑driven culture. Responsibilities Design and implement a distributed KV cache system to store and retrieve intermediate states (e.g., attention keys/values) for transformer‑based LLMs across GPUs or nodes. Optimize low‑latency access and eviction policies for caching long‑context LLM inputs, token streams, and reused embeddings. Collaborate with inference and serving teams to integrate the cache with token streaming pipelines, batched decoding, and model parallelism. Develop cache consistency and synchronization protocols for multi‑tenant, multi‑request environments. Implement memory‑aware sharding, eviction (e.g., windowed LRU, TTL), and replication strategies across GPUs or distributed memory backends. Monitor system performance and iterate on caching algorithms to reduce compute costs and response time for inference workloads. Evaluate and, where needed, extend open‑source KV stores or build custom GPU‑aware caching layers (e.g., CUDA, Triton, shared memory, RDMA). Qualifications Minimum Qualifications: PhD in Computer Science, Applied Mathematics, Electrical Engineering, or a related technical field. Strong understanding of transformer‑based model internals and how KV caching affects autoregressive decoding. Experience with distributed systems, memory management, and low‑latency serving (RPC, gRPC, CUDA‑aware networking). Familiarity with high‑performance compute environments (NVIDIA GPUs, TensorRT, Triton Inference Server). Proficiency in languages like C++, Rust, Go, or CUDA for systems‑level development. Preferred Qualifications: Prior experience building inference‑serving systems for LLMs (e.g., vLLM, SGLang, FasterTransformer, DeepSpeed, Hugging Face Text Generation Inference). Experience with memory hierarchy optimization (HBM, NUMA, NVLink) and GPU‑to‑GPU communication (NCCL, GDR, GDS, InfiniBand). Exposure to cache‑aware scheduling, batching, and prefetching strategies in model serving. Job Information The base salary range for this position in the selected city is $202160 - $368220 annually. Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units. Benefits Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short‑term and long‑term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure). The Company reserves the right to modify or change these benefits programs at any time, with or without notice. Equal Opportunity Statement For Los Angeles County (unincorporated) Candidates: Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state, and local laws including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Our company believes that criminal history may have a direct, adverse and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment: Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues; Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems; Exercising sound judgment. #J-18808-Ljbffr ByteDance
- ByteDance in Seattle is seeking a Senior Research Engineer/Scientist for Storage for LLM to design and maintain a high-performance KV cache layer for LLM inference across GPUs and nodes. This role focuses on reducing latency, increasing throughput, and lowering cost for...Senior
$202.16k - $368.22k
Senior Research Engineer / Scientist - AI for Databases Location: Seattle Team: Infrastructure Employment Type... ...such as query planning, indexing, storage management, and workload prediction/... ...skills. Familiarity with LLM, reinforcement learning, neural architecture...SeniorTemporary workLocal area- About the TeamThe Vision-Applied Research team focuses on applied research in Generative... ....The team is looking for a Research Engineer / Scientist who can take initiatives in designing and... ...Proficiency in training generative AI or LLM models using widely adopted frameworks...Senior
$104.17k
...opportunity for a Scientific Director to join their team.About this OpportunityReporting to the Vice Chair of Research, the Scientific Director serves as a senior scientific leader within the Department of Emergency Medicine, providing advanced research expertise and...SeniorFull timeTemporary workTraineeshipWork at officeShift work$165k - $310k
...AI systems—designed to take ideas from research to production with less friction.... ...We're are looking for an experienced Senior Research Engineer who has built, trained, and optimized... ...PyTorch. Experience with modern LLM training and post-training techniques...SeniorFull timeWork at officeRemote workWork from homeFlexible hours2 days per week$202.16k - $368.22k
Senior Research Scientist - DPU & AI Infra Location: Seattle Team: Infrastructure Employment Type: Regular... ...for ByteDance and Volcano Engine Public Cloud. Our mission is to advance... ...technologies across compute, networking, and storage for cloud and AI computing. Our...SeniorTemporary workLocal area$136k
...large language models (LLM), generative audio (... ...-knit team of applied scientists and product managers who... ...bring cutting edge research to raise the bar within... ...model fine-tune to prompt engineering. A strong developing... ...explain methods to senior leadership. Willingness...Senior$146.88k - $220.32k
...Research Engineer Ai2 is a Seattle based non-profit AI research institute founded in 2014 by the late Paul Allen. Our mission is building... ...responsibilities: Building and optimizing infrastructure for LLM, multimodal, and agentic research — including training/...SeniorWork at officeWeekend work$192k - $304.75k
...now looking for an Applied Deep Learning Research Scientist, Efficiency!Join our ADLR - Efficiency... ...in AI, computer science, computer engineering, math or a related field or equivalent... ...network architectures, optimizers and LLM training.Experience with modern DL training...SeniorFull time$232.56k - $427.5k
Research Engineer - LLM/VLM Inference Optimization (Seed Infra) Location: Seattle Team: Technology Employment Type: Regular Job Code: A236224 Responsibilities Design, develop, and optimize high-performance inference systems for large-scale LLMs and VLMs, covering inference...Temporary workLocal area$182k - $242k
...ablations after you read it. This is an applied research role. You will be expected to generate... ...qualifications, we’re looking for strong engineers with great taste. The most important... ...everything from CUDA kernels to high-performance LLM tracing dashboards, and you will have an...SeniorPermanent employmentTemporary workCasual workWork at officeFlexible hours$167.1k - $226.1k
...into fulfillment center storage bins, alongside human... ...that need experienced scientists to drive from... ...job responsibilities- Research, propose, architect, and... ...safety, operations, vendor engineering) to deliver novel, synergistic... ...and serve as the senior technical interface to...SeniorFull timeTemporary workSeasonal workFlexible hours$80.24k
Job DescriptionThe Division of Metabolism, Endocrinology and Nutrition (MET) has an outstanding opportunity for a Research Scientist/Engineer 3 position to support studies within the Pyle and Bjornstad Laboratories.Housed within the University of Washington Medicine Diabetes...Full timeTemporary workTraineeshipWork at officeShift work$120.7k - $238.6k
...from everywhere in the organization, and we know the next big idea could be yours!The OpportunityPhotoshop ART is seeking a Research Scientist to join our inpainting R&D team focused on making significant progress in image generation/restoration, low level vision, image...Full timeTemporary workInternshipLocal areaWorldwide$88.8k
...Division of Hematology & Oncology has an outstanding opportunity for a Research Scientist to join their Cancer Vaccine Institute (CVI) team.About this OpportunityReporting to the Research Scientist Senior at the Cancer Vaccine Institute (CVI), the Research Scientist 3 is...Full timeTemporary workTraineeshipWork at officeFlexible hoursShift workAfternoon shift- Axon seeks a Senior AI Research Scientist to join a new team focusing on agentic video and multimodal reasoning... ...with product managers and engineers to train models and deploy cutting‑edge... ...understanding, multimodal data, and LLM grounding, driving practical impact in...Senior
- CoreWeave's OpenPipe team is seeking an applied research engineer to advance continuous learning for self-improving agents. You will generate... ...8+ years in ML or a PhD with 4+ years, with deep expertise in LLM training methods, supervision, RL, and policy optimization, and...Senior
$125.5k - $169.8k
...experienced, innovative, hands-on, and customer-obsessed Sr. Innovation Engineer to lead the application of new automation technologies into our... .... Basic qualifications- 5+ years of design & innovation, research & development work experience- Bachelor's degree in Engineering...SeniorWork experience placementFlexible hours$132.1k - $178.8k
Last Mile Engineering is searching for an innovative and solutions-oriented engineering project... ...to be able to influence and work with senior leaders across multiple organizations... ...equivalent- 5+ years of design & innovation, research & development work experience-...SeniorWork experience placementLocal areaWorldwideFlexible hours$69.6k
Job Description The Institute for Protein Design has an outstanding opportunity for a Research Scientist/Engineer 2 to join their team. About This Opportunity As a member of the IPD Core R&D Labs, you will make important research contributions that accelerate progress...Full timeShift workDay shift$113.7k - $211.9k
Job responsibilities Conduct cutting‑edge research and development in Generative AI Develop and transfer novel technologies to Adobe... ...Generative AI Collaborate with world‑class researchers and engineers to bring research ideas to production Publish and present your...Temporary work$132.1k - $178.8k
...performance insights and drive engineering excellence? Does the... ..., analytically-minded Senior Performance Engineer... ..., and simulation scientists with a critical business... ..., automated storage and vision systems to... ...design & innovation, research & development work experience...SeniorWork experience placementWorldwideFlexible hours$80.24k - $108k
...Disease Diagnostics within the University of Washington Department of Laboratory Medicine and Pathology is seeking a talented Research Scientist/Engineer 3 (one-year, full-time appointment) to join our Research and Development Laboratory at the UW Central Lab in Renton, WA....Full timeWork at officeDay shift$68.74k
Research Scientist/Engineer2 in the Dr. Cory Simpson Lab at the South Lake Union campus of the University of Washington. This position involves wet‑lab work, laboratory organization, and training of junior members while contributing to the development of novel therapies...Full timeMonday to FridayDay shift- ByteDance is seeking a Sr. Research Engineer/Scientist (all levels) for efficient models in Seattle. The role focuses on designing and implementing compact, scalable models for large-scale generative AI, including distillation, compression, and hardware-aware inference....Senior
$57 per hour
...with the opportunity to actively contribute to our products and research, as well as to the organization's future plans and emerging technologies... ..., ViT-G, ViT-22B, EVA-enormous, etc) Explore the application of LLM in our business scenarios, like pre-training, zero-shot/few-shot...Hourly payInternshipLocal area- Research Scientist, LLM Evaluation & Post-Training page is loaded## Research Scientist, LLM Evaluation & Post-Traininglocations: Remote Work(... ...scientists, along with more than 4,000 AI practitioners and engineers. We harness the power of an integrated solution ecosystem—comprising...Full timeRemote work
- Centific Global Solutions, Inc. is seeking a Research Scientist for LLM Evaluation & Post-Training in Seattle, WA. This full-time role involves leading research initiatives to improve model evaluation methodologies. Responsibilities include developing evaluation frameworks...Full time
$78k - $185k
...Development team supports Portfolio Management, Strategy and Research with equity models and tools to guide our investment... ...tools.As part of the Quantitative Development team, the Senior Research Software Engineer will be responsible for modernizing existing...SeniorTemporary workWork at officeLocal areaRemote workWorldwide3 days per week$68.74k
...Climate Impacts Group (CIG) has an outstanding opportunity for a Research Scientist 2 - Social Scientist to join their team. About This... ...experience who can add breadth to the work we do and support CIG’s senior researchers on climate change adaptation projects with our...Full timeTemporary workWork at officeLocal areaRemote workShift workDay shift2 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Research Engineer / Scientist - Storage for LLM. Be the first to apply!
- deep learning research engineer Seattle, WA
- research engineer Seattle, WA
- research programmer Seattle, WA
- scientist ii Seattle, WA
- scientist 1 Seattle, WA
- image scientist Seattle, WA
- qc scientist Seattle, WA
- research scientist Seattle, WA
- analytical scientist Seattle, WA
- research scientist - biology Seattle, WA

