Research Engineer / Scientist - Storage for LLM Technology - Infrastructure San Jose Regular
$156k - $387.6kByteDance
Research Engineer / Scientist - Storage for LLM Location: San Jose Team: Infrastructure Employment Type: Regular Job Code: A244964 Responsibilities Design and implement a distributed KV cache system to store and retrieve intermediate states (e.g., attention keys/values) for transformer-based LLMs across GPUs or nodes. Optimize low‑latency access and eviction policies for caching long‑context LLM inputs, token streams, and reused embeddings. Collaborate with inference and serving teams to integrate the cache with token streaming pipelines, batched decoding, and model parallelism. Develop cache consistency and synchronization protocols for multi‑tenant, multi‑request environments. Implement memory‑aware sharding, eviction (e.g., windowed LRU, TTL), and replication strategies across GPUs or distributed memory backends. Monitor system performance and iterate on caching algorithms to reduce compute costs and response time for inference workloads. Evaluate and, where needed, extend open‑source KV stores or build custom GPU‑aware caching layers (e.g., CUDA, Triton, shared memory, RDMA). Qualifications Minimum Qualifications PhD in Computer Science, Applied Mathematics, Electrical Engineering, or a related technical field. Strong understanding of transformer‑based model internals and how KV caching affects autoregressive decoding. Experience with distributed systems, memory management, and low‑latency serving (RPC, gRPC, CUDA‑aware networking). Familiarity with high‑performance compute environments (NVIDIA GPUs, TensorRT, Triton Inference Server). Proficiency in languages such as C++, Rust, Go, or CUDA for systems‑level development. Preferred Qualifications Prior experience building inference‑serving systems for LLMs (e.g., vLLM, SGLang, FasterTransformer, DeepSpeed, Hugging Face Text Generation Inference). Experience with memory hierarchy optimization (HBM, NUMA, NVLink) and GPU‑to‑GPU communication (NCCL, GDR, GDS, InfiniBand). Exposure to cache‑aware scheduling, batching, and prefetching strategies in model serving. Job Information The base salary range for this position in the selected city is $156,000 – $387,600 annually. Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives and restricted stock units. Benefits Employees have day‑one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short‑term and long‑term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year, and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure). The Company reserves the right to modify or change these benefits programs at any time, with or without notice. Equal Opportunity Statement For Los Angeles County (unincorporated) Candidates: Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues; Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems; Exercising sound judgment. Reasonable Accommodation ByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at #J-18808-Ljbffr ByteDance
$244.8k
Research Engineer - LLM/VLM Inference Optimization (Seed Infra) Location San Jose Team Technology Employment Type Regular Job Code A02632A Responsibilities About the Team The Seed Infrastructures team oversees the distributed training, reinforcement learning framework...SuggestedTemporary workLocal area$212.8k - $387.6k
ByteDance in San Jose is seeking talented individuals to join our team focusing on AI infrastructure and LLM technologies. As a graduate, you will tackle complex challenges and have opportunities to innovate. Successful candidates must have a PhD in a relevant field and...Suggested$60 per hour
...26 TikTok Algorithm Research Scientist Intern (TikTok-Recommendation... ...(PhD) Location: San Jose Employment Type:... ...researchers and engineers, who support and innovate... ...plans and emerging technologies. Our dynamic... ...optimize algorithms and infrastructure to improve recommendation...SuggestedHourly payInternshipLocal area$156k - $387.6k
ByteDance in San Jose, California is seeking candidates for a research role in LLM/AI and infrastructure systems. The position requires a PhD in a relevant field and strong programming skills in languages such as C/C++, Go, Java, or Python. Successful candidates will engage...Suggested$60 per hour
Applied Scientist Intern (Recommendation AI... ...(PhD) Location: San Jose Employment Type:... ...focusing on generative technologies in search,... ...Science, Computer Engineering, or a related technical... ...as single-modal LLM application and... ...candidates with research results and extensive...SuggestedHourly payInternshipLocal area- ...creativity and enrich life around the globe. About the TeamThe Macro Research team operates as ByteDance's internal knowledge and advisory... ...research on economic and philosophical studies, geopolitics, technology governance, international affairs and history. The team tracks...Local area
$224k - $356.5k
...a senior or principal engineer who specializes in building cutting-edge infrastructure for large-scale foundation... ...Embodied Agent Research (GEAR) group. Our team... ...models and full-stack technology for humanoid robots.You... ...building large-scale LLM and multimodal LLM training...Full time- ByteDance Technology in San Jose seeks PhD-level researchers for a Compute division role. You will help build scalable AI infra, spanning training and inference... ..., and cross-functional collaboration to advance cloud infrastructure for global products. #J-18808-Ljbffr ByteDance
$212.8k - $387.6k
...Bytedance's system infrastructure is currently at... ..., focusing on LLM/AI + Infrastructure technologies, which includes... .... Topics & Research Areas With The... ...vector indexing engine to support ultra... ...full-text and regular SQL query processing... ...stability. Storage Systems:...Temporary workLocal area$187.04k - $359.72k
Senior Research Scientist, Foundation Model (LLM/ VLLM), TikTok - Trust and Safety... ...machine learning technologies and scale them to detect... ...machine learning engineers who can take... ...Analytical Chemistry San Mateo County, CA $5... ....00 1 week ago San Jose, CA $185,900.00-$22...Full timeTemporary workLocal area- ...Together, we advance your career. THE ROLE:We are hiring a AI Research Scientist - Infrastructure Engineer, Reinforcement Learning, to own reinforcement learning... ...orchestrationPrior ownership of RL training infra, LLM post-training pipelines, or large-scale experiment...
- ByteDance is seeking a Research Engineer / Scientist for their San Jose location. The role involves designing and optimizing distributed KV caching systems for transformer-based LLMs. Candidates must have a PhD in a relevant field and expertise in distributed systems and...
$162k - $316.8k
...excites you every day. @2026 TikTok Technology Machine Learning/Research Engineer Graduate (Monetization... ...) - 2027 Start (PhD) Location: San Jose Employment Type: Regular Job Code: A24339 Responsibilities... ...technologies, including ML/DL, RL, LLM, and scaling laws in ad...Temporary workLocal area$95.4k - $152.6k
...specialists, powering next-generation technologies through sophisticated solutions. Behind... ...environment and enjoy partnering with engineering teams to bring innovative products to market... ...Engineer at our office located in San Jose, CA. This role will support the day-to-...Work at officeImmediate startFlexible hours- Innovation Engineer - Market Development (Optoelectronics)About the RoleWe are seeking a... ...application growth analysis, and collaborate regularly with teams in Europe and Asia to scale... ...contact ****@*****.*** assistanceSummaryLocation: San Jose, CAType: Full timeFull timeWork from home
$139.7k - $272.8k
...specialists, powering next-generation technologies through sophisticated solutions. Behind... ...activities, and collaborate closely with engineering and product teams. You will help... ...is an on-site position (not remote) in San Jose, CA.We are only considering candidates...Local areaRelocationFlexible hours$122.5k - $196k
...specialists, powering next-generation technologies through sophisticated solutions.... ..., experienced Reliability Engineer to join our Memory Test Division in San Jose, California. In this role, you will... ..., wireless products, data storage and complex electronic systems,...Local areaRelocationFlexible hours$170.5k - $272.8k
...specialists, powering next-generation technologies through sophisticated solutions. Behind... ...technical expert supporting customers in the San Jose CA region on Testinsight software... ...need to collaborate with EDA and test engineers at the customer site and work closely with...Flexible hours- ...life around the globe. Location: San Jose Team: Technology Employment Type: Intern Job... ...Responsibilities About the teamThe Seed Infrastructures team oversees the distributed... ...productivity Collaborate with researchers and engineers to translate model requirements into...Temporary workInternship
$150k - $200k
...models. Responsibilities As a Research Scientist/ ML Engineer, you will play a crucial... ...with proprietary auto-align technology and powered by state-of-the... .../ Process Developer San Francisco Bay Area $130,000... ...Scientist Training Program San Jose, CA $135,000.00-$143,000.00...Full time- ByteDance invites a campus intern in San Jose, CA to join a lab focusing on LLM/AI infrastructure, distributed systems, and new system design. You will explore research questions, develop prototypes, and measure performance in a scale-driven environment. You will collaborate...Internship
$156k - $387.6k
Research Scientist Graduate (Compute Platform -... ...PhD) Location: San Jose Employment Type: Regular Job Code: A727... ...in big data infrastructure and data products... ...and big data engineering, which... ...cutting-edge technology evolvement in... ...optimization, storage system, hardware...Temporary workWork experience placementLocal area$156k - $387.6k
...ByteDance's system infrastructure is currently at a... ...projects, focusing on LLM/AI + Infrastructure technologies, which includes... ...Conduct research surveys and propose... ...Science, Computer Engineering, or a related technical... ...such as distributed storage and database systems...Temporary workLocal area$212.8k - $387.6k
Senior Research Scientist - Machine Learning System Location: San Jose Team: Technology Employment Type: Regular Job Code: A103221A Responsibilities... ...combines system engineering and the art of... ...systems for LLM/AIGC/AGI. In our... ...with GPU/NPU/RDMA/Storage and keep it...Temporary workLocal area$136.8k - $259.2k
Overview Research Engineer (LLM/ML/RL) - TikTok Ads Core ML, Ranking. Base pay range: $136,800.00... ...advertising delivery system using frontier technologies, including ML/DL, RL, LLM, and... ...notified about new Research Engineer jobs in San Jose, CA. #J-18808-Ljbffr TikTokFull timeWorldwide$155.5k - $248.9k
...specialists, powering next-generation technologies through sophisticated solutions. Behind... ...environment and enjoy partnering with engineering teams to bring innovative products to market... ...Engineer to join our team in San Jose, California. In this highly visible role...Flexible hours$254.4k
Tech Lead Research Scientist/Engineer, Neural Graphics and World Models - TikTok Location: San Jose Employment Type: Regular Job Code: A138238 Share this listing: Responsibilities About... ...and access to information technology systems; and Exercising sound judgement...Temporary work$170.5k - $272.8k
...specialists, powering next-generation technologies through sophisticated solutions. Behind... ...with Sales, Marketing, Applications, and Engineering to shape proposals, influence product... ...Location Customer engagement is centered in San Jose or Irvine, CA. Travel is 1-2 weeks/...Flexible hours$212.8k
Senior Research Scientist (Multimodal Large Language Model) - PICO Location: San Jose Team: Technology Employment Type: Regular Job Code: A57637 About the Team PICO... ...(including software engineering, product design, and... ...experience in LLM tool use, reinforcement...Temporary workLocal area$162k - $316.8k
Research Scientist Graduate (Seed Quantum Chemistry and Machine Learning) - 2027 Start (PhD) Location: San Jose Team: Technology Employment Type: Regular Job Code: A99246 Share this listing: Responsibilities... ...-scale computational infrastructure. We are looking for...Temporary workLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer / Scientist - Storage for LLM Technology - Infrastructure San Jose Regular. Be the first to apply!
- research programmer San Jose, CA
- deep learning research engineer San Jose, CA
- research engineer San Jose, CA
- molecular biology scientist San Jose, CA
- water quality scientist San Jose, CA
- machine learning scientist San Jose, CA
- image scientist San Jose, CA
- machine learning research scientist San Jose, CA
- materials scientist San Jose, CA
- health scientist San Jose, CA
