Research Scientist: Multimodal World Models & Pre-Training
$187.04kByteDance
A leading tech company is seeking a Research Scientist for Multimodal Interaction and World Model. This role focuses on enhancing models for multimodal data understanding, requiring expertise in AI and hands-on experience with frameworks like PyTorch/JAX. The position offers a competitive salary between $187,040 and $438,000 annually, plus benefits that include medical, dental, and vision insurance, a 401(k) plan, and paid personal time. #J-18808-Ljbffr ByteDance
$187.04k
Research Scientist - Multimodal Interaction and World Model - Pre-Training Location: San Jose Responsibilities About Seed Team: Established in 2023, the ByteDance Seed team is dedicated to discovering new approaches to general intelligence, and pursuing the edge of intelligence...TrainingTemporary workLocal area$212.8k - $387.6k
Team Overview The Seed Multimodal Interaction and World Model team is dedicated to developing models that have human-level multimodal understanding and... ...for reasoning, planning, and interaction. Build training pipelines including data curation, alignment, and reinforcement...TrainingTemporary workInternshipLocal area$244.8k
...artificial general intelligence, with research spanning MLLM, GenMedia, AI for Science... ...industry-leading general foundation models and multimodal capabilities, powering over 50 application... ...in PyTorch/JAX and distributed training frameworks Familiar with state-of-the-...Training$192k - $304.75k
We are now looking for a Senior Research Scientist focused on Multimodal Foundation Models and Robotics! NVIDIA is searching... ...the virtual and the physical world.You will work with an amazing and... ...embodied agents;Develop large-scale AI training and inference methods for...TrainingFull time$184k - $287.5k
...AI. We are seeking passionate researchers to advance longitudinal multimodal foundation models for healthcare. Recent progress... ...foundation model architectures and training strategies for disease... ...to translate research into real-world healthcare applications.Background...TrainingFull time$60 per hour
...excites you every day. Research Scientist Intern (TikTok-Neural Graphics and World Models) - 2027 Start (PhD) Location... .... Build components of training, inference, data, or... ...autoregressive models, multimodal transformers, long-context modeling, pre-training, or post-...TrainingHourly paySummer workInternshipLocal area- ...Institute of Foundation Models We are a dedicated research lab for building,... ...foundation model training, alongside world-class researchers, data scientists, and engineers,... ...infrastructure for training multimodal LLMs and video... ...efficient video pre-training strategies...TrainingVisa sponsorship
$162k - $316.8k
World Model Research Scientist (Intelligent Creation) - Global Frontier Tech Recruitment Program - 2027... ...and development in generative AI and multimodal models (e.g., image, video).... ...foundation models (LLMs/VLMs), through post-training techniques. Design and build models...TrainingTemporary workLocal area$212.8k
Senior Research Scientist (Multimodal Large Language Model) - PICO Location: San Jose Team: Technology Employment Type... ...system enhancement, and end-to-end training/inference acceleration. Drive the... ...in multimodal large model pre-training, post-training, fine-tuning...TrainingTemporary workLocal area$192k - $304.75k
We are seeking an outstanding Research Scientist or Research Engineer with a passion for... ...generation and its application to training autonomous driving models of tomorrow. As part of NVIDIA's Physical... ...technologies, including generative world models, end-to-end driving,...TrainingFull time- ...is to entertain the world. Together, we are writing... ...Representation Models team creates a single... ...RoleWe are looking for a Research Scientist specializing in... ...Semantic IDs, continuous pre-training, novel representation... ...in computer vision or multimodal AIIndustry experience...TrainingHourly payFull timeImmediate startFlexible hours
$254.4k
Tech Lead Research Scientist/Engineer, Neural Graphics and World Models - TikTok Location: San Jose Employment... ...rendering. Develop, train, adapt, and evaluate AI... ...models, or multimodal transformers. Engineering... ...distributed training, and pre-training and post-training...TrainingTemporary work$190k - $250k
...'t just perceive the world, it learns how the physics... ...generative world models that learn to predict... ...scalable closed-loop training, validation, and long... ...We are looking for a research scientist to lead the design... ...approachesBuild methods for joint multimodal generation that...TrainingTemporary workWork at officeVisa sponsorship$244.8k
...our Company. Responsibilities Deeply involve in post‑training (SFT/RL) of Seed multimodal models and LLM. Participate in unified modeling for image... ...capabilities of agentic foundation models, and conduct in‑depth research on agentic RL. Develop large‑scale, diverse, and...TrainingTemporary workLocal area$145k - $250k
Research Scientist Graduate- CV/NLP/Multimodal LLM,(Trust and Safety) - 2026 Start(PhD... ...and multimodality models and algorithms to protect... ...to everyone in the world. Responsibilities... ...distributed model training framework... ...business scenarios, like pre-training, zero-shot...TrainingFull timeTemporary workFixed term contractSummer workInternshipLocal areaFlexible hours- Institute of Foundation Models, operating the AllWorld Team at MBZUAI, seeks researchers to develop the PAN world models that simulate physical environments. You will train on large clusters, design benchmarks... ...to push state-of-the-art multimodal AI. Candidate requirements...
- ByteDance in San Jose seeks a PhD candidate to join the Seed Multimodal Interaction team, focused on developing innovative multimodal models that integrate vision, language, and more. The successful candidate will work on improving decision-making and reasoning capabilities...
$244.8k
...focuses on foundational models for visual generation, developing multimodal generative models, and carrying out leading research and application development... ...Design data pipelines, pre-training strategies, and post-... ...frameworks. Explore real‑world applications of vision...TrainingTemporary workLocal area$60 per hour
...building machine learning models and systems to protect... ...to everyone in the world. Project Overview,... ...complexity in multilingual and multimodal content, and upgraded... ...MoE architecture training and routing optimization... ...direction and produce research outcomes with...TrainingHourly paySummer workInternshipLocal areaFlexible hoursShift work- Responsibilities Conduct research on multimodal foundation models and related systems. Explore methods to improve model capabilities across modalities, including areas such as vision-language modeling, world modeling, and representation learning. Design and prototype algorithms...
$204k - $259k
...the mission to be the world's most trusted driver.... ...collaborations with other research teams in Alphabet. AI... ..., generative modeling, Bayesian inference, hierarchical... ...report to a Principal Scientist. You will:... ...Foundation World Model post-training and evaluation...TrainingFull timeTemporary workRemote work$165k - $185k
Company DescriptionThe Bosch Research and Technology Center North America... ...Valley focuses on Foundation Models, Big Data Visual Analytics,... ...Bosch teams around the world. Their creativity is the key... ...foundation models, including training, fine-tuning, and promptingIn...TrainingWork experience placementWorldwide- ...Number: P25F11 Honda Research Institute USA (HRI-... ...seeking a Research Scientist to push the frontiers... ...for embodied, real-world intelligence. This is... ...will develop multimodal and foundation-model-based approaches that... ...embodied intelligence. Train, fine-tune, and evaluate...TrainingWork experience placementShift work
- Research Scientist: Robotics Foundation Models - Honda Research Institute USA Honda Research Institute... ...AI systems that combine multimodal foundation models (e.g.,... ...behaviors in real‑world environments. San Jose,... ...task failure reasoning. Train, fine‑tune, and design vision...TrainingWork experience placement
- ...The Content Representation Models team develops unified foundation... ...the Role We are seeking a Research Scientist specializing in embeddings... ...in computer vision or multimodal AI. Industry experience in... ...retrieval. Experience with LLM pre‑training, fine‑tuning, or...TrainingFull timeFlexible hours
$244.8k
...GenAI team focuses on applied research in Generative AI,... ...with ease. The team works on multimodal foundation models, image and video generation... ...efficient model design, and world models. About The Role With... ...models through large‑scale training and post‑training techniques...TrainingTemporary workLocal area- ...Vision team focused on visual generation models. Responsibilities include developing and... ...optimizing architectures, and exploring real-world applications. Ideal candidates have... ...Python, experience in computer vision or multimodal learning, and a strong understanding of deep...
- ...Devices (AMD) is seeking an Applied Research Scientist to advance large language and multimodal models, including image/video generation. You will train, fine‑tune, and align LLMs, LMMs, and... ...strategy. The role requires handling pre-training, fine-tuning, and alignment...Training
- ...Institute of Foundation Models We are a dedicated research lab for building,... ...foundation model training, alongside world-class researchers, data scientists, and engineers,... ...advancing state-of-the-art multimodal foundation models... ..., data recipes for pre-training and post-...Training
$156k - $316.8k
Research Scientist — Privacy-Preserving Large-Scale Model Training & Architecture Optimization Location: San Jose Employment Type... ...efficiency under real-world production constraints. Optimize... ...shared context, long-sequence multimodal reasoning, and scalable training...TrainingTemporary workLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist: Multimodal World Models & Pre-Training. Be the first to apply!
- molecular biology scientist San Jose, CA
- water quality scientist San Jose, CA
- machine learning scientist San Jose, CA
- image scientist San Jose, CA
- machine learning research scientist San Jose, CA
- materials scientist San Jose, CA
- health scientist San Jose, CA
- scientist San Jose, CA
- quality control scientist San Jose, CA
- scientist biology San Jose, CA


