Research Scientist - VLM Pretraining
Epsilon Labs, Inc.
About Us We're tackling one of healthcare's most critical challenges in medical imaging and diagnostics. Our company operates at the intersection of cutting-edge AI and clinical practice, building technology that directly impacts patient outcomes. We've assembled one of the industry's most comprehensive and diverse medical imaging datasets and have a proven product-market fit with a substantial customer pipeline already in place. Role Overview We're seeking a Research Scientist with deep expertise in large-scale vision-language pretraining to join our ML Research team . You'll be at the forefront of developing state-of-the-art multimodal models for clinical use in radiology settings. This role owns the pretraining stage of our radiology report generation model: VLM architecture design, multimodal data and task mixtures, and the large-scale training runs that build grounded visual understanding across X-rays, CT scans, and MRI. You'll work with one of the largest and most diverse medical imaging datasets in the industry, paired with the reports that make multimodal pretraining at this scale possible, while maintaining the clinical rigor required for healthcare deployment. Post-training and RL are owned by a partner role you'll collaborate with closely. Key Responsibilities Design, train, and scale vision-language foundation models for radiology applications, owning the pretraining stage end to end. Develop VLM architectures suited to medical imaging, including native and variable resolution handling, high-resolution tiling, connector design, and token budgets for volumetric studies. Build and tune multimodal pretraining mixtures across captioning, VQA, grounding, and retrieval tasks, balancing data sources to avoid regressions in language capability. Develop fine-grained visual grounding during pretraining, enabling models to localize findings within medical images using bounding boxes or segmentation masks. Own pretraining evaluation (zero- and few-shot transfer, probing, and downstream fine-tunability) — as the signal for base model quality. Train joint vision-language embedding spaces using contrastive and generative objectives, including region- and sentence-level alignment between images and reports. Contribute hands‑on to all stages of pretraining including dataset curation, architecture design, distributed training, and handoff of base checkpoints to post‑training. Stay current with cutting‑edge research in vision‑language modeling and large‑scale multimodal pretraining. Drive research and technical excellence through conference publications and technical blog posts, establishing best practices for pretraining medical VLMs at scale. Qualifications 6+ years of academia/industry experience in vision-language modeling, multimodal learning, or related fields Deep expertise in pretraining large vision-language models (e.g., LLaVA, Flamingo, CogVLM, Qwen-VL, InternVL, or similar architectures) Strong foundation in modern VLM pretraining techniques including: Vision-language connector and fusion architectures (projection, cross-attention, resampler-based) Variable and high-resolution image handling (native resolution, dynamic tiling, token compression) Contrastive and generative objectives for learning joint vision-language embedding spaces Data and task mixture design, including curriculum and mixture-ratio ablations Experience with fine-grained visual grounding (referring expression comprehension, phrase grounding, box or mask prediction) Track record of implementing complex models from research papers and adapting them to new domains Proficiency in PyTorch or JAX, with experience training large models on multi-GPU/distributed systems Experience with autoregressive language modeling and long-context training Hands‑on experience with medical imaging applications, particularly radiology report generation Strong software engineering skills and ability to write production-quality code Preferred Qualifications Publications at top-tier conferences (NeurIPS, ICML, ICLR, CVPR, ACL, EMNLP, MICCAI) Experience training vision encoders from scratch, or co‑designing them with a downstream VLM Experience with interleaved image-text pretraining and synthetic recaptioning pipelines Experience with 3D medical image processing and temporal modeling Familiarity with clinical NLP and medical knowledge representation Knowledge of evaluation methodologies for long‑form generation, including factuality assessment and hallucination detection Experience with model interpretability, explainability, and uncertainty quantification in safety‑critical applications #J-18808-Ljbffr Epsilon Labs, Inc.
- Epsilon Health is seeking a Research Scientist to lead pretraining of large vision-language models for radiology report generation. You will work on scale-ready VLM architectures, multimodal data mixtures, and grounding in medical images like X-rays, CTs, and MRIs. The...Suggested
- ...environments. You will “live and breathe” all forms of robot data. You’ll be responsible for: Designing and executing large‑scale pretraining runs for robot foundation models (transformer‑and diffusion‑based architectures) Defining model architectures, objectives, and...Suggested
$250k
Research Scientist - Vision Data Infrastructure This role is offered by Storm3. Your actual pay will... ...and video datasets, enabling efficient pretraining, alignment, and simulation-based data... ...data. Create specialized datasets for VLM training, visual reasoning, and agent...SuggestedFull time- Epsilon Labs, Inc. is seeking a Research Scientist to lead pretraining and scaling of vision-language models for radiology, across X-ray, CT, and MRI data. You will own the pretraining stage, design architectures, manage data mixtures, and drive model quality through evaluation...Suggested
- ...base intelligence layer for robotics models in San Francisco. The role involves training large-scale transformer models, designing pretraining runs, and collaborating on data collection operations. Ideal candidates will have extensive experience in distributed training...Suggested
$155k - $269k
...realistic, scalable, controllable, and efficient simulation. As a Research Scientist in World Models, you will develop algorithms and... ...generation, including video models, multimodal generative models, LLM/VLM/VLA models, and predictive models of traffic participants and...Full timeWork at officeWork from homeFlexible hours$140k - $200k
...powers breakthrough AI models at leading research labs and enterprises. Since 2018, we’ve... .... This is not a traditional research scientist role. You will not spend months... ...understanding of LLM training pipelines — pretraining, supervised fine-tuning, RLHF/DPO, and...Work at officeFlexible hours2 days per week$350k
...as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together... ...Architectures project that involves collaborating with Pretraining. Responsibilities Develop methods for understanding LLMs by...Temporary workWork at officeRemote workVisa sponsorshipFlexible hours$250k - $325k
Research Scientist (Generative Modeling) About World Labs: We build foundational world models that can perceive, generate, reason, and interact... ...the following areas is a strong plus : Data curation for pretraining or post-training Tokenizers and VAEs for image, video, or 3...- ...customer pipeline already in place. Role Overview We're seeking a Research Scientist with deep expertise in vision foundation models to join our... ...for medical imaging applications. This role focuses on pretraining and scaling vision encoders for radiology diagnosis across...
- Anthropic in San Francisco seeks a Research Scientist to measure recursive-self-improvement in large models. You will design evaluations,... ...hands-on work with strategic planning, collaborating across pretraining, RL, and policy teams to advance safe and reliable AI...
$350k
...Our team is a quickly growing group of committed researchers, engineers, policy experts, and business... ...systems. About the role We're looking for a Research Scientist who has done hands-on research on large models (pretraining, fine-tuning, RL, evals, or agents scaffolds)...Work at officeVisa sponsorshipFlexible hours- About the Role Pretraining gives us a general model. Post-training makes it useful, controllable, safe, and performant in the real world... ...ML or controls, and all the places in between. This is where research meets reality. You’ll be responsible for: Designing fine-tuning...
- Epsilon Labs, Inc. is seeking a Research Scientist with deep expertise in post-training and reinforcement learning to advance multimodal... ...models for clinical radiology use. You will own stages after pretraining, including supervised fine-tuning, reward modeling, and...
- ...pipeline already in place. Role Overview We're seeking a Research Scientist with deep expertise in post‑training and reinforcement learning... ...in radiology settings. This role owns every stage after pretraining: supervised fine‑tuning, reward modeling, reinforcement learning...
$114.2k - $306.6k
...are not duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques,... ...Research is looking for outstanding AI Research Scientists / Research Engineers.**Our team discovers... ...Vision:** Vision-language models (VLM), video understanding, and visual grounding...Full time$117.2k - $313.7k
Salesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers to discover new research problems, develop novel models, and bridge... ...and Computer Vision: Vision‑language models (VLM), video understanding, visual grounding for agents,...- University of California, San Francisco is seeking a Professional Researcher to join the Department of Pediatrics. Appointment at Assistant... ...conducting research in the lab of an established Principal Scientist, collaborating with staff and postdocs, designing experiments,...
$120.7k - $238.6k
The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research team. We welcome outstanding candidates at all levels (new graduates, experienced, principal) in all related technical fields, such as Machine Learning, Deep...Full timeTemporary workLocal areaWorldwide$127k - $333.7k
THE DEPARTMENT OF MEDICINE AT THE UNIVERSITY OF CALIFORNIA SAN FRANCISCO (UCSF) is recruiting for the position of Research Scientist in the Division of Geriatrics. The appointment will be made at the level of Assistant, Associate, or Full Adjunct Professor. Candidates...- We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain... ....Role overview:We are seeking an exceptional Research Scientist to join our team, focusing on developing cutting-edge diffusion...Relocation
$150k - $250k
...manufacturing, consumer goods, and global social organizations.We research and deploy technologies that power AI-native operations — both... ...through data curation, reward modeling, or continual pretraining.Experience Building with Models, Not Just Building Models: We...Work at office3 days per week- ...is to make those benefits real. We work across the full model stack—pretraining, midtraining, reinforcement learning, post‑training, evaluations, harnessing, and deployment—and connect that research to the patients, clinicians, and real‑world outcomes we aim to improve...Work at officeRelocation package
- ...Machine Learning Scientist Sesame believes in a future where computers are lifelike - with the ability to see, hear, and collaborate... ...goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or Computer...Full timeContract workFlexible hours
- Overview Join our core R&D team building end-to-end automated research systems. Zochi publishes the first fully-AI generated A* conference paper Locus becomes the first AI-system to outperform human experts at AI R&D Key Responsibilities Design & implement novel architectures...
- ...Researcher Opportunities We are looking for researchers with a record of excellent research results in the fields of machine learning and robotics, at all levels. Successful candidates will have both excellent fundamentals and excellent implementation skills, with...
$225k - $300k
...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access, and those with physicians in their immediate family. For everyone else, care is fragmented and impersonal. Medical history...Work at officeImmediate start- ...Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our...Full time
$158k - $269k
...Research Scientist Opportunity At Waabi Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class team, we're unlocking the next era of autonomous transportation with technology that's powering commercial autonomous trucks and...Full timeWork experience placementInternshipWork at officeWork from homeFlexible hours$127k - $333.7k
...description THE DEPARTMENT OF MEDICINE AT THE UNIVERSITY OF CALIFORNIA SAN FRANCISCO (UCSF) is recruiting for the position of Research Scientist in the Division of Geriatrics. The appointment will be made at the level of Assistant, Associate, or Full Adjunct Professor....Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist - VLM Pretraining. Be the first to apply!
- materials scientist San Francisco, CA
- scientist assay development San Francisco, CA
- entry level research scientist San Francisco, CA
- health scientist San Francisco, CA
- quality control scientist San Francisco, CA
- deep learning scientist San Francisco, CA
- research associate scientist San Francisco, CA
- application scientist San Francisco, CA
- scientist antibody discovery San Francisco, CA
- senior analytical scientist San Francisco, CA

