Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist - VLM Pretraining

Epsilon Labs, Inc.

About Us We're tackling one of healthcare's most critical challenges in medical imaging and diagnostics. Our company operates at the intersection of cutting-edge AI and clinical practice, building technology that directly impacts patient outcomes. We've assembled one of the industry's most comprehensive and diverse medical imaging datasets and have a proven product-market fit with a substantial customer pipeline already in place. Role Overview We're seeking a Research Scientist with deep expertise in large-scale vision-language pretraining to join our ML Research team . You'll be at the forefront of developing state-of-the-art multimodal models for clinical use in radiology settings. This role owns the pretraining stage of our radiology report generation model: VLM architecture design, multimodal data and task mixtures, and the large-scale training runs that build grounded visual understanding across X-rays, CT scans, and MRI. You'll work with one of the largest and most diverse medical imaging datasets in the industry, paired with the reports that make multimodal pretraining at this scale possible, while maintaining the clinical rigor required for healthcare deployment. Post-training and RL are owned by a partner role you'll collaborate with closely. Key Responsibilities Design, train, and scale vision-language foundation models for radiology applications, owning the pretraining stage end to end. Develop VLM architectures suited to medical imaging, including native and variable resolution handling, high-resolution tiling, connector design, and token budgets for volumetric studies. Build and tune multimodal pretraining mixtures across captioning, VQA, grounding, and retrieval tasks, balancing data sources to avoid regressions in language capability. Develop fine-grained visual grounding during pretraining, enabling models to localize findings within medical images using bounding boxes or segmentation masks. Own pretraining evaluation (zero- and few-shot transfer, probing, and downstream fine-tunability) — as the signal for base model quality. Train joint vision-language embedding spaces using contrastive and generative objectives, including region- and sentence-level alignment between images and reports. Contribute hands‑on to all stages of pretraining including dataset curation, architecture design, distributed training, and handoff of base checkpoints to post‑training. Stay current with cutting‑edge research in vision‑language modeling and large‑scale multimodal pretraining. Drive research and technical excellence through conference publications and technical blog posts, establishing best practices for pretraining medical VLMs at scale. Qualifications 6+ years of academia/industry experience in vision-language modeling, multimodal learning, or related fields Deep expertise in pretraining large vision-language models (e.g., LLaVA, Flamingo, CogVLM, Qwen-VL, InternVL, or similar architectures) Strong foundation in modern VLM pretraining techniques including: Vision-language connector and fusion architectures (projection, cross-attention, resampler-based) Variable and high-resolution image handling (native resolution, dynamic tiling, token compression) Contrastive and generative objectives for learning joint vision-language embedding spaces Data and task mixture design, including curriculum and mixture-ratio ablations Experience with fine-grained visual grounding (referring expression comprehension, phrase grounding, box or mask prediction) Track record of implementing complex models from research papers and adapting them to new domains Proficiency in PyTorch or JAX, with experience training large models on multi-GPU/distributed systems Experience with autoregressive language modeling and long-context training Hands‑on experience with medical imaging applications, particularly radiology report generation Strong software engineering skills and ability to write production-quality code Preferred Qualifications Publications at top-tier conferences (NeurIPS, ICML, ICLR, CVPR, ACL, EMNLP, MICCAI) Experience training vision encoders from scratch, or co‑designing them with a downstream VLM Experience with interleaved image-text pretraining and synthetic recaptioning pipelines Experience with 3D medical image processing and temporal modeling Familiarity with clinical NLP and medical knowledge representation Knowledge of evaluation methodologies for long‑form generation, including factuality assessment and hallucination detection Experience with model interpretability, explainability, and uncertainty quantification in safety‑critical applications #J-18808-Ljbffr Epsilon Labs, Inc.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Research Scientist - VLM Pretraining in San Francisco, CA vacancy
  •  ...environments. You will “live and breathe” all forms of robot data. You’ll be responsible for: Designing and executing large‑scale pretraining runs for robot foundation models (transformer‑and diffusion‑based architectures) Defining model architectures, objectives, and... 
    Suggested

    Generalist

    San Francisco, CA
    4 days ago
  • $250k

    Research Scientist - Vision Data Infrastructure This role is offered by Storm3. Your actual pay will...  ...and video datasets, enabling efficient pretraining, alignment, and simulation-based data...  ...data. Create specialized datasets for VLM training, visual reasoning, and agent... 
    Suggested
    Full time

    Storm3

    San Francisco, CA
    2 days ago
  • Epsilon Labs, Inc. is seeking a Research Scientist to lead pretraining and scaling of vision-language models for radiology, across X-ray, CT, and MRI data. You will own the pretraining stage, design architectures, manage data mixtures, and drive model quality through evaluation... 
    Suggested

    Epsilon Labs, Inc.

    San Francisco, CA
    3 days ago
  •  ...base intelligence layer for robotics models in San Francisco. The role involves training large-scale transformer models, designing pretraining runs, and collaborating on data collection operations. Ideal candidates will have extensive experience in distributed training... 
    Suggested

    Generalist

    San Francisco, CA
    4 days ago
  • $155k - $269k

     ...realistic, scalable, controllable, and efficient simulation. As a Research Scientist in World Models, you will develop algorithms and...  ...generation, including video models, multimodal generative models, LLM/VLM/VLA models, and predictive models of traffic participants and... 
    Suggested
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    3 days ago
  • $117.2k - $313.7k

     ...the future of Salesforce.The ExperienceSalesforce AI Research is looking for outstanding AI Research Scientists and Research Engineers. Our team discovers new...  ...Multimodal and Computer Vision: Vision-language models (VLM), video understanding, visual grounding for agents,... 
    Full time

    Salesforce

    San Francisco, CA
    1 day ago
  • $140k - $200k

     ...powers breakthrough AI models at leading research labs and enterprises. Since 2018, we’ve...  .... This is not a traditional research scientist role. You will not spend months...  ...understanding of LLM training pipelines — pretraining, supervised fine-tuning, RLHF/DPO, and... 
    Work at office
    Flexible hours
    2 days per week

    Labelbox

    San Francisco, CA
    1 day ago
  • $350k

     ...as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together...  ...Architectures project that involves collaborating with Pretraining. Responsibilities Develop methods for understanding LLMs by... 
    Temporary work
    Work at office
    Remote work
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    1 day ago
  • $250k - $325k

    Research Scientist (Generative Modeling) About World Labs: We build foundational world models that can perceive, generate, reason, and interact...  ...the following areas is a strong plus : Data curation for pretraining or post-training Tokenizers and VAEs for image, video, or 3... 

    World Labs Inc.

    San Francisco, CA
    1 day ago
  • Epsilon Labs, Inc. is seeking a Research Scientist with deep expertise in post-training and reinforcement learning to advance multimodal...  ...models for clinical radiology use. You will own stages after pretraining, including supervised fine-tuning, reward modeling, and... 

    Epsilon Labs, Inc.

    San Francisco, CA
    2 days ago
  • About the Role Pretraining gives us a general model. Post-training makes it useful, controllable, safe, and performant in the real world...  ...ML or controls, and all the places in between. This is where research meets reality. You’ll be responsible for: Designing fine-tuning... 

    Generalist

    San Francisco, CA
    4 days ago
  • $350k

     ...Our team is a quickly growing group of committed researchers, engineers, policy experts, and business...  ...systems. About the role We're looking for a Research Scientist who has done hands-on research on large models (pretraining, fine-tuning, RL, evals, or agents scaffolds)... 
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    3 days ago
  •  ...pipeline already in place. Role Overview We're seeking a Research Scientist with deep expertise in post‑training and reinforcement learning...  ...in radiology settings. This role owns every stage after pretraining: supervised fine‑tuning, reward modeling, reinforcement learning... 

    Epsilon Labs, Inc.

    San Francisco, CA
    4 days ago
  • $114.2k - $306.6k

     ...are not duplicating efforts.*Salesforce Research advances state-of-the-art AI techniques,...  ...Research is looking for outstanding AI Research Scientists / Research Engineers.**Our team discovers...  ...Vision:** Vision-language models (VLM), video understanding, and visual grounding... 
    Full time

    Niebles

    San Francisco, CA
    2 days ago
  • $120.7k - $238.6k

    The OpportunityAdobe Research is looking for research scientists in Generative AI to join a world-class research team. We welcome outstanding candidates at all levels (new graduates, experienced, principal) in all related technical fields, such as Machine Learning, Deep... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    2 days ago
  • $127k - $333.7k

    THE DEPARTMENT OF MEDICINE AT THE UNIVERSITY OF CALIFORNIA SAN FRANCISCO (UCSF) is recruiting for the position of Research Scientist in the Division of Geriatrics. The appointment will be made at the level of Assistant, Associate, or Full Adjunct Professor. Candidates... 

    University of California, San Francisco

    San Francisco, CA
    1 day ago
  • We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain...  ....Role overview:We are seeking an exceptional Research Scientist to join our team, focusing on developing cutting-edge diffusion... 
    Relocation

    Genmo

    San Francisco, CA
    1 day ago
  • University of California, San Francisco is seeking a Professional Researcher to join the Department of Pediatrics. Appointment at Assistant...  ...conducting research in the lab of an established Principal Scientist, collaborating with staff and postdocs, designing experiments,... 

    University of California, San Francisco

    San Francisco, CA
    5 hours ago
  • $150k - $250k

     ...manufacturing, consumer goods, and global social organizations.We research and deploy technologies that power AI-native operations — both...  ...through data curation, reward modeling, or continual pretraining.Experience Building with Models, Not Just Building Models: We... 
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    1 day ago
  •  ...Research Scientist Engineering · Full-time · San Francisco; New York Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our... 
    Full time

    Anysphere

    San Francisco, CA
    15 hours ago
  • $158k - $269k

     ...Research Scientist Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class team, we're unlocking the next era of autonomous transportation with technology that's powering commercial autonomous trucks and robotaxis. Waabi is backed... 
    Full time
    Work experience placement
    Internship
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    2 days ago
  •  ...Phonic Phonic is a product and research lab focused on powering the most realistic, human-like voice AI conversations. We've re-thought...  ...over $30M from tier 1 VCs. About The Role As a Research Scientist at Phonic, you'll drive original research that pushes the... 
    Work at office

    Phonic

    San Francisco, CA
    3 days ago
  •  ...Overview Join our core R&D team building end-to-end automated research systems . Zochi publishes the first fully-AI generated A* conference paper Locus becomes the first AI-system to outperform human experts at AI R&D Key Responsibilities Design... 

    Intology

    San Francisco, CA
    20 hours ago
  •  ...Machine Learning Scientist Sesame believes in a future where computers are lifelike - with the ability to see, hear, and collaborate...  ...goals. As a Machine Learning Scientist at Sesame, you are a research-oriented person with experience in NLP, Speech, and/or Computer... 
    Full time
    Contract work
    Flexible hours

    SESAME

    San Francisco, CA
    1 day ago
  •  ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team to push the boundaries of action-conditioned world models and generative AI for physical systems. As a researcher, you will drive original research into the intersection of generative... 
    Work at office

    HEDRA INC

    San Francisco, CA
    15 hours ago
  •  ..., which is part of why we think it is worth doing. But it is also a very practical engineering and research problem. The role We are hiring a Research Scientist to help build the computational and physics-based modeling core of the company. This is a research... 
    Immediate start

    Tabula Inc

    San Francisco, CA
    1 day ago
  •  ...Research Scientist ThirdLayer is solving one of the hardest problems in deploying agents: AI models are generic, but people's work and processes are specific. Our product, Dex, embeds within your computer and team, ingesting a continuous stream of data across your... 

    Thirdlayer (yc W25)

    San Francisco, CA
    4 days ago
  •  ...By applying to this role, you will be considered for Research Scientist roles across all teams at OpenAI. About the Role As a Research Scientist here, you will develop innovative machine learning techniques and advance the research agenda of the team you work... 

    Openai

    San Francisco, CA
    4 days ago
  •  ...Research Scientist Focused On Rodent Behavior And Neural Imaging At Nudge, our mission is to develop the best technology for interfacing with the brain to improve people's lives. We're starting with an approach that we believe can help the most people the fastest,... 

    Nudge Inc.

    San Francisco, CA
    2 days ago
  •  ...Researcher Opportunities We are looking for researchers with a record of excellent research results in the fields of machine learning and robotics, at all levels. Successful candidates will have both excellent fundamentals and excellent implementation skills, with... 

    Physical Intelligence

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist - VLM Pretraining. Be the first to apply!