Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Multimodal AI Researcher: Generative Models, Realtime Vision

Socket.dev

Apple in Sunnyvale is seeking a Multimodal AI Researcher to push the boundaries of foundation models for real-time multimodal data, including video, audio, and text. You will work on interactive models, audio-to-audio modeling, and streaming multimodal systems, driving data requirements, validation strategies, and delivering research that informs product features. Ideal candidates have a BS with 3+ years of experience, hands-on work with LLMs and VLMs, and strong Python/PyTorch skills. #J-18808-Ljbffr Socket.dev

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Multimodal AI Researcher: Generative Models, Realtime Vision in Sunnyvale, CA vacancy
  • The Video Computer Vision organization is working on...  ...delivers cutting-edge AI, machine learning,...  ...perception, digital humans, multimodal generative AI, and agents. Our...  ...products. We are an applied research group, we push the...  ...developing foundation models for generative AI and... 
    Suggested

    Socket.dev

    Sunnyvale, CA
    3 days ago
  • $192k - $304.75k

     ...seeking an outstanding Research Scientist or Research Engineer...  ...for synthetic data generation and its application to...  ...autonomous driving models of tomorrow. As part of NVIDIA's Physical AI Platform Research team,...  ...driving, reasoning, and vision-language models. The ideal... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • A leading technology company in Santa Clara seeks a Machine Learning Researcher for its AIML Multimodal Foundation Model Team. You will develop advanced multimodal foundation models and agent capabilities for Apple's products. Ideal candidates possess a PhD or MS in a relevant... 
    Suggested

    Apple Inc.

    Santa Clara, CA
    4 days ago
  • $165k - $195k

    Company DescriptionThe Bosch Research and Technology Center...  ...global research, our AI research in Silicon Valley...  ...focuses on Foundation Models, Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics...  ...AI, synthetic data generation, agentic AIProficiency... 
    Suggested
    Full time
    Work experience placement
    Local area
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    2 days ago
  •  ...Postdoctoral Researcher At Toyota Research Institute...  ...the state of the art in AI, robotics, driving, and...  ...project exploring next-generation extended reality (XR) experiences...  ...especially in computer vision. Extensive hands-on...  ...image/video generative models. Strong programming... 
    Suggested
    Work at office
    Shift work

    Toyota Research Institute

    Los Altos, CA
    1 day ago
  • Ifm Us is seeking a Research Scientist to advance Vision Language Models. This role focuses on research and development in multimodal AI, integrating visual understanding with language reasoning. The successful candidate will contribute to technical reports, mentor junior... 

    Ifm Us

    Sunnyvale, CA
    5 days ago
  • Institute of Foundation Models, operating the AllWorld Team at MBZUAI, seeks researchers to develop the PAN world models that...  ...research to push state-of-the-art multimodal AI. Candidate requirements...  ...hands-on experience with video generative models, and proficiency in #J... 

    Ifm Us

    Sunnyvale, CA
    5 days ago
  • $192k - $304.75k

     ...are now looking for a Senior Research Scientist focused on Multimodal Foundation Models and Robotics! NVIDIA is...  ...scale robot learning, game AI, and physical simulation. Our...  ...following topics: LLMs; Large vision-language models; Video generative models and diffusion... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $165k - $185k

    Company DescriptionThe Bosch Research and Technology Center...  ...global research, our AI research in Silicon...  ...focuses on Foundation Models, Big Data Visual Analytics...  ...Language Processing, Computer Vision & Mixed Reality, Cloud...  ...building and applying multimodal transformer-based... 
    Work experience placement
    Local area
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    8 hours ago
  • $238.9k - $305.5k

     ...Group (ATG) is the research division of the...  ...engineering, such as AI/ML, algorithms,...  ...processing, computer vision, data science &...  ...and advance next-generation image/video capture...  ...processing, spectral/3D modeling, geometry, and...  ...vision, audio, or multimodal domains (e.g., source... 
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby

    Sunnyvale, CA
    3 days ago
  • $190k - $250k

     ...artificial intelligence (AI) powered technology...  ...large-scale generative world models that learn to predict...  ...We are looking for a research scientist to lead the...  ...approachesBuild methods for joint multimodal generation that...  ...Medical, Dental, and Vision plans through Kaiser... 
    Temporary work
    Work at office
    Visa sponsorship

    Kodiak Robotics

    Mountain View, CA
    2 days ago
  • $193.93k - $352.29k

     ...profound opportunity for AI to drive positive...  ..., partner-led business model, Nuro is working toward...  ...collaborate closely with researchers and engineers on the Learned...  ...teams to tackle plan generation challenges in...  ...driving. Experiences in vision-language-action models,... 
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    3 days ago
  •  ...Institute of Foundation Models We are a dedicated research lab for building, understanding...  ..., nurture the next generation of AI builders, and drive...  ...Research Scientist in the Vision Language Model (VLM) team...  ...advancing state-of-the-art multimodal foundation models that... 

    Institute of Foundation Models

    Sunnyvale, CA
    more than 2 months ago
  • $192.2k - $260k

     ...ownership of the product, related research and experimentation,...  ...learning techniques in computer vision (CV), Generative AI, multimedia understanding...  ...and develop generative models for controllable synthesis...  ...production pipelines. Design multimodal GenAI workflows including... 
    Local area
    Flexible hours

    Socket.dev

    Sunnyvale, CA
    4 days ago
  •  ...We are seeking a Contract AI Researcher - Efficient AI to join LG's...  ...that make modern LLMs, VLMs, multimodal models, and AI agents faster,...  ...reasoning optimization, and next-generation AI architectures. Your work...  ...on standardized language, vision, reasoning, and agentic... 
    Contract work
    For contractors

    PVH (Tommy Hilfiger/Calvin Klein)

    Santa Clara, CA
    3 days ago
  • Samsung Research America is seeking a highly skilled Robotics Researcher...  ...intelligence and foundational AI for real‑world robotic tasks....  ...on robotics foundation models, VLA, and VLMs, with collaboration...  ..., dexterous manipulation, and multimodal data fusion, while producing patents... 

    Samsung Electronics GmbH

    Mountain View, CA
    5 days ago
  • $163.8k - $307.6k

     ...in Palo Alto, California seeks candidates for a role focused on Omni multimodal large models. Responsibilities include conducting R&D, analyzing performance bottlenecks, and exploring next-generation architectures. Ideal candidates will have a Bachelor's degree in Computer... 

    Lightspeed Studios

    Palo Alto, CA
    4 days ago
  •  ...USA's Emerging Technology Lab in Santa Clara, CA seeks a Contract AI Researcher focused on Efficient AI. You will explore model compression, quantization, and on-device inference to make LLMs and multimodal models faster and lighter for real-world applications. The role... 
    Contract work

    LG Electronics USA

    Santa Clara, CA
    5 days ago
  •  ...About the Role As an AI Researcher for Computer Vision & Autonomous Robots at TCS,...  ...aspire to build the next generation of intelligent robotic systems...  ...perception and SLAM to multimodal sensor fusion and...  ...Transformers, Diffusion Models, Agentic AI frameworks).... 
    Full time

    Tata Consultancy Services

    Santa Clara, CA
    2 days ago
  • A leading technology firm in California is seeking a passionate Research Scientist to advance next-generation AI hardware platforms. The role involves developing multimodal intelligence models, benchmarking innovative LLM architectures, and collaborating across teams to... 

    Jobleads-US

    Palo Alto, CA
    4 days ago
  • $184k - $253k

     ...that literally connect our world - like AI and IoT. What We Offer Salary: $184,000...  ...pretrain, fine‑tune, and align LLMs and generative models tailored for scientific and materials science...  ...materials science, and publish original research in top venues. Mentor junior team... 

    Applied Materials

    Santa Clara, CA
    2 days ago
  • $147k - $211k

    Google DeepMind in Mountain View, CA is seeking a Research Scientist to advance multimodal AI research. You will set up large-scale tests, prototype ideas...  ...concepts across domains, from core ML to NLP and vision. You will publish findings, collaborate with world-class... 

    Google DeepMind

    Mountain View, CA
    4 days ago
  • $184k - $287.5k

     ...healthcare through accelerated computing and AI. We are seeking passionate researchers to advance longitudinal multimodal foundation models for healthcare. Recent progress in...  ...prediction, temporal reasoning, and multimodal generative modeling.Build large-scale datasets,... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $147k - $211k

    SnapshotWe are seeking strong Research Scientists with expertise in AI research and experience in...  ...interdisciplinary sociotechnical modeling to join a multimodal safety research effort within...  ...in deep learning, computer vision, and generative architectures. This role requires... 
    Full time

    DeepMind

    Mountain View, CA
    3 days ago
  • $192k - $304.75k

    We are now looking for a Senior Research Scientist for Generative AI!NVIDIA is searching for a world-class researcher...  ...great impacts with generative AI models. You will be building research...  ...practice of deep learning, computer vision, natural language processing, or computer... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $192.2k - $260k

     ...takes you!Key job responsibilities• Build generative AI models that create production-ready content,...  ..., model validation and serving.• Research new and innovative machine learning approaches...  ...health insurance (medical, dental, vision, prescription, Basic Life & AD&D... 
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    4 days ago
  • $200k - $287.5k

     ...-time /HybridAt Toyota Research Institute (TRI), we’re...  ...the state of the art in AI, robotics, driving, and...  ...Policy and Large Behavior Models (LBM).The OpportunityWe...  ...new capabilities in generative AI (e.g., recent results...  ...a focus on computer vision as the primary sensing... 
    Full time
    Local area
    Shift work

    Toyota Research Institute

    Los Altos, CA
    1 day ago
  • $140.4k - $264k

     ....What the Role Entails1.Conduct research and development of Omni multimodal large models, including the design and construction...  ...Omni-modal understanding and generation capabilities, research next-...  ...also eligible for medical, dental, vision, life and disability benefits,... 
    Full time
    Relocation package

    Tencent

    Palo Alto, CA
    3 days ago
  • RoboForce is an AI robotics company developing...  ...a Senior / Staff AI Research Scientist, Foundation Models to advance robotic...  ...Design and deploy vision-language(-action) models...  ...action-conditioned generative modeling for robot...  ...Decent understanding of multimodal models, modern ML... 
    Work at office
    Visa sponsorship

    RoboForce

    Milpitas, CA
    4 days ago
  • $185k - $400k

     ...pioneering the next generation of creative...  ...around real-time, multimodal generation and intelligent...  ...seeking accomplished Research Scientists in Foundation Models with expertise in...  ...real-time multimodal AI research.What We’re...  ...platforms. Our vision is to break down technical... 
    Remote work

    Pika

    Palo Alto, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Multimodal AI Researcher: Generative Models, Realtime Vision. Be the first to apply!