Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Multimodal AI Researcher

Socket.dev

The Video Computer Vision organization is working on breakthrough technologies for future Apple products. Our team delivers cutting-edge AI, machine learning, computer vision and graphics algorithms that power technologies including human understanding, perception, digital humans, multimodal generative AI, and agents. Our algorithms ship across a range of Apple products, including iPhone and Apple Vision Pro, where our work has contributed to technologies like Personalized Spatial Audio, EyeSight, and Persona as well as future Apple products. We are an applied research group, we push the state of the art and then bring it to product. In this role, you will collaborate with world-class experts in AI, ML, Software, and Hardware to tackle fundamental challenges in human-centric solutions that will impact millions of users across Apple's ecosystem. Description We are looking for a Multimodal AI Researcher with a strong background in developing foundation models for generative AI and multimodal systems that integrate various types of real-time sensor data such as video and audio with other modalities like text. Our ongoing investigations include interactive models, audio-to-audio modeling and systems. You will work on hard, open research problems in multimodal generative AI and agents, and you will see that work through to real features used by millions of people. You will collaborate with others to drive data requirements, validation strategies, and key performance indicators, and conduct algorithm research and development that serves product needs. We hire researchers who are highly motivated and deeply care about shipping. A successful candidate will stay up-to-date with the latest advancements in multimodal foundations models and applying this knowledge to drive innovation, but also take a practical approach to problem solving and software engineering. Minimum Qualifications BS and a minimum of 3 years relevant industry experience. Experience building models for multimodal perception systems. Experience working with LLMs and VLMs. Software engineering skills and proficiency in Python and PyTorch. Curiosity and willingness to learn new things in order to improve the quality of their solutions. Preferred Qualifications MS or PhD in computer vision, computer graphics, machine learning, computer science, computer engineering or related fields. Experience in developing, training/tuning foundation models and multimodal LLMs. Experience with training and troubleshooting generative architectures such as diffusion, reinforcement learning, flow matching or normalizing flow at scale. Experience with real-time or streaming multimodal models. Experience with speech understanding and generation. Experience applying reinforcement learning to help post-train foundation models. Excellent communication and experience working with multi-functional teams. Self-motivated with proven track record to optimally prioritize and deliver tasks on schedule. #J-18808-Ljbffr Socket.dev

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Multimodal AI Researcher in Sunnyvale, CA vacancy
  • Samsung Research America is seeking a highly skilled Robotics Researcher to advance embodied intelligence and foundational AI for real‑world robotic tasks. The role emphasizes leading research...  ..., dexterous manipulation, and multimodal data fusion, while producing patents... 
    Suggested

    Samsung Electronics GmbH

    Mountain View, CA
    4 days ago
  • Apple in Sunnyvale is seeking a Multimodal AI Researcher to push the boundaries of foundation models for real-time multimodal data, including video, audio, and text. You will work on interactive models, audio-to-audio modeling, and streaming multimodal systems, driving... 
    Suggested

    Socket.dev

    Sunnyvale, CA
    2 days ago
  • A leading technology company in Santa Clara seeks a Machine Learning Researcher for its AIML Multimodal Foundation Model Team. You will develop advanced multimodal foundation models and agent capabilities for Apple's products. Ideal candidates possess a PhD or MS in a... 
    Suggested

    Apple Inc.

    Santa Clara, CA
    3 days ago
  • $163.8k - $307.6k

    Lightspeed Studios in Palo Alto, California seeks candidates for a role focused on Omni multimodal large models. Responsibilities include conducting R&D, analyzing performance bottlenecks, and exploring next-generation architectures. Ideal candidates will have a Bachelor... 
    Suggested

    Lightspeed Studios

    Palo Alto, CA
    3 days ago
  •  ...About the Role As an AI Researcher for Computer Vision & Autonomous Robots at TCS, you’ll work on the frontier of applied artificial intelligence...  ...and intelligent machines. From visual perception and SLAM to multimodal sensor fusion and reinforcement learning, you’ll be pushing... 
    Suggested
    Full time

    Tata Consultancy Services

    Santa Clara, CA
    1 day ago
  • NVIDIA in Santa Clara, CA is seeking a Senior Deep Learning Scientist to advance streaming and agentic multimodal AI. You will develop models capable of reasoning, planning, and acting across modalities using PyTorch and scalable Python pipelines. Responsibilities include... 

    NVIDIA AI

    Santa Clara, CA
    3 days ago
  •  ...synthetic data generation for training large language models. This role involves building data generation pipelines and advancing multimodal data generation, collaborating across various teams. The ideal candidate holds a PhD and has a strong background in synthetic... 

    NVIDIA

    Santa Clara, CA
    17 hours ago
  • Ifm Us is seeking a Research Scientist to advance Vision Language Models. This role focuses on research and development in multimodal AI, integrating visual understanding with language reasoning. The successful candidate will contribute to technical reports, mentor junior... 

    Ifm Us

    Sunnyvale, CA
    4 days ago
  • $192k - $356.5k

    NVIDIA is seeking a Senior Research Scientist focused on Multimodal Foundation Models and Robotics in Santa Clara, CA. The ideal candidate will design AI algorithms for humanoid robots, develop training methods, and collaborate with multidisciplinary teams. Candidates... 

    NVIDIA

    Santa Clara, CA
    2 days ago
  •  ...Foundation Models, operating the AllWorld Team at MBZUAI, seeks researchers to develop the PAN world models that simulate physical...  ...collaboration across engineering and research to push state-of-the-art multimodal AI. Candidate requirements include MSc/PhD in ML or CS, hands-on... 

    Ifm Us

    Sunnyvale, CA
    4 days ago
  • LG Electronics is seeking a Contract AI Researcher for the Emerging Technology Lab in Santa Clara, CA (hybrid). The role focuses on making modern LLMs, VLMs, and multimodal AI faster, smaller, and deployable in real-world environments. The ideal candidate will bridge research... 
    Contract work

    PVH (Tommy Hilfiger/Calvin Klein)

    Santa Clara, CA
    2 days ago
  •  ...USA's Emerging Technology Lab in Santa Clara, CA seeks a Contract AI Researcher focused on Efficient AI. You will explore model compression, quantization, and on-device inference to make LLMs and multimodal models faster and lighter for real-world applications. The role... 
    Contract work

    LG Electronics USA

    Santa Clara, CA
    4 days ago
  • $84.13 - $91.34 per hour

    AI Researcher - Efficient AI (Contractor) Step into the innovative world of LG Electronics. As a global leader in technology, LG Electronics...  ..., developing technologies that make modern LLMs, VLMs, multimodal models, and AI agents faster, smaller, and more deployable in... 
    Full time
    Contract work
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    Santa Clara, CA
    4 days ago
  • Apple Inc. in Cupertino is seeking a Senior Applied ML Researcher to design, train, and deploy state‑of‑the‑art models for video, audio, and multimodal tasks. You will work closely with research scientists, engineers, and product teams to enable intelligent systems that... 

    Apple

    Cupertino, CA
    4 days ago
  • LG Electronics is seeking a Contract AI Researcher for the Emerging Technology Lab in Santa Clara, CA. The role focuses on making LLMs, VLMs, and multimodal AI faster, smaller, and more deployable on edge devices. Work spans compression, quantization, efficient inference... 
    Contract work

    LG Electronics North America

    Santa Clara, CA
    3 days ago
  • Google DeepMind in Mountain View is seeking a Research Scientist in Multimodal Alignment, Safety and Fairness to advance foundational AI research and practical deployments. You will design experiments, develop scalable models, and collaborate across teams to push the boundaries... 

    Google

    Mountain View, CA
    4 days ago
  • $147k - $211k

    Google DeepMind in Mountain View, CA is seeking a Research Scientist to advance multimodal AI research. You will set up large-scale tests, prototype ideas, and deploy promising concepts across domains, from core ML to NLP and vision. You will publish findings, collaborate... 

    Google DeepMind

    Mountain View, CA
    3 days ago
  • $147k - $211k

    SnapshotWe are seeking strong Research Scientists with expertise in AI research and experience in interdisciplinary sociotechnical modeling to join a multimodal safety research effort within Google DeepMind's Frontier AI unit. This role requires a passion for understanding... 
    Full time

    DeepMind

    Mountain View, CA
    2 days ago
  • $140.4k - $264k

     ...source collaboration, constructing new platforms and supporting business innovation.What the Role Entails1.Conduct research and development of Omni multimodal large models, including the design and construction of training data, foundational model algorithm design,... 
    Full time
    Relocation package

    Tencent

    Palo Alto, CA
    2 days ago
  • Nuro in the United States is seeking a researcher-engineer to own the evaluation pipeline end to end for our autonomous systems. You will...  ..., hands-on post-training experience, and comfort with Python, production systems, and multimodal models. #J-18808-Ljbffr Nuro

    Nuro

    Mountain View, CA
    1 day ago
  • $165k - $195k

    Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania...  ...well as advanced MEMS design.As a part of the global research, our AI research in Silicon Valley focuses on Foundation Models, Natural... 
    Full time
    Work experience placement
    Local area
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    1 day ago
  • $162.7k - $263.18k

     ...Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the...  ...great outcomes.Job SummaryYour CareerAs a Principal Security Researcher, you will work at the forefront of AI-assisted vulnerability... 
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    2 days ago
  • $188k - $274k

     ...across organizations to gain support for research-based, user-centric solutions.Own project...  ...exclusive internal tools.As part of Lens on the Multimodal Search team, we're evolving Search to...  ...any way. As Search goes through a deeper AI evolution with AI Mode and other emergent... 

    Google

    Mountain View, CA
    3 days ago
  • $207k - $300k

     ...We approach projects that have the aspiration and riskiness of research with the speed and ambition of a startup.About the teamWe're a small...  ...on a mission to tackle climate change by developing novel AI reasoning capabilities that enable stakeholders to target their... 
    Full time

    X Company

    Mountain View, CA
    2 days ago
  •  ...Postdoctoral Researcher At Toyota Research Institute (TRI), we're on a mission to improve the quality of human life. We're developing new...  ...'ve built a world-class team advancing the state of the art in AI, robotics, driving, and material sciences. The Mission Our... 
    Work at office
    Shift work

    Toyota Research Institute

    Los Altos, CA
    1 day ago
  • NVIDIA is seeking an outstanding Senior Agentic AI Applied Researcher to build groundbreaking multi-modal agentic AI solutions for data science and machine learning. You will collaborate across internal teams to apply AI to data pipelines, experimentation workflows, and... 
    Worldwide

    NVIDIA

    Santa Clara, CA
    17 hours ago
  • Toyota Research Institute in California is seeking a Machine Learning Researcher to advance interpretable AI methods for end-to-end learned automated driving systems. You will work with senior researchers to develop glass-box representations, improving diagnosability and... 

    Tri

    Los Altos, CA
    1 day ago
  • Metamorphic is seeking Research Scientists to join our growing AI research team. You will push the boundaries of multimodal intelligence, write production-quality code, and help establish the benchmarks that will define this new era of intelligent systems. You will collaborate... 

    Metamorphic

    Palo Alto, CA
    17 hours ago
  • Simular is seeking a Research Scientist to push the boundaries of AI research across planning, reinforcement learning, and multimodal reasoning. You will drive end-to-end experiments, from data collection to model evaluation, and collaborate with engineers to bring research... 

    Simular

    Palo Alto, CA
    4 days ago
  • Job Description: Articul8 AI is seeking an Applied AI Researcher to advance our domain‑specific GenAI platform. You will design and run experiments,...  ...This role spans model training, reinforcement learning, multimodal understanding, and knowledge representation. Responsibilities... 

    Articul8

    Palo Alto, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Multimodal AI Researcher. Be the first to apply!