Senior Research Scientist- Vision-Language-Action (VLA) Models
$185k - $215kBosch Group Inc
Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group ( a company with over 70 billion euro revenue, 400,000 employees worldwide, a very diverse product portfolio, and a history spanning over 125 years. The Research and Technology Center North America (RTC-NA) is dedicated to providing technologies and system solutions for various Bosch business fields, primarily in the field of artificial intelligence, energy technologies, internet technologies, circuit design, semiconductors and wireless, as well as advanced MEMS design. As a part of the global research, our AI research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data Science, AI System Engineering, Time-series Analysis. We develop scalable, intelligent, and trustworthy AIoT solutions for Bosch products and services in application areas such as automated driving, advanced driver assistance systems (ADAS), robotics, smart manufacturing, enterprise AI, health care, smart home and building solutions. Originating from the AI research in Silicon Valley, our Intelligent Autonomous Systems group is responsible for enabling future autonomous Bosch products by pushing the boundaries of automated driving, advanced driver assistance systems (ADAS), robotics and automation through key innovations that encompass system architecture and AI components. These include methods for motion planning, high level task planning and decision making as well as systems for making these technologies work on real products by building frameworks that take advantage of technologies in the field of reliable distributed computing. We work with internal partners of different Bosch business units to transfer our solutions into future products. We also actively collaborate with leading groups in academia and industry to promote research ideas and publish research findings in internationally renowned conferences and journals such as CVPR, ICRA, IROS, RSS, NeurIPS and CoRL. Job Description As a Senior Research Scientist- Vision-Language-Action (VLA) Models, you contribute to research projects at the forefront of the ADAS/AD industry. Key responsibilities include:
- Conduct research and engineering in core AI and machine learning fields to enable Embodied AI (including computer vision, autonomous planning, open-world learning, and so on) for related business domains of ADAS/AD, industrial automation, robotics etc.
- Push the boundaries in (modular) end-to-end perception and planning for ADAS/AD, incorporating advancements in large vision-language-(action) models to aid reasoning capabilities and explainability.
- Collaborate cross-functionally with global research and engineering teams to ensure seamless technology transfer and system integration.
- Implement research results to solve real-world challenges, ensuring high-quality system integration within Bosch's existing platforms.
- Stay at the forefront of innovation by actively engaging with academic and industry communities through conferences, workshops, and technical events.
- Document and disseminate research findings through high-caliber publications and/or patent submissions.
- Ph.D. in Computer Science, Robotics or a related discipline or Master's degree with >= 2/4 years industry experience after graduation.
- A minimum of 5 years of R&D experience, or an equivalent graduate research background, primarily in AI technologies including Computer Vision and Robotic or Automotive Motion and Behavioral Planning.
- Proficiency in one or more programming languages commonly used in machine learning (e.g., Python, C++, Rust).
- Strong interpersonal, communication, and teamwork capabilities.
- Knowledge of major machine learning frameworks like TensorFlow or PyTorch.
- Hands-on experience in reinforcement learning for behavior or motion planning or other applicable contexts and familiarity with common RL techniques (e.g. PPO, DQN, DDPG).
- A strong portfolio of publications in premier machine learning, deep learning, robotics and computer vision journals and conferences.
- Experience with real-world product development and deployment of autonomous systems.
- Hands-on experience building and applying multimodal transformer-based sequence-to-sequence models, especially multimodal vision-language-action models.
- Hands-on experience in computer vision and deep learning, with work in any of the following areas: multimodal transformers, multimodal language models, diffusion models, NeRF, gaussian splatting, object detection / segmentation, 3D scene understanding, sensor calibration, SfM, voxel/BEV grid-based feature representation.
Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Senior Research Scientist- Vision-Language-Action (VLA) Models in Sunnyvale, CA vacancy
$185k - $215k
...DescriptionThe Bosch Research and Technology... ...focuses on Foundation Models, Big Data Visual Analytics... ...AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality,... ...Job DescriptionAs a Senior Research Scientist- Vision-Language-Action (VLA) Models, you contribute...SeniorLanguageWork experience placementLocal areaWorldwide$170k - $260k
...CAResearch and Development - Computer Vision and Deep Learning /Full-time /... ...decision-making, building the Vision-Language-Action (VLA) models that form SuperDrive's reasoning layer... ...platform teams to bring models from research to productionRequired qualificationsM...SeniorLanguageFull time$192k - $304.75k
We are now looking for a Senior Research Scientist focused on Multimodal Foundation Models and Robotics! NVIDIA is searching for... ...following topics: LLMs; Large vision-language models; Video generative... ...and diffusion algorithms; or Action-based transformers.Outstanding...SeniorLanguageFull time$193.93k - $352.29k
...flexible, partner-led business model, Nuro is working toward a... ...collaborate closely with researchers and engineers on the Learned... ...models. Leverage large language models and world foundation... ...autonomous driving. Experiences in vision-language-action models, reinforcement...SeniorLanguageImmediate startFlexible hours- ...office collaboration. We are looking for a Senior / Staff AI Research Scientist, Foundation Models to advance robotic embodied intelligence.... ...tasks. Responsibilities Design and deploy vision-language(-action) models (VLM/VLA) for contextual understanding and generalized...SeniorLanguageWork at officeVisa sponsorship
$165k - $195k
Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California,... ...AI research in Silicon Valley focuses on Foundation Models, Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Big Data Visual Analytics...SeniorLanguageFull timeWork experience placementLocal areaWorldwide$165k - $185k
Company DescriptionThe Bosch Research and Technology Center North America with offices... ...in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data Science...SeniorLanguageWork experience placementWorldwide$244.14k - $413.16k
...stacks, we are developing large-scale Vision-Language-Action (VLA) models and World Models to handle the... ...tail scenarios of global driving. As a Senior Staff Machine Learning Engineer, you... ...The ability to balance cutting-edge research with the deterministic requirements...SeniorLanguageFull timeOverseas- ...About the Institute of Foundation Models We are a dedicated research lab for building, understanding, using... ...world-class researchers, data scientists, and engineers, tackling the most... ...Summary As a Research Scientist in the Vision Language Model (VLM) team, your role will...Language
$185k - $215k
...DescriptionThe Bosch Research and Technology... ...focuses on Foundation Models, Big Data Visual... ...AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality,... ...Job DescriptionAs a Senior Research Scientist- Robotics AI, you... ...vision-language-(action) models to aid reasoning...SeniorLanguageWork experience placementLocal areaWorldwide$192.2k - $260k
...practical experience to join the Modeling and Optimization (MOP)... ...learning, robotics, operations research, statistics, mathematics or equivalent... ...- Knowledge of programming languages such as C/C++, Python, Java... ...insurance (medical, dental, vision, prescription, Basic Life &...SeniorLanguageLocal areaFlexible hours$174.72k - $295.68k
...connectivity.We are looking for a full-time Machine Learning Engineer / Research Scientist to drive the modeling and algorithmic development of XPENG’s next-generation Vision-Language-Action (VLA) Foundation Model — the core brain that powers our end-to-end autonomous...SeniorLanguageFull time$174.72k - $295.68k
...expertise in generative modeling and large-scale deep... ...this role, you will research, implement, and... ...evolves under an agent's actions, and serving as a learned... ...experts in computer vision, generative AI, and... ...training to improve Vision-Language-Action (VLA) driving performance....SeniorLanguageFull time- ...robotic platforms. As a Senior AI/ML Research Engineer, you will... ...the foundation models—VFMs, VLMs, and VLA models—that let our... ...adapting large pretrained vision and multimodal... ..., instruments, actions, and context from intraoperative... ...(VA), vision-language-action (VLA), or...SeniorLanguageLocal areaWorldwideFlexible hours
$192.2k - $260k
...Delivery Foundation Model team, where you'll... ...world-class scientists and engineers to pioneer... ...an exceptional Senior Applied Scientist... ...direction for specific research initiatives,... ...ambitious research vision with real-world impact... ..., C++ or other languages- Strong publication...SeniorLanguageLocal areaWorldwideFlexible hours$200k - $287.5k
.../HybridAt Toyota Research Institute (TRI), we... ...and Large Behavior Models (LBM).The... ...are looking for a Senior Machine Learning Researcher... ...-art, pixels-to-action, end-to-end system... ...integrating visual-language-action modalities.... ...focus on computer vision as the primary sensing...SeniorLanguageFull timeLocal areaShift work$262k - $364k
...Senior Staff Research Scientist, Gemini Research, DeepMind Share Senior... ...in our Gemini models. Our team is deeply... ...experienced in building large language models end-to-end,... ...conferences and visioning activities.... ...opportunity and affirmative action employer. We are...SeniorLanguage$218.8k - $335.3k
...ready to redefine mobility and shape the future of autonomous transportation? As a Staff Research Scientist specializing in Vision-Language Models (VLMs), Vision-Language-Action models (VLAs), and Onboard Foundational Models, you will advance the frontier of artificial...LanguageFull timeLocal areaRemote workWork from homeRelocation packageFlexible hours$192k - $304.75k
We are now looking for a Senior Research Scientist for Generative AI!NVIDIA is searching for a world... ...great impacts with generative AI models. You will be building research... ...practice of deep learning, computer vision, natural language processing, or computer graphicsTrack...SeniorLanguageFull time$262k - $364k
...Senior Staff Research Scientist/Engineer, Agentic Post Training, DeepMind Share... ...include: Health, dental, vision, life, disability insurance... ...~ Experience with large language model, reinforcement learning, and... ...and affirmative action employer. We are committed...SeniorLanguageTemporary work$174k - $252k
Conduct AI research and development in the field of health... ...for multimodal AI models and set up tests to... ...of work. As a Research Scientist, you'll setup large-scale... ...data mining, natural language processing, hardware... ...insight from it, and take action toward their health...SeniorLanguage$184k - $287.5k
...software and applied research engineers focused on geometric... ...raw sensor data into actionable world understanding,... ...training foundation models.Advance the state of... ...in geometric computer vision, combining classical multi... ...and evaluate vision-language models on skills...SeniorLanguageFull time$192.2k - $260k
...exceptional Sr. Applied Scientist to lead the... ...where conventional vision alone falls short.... ...insufficient. You will lead research that translates... ...deep learning models for object... ...rigor, and a bias for action. We believe in learning... ...Python or related language ~ PhD in...SeniorLanguageLocal areaFlexible hoursNight shift$224k - $356.5k
...understood using advanced computer vision and deep learning. Our team... ..., image, and 3D data into actionable insights. You will... ...perception, simulation, and large models to bring research into production at scale.... ...with a focus on Vision-Language Model (VLM).Experience building...SeniorLanguageFull time$174k - $252k
...combination of engineering and research expertise to advance Gemini'... ...state-of-the-art Large Language Models (LLM's) core image/video understanding... ...closely with Research Scientists and Software engineers (SWEs... ...g., LLMs, multimodal, large vision models) or with genAI-...Language$262k - $365k
Author research papers to share and generate impact of... ...conferences and visioning activities. Deliver full... ...of work. As a Research Scientist, you'll setup large-scale... ...data mining, natural language processing, hardware and... ...our advanced AI models, delivers computing power...SeniorLanguageWorldwide$180k - $258.75k
.../HybridAt Toyota Research Institute (TRI), we... ...world foundation models that leverage large... ...flow, semantics, actions, tactile, audio, etc... ..., and Video-Language-Action models, with... ...alongside research scientists, understand the research... ..., dental, and vision insurance, 401(k)...SeniorLanguageFull timeLocal areaShift work$174k - $252k
Define and lead research agendas, experimental designs, and evaluation... ..., audio, and multimodal modeling.Co-design and optimize models... ...types of work. As a Research Scientist, you'll setup large-scale tests... ...learning, data mining, natural language processing, hardware and...SeniorLanguage$166k - $259k
...multidisciplinary team of AI researchers, software engineers, power systems... ...systems or power electronics modeling and simulations.Experience in... ...general-purpose programming languages such as Python, Java, or C/... ...equityMedical, dental, and vision coverageGenerous PTO and flexible...SeniorLanguageFull timeFlexible hours- ...Sunnyvale, CA for a Senior Machine Learning Engineer... ...focus on Computer Vision, Deep Learning,... ...initiate and drive new research directions, and have... ...multimodal data, and action models Contribute to surgical... ...(VA), and vision language action (VLA) models Lead development...SeniorLanguageImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Research Scientist- Vision-Language-Action (VLA) Models. Be the first to apply!
Related searches
- analytical scientist Sunnyvale, CA
- drug safety scientist Sunnyvale, CA
- qc scientist Sunnyvale, CA
- operations research scientist Sunnyvale, CA
- validation scientist Sunnyvale, CA
- image scientist Sunnyvale, CA
- research scientist Sunnyvale, CA
- application scientist Sunnyvale, CA
- scientist Sunnyvale, CA
- research scientist - biology Sunnyvale, CA



