Senior Research Scientist- Vision-Language-Action (VLA) Models
$185k - $215kBosch Group
Job Description
Job Description
Company Description
The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group ( a company with over 70 billion euro revenue, 400,000 employees worldwide, a very diverse product portfolio, and a history spanning over 125 years. The Research and Technology Center North America (RTC-NA) is dedicated to providing technologies and system solutions for various Bosch business fields, primarily in the field of artificial intelligence, energy technologies, internet technologies, circuit design, semiconductors and wireless, as well as advanced MEMS design.
As a part of the global research, our AI research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data Science, AI System Engineering, Time-series Analysis. We develop scalable, intelligent, and trustworthy AIoT solutions for Bosch products and services in application areas such as automated driving, advanced driver assistance systems (ADAS), robotics, smart manufacturing, enterprise AI, health care, smart home and building solutions.
Originating from the AI research in Silicon Valley, our Intelligent Autonomous Systems group is responsible for enabling future autonomous Bosch products by pushing the boundaries of automated driving, advanced driver assistance systems (ADAS), robotics and automation through key innovations that encompass system architecture and AI components. These include methods for motion planning, high level task planning and decision making as well as systems for making these technologies work on real products by building frameworks that take advantage of technologies in the field of reliable distributed computing. We work with internal partners of different Bosch business units to transfer our solutions into future products. We also actively collaborate with leading groups in academia and industry to promote research ideas and publish research findings in internationally renowned conferences and journals such as CVPR, ICRA, IROS, RSS, NeurIPS and CoRL.
Job DescriptionAs a Senior Research Scientist- Vision-Language-Action (VLA) Models, you contribute to research projects at the forefront of the ADAS/AD industry. Key responsibilities include:
- Conduct research and engineering in core AI and machine learning fields to enable Embodied AI (including computer vision, autonomous planning, open-world learning, and so on) for related business domains of ADAS/AD, industrial automation, robotics etc.
- Push the boundaries in (modular) end-to-end perception and planning for ADAS/AD, incorporating advancements in large vision-language-(action) models to aid reasoning capabilities and explainability.
- Collaborate cross-functionally with global research and engineering teams to ensure seamless technology transfer and system integration.
- Implement research results to solve real-world challenges, ensuring high-quality system integration within Bosch's existing platforms.
- Stay at the forefront of innovation by actively engaging with academic and industry communities through conferences, workshops, and technical events.
- Document and disseminate research findings through high-caliber publications and/or patent submissions.
Basic Qualifications
- Ph.D. in Computer Science, Robotics or a related discipline or Master's degree with >= 2/4 years industry experience after graduation.
- A minimum of 5 years of R&D experience, or an equivalent graduate research background, primarily in AI technologies including Computer Vision and Robotic or Automotive Motion and Behavioral Planning.
- Proficiency in one or more programming languages commonly used in machine learning (e.g., Python, C++, Rust).
- Strong interpersonal, communication, and teamwork capabilities.
- Knowledge of major machine learning frameworks like TensorFlow or PyTorch.
- Hands-on experience in reinforcement learning for behavior or motion planning or other applicable contexts and familiarity with common RL techniques (e.g. PPO, DQN, DDPG).
- A strong portfolio of publications in premier machine learning, deep learning, robotics and computer vision journals and conferences.
Preferred Qualifications
- Experience with real-world product development and deployment of autonomous systems.
- Hands-on experience building and applying multimodal transformer-based sequence-to-sequence models, especially multimodal vision-language-action models.
- Hands-on experience in computer vision and deep learning, with work in any of the following areas: multimodal transformers, multimodal language models, diffusion models, NeRF, gaussian splatting, object detection / segmentation, 3D scene understanding, sensor calibration, SfM, voxel/BEV grid-based feature representation.
We offer a competitive base salary for this position with a range in US-California of --$185,000 - $215,000 along with an annual corporate bonus, and a long-term incentive bonus designed to reward sustained impact and contribution over time. Within the salary range, the individual pay is determined based on several factors, including, but not limited to, work experience and job knowledge, complexity of the role, job location, etc.
Your well-being matters at Bosch! We offer a a benefits package designed to empower you in every area of your life. This includes premium health coverage, a 401(k) with generous matching, resources for financial planning and goal setting, ample paid time off, parental leave, and comprehensive life and disability protection. Your Recruiter can share more details for this position during the interview process.
Learn more about our full benefits offerings by visiting:
Equal Opportunity Employer, including disability / veterans.
*Bosch adheres to Federal, State, and Local laws regarding drug-testing. Employment is contingent upon the successful completion of a drug screen and background check. Candidates who have been offered the position must pass both screenings before their start date.
#LI-JM1
$185k - $215k
...Description The Bosch Research and Technology... ...focuses on Foundation Models, Big Data Visual Analytics... ...AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality,... ...Job Description As a Senior Research Scientist – Vision‑Language‑Action (VLA) Models, you contribute...SeniorLanguageWork experience placementLocal areaWorldwide$39 - $66 per hour
...Description The Bosch Research and Technology Center... ...Valley focuses on Foundation Models, Big Data Visual... ...Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud... ...Intern for World-Action Models / VLA for Autonomous Driving,...LanguageWork experience placementInternshipLocal areaWorldwide$193.93k - $352.29k
...flexible, partner-led business model, Nuro is working toward a... ...collaborate closely with researchers and engineers on the Learned... ...models. Leverage large language models and world foundation... ...autonomous driving. Experiences in vision-language-action models, reinforcement...SeniorLanguageImmediate startFlexible hours$185k - $215k
# Senior Research Scientist- Robotics AI## Bosch GroupPublished 09 Jun 2026### Share this jobSunnyvale... ...Time## Role Highlights### Languages usedPythonC++Rust### Key skillsAIDeep... ...advancements in large vision-language-action models to improve reasoning and explainability...SeniorLanguage- The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania... ...AI research in Silicon Valley focuses on Foundation Models, Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Big Data Visual...SeniorLanguageWorldwide
$165k - $185k
...Description Company Description The Bosch Research and Technology Center North America... ...Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data...SeniorLanguageWork experience placementWorldwide- ...About the Institute of Foundation Models We are a dedicated research lab for building, understanding, using... ...world-class researchers, data scientists, and engineers, tackling the most... ...Summary As a Research Scientist in the Vision Language Model (VLM) team, your role will...Language
- ...innovative and driven Applied Researchers in deep learning. Ideal... ...foundational and generative AI models plus agentic AI systems. It strives... ...-scale models like Large Language Models (LLMs), Transformers,... ...multimodal models that integrate vision, language, and structured/...SeniorLanguage
$167.1k - $226.1k
...resourceful Applied Scientist in the field of Large Language Models (LLMs), Artificial... ...evaluation. As a senior member of the team... ...to set the research agenda for how we... ...and coupled with actionable conclusions. Partner... ...(medical, dental, vision, prescription, Basic...SeniorLanguageLocal areaFlexible hours$126k - $248k
...fine-tuned embedding models and rerankers to... ...strong team of AI researchers from Stanford, MIT,... ...We are seeking a Senior Research Scientist to join our team and... ...learning, and natural language processing * Familiarity... ...to utilize research vision to innovate the entire...SeniorLanguageFull timeLocal areaWorldwideFlexible hours$204k - $259k
...ML Frameworks & Efficiency team partners with Research and Production teams across Waymo to develop models in Perception and Planning that are core to... ...Stay current with the latest research in RL, Vision-Language-Action (VLA) models, and World models to inform and inspire...SeniorLanguageFull timeRemote work$200k - $287.5k
...Description At Toyota Research Institute (TRI),... ...Large Behavior Models (LBM). The Opportunity... ...are looking for a Senior Machine Learning... ...-art, pixels-to-action, end-to-end system... ...visual-language-action modalities.... ...focus on computer vision as the primary sensing...SeniorLanguageLocal areaShift work$240.1k
...Description We are looking for a senior AI/ML researcher who can architect and guide... ...healthcare foundation model and a high-reliability inference... ...will define the technical vision for how our organization should... ...in related AI domains. 5+ Language/multimodal model tuning and...SeniorLanguageLocal area$192.2k - $260k
...LLC Overview Are you a passionate scientist in the computer vision area who aspires to apply your skills... ...Multi‑modal LLMs and/or Vision Language Models and collaborating with different Amazon... ...for computer vision applications. Research and implement state‑of‑the‑art computer...SeniorLanguageFlexible hours$192.2k - $260k
...an experienced Applied Scientist who will join a team... ...you can pursue applied research, with many peta-bytes... ...and evaluation of AI models for predictive learning... ...C++, Python or related language - Experience with neural... ...(medical, dental, vision, prescription, Basic Life...SeniorLanguageLocal areaFlexible hours$143.63k - $299.38k
...many challenges in the areas of natural language processing, machine learning techniques,... ...About You You are a seasoned Applied ML Researcher who thrives at the intersection of theoretical... ...distillation. You believe that a model is only as good as its evaluation framework...SeniorLanguageWork at officeFlexible hours$192.2k - $260k
...Delivery Foundation Model team, where you'll... ...world-class scientists and engineers to pioneer... ...an exceptional Senior Applied Scientist... ...direction for specific research initiatives,... ...ambitious research vision with real‑world impact... ..., C++ or other languages Strong publication...SeniorLanguageLocal areaWorldwideFlexible hours$197.8k - $296.6k
...Lab Summary The Robot Intelligence Lab at Samsung Research America is a new facility dedicated to advancing the field of... ...develop advanced technologies such as robotics foundation models, vision‑language‑action (VLA) models, vision language models (VLMs), reinforcement...SeniorLanguage- ...software and foundation models enable vehicles to... ...systems. Our vision is to create... ...looking for Applied Scientists to join Wayve Labs... ...a high‑conviction research team with the strategic... ...and costs of actions. Representation Learning... ..., using vision, language, and active...LanguageFull timeWork at officeWork from homeVisa sponsorshipRelocation packageFlexible hours
$126k - $423k
...looking for multiple passionate Research Scientists to join the Research Group... ...on pretraining world‑action foundation model with various world modalities including vision and physics associated with... ..., human data incorporation, language modality, and spatial reasoning...LanguageFull timeFor contractorsFor subcontractorCasual workWork at officeImmediate startRemote workDay shift$126k - $423k
...looking for multiple passionate Research Scientists to join the Research Group... ...on pretraining world‑action foundation model with various world modalities including vision and physics associated with... ..., human data incorporation, language modality, and spatial reasoning...LanguageFull timeFor contractorsFor subcontractorCasual workWork at officeImmediate startRemote workDay shift$117k - $234k
...summary: As a Senior Data Scientist at Walmart, you will... ...data sourcing, model development, validation... ...challenges into actionable insights, guiding... ...programming languages such as Python and... ...include medical, vision and dental... ...Technology, Operations Research, Statistics, Applied...SeniorLanguageFull timeTemporary workPart time$85 - $90 per hour
...the development of should cost model and multi-sourcing Drive ECO... ...) and implement corrective actions Establish Traceability on components... ...~MATERIAL MANAGEMENT Languages: English ( Speak, Read,... ..., including medical, dental, vision, and 401K contributions, as well...SeniorLanguageFull timeContract work$138k - $185k
...was founded with a vision to solve for... ...are looking for a Senior Product Marketing... ...conceptually, what actions it can take, when... ...Stay ahead of AI research trends, product patterns... ...clear, intuitive language for multiple audiences... ...a hybrid work model that aims to boost...SeniorLanguageRemote workFlexible hoursShift work$301.75k - $355k
...Crusoe. About This Role The Senior Director for the Model LifeCycle team will... ...Learning models, including Large Language Models (LLMs). What You’... ...strongly preferred. Research publications at NeurIPS, ICML... ...Comprehensive health, dental & vision insurance Employer...SeniorLanguageTemporary work$176k - $253k
...Job Description At Toyota Research Institute (TRI), we’re on a... ...are looking for a Research Scientist to join us in building intelligent... ...to explore how large language models and agentic infrastructure can... ...with large language models, vision-language models, or agentic...LanguageWork experience placementInternshipLocal areaShift work- ...We are now looking for a Research Scientist with a focus in System Software and I/O! NVIDIA is seeking Research Scientists with a focus in... ...release. Experience with C, C++, CUDA, Python, and scripting languages. MPI and NACL would be a plus. Strong interpersonal skills are...SeniorLanguageWork experience placement
- ...era of mobility and logistics by leading the technical vision for Vision‑Language‑Action (VLA) models. You will identify weaknesses in state‑of‑the‑art... ...in AI, Computer Science, or Robotics from a top‑tier research laboratory. Deep expertise in Vision‑Language Models...LanguageLocal area
$180k - $258.75k
...Description At Toyota Research Institute (TRI),... ...world foundation models that leverage... ...flow, semantics, actions, tactile, audio, etc... ...Augmentation, and Video-Language-Action models,... ...alongside research scientists, understand the... ...medical, dental, and vision insurance, 401(k)...SeniorLanguageLocal areaShift work$192.2k - $260k
...Senior Applied Scientist, Sponsored Products and Brands -- Offsite... ...a long‑term science vision and roadmap for the Offsite... ...and the latest modeling techniques in the field... ...6+ years of applied research experience. Experience... ...C++, Python or related languages. Experience with...SeniorLanguageFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Research Scientist- Vision-Language-Action (VLA) Models. Be the first to apply!
- drug safety scientist Sunnyvale, CA
- molecular biology scientist Sunnyvale, CA
- safety scientist Sunnyvale, CA
- validation scientist Sunnyvale, CA
- support scientist Sunnyvale, CA
- water quality scientist Sunnyvale, CA
- machine learning research scientist Sunnyvale, CA
- manufacturing scientist Sunnyvale, CA
- senior principal scientist Sunnyvale, CA
- research scientist - biology Sunnyvale, CA


