Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Research Scientist- Vision-Language-Action (VLA) Models

$185k - $215k

Bosch USA

Company Description The Bosch Research and Technology Center North America with offices in Sunnyvale, California; Pittsburgh, Pennsylvania; and Cambridge, Massachusetts is a part of the global Bosch Group ( a company with over 70 billion euro revenue, 400,000 employees worldwide, a very diverse product portfolio, and a history spanning over 125 years. The Research and Technology Center North America (RTC‑NA) is dedicated to providing technologies and system solutions for various Bosch business fields, primarily in the fields of artificial intelligence, energy technologies, internet technologies, circuit design, semiconductors and wireless, as well as advanced MEMS design. The AI research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data Science, AI System Engineering, and Time‑series Analysis. We develop scalable, intelligent, and trustworthy AIoT solutions for Bosch products and services in application areas such as automated driving, advanced driver assistance systems (ADAS), robotics, smart manufacturing, enterprise AI, health care, smart home and building solutions. The Intelligent Autonomous Systems group works to enable future autonomous Bosch products by pushing the boundaries of automated driving, advanced driver assistance systems, robotics and automation through key innovations that encompass system architecture and AI components. These include methods for motion planning, high‑level task planning and decision making, as well as systems for making these technologies work on real products by building frameworks that take advantage of technologies in the field of reliable distributed computing. We collaborate with internal partners of different Bosch business units to transfer our solutions into future products and actively collaborate with leading groups in academia and industry to promote research ideas and publish research findings in internationally renowned conferences and journals such as CVPR, ICRA, IROS, RSS, NeurIPS and CoRL. Job Description As a Senior Research Scientist – Vision‑Language‑Action (VLA) Models, you contribute to research projects at the forefront of the ADAS/AD industry. Conduct research and engineering in core AI and machine learning fields to enable Embodied AI (including computer vision, autonomous planning, open‑world learning, and so on) for related business domains of ADAS/AD, industrial automation, robotics, etc. Push the boundaries in (modular) end‑to‑end perception and planning for ADAS/AD, incorporating advancements in large vision‑language‑action models to aid reasoning capabilities and explainability. Collaborate cross‑functionally with global research and engineering teams to ensure seamless technology transfer and system integration. Implement research results to solve real‑world challenges, ensuring high‑quality system integration within Bosch's existing platforms. Stay at the forefront of innovation by actively engaging with academic and industry communities through conferences, workshops, and technical events. Document and disseminate research findings through high‑caliber publications and/or patent submissions. Qualifications Basic Qualifications Ph.D. in Computer Science, Robotics or a related discipline or Master's degree with ≥ 2/4 years industry experience after graduation. A minimum of 5 years of R&D experience, or an equivalent graduate research background, primarily in AI technologies including Computer Vision and Robotic or Automotive Motion and Behavioral Planning. Proficiency in one or more programming languages commonly used in machine learning (e.g., Python, C++, Rust). Strong interpersonal, communication, and teamwork capabilities. Knowledge of major machine learning frameworks like TensorFlow or PyTorch. Hands‑on experience in reinforcement learning for behavior or motion planning or other applicable contexts and familiarity with common RL techniques (e.g., PPO, DQN, DDPG). A strong portfolio of publications in premier machine learning, deep learning, robotics and computer vision journals and conferences. Preferred Qualifications Experience with real‑world product development and deployment of autonomous systems. Hands‑on experience building and applying multimodal transformer‑based sequence‑to‑sequence models, especially multimodal vision‑language‑action models. Hands‑on experience in computer vision and deep learning, with work in any of the following areas: multimodal transformers, multimodal language models, diffusion models, NeRF, gaussian splatting, object detection / segmentation, 3D scene understanding, sensor calibration, SfM, voxel/BEV grid‑based feature representation. Additional Information We offer a competitive base salary for this position with a range in US‑California of $185,000 – $215,000 along with an annual corporate bonus, and a long‑term incentive bonus designed to reward sustained impact and contribution over time. Within the salary range, the individual pay is determined based on several factors, including, but not limited to, work experience and job knowledge, complexity of the role, job location, etc. Your well‑being matters at Bosch! We offer a benefits package designed to empower you in every area of your life. This includes premium health coverage, a 401(k) with generous matching, resources for financial planning and goal setting, ample paid time off, parental leave, and comprehensive life and disability protection. Your Recruiter can share more details for this position during the interview process. Learn more about our full benefits offerings by visiting: Equal Opportunity Employer Bosch adheres to Federal, State, and Local laws regarding drug‑testing. Employment is contingent upon the successful completion of a drug screen and background check. Candidates who have been offered the position must pass both screenings before their start date. #J-18808-Ljbffr Bosch USA

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Research Scientist- Vision-Language-Action (VLA) Models in Sunnyvale, CA vacancy
  • $39 - $66 per hour

     ...Description The Bosch Research and Technology Center...  ...Valley focuses on Foundation Models, Big Data Visual...  ...Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud...  ...Intern for World-Action Models / VLA for Autonomous Driving,... 
    Language
    Work experience placement
    Internship
    Local area
    Worldwide

    Bosch USA

    Sunnyvale, CA
    3 days ago
  • $192k - $304.75k

    We are now looking for a Senior Research Scientist focused on Multimodal Foundation Models and Robotics! NVIDIA is searching for...  ...following topics: LLMs; Large vision-language models; Video generative...  ...and diffusion algorithms; or Action-based transformers.Outstanding... 
    Senior
    Language
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...is built. We are seeking a senior vision language model engineer to design and build...  ...be doing:Partner with our researchers to develop and evaluate prototypes...  ...., video, sensor, language/action traces) tailored for end‑to...  ..., and multimodal VLM/VLA or foundation models.... 
    Senior
    Language
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $193.93k - $352.29k

     ...flexible, partner-led business model, Nuro is working toward a...  ...collaborate closely with researchers and engineers on the Learned...  ...models. Leverage large language models and world foundation...  ...autonomous driving. Experiences in vision-language-action models, reinforcement... 
    Senior
    Language
    Immediate start
    Flexible hours

    Nuro

    Mountain View, CA
    2 days ago
  • $192k - $304.75k

    We're now looking for a Senior Research Scientist, Multi-Modal Language Models!NVIDIA is seeking a Senior Research Scientist passionate about multi modal language...  ...related areas.4+ years of experiences in computer vision, especially multi-modal LLMs.Proficiency in Python... 
    Senior
    Language
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $192k - $304.75k

    We are seeking an outstanding Research Scientist or Research Engineer with a passion for synthetic...  ...to training autonomous driving models of tomorrow. As part of NVIDIA's Physical...  ...models, end-to-end driving, reasoning, and vision-language models. The ideal candidate will have... 
    Senior
    Language
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $165k - $195k

    Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California,...  ...AI research in Silicon Valley focuses on Foundation Models, Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Big Data Visual Analytics... 
    Senior
    Language
    Full time
    Work experience placement
    Local area
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    4 hours ago
  • $200k - $287.5k

     ...Senior Machine Learning Researcher At Toyota Research Institute (TRI),...  ...and Large Behavior Models (LBM). We are looking...  ...-the-art, pixels-to-action, end-to-end system...  ...integrating visual-language-action modalities. Beyond...  ...a focus on computer vision as the primary... 
    Senior
    Language
    Local area
    Shift work

    Toyota Research Institute

    Los Altos, CA
    5 days ago
  • $165k - $185k

    Company DescriptionThe Bosch Research and Technology Center North America with offices...  ...in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data Science... 
    Senior
    Language
    Work experience placement
    Worldwide

    Robert Bosch

    Sunnyvale, CA
    1 day ago
  •  ...About the Institute of Foundation Models We are a dedicated research lab for building, understanding, using...  ...world-class researchers, data scientists, and engineers, tackling the most...  ...Summary As a Research Scientist in the Vision Language Model (VLM) team, your role will... 
    Language

    Institute of Foundation Models

    Sunnyvale, CA
    9 days ago
  • $167.1k - $226.1k

     ...resourceful Applied Scientist in the field of Large Language Models (LLMs), Artificial...  ...evaluation. As a senior member of the team...  ...to set the research agenda for how we...  ...and coupled with actionable conclusions. Partner...  ...(medical, dental, vision, prescription, Basic... 
    Senior
    Language
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    2 days ago
  • $244.14k - $413.16k

     ...stacks, we are developing large-scale Vision-Language-Action (VLA) models and World Models to handle the...  ...tail scenarios of global driving. As a Senior Staff Machine Learning Engineer, you...  ...The ability to balance cutting-edge research with the deterministic requirements... 
    Senior
    Language
    Full time
    Overseas

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $208k - $327.75k

     ...architectures.We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for autonomous...  ...of frontier AI research, hardware architecture, systems...  ...autonomous vehicle stack, including Vision-Language-Action (VLA) models, Multimodal... 
    Senior
    Language
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $192.2k - $260k

     ...practical experience to join the Modeling and Optimization (MOP)...  ...learning, robotics, operations research, statistics, mathematics or equivalent...  ...- Knowledge of programming languages such as C/C++, Python, Java...  ...insurance (medical, dental, vision, prescription, Basic Life &... 
    Senior
    Language
    Local area
    Flexible hours

    Amazon

    Santa Clara, CA
    2 days ago
  • $174.72k - $295.68k

     ...connectivity.We are looking for a full-time Machine Learning Engineer / Research Scientist to drive the modeling and algorithmic development of XPENG’s next-generation Vision-Language-Action (VLA) Foundation Model — the core brain that powers our end-to-end autonomous... 
    Senior
    Language
    Full time

    XPENG Motors

    Santa Clara, CA
    2 days ago
  • $192.2k - $260k

     ...Delivery Foundation Model team, where you'll...  ...world-class scientists and engineers to pioneer...  ...an exceptional Senior Applied Scientist...  ...direction for specific research initiatives,...  ...ambitious research vision with real-world impact...  ..., C++ or other languages- Strong publication... 
    Senior
    Language
    Local area
    Worldwide
    Flexible hours

    Amazon

    Santa Clara, CA
    2 days ago
  •  ...robotic platforms. As a Senior AI/ML Research Engineer, you will...  ...the foundation models—VFMs, VLMs, and VLA models—that let our...  ...adapting large pretrained vision and multimodal...  ..., instruments, actions, and context from intraoperative...  ...(VA), vision-language-action (VLA), or... 
    Senior
    Language
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    Sunnyvale, CA
    4 hours ago
  • $204k - $259k

     ...ML Frameworks & Efficiency team partners with Research and Production teams across Waymo to develop models in Perception and Planning that are core to...  ...Stay current with the latest research in RL, Vision-Language-Action (VLA) models, and World models to inform and inspire... 
    Senior
    Language
    Full time
    Remote work

    Waymo

    Mountain View, CA
    8 hours ago
  • $240.1k

     ...Description We are looking for a senior AI/ML researcher who can architect and guide...  ...healthcare foundation model and a high-reliability inference...  ...will define the technical vision for how our organization should...  ...in related AI domains. 5+ Language/multimodal model tuning and... 
    Senior
    Language
    Local area

    Amazon

    Santa Clara, CA
    2 days ago
  • $192.2k - $260k

     ...LLC Overview Are you a passionate scientist in the computer vision area who aspires to apply your skills...  ...Multi‑modal LLMs and/or Vision Language Models and collaborating with different Amazon...  ...for computer vision applications. Research and implement state‑of‑the‑art computer... 
    Senior
    Language
    Flexible hours

    Amazon

    Sunnyvale, CA
    2 days ago
  • $126k - $248k

     ...fine‑tuned embedding models and rerankers to...  ...a strong team of AI researchers from Stanford, MIT,...  ...Overview We are seeking a Senior Research Scientist to join our team and...  ..., and natural language processing. Familiarity...  ...to utilize research vision to innovate the entire... 
    Senior
    Language
    Local area

    MongoDB

    Palo Alto, CA
    4 days ago
  • $192.2k - $260k

     ...and experienced Applied Scientist to support adoption...  ...services and tools for model customization, including...  ...across large language models. As an Applied...  ...and 6+ years of applied research experience Experience...  ...insurance (medical, dental, vision, prescription, Basic Life... 
    Senior
    Language
    Local area
    Flexible hours

    Socket

    Sunnyvale, CA
    3 days ago
  • $192.2k - $260k

     ...an experienced Applied Scientist who will join a team...  ...you can pursue applied research, with many peta-bytes...  ...and evaluation of AI models for predictive learning...  ...C++, Python or related language - Experience with neural...  ...(medical, dental, vision, prescription, Basic Life... 
    Senior
    Language
    Local area
    Flexible hours

    Amazon

    Sunnyvale, CA
    2 days ago
  • $171.6k - $222.2k

     ...We are seeking a Senior Applied Scientist to join our team in developing pioneering AI research, Generative AI, Agentic AI, Large Language Models (LLMs), Diffusion and Flow Models, and other advanced...  ..., Multi-modality Computer Vision, Diffusion Models, Reinforcement... 
    Senior
    Language
    Worldwide
    Flexible hours

    PVH (Tommy Hilfiger/Calvin Klein)

    Sunnyvale, CA
    2 days ago
  • $192.2k - $260k

     ...Senior Applied Scientist, Data Processing Agents Science Job...  ...databases, programming languages, and distributed...  ...management, data governance, actions, and experiences,...  ...features; researching the state of the art...  ...insurance (medical, dental, vision, prescription, Basic... 
    Senior
    Language
    Local area
    Flexible hours

    Amazon

    Santa Clara, CA
    1 day ago
  • $213k - $263k

     ...Senior Machine Learning Engineer, Computer Vision/VLM Waymo is an autonomous driving technology company with the mission...  ...-art computer vision / multimodal models (e.g., Gemini) to extract the rich...  ...prompting strategies for Vision-Language Models (VLMs) to elicit complex,... 
    Senior
    Language
    Full time
    Remote work

    Waymo

    Mountain View, CA
    4 days ago
  • $143.63k - $299.38k

     ...many challenges in the areas of natural language processing, machine learning techniques,...  ...About You You are a seasoned Applied ML Researcher who thrives at the intersection of theoretical...  ...distillation. You believe that a model is only as good as its evaluation framework... 
    Senior
    Language
    Work at office
    Flexible hours

    Yahoo Holdings Inc.

    Mountain View, CA
    1 day ago
  • $192k - $304.75k

     ...searching for an outstanding Senior Researcher working on efficient deep...  ...about methods for post-training model optimization (pruning,...  ...the top venues in computer vision and machine learning. Our existing...  ....Experience with large language models and large vision-language... 
    Senior
    Language
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $192k - $304.75k

    We are now looking for a Senior Research Scientist for Generative AI!NVIDIA is searching for a world...  ...great impacts with generative AI models. You will be building research...  ...practice of deep learning, computer vision, natural language processing, or computer graphicsTrack... 
    Senior
    Language
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $218.8k - $335.3k

     ...ready to redefine mobility and shape the future of autonomous transportation? As a Staff Research Scientist specializing in Vision-Language Models (VLMs), Vision-Language-Action models (VLAs), and Onboard Foundational Models, you will advance the frontier of artificial... 
    Language
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    9 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Research Scientist- Vision-Language-Action (VLA) Models. Be the first to apply!