Visual Generation & Multimodal Evaluation Researcher Graduate (AML-Ark-US) - 2027 Start (PhD)
$153.9k - $300.96kByteDance
Responsibilities The Applied Machine Learning Ark team combines system engineering and machine learning to develop and operate Large Language Model (LLM) service platforms that offer businesses Model-as-a-Service (MaaS) solutions, serving both large model providers and downstream users. The US team drives the design, development, and operation of MaaS solutions across the US and international markets outside mainland China. We are building full-stack, end-to-end solutions spanning text and multimodal LLM algorithms, LLM training/fine-tuning/inference frameworks, prompt engineering, model alignment, and intelligent agent systems. Beyond model serving, we operate large-scale log analytics pipelines that process massive volumes of invocation logs from text models, multimodal models, and agent systems — extracting usage patterns, quality signals, and actionable insights to inform model improvement, system optimization, and product decisions through continuous, data-driven feedback loops. We are actively seeking talented engineers and researchers specializing in Large Language Models and AI Agent systems to join our dynamic team. We are looking for talented individuals to join our team. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Successful candidates must be able to commit to an onboarding date by the end of the year. Please state your availability and graduation date clearly in your resume. Build evaluation systems for image and video models/agents, covering generation quality, instruction following, multimodal understanding, and safety. Develop automated metrics and model-based evaluators, and design reproducible human evaluation protocols. Design and develop video generation/debugging agents that orchestrate multi-step creative workflows. Build large-scale image and video data pipelines, and turn evaluation findings into model and product improvements. Qualifications Minimum Qualifications: Individuals who are completing or have recently completed a PhD's degree in Computer Science, Artificial Intelligence, Machine Learning, Computer Vision, or a related field. Solid foundation in deep learning and computer vision, including generative modeling fundamentals. Practical experience in at least one of: visual generation, multimodal LLMs, video understanding, or visual quality assessment. Strong Python skills and proficiency with PyTorch or an equivalent framework, or multimodal evaluation framework. Demonstrated research or engineering ability through publications, substantial projects, internships, or open-source work. Preferred Qualifications Publications at top-tier vision or ML venues, e.g., NeurIPS, ICML, CVPR, ICCV, ECCV, etc. Hands-on experience with modern visual generation stacks, including diffusion-based models and their post-training. Familiarity with visual generation benchmarks, or experience building evaluation frameworks. Experience applying agent frameworks to creative workflows, or working with large-scale video data infrastructure. Job Information 【For Pay Transparency】Compensation Description (Annually) The base salary range for this position in the selected city is $153900 - $300960 annually. Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units. Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure). The Company reserves the right to modify or change these benefits programs at any time, with or without notice. For Los Angeles County (unincorporated) Candidates Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state, and local laws including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Our company believes that criminal history may have a direct, adverse and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment: Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues; Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems; and Exercising sound judgment. About Us Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Lemon8, CapCut and Pico as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content. Why Join ByteDance Inspiring creativity is at the core of ByteDance's mission. Our innovative products are built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and enrich life - a mission we work towards every day. As ByteDancers, we strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. By constantly iterating and fostering an \"Always Day 1\" mindset, we achieve meaningful breakthroughs for ourselves, our Company, and our users. When we create and grow together, the possibilities are limitless. Join us. Diversity & Inclusion ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too. Reasonable Accommodation ByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at #J-18808-Ljbffr ByteDance
$121.6k - $243.2k
...Applied Machine Learning Ark team combines system... ...downstream users. The US team drives the design... ...spanning text and multimodal LLM algorithms, LLM training... ...engineers and researchers specializing in Large... ...Responsibilities Design evaluation systems for LLM-based...For graduatesTemporary workInternship- ...join our team in 2027. As a graduate, you will get... ...of data generated on the platform... ...multilingual and multimodal content, and upgraded... ...millions of US dollars while... ...and produce research outcomes with... ...completed a PhD degree in Computer... ...with evaluation of AI systems,...For graduatesFor phdFlexible hoursShift work
- ...individuals to join our team. As a graduate, you will get opportunities... ...team focuses on applied research in Generative AI, delivering intelligent... ...-edge areas including multimodal foundation models, image and... ...have recently completed a PhD degree in Software Development...For graduatesFor phd
- ...join our team in 2027. As a graduate, you will get... ...signals, and multimodal product... ...temporal and visual/textual representations... ...as a generative problem of producing... ...applications.5. Agent Evaluation, Safety &... ...to the research community via... ...recently completed a PhD in Software...VisualFor graduatesFor phd
$64.88 - $121.85 per hour
...Entails Conduct research and development of Omni multimodal large models... ...capability evaluation, and... ...understanding and generation capabilities... ...fields; graduate degrees are... ...e.g., audio‑visual) research is... ...Location: State(s) US-California-... ...those who start working during...VisualFor graduatesHourly payFull timeRelocation package- ...join our team in 2027. As a graduate, you will get... ...the next‑generation AI‑native computing... ...systems, multimodal serving, advanced... ...workloads Bridging research breakthroughs... ...number of PhD candidates (... ...including deployment, evaluation, and iteration... ...reach out to us at #J-18808-...For graduatesFor phdTemporary workLocal area
- ...gets done. Location: US-WA-Bellevue... .../ Hybrid Team: AI Research Business Unit: Engineering... ...for recent PhD graduates looking to deepen... ...expected completion by start date) in Computer... ..., fine-tuning, or evaluating LLMs or foundation... ...retrieval-augmented generation (RAG), knowledge...For graduatesFor phdContract workFixed term contractInternship
$142.7k - $270.95k
...Agentic and other Generative AI scenarios... ..., NLP, and multimodal learning.... ...to evaluation, optimization... ...with Adobe Research, engineering... ...QualificationsMasters, or PhD in Computer... .../ visual transformers... ...you can help us advance our... ...sales roles starting salaries are...VisualFor phdFull timeTemporary workLocal areaWorldwide$57 per hour
...understanding, multimodal understanding,... ...are looking for Research Interns in the... ...individuals to join us for an... ...at ByteDance. PhD Internship PhD... ...in your resume (Start date, End date)... ...Build the next‑generation monetization platform... ...Graduating December 2025 onwards...For graduatesFor phdHourly payWork experience placementSummer workInternshipLocal area- ...is seeking engineers and researchers to advance Large Language... ...and data pipelines across US and international markets,... ...end-to-end video and image generation, evaluation, and multimodal capabilities. The role emphasizes... ...to cutting-edge visual AI research. #J-18808-Ljbffr...Visual
$232.56k - $427.5k
Research Scientist in AI Foundation Model Infrastructure - Seed - Graduates - 2027 Start (PhD) Location: Seattle Team: Technology Employment Type:... ...large-scale model training, evaluation, and inference. Optimize... ...accommodation, please reach out to us at #J-18808-Ljbffr...For graduatesFor phdTemporary workLocal area$129.6k - $219.6k
...is seeking researchers in artificial... ...leaders in generative AI, large language... ...learning, multimodal reasoning... ...reusable code, run evaluations, and analyze... ...obtaining a PhD degree in AI... ...in both new graduates and those... ...those who start working... ...innovation and allow us to better...For graduatesFor phdRelocation package- Research Scientist Graduate (Foundation Model, Generative AI) - 2025 Start (PhD) Join ByteDance as a Research Scientist Graduate, focusing on foundation models and generative... ...and development in foundation models and multimodal machine learning, especially in generative...For graduatesFor phd
$153.9k - $300.96k
...computer vision, NLP and multimodality models and... ...join our team. As a graduate, you will get... ...recently completed a PhD degree in computer... ...discipline. Related Research Experience at... ...of work from the start, so our candidates... ...employee's journey with us. We believe that...For graduatesFor phdTemporary work$155k - $175k
...Engineer - LLMs & Generative AITruveta is... ...to enable researchers to find... ...enthusiastic new graduates who are... ...perfect place to start.This... ...language and multimodal foundation models... ....Design and evaluate models for... ...years with a PhD).Experience... ...between. Join us as we build...For graduatesFor phdFor contractorsVisa sponsorshipWork visaFlexible hours$50 - $60 per hour
ML Postdoc Researcher - Healthcare AI Innovation... ...over a third of the US population), we are... ...designed for recent PhD graduates who are passionate... ...Design, develop, and evaluate state‑of‑the‑art... ...foundation models, multimodal learning, causal ML, generative AI) Work Across Modalities...For graduatesFor phdHourly payFull timeFor contractorsRemote work$57 per hour
Student Researcher (AI Foundation Model Infrastructure - Seed) - 2027 Start (PhD) Location: Seattle Team: Technology Employment Type: Intern Job Code: A258974A Responsibilities Conduct research on infrastructure and systems for large‑scale models. Explore methods...For phdHourly payInternshipLocal area$153.9k - $300.96k
...join our team. As a graduate, you will get... ...Role This full-time PhD Graduate role provides early-career researchers and engineers the... ...Search research and evaluation: developing new retrieval... ...understanding, multimodal search, or LLM... ...and Tokyo. Why Join Us Inspiring...For graduatesFor phdFull timeTemporary workLocal areaWorldwide$57 per hour
...of-the-art computer vision, NLP and multimodality models and algorithms to protect our... ...looking for talented individuals to join us for an internship. PhD internships at Our Company provide... ...contribute to our products and research, as well as to the organization's future...For phdHourly payInternshipLocal area$229k - $343k
...operates Snapchat, a visual messaging app... ....Snap’s Generative ML Platform team... ...worldwide. From multimodal LLMs and video... ...up-to-date with research and are excited... ...experience; or a PhD in a related... ...shy and provide us some information... ...successful candidate’s starting pay will be...VisualFor phdFull timeLive inWork at officeLocal areaWorldwide- ...roughly 20 cheaper and visual models served... ...infrastructure for Generative AI at DoorDash, leading... ...…B.S., M.S., or PhD. in Computer Science... ...AWQ/GPTQ), or cold-start/throughput tuningExperience... ...recruiting team by evaluating job related... ...members who can help us go from a company...VisualFor phdHourly payWork at officeLocal areaRemote workFlexible hours
$102.6k - $143.64k
...US Commercial Infrastructure Management Graduate (Global Procurement) - 2027 Start (MBA) Location: Seattle Employment Type: Regular Job Code: A174505 Responsibilities... ...financial models (TCO, ROI, NPV) to evaluate investment scenarios, driving data-driven asset...For graduatesContract workTemporary workFor contractors$202.16k - $368.22k
...scale model training, evaluation, and inference.... ...field, with an expected graduation date in 2027 and the ability to commit... ..., the Seed team's research spans MLLM, GenMedia,... ...models and cutting‑edge multimodal capabilities. Our... ...please reach out to us at #J-18808-Ljbffr...For graduatesTemporary workInternshipLocal area$130k - $180k
...teams across Naver Search US, BAND, ThingsBook, and... ...Models (LLMs) and generative AI systems Build and optimize... ...conduct experiments to evaluate model performance,... ...A Master’s degree or PhD in Computer Science, Machine... ...to translating research concepts into tangible...For phdFull timeWork visa$247k - $396k
...We're searching for a PhD to serve as our technical... ...art models—and helping us rediscover and rebuild... ...Qualifications: PhD Graduate in AI, Computer Science... ...or Robotics (Top-tier research lab background) Deep... ...successful candidate's starting base pay will be...For graduatesFor phdWork at officeLocal area3 days per week- Research Scientist, LLM Evaluation & Post-Training page is loaded##... ...define next-generation evaluation-driven... ...for LLM and multimodal systems, covering... ...Education:** MS or PhD in Computer... ...foundation models (graduate research counts... ...analysis, and visualization. Hands-on...For graduatesFor phdFull timeRemote work
$202.16k - $368.22k
...signals, and multimodal product... ...temporal and visual/textual representations... ...as a generative problem of... ...applications. Agent Evaluation, Safety &... ...to the research community via... ...end of year 2027. Please... ...availability and graduation date clearly... ...completed a PhD in Software...VisualFor graduatesFor phdTemporary workLocal area- Join us as we work together to inspire creativity and enrich life... ...on the development and research of the compute infrastructure... ...individuals to join our team. As a graduate, you will get opportunities to... ...or have recently completed a PhD degree in Computer Science, Computer...For graduatesFor phd
$121.6k - $243.2k
...Learning Engineer Graduate (E-Commerce Recommendation... ...Alliance) - 2027 Start Location: Seattle... ...LLMs , NLP, and multimodal machine learning to... ...AI solutions from research to production. - Stay... ...model iteration, evaluation, and deployment.... ...and Tokyo. Why Join Us Inspiring...For graduatesTemporary workWork experience placementInternship$127.4k - $191.2k
...are building the next generation of these systems. The... ...do. We are looking for PhD graduates who have worked at that... ..., build, and evaluate next-generation ranking... ...engineering to bring research ideas into production,... ...strengths that enable us to support the growing...For graduatesFor phdFull timeInternshipWork at officeWorldwideShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Visual Generation & Multimodal Evaluation Researcher Graduate (AML-Ark-US) - 2027 Start (PhD). Be the first to apply!
- senior researcher Seattle, WA
- machine learning researcher Seattle, WA
- researcher Seattle, WA
- senior design researcher Seattle, WA
- design researcher Seattle, WA
- qualitative researcher Seattle, WA
- data collection researcher Seattle, WA
- product researcher Seattle, WA
- survey researcher Seattle, WA
- legal researcher Seattle, WA

