Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Finance AI Specialist: LLM Evaluation & Training

Obsidian

Cincinnatus LLC is recruiting for a finance SME to join a leading AI lab's GenAI team, evaluating model outputs against rubrics and guiding financial judgment in AI training data. This is a W-2 employment placement with the option to be placed at a premier AI Lab as part of the extended workforce.

The role emphasizes rigorous financial analysis, collaboration with researchers, and developing scoring guidelines, with a requirement of 35 hours per week on weekdays and a proven track record at top

#J-18808-Ljbffr
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Finance AI Specialist: LLM Evaluation & Training in San Francisco, CA vacancy
  • Turing, based in San Francisco, invites experts in finance to collaborate with AI researchers on evaluating model performance across capital markets, portfolio...  ...management, and related areas. You will help shape training methods, evaluation strategies, and benchmarks, with... 
    Training
    Remote job
    Hourly pay
    Flexible hours

    Turing

    San Francisco, CA
    3 days ago
  • $150k - $225k

    AI Software Engineer (Top AI/LLM Startup) AI Software Engineer (Top AI/LLM Startup) Focus Capital Markets...  ...Million in funding with a billion dollar evaluation and over $100M in revenue since it's...  ...ago Software Engineer, Python - AI Training (Freelance, Remote) Machine Learning... 
    Training
    Full time
    Freelance
    Remote work
    2 days per week

    Focus Capital Markets

    San Francisco, CA
    2 days ago
  •  ...seeking a senior Insurance SME to join a leading AI lab's GenAI team in San Francisco. You will guide underwriting-focused evaluation of AI model outputs, develop scoring rubrics...  ...-functional teams to ensure high-quality training data for advanced insurance tasks. This W-2... 
    Training

    Mercor

    San Francisco, CA
    2 days ago
  • Mercor is partnering with a leading AI research organization to engage experienced investment banking professionals for a project focused on evaluating how well AI systems perform real-world banking work. You will define what excellent work looks like by designing task-... 
    Training
    Remote job

    Mercor

    San Francisco, CA
    3 days ago
  • HeyMilo AI is hiring a Research Engineer to join our Applied AI team in San Francisco...  ...creating simulators, reward functions, and evaluation harnesses to measure model performance....  ...definition to reproducible environments for training and evaluation. #J-18808-Ljbffr HeyMilo... 
    Training

    HeyMilo AI

    San Francisco, CA
    3 days ago
  • $100k - $160k

    Applied AI Consultant: Build, Teach, Transform.Location: London...  ...prototype AI-powered solutions using LLM APIs (Claude, GPT, Gemini),...  ...and deliver hands-on AI training sessions for technical audiences...  ...servicesStay at the frontier: evaluate emerging models, tools, and architectural... 
    Training
    Full time
    Flexible hours
    Shift work

    BTS Group

    San Francisco, CA
    1 day ago
  • Cohere, a leading security-first enterprise AI company, is hiring a Member of Technical Staff in Data Analysis and Evaluation. You will design data-collection tasks, apply...  ...model performance, including distributed training of LLMs. The role emphasizes rigorous experimentation... 
    Training

    Cohere

    San Francisco, CA
    1 day ago
  • OpenAI is seeking a researcher to advance frontier evaluations and environments for safe AGI/ASI. You will help design north star model environments and steer major training runs so that research outputs translate into real-world products. Collaborate with researchers,... 
    Training

    Neura Market

    San Francisco, CA
    5 days ago
  •  ...software development with AI-powered formal...  ...for scaled distributed training. You'll be at the forefront...  ...team of AI experts, EBM specialists, formal verification engineers...  ...and models Evaluate reasoning approaches, including...  ...Optimizing and scaling LLM pipelines Adjust... 
    Training
    Full time
    Contract work

    Logical Intelligence

    San Francisco, CA
    1 day ago
  •  ...Meet Eloquent AI At Eloquent AI, we’re building...  ...of people manage their finances every day. From automating...  ...next-gen multimodal LLM architectures (LLMs, speech...  ...solutions ~​Refine training paradigms for real-...  ...via user simulations and evaluations. Requirements ~3... 
    Training
    Full time

    Eloquent AI

    San Francisco, CA
    1 day ago
  • $180k - $300k

     ...The Role You'll own the core AI systems that power Gamma: the...  ...job is to elevate quality, evaluate new frontier models, and push...  ...existing foundation models, not training new ones. You'll focus on prompting...  ...You'll Do Own Gamma's LLM and image prompts, measuring... 
    Training
    Full time
    Work at office
    Immediate start
    Work from home

    Gamma

    San Francisco, CA
    1 day ago
  •  ...Arena Intelligence Arena Intelligence is the open platform for evaluating how AI models perform in the real world. Created by researchers...  ...meaningful home here. We’re looking for: Hands-on experience training large-scale models, including reward models, preference models... 
    Training
    Permanent employment
    Work at office

    Arena

    San Francisco, CA
    2 days ago
  • $7.5k

     ...AI Engineer Location: San Francisco, CA or Phoenix, AZ (In-Office...  ...accuracy and latency: tune LLM and VLM pipelines, and...  ...Create robust evals: build evaluation frameworks that make AI behaviour...  ...stages, with scope to fine-tune or train models from scratch as individual... 
    Training
    Full time
    Work at office
    Relocation
    Visa sponsorship
    Relocation package

    Eql Tech

    San Francisco, CA
    1 day ago
  • $314.8k - $359.3k

     ...description": "Senior Distinguished AI Engineer At Capital...  ...including foundation model training, large language model...  ...similarity search, guardrails, model evaluation, experimentation, governance,...  ...and introduce state-of-the-art LLM optimization techniques to improve... 
    Training
    Full time
    Part time
    Local area

    Capital One Financial Corporation

    San Francisco, CA
    1 day ago
  • $286.2k - $326.7k

    Sr. Distinguished AI Engineer (Remote Eligible) At Capital...  ...including foundation model training, large language model...  ...similarity search, guardrails, model evaluation, experimentation, governance,...  ...and introduce state-of-the-art LLM optimization techniques to improve... 
    Training
    Full time
    Part time
    Local area
    Remote work

    Capital One Financial Corporation

    San Francisco, CA
    1 day ago
  • $300k - $400k

     ...reinvent how designers work in the AI era. We’re backed by top...  ...build and ship production-grade, LLM-powered features, define technical...  ...intersection of research, model training, and product—you’ll design the pipelines, evaluate tradeoffs, and lead execution end... 
    Training
    Full time

    Noon S.r.o

    San Francisco, CA
    1 day ago
  •  ...We are looking for a Staff AI Engineer to join the GenAI + Discovery...  ..., search and retrieval to evaluation and ROI frameworks. This is a...  ...and the shared genAI platform: LLM, and workflow orchestration,...  ...termination, promotional and training opportunities, without regard... 
    Training
    Full time
    Work at office
    Worldwide
    Flexible hours
    3 days per week

    Strava

    San Francisco, CA
    1 day ago
  •  ...Yutori is reimagining how people interact with the web by building AI agents that can reliably do everyday digital tasks. We are building the entire stack to be agent‑first, from training our own models to generative product interfaces. Towards this goal, we are looking... 
    Training
    Work at office
    Relocation
    Visa sponsorship

    Yutori

    San Francisco, CA
    3 days ago
  • $148.5k - $223.9k

     ...SalesforceSalesforce is the #1 AI CRM, where humans with agents...  ...workflows and integrates with LLM backends. You will own scalable...  ...our recruiters assess and evaluate candidates’ resumes and qualifications...  ..., promotion, benefits, training, assessment of job performance... 
    Training
    Full time

    Salesforce

    San Francisco, CA
    3 days ago
  • $110.7k - $372.9k

     ...need. Deloitte has a new AI-first effort, backed by...  ...and operationalize the LLM- and SLM-powered...  ...right time. Reliability, evaluation & safety • Implement observability...  ...our modeling and post-training engineers to improve...  ..., clinical and domain specialists, and product leaders to... 
    Training
    Local area
    Visa sponsorship

    Deloitte

    San Francisco, CA
    2 days ago
  • $2,000 per month

    Amplitude is the leading AI analytics platform, helping over 4,...  ...'s AI systems: orchestration, evaluation, retrieval, context management...  ...2+ years building production LLM or applied AI systemsExperience...  ...mentorship programs, management training, and wellness initiatives. We... 
    Training
    Home office

    Amplitude

    San Francisco, CA
    5 days ago
  • $146.37k - $246k

     ...About the role:Samsara's Finance & Strategy team is...  ...’s data foundation and AI systems. Our team drives...  ...through documentation, training, technical support, and...  ...large language model (LLM) applications, including...  ...frameworks, tool use, and evaluation techniques beyond... 
    Training
    Full time
    Remote work
    Flexible hours

    Samsara

    San Francisco, CA
    3 days ago
  •  ...About Orum Orum ’s AI-powered suite frees salespeople to do what...  ...idea to production Build LLM-based systems for coaching insights...  ...AI use cases Establish evaluation, monitoring, and feedback...  ...and deploying ML pipelines for training, inference, monitoring, and continuous... 
    Training
    Remote job
    Full time

    Orum

    San Francisco, CA
    1 day ago
  •  ...is building the world's first AI physical product creation platform...  ...AI scale: the pipelines, the evaluation infrastructure, and the...  ...best-practice rigor to how we train, deploy, evaluate, and improve...  ...diffusion / text-to-image models and LLM-based applications, including... 
    Training
    Full time

    Arcade

    San Francisco, CA
    1 day ago
  • $55k - $151.47k

     ...OpportunityAs part of the People Tech & AI team you will develop, test, and validate...  ...the role, three years of specialized training and/or progressively responsible work experience...  ...with Power Automate platform- Executing LLM evaluation frameworks using defined metrics-... 
    Training
    Full time
    Work experience placement
    H1b
    Remote work

    PwC

    San Francisco, CA
    4 days ago
  • $203.5k

     ...Information Job Title Lead, AI Engineering Job ID 10264...  ...stack: model experimentation, evaluation design, and production system...  ...data labeling strategies for LLM applicationsExperience with:...  ...education, licensure/certifications, training and skill level. Annual... 
    Training
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    San Francisco, CA
    3 days ago
  • $180k - $215k

     ...The flagship product—an AI-driven, non-invasive...  ...Implement advanced guardrails, evaluation frameworks, and...  ...industries (Healthcare, Finance, etc.) or navigating HIPAA...  ...to open-source LLM or Agentic AI frameworks...  ...including recruitment, hiring, training, relocation, promotion,... 
    Training
    Local area
    Worldwide
    Relocation

    HeartFlow

    San Francisco, CA
    1 day ago
  • $150k - $250k

    About Distyl AI Distyl is an applied AI technology company partnering with the world’s...  ...Looking ForAt Distyl, we build AI systems using Evaluation-Driven Development—an approach where...  ...intuition aloneDefine, calibrate, and operate LLM-based graders, aligning automated... 
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    3 days ago
  • $154.39k - $247.02k

     ...at a company where you matter.AI Infrastructure Engineer, Corporate...  ...role. You do not need to train models or develop novel ML techniques...  ...workflows, model APIs, evaluations, and AI safety considerations....  ...application concepts such as LLM APIs, prompt engineering, RAG,... 
    Training
    Work experience placement

    Axon

    San Francisco, CA
    21 hours ago
  • $293.5k

     ...Title Expert Senior Manager, AI Engineering Job ID 10433...  ..., model experimentation, and evaluation design to production system...  ...data labeling strategies for LLM appsDeep experience with:Advanced...  ...education, licensure/certifications, training and skill level. Annual... 
    Training
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Finance AI Specialist: LLM Evaluation & Training. Be the first to apply!