Finance AI Specialist: LLM Evaluation & Training
Obsidian
Cincinnatus LLC is recruiting for a finance SME to join a leading AI lab's GenAI team, evaluating model outputs against rubrics and guiding financial judgment in AI training data. This is a W-2 employment placement with the option to be placed at a premier AI Lab as part of the extended workforce.
The role emphasizes rigorous financial analysis, collaboration with researchers, and developing scoring guidelines, with a requirement of 35 hours per week on weekdays and a proven track record at top
#J-18808-Ljbffr- Turing, based in San Francisco, invites experts in finance to collaborate with AI researchers on evaluating model performance across capital markets, portfolio... ...management, and related areas. You will help shape training methods, evaluation strategies, and benchmarks, with...TrainingRemote jobHourly payFlexible hours
$150k - $225k
AI Software Engineer (Top AI/LLM Startup) AI Software Engineer (Top AI/LLM Startup) Focus Capital Markets... ...Million in funding with a billion dollar evaluation and over $100M in revenue since it's... ...ago Software Engineer, Python - AI Training (Freelance, Remote) Machine Learning...TrainingFull timeFreelanceRemote work2 days per week- ...seeking a senior Insurance SME to join a leading AI lab's GenAI team in San Francisco. You will guide underwriting-focused evaluation of AI model outputs, develop scoring rubrics... ...-functional teams to ensure high-quality training data for advanced insurance tasks. This W-2...Training
- Mercor is partnering with a leading AI research organization to engage experienced investment banking professionals for a project focused on evaluating how well AI systems perform real-world banking work. You will define what excellent work looks like by designing task-...TrainingRemote job
- HeyMilo AI is hiring a Research Engineer to join our Applied AI team in San Francisco... ...creating simulators, reward functions, and evaluation harnesses to measure model performance.... ...definition to reproducible environments for training and evaluation. #J-18808-Ljbffr HeyMilo...Training
$100k - $160k
Applied AI Consultant: Build, Teach, Transform.Location: London... ...prototype AI-powered solutions using LLM APIs (Claude, GPT, Gemini),... ...and deliver hands-on AI training sessions for technical audiences... ...servicesStay at the frontier: evaluate emerging models, tools, and architectural...TrainingFull timeFlexible hoursShift work- Cohere, a leading security-first enterprise AI company, is hiring a Member of Technical Staff in Data Analysis and Evaluation. You will design data-collection tasks, apply... ...model performance, including distributed training of LLMs. The role emphasizes rigorous experimentation...Training
- OpenAI is seeking a researcher to advance frontier evaluations and environments for safe AGI/ASI. You will help design north star model environments and steer major training runs so that research outputs translate into real-world products. Collaborate with researchers,...Training
- ...software development with AI-powered formal... ...for scaled distributed training. You'll be at the forefront... ...team of AI experts, EBM specialists, formal verification engineers... ...and models Evaluate reasoning approaches, including... ...Optimizing and scaling LLM pipelines Adjust...TrainingFull timeContract work
- ...Meet Eloquent AI At Eloquent AI, we’re building... ...of people manage their finances every day. From automating... ...next-gen multimodal LLM architectures (LLMs, speech... ...solutions ~Refine training paradigms for real-... ...via user simulations and evaluations. Requirements ~3...TrainingFull time
$180k - $300k
...The Role You'll own the core AI systems that power Gamma: the... ...job is to elevate quality, evaluate new frontier models, and push... ...existing foundation models, not training new ones. You'll focus on prompting... ...You'll Do Own Gamma's LLM and image prompts, measuring...TrainingFull timeWork at officeImmediate startWork from home- ...Arena Intelligence Arena Intelligence is the open platform for evaluating how AI models perform in the real world. Created by researchers... ...meaningful home here. We’re looking for: Hands-on experience training large-scale models, including reward models, preference models...TrainingPermanent employmentWork at office
$7.5k
...AI Engineer Location: San Francisco, CA or Phoenix, AZ (In-Office... ...accuracy and latency: tune LLM and VLM pipelines, and... ...Create robust evals: build evaluation frameworks that make AI behaviour... ...stages, with scope to fine-tune or train models from scratch as individual...TrainingFull timeWork at officeRelocationVisa sponsorshipRelocation package$314.8k - $359.3k
...description": "Senior Distinguished AI Engineer At Capital... ...including foundation model training, large language model... ...similarity search, guardrails, model evaluation, experimentation, governance,... ...and introduce state-of-the-art LLM optimization techniques to improve...TrainingFull timePart timeLocal area$286.2k - $326.7k
Sr. Distinguished AI Engineer (Remote Eligible) At Capital... ...including foundation model training, large language model... ...similarity search, guardrails, model evaluation, experimentation, governance,... ...and introduce state-of-the-art LLM optimization techniques to improve...TrainingFull timePart timeLocal areaRemote work$300k - $400k
...reinvent how designers work in the AI era. We’re backed by top... ...build and ship production-grade, LLM-powered features, define technical... ...intersection of research, model training, and product—you’ll design the pipelines, evaluate tradeoffs, and lead execution end...TrainingFull time- ...We are looking for a Staff AI Engineer to join the GenAI + Discovery... ..., search and retrieval to evaluation and ROI frameworks. This is a... ...and the shared genAI platform: LLM, and workflow orchestration,... ...termination, promotional and training opportunities, without regard...TrainingFull timeWork at officeWorldwideFlexible hours3 days per week
- ...Yutori is reimagining how people interact with the web by building AI agents that can reliably do everyday digital tasks. We are building the entire stack to be agent‑first, from training our own models to generative product interfaces. Towards this goal, we are looking...TrainingWork at officeRelocationVisa sponsorship
$148.5k - $223.9k
...SalesforceSalesforce is the #1 AI CRM, where humans with agents... ...workflows and integrates with LLM backends. You will own scalable... ...our recruiters assess and evaluate candidates’ resumes and qualifications... ..., promotion, benefits, training, assessment of job performance...TrainingFull time$110.7k - $372.9k
...need. Deloitte has a new AI-first effort, backed by... ...and operationalize the LLM- and SLM-powered... ...right time. Reliability, evaluation & safety • Implement observability... ...our modeling and post-training engineers to improve... ..., clinical and domain specialists, and product leaders to...TrainingLocal areaVisa sponsorship$2,000 per month
Amplitude is the leading AI analytics platform, helping over 4,... ...'s AI systems: orchestration, evaluation, retrieval, context management... ...2+ years building production LLM or applied AI systemsExperience... ...mentorship programs, management training, and wellness initiatives. We...TrainingHome office$146.37k - $246k
...About the role:Samsara's Finance & Strategy team is... ...’s data foundation and AI systems. Our team drives... ...through documentation, training, technical support, and... ...large language model (LLM) applications, including... ...frameworks, tool use, and evaluation techniques beyond...TrainingFull timeRemote workFlexible hours- ...About Orum Orum ’s AI-powered suite frees salespeople to do what... ...idea to production Build LLM-based systems for coaching insights... ...AI use cases Establish evaluation, monitoring, and feedback... ...and deploying ML pipelines for training, inference, monitoring, and continuous...TrainingRemote jobFull time
- ...is building the world's first AI physical product creation platform... ...AI scale: the pipelines, the evaluation infrastructure, and the... ...best-practice rigor to how we train, deploy, evaluate, and improve... ...diffusion / text-to-image models and LLM-based applications, including...TrainingFull time
$55k - $151.47k
...OpportunityAs part of the People Tech & AI team you will develop, test, and validate... ...the role, three years of specialized training and/or progressively responsible work experience... ...with Power Automate platform- Executing LLM evaluation frameworks using defined metrics-...TrainingFull timeWork experience placementH1bRemote work$203.5k
...Information Job Title Lead, AI Engineering Job ID 10264... ...stack: model experimentation, evaluation design, and production system... ...data labeling strategies for LLM applicationsExperience with:... ...education, licensure/certifications, training and skill level. Annual...TrainingPermanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week$180k - $215k
...The flagship product—an AI-driven, non-invasive... ...Implement advanced guardrails, evaluation frameworks, and... ...industries (Healthcare, Finance, etc.) or navigating HIPAA... ...to open-source LLM or Agentic AI frameworks... ...including recruitment, hiring, training, relocation, promotion,...TrainingLocal areaWorldwideRelocation$150k - $250k
About Distyl AI Distyl is an applied AI technology company partnering with the world’s... ...Looking ForAt Distyl, we build AI systems using Evaluation-Driven Development—an approach where... ...intuition aloneDefine, calibrate, and operate LLM-based graders, aligning automated...Work at office3 days per week$154.39k - $247.02k
...at a company where you matter.AI Infrastructure Engineer, Corporate... ...role. You do not need to train models or develop novel ML techniques... ...workflows, model APIs, evaluations, and AI safety considerations.... ...application concepts such as LLM APIs, prompt engineering, RAG,...TrainingWork experience placement$293.5k
...Title Expert Senior Manager, AI Engineering Job ID 10433... ..., model experimentation, and evaluation design to production system... ...data labeling strategies for LLM appsDeep experience with:Advanced... ...education, licensure/certifications, training and skill level. Annual...TrainingPermanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Finance AI Specialist: LLM Evaluation & Training. Be the first to apply!
- financial operations specialist San Francisco, CA
- financial crimes specialist San Francisco, CA
- finance specialist San Francisco, CA
- financial management specialist San Francisco, CA
- ai scientist San Francisco, CA
- ai data scientist San Francisco, CA
- visa sponsorship finance jobs San Francisco, CA
- patient financial services San Francisco, CA
- asset finance San Francisco, CA
- finance no experience San Francisco, CA


