Remote AI Philosophy Benchmark Architect
Mercor Inc
New York, NY
- Remote job
Mercor seeks expert philosophers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous philosophy questions across core domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. This fully remote, asynchronous role requires 10+ hours per week. You will be assigned to Question Authoring or Question Verification, and must provide edits, difficulty ratings, and 1 correct #J-18808-Ljbffr Mercor Inc
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote AI Philosophy Benchmark Architect in New York, NY vacancy
- ...author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous MCQs... ...domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will be assigned one of two...Remote job
- ...review academic assessment content for an AI research initiative. You will create and... ...domains and help establish gold-standard benchmarks for AI evaluation. Role focuses on expert... ...validation, with 10+ hours per week, fully remote, asynchronous collaboration. #J-18808-...Remote job10 hours per week
- Mercor is seeking AI experts for a benchmark dataset project evaluating models on visual document understanding and instruction-following in the Education domain. This remote role offers ~15-20 hours per week with flexible scheduling from the US or Canada. Payments are...Remote jobFlexible hours
- ...quality academic assessment content for an AI research initiative. You will craft and... ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will... ...Authoring or Question Verification, with remote collaboration and flexible scheduling....Remote jobFlexible hours
- Mercor seeks contributors for a benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Architecture domain. This is a fully remote independent contractor role, roughly 15-20 hours per week, available to candidates...Remote jobWeekly payFor contractorsFlexible hours
- A leading AI research accelerator is seeking a Biomedical Informatics Subject Matter Expert to design challenging evaluation questions... ...rubrics to measure outcomes in bioinformatics domains. This remote position requires a Master’s degree or PhD in Biomedical Informatics...Remote workContract work
- A leading AI research accelerator is seeking a highly skilled Biomedical Informatics Subject Matter Expert to develop rigorous evaluation... ...skills in Python or R. The successful candidate will work remotely on cutting-edge AI projects, with a commitment of at least 30 hours...Remote workContract work
- A top AI research partner is seeking a Biomedical Informatics Subject Matter Expert to develop rigorous evaluation questions for advanced... ..., and practical experience in the field. This position is fully remote and offers the opportunity to work on exciting AI projects with...Remote work
- ...delivery. The role focuses on authoring and verifying challenging medical questions and establishing rigorous evaluation standards for AI systems. Engagement is fully remote and asynchronous, with a minimum expectation of 10 hours per week. #J-18808-Ljbffr 24-Mag LlcRemote jobPart timeFor contractors10 hours per week
- ...Astrophysics & Cosmology Expert to design graduate-level problems that test AI systems using real scientific software, including simulations and... ...and Linux familiarity are essential; 15-20 hours per week work is expected, with remote compute sandbox access. #J-18808-Ljbffr ObsidianRemote work
- YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to... ...synthesize evidence, and document rationales with citations. This is a remote contractor role emphasizing high-quality medical content. #J-188...Remote jobFor contractors
$110.6k - $178k
Product Architect - AI & Digital ProductsCome make the world and accelerate... ...Product team working as a remote employee. We're building a... ...agents.Experience evaluating, benchmarking, and optimizing Large... ...Development: Our lifelong learning philosophy means you’ll have access to...Remote workFull timeLocal area- ...your goals. EisnerAmper is seeking an AI Platform Architect, Strategy & Enablement who will serve... ..., deprecated, retired) and evaluation benchmarks across the firm's multi-LLM environment... ...Conshohocken; Raleigh; Boston; Owings Mills; Dallas; Remote; PrincetonType: Full timeRemote workFull timeLocal areaWork visa
- YO AI Labs seeks a Product Management Expert on a remote, US-based contractor basis to shape AI-driven product workflows. You will create executive-ready deliverables, evaluate and benchmark PM outputs, and design realistic product scenarios reflecting senior-level thinking...Remote workFor contractors
- YO AI Labs is seeking Medical Evaluation Specialists to craft and validate challenging... ...questions and answers for AI evaluation. Remote contractor role focusing on clinical... ...ensuring accuracy and defensible explanations for AI benchmarks. #J-18808-Ljbffr YO AI LabsRemote jobFor contractors
- ...author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple... ..., evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will be assigned one of two...Remote job
- ...quality academic assessment content for an AI research initiative. You will write and... ...quality, and help establish gold-standard benchmarks for advancing AI capabilities. You may... ...with 10+ hours per week and asynchronous, fully remote work. #J-18808-Ljbffr Mercor IncRemote job10 hours per week
- ...author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple... ..., evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will be assigned one of two...Remote job
- ...author and review high-quality academic assessment content for an AI research initiative. You will craft and verify rigorous MCQs... ...domains, evaluate solution quality, and help establish gold-standard benchmarks for advancing AI capabilities. You will choose between Question...Remote job10 hours per week
- ...high-quality academic assessment content for an AI research initiative. You will write and verify rigorous... ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will join a fully remote, asynchronous team with a flexible 10+ hours per...Remote job10 hours per weekFlexible hours
- ...quality academic assessment content for an AI research initiative. You will write and... ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will... ...or question verification, with a remote, asynchronous schedule. #J-18808-Ljbffr...Remote job
- ...quality academic assessment content for an AI research initiative. You will write and... ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will... ...Authoring or Question Verification, with remote, asynchronous collaboration and a...Remote job10 hours per week
- Mercor is seeking a researcher to work on a benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Architecture domain. The role is fully remote and can be completed on your own schedule, with ~15-20 hours per week...Remote jobWeekly payContract workPart timeFor contractors
$50 - $63 per hour
...Develop and validate rigorous academic philosophy assessments that help benchmark AI capabilities. This fully remote role focuses on creating and reviewing challenging multiple-choice questions across Formal Ontology and Knowledge Representation, AI Ethics, Applied Epistemology...Remote workHourly pay10 hours per week- Mercor is seeking a remote independent contractor to work on a benchmark dataset project evaluating AI models for visual document understanding and instruction-following within the architecture domain. The role offers flexible scheduling and can be completed on your own...Remote jobWeekly payPart timeFor contractorsFlexible hours
- ...The Principal Architect AI SME in Delivery leads architecture for large Azure programs that... ...qualification and estimating using playbooks, benchmarks, and templates so the organization... .... Experience leading hybrid and remote teams and working with multiple stakeholders...Remote workFull timeTemporary workVisa sponsorshipWork visa
- Mercor is hiring PhD and Master's scientists to author AI evaluation tasks for a new benchmark in scientific computing. You will create original, executable research problems that today’s frontier models cannot solve, sourced from papers, datasets, or your own scenarios...Part time
- Vals AI, Inc. in San Francisco is seeking exceptional researchers and research engineers to design and build the next generation of AI benchmarks that evaluate real-world LLM capabilities. You will lead development of novel benchmarks shaping how foundation models are...
- ...design a central evaluation framework and reusable pipelines that span models and teams. You will implement evaluation pipelines, benchmark suites, baselines, and visualization tools to turn results into shared, actionable understanding. Strong SWE and statistics are essential...
- Full Time| Monday-Friday 8:00AM-4:30PM| Weekly Earned Wage Access is an option for this position. 60/40 onsite/remote Job Purpose or Goals: The AI architect defines the technical architecture for enterprise‑level AI systems and their integrations, ensuring they are designed...Remote workFull timeMonday to Friday
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote AI Philosophy Benchmark Architect. Be the first to apply!
Related searches
- software engineering manager remote New York, NY
- system engineer remote New York, NY
- remote technical project manager New York, NY
- remote recruitment consultant New York, NY
- remote property manager New York, NY
- remote reviewer New York, NY
- remote insurance agent New York, NY
- remote no experience New York, NY
- remote life underwriter New York, NY
- remote medical coding supervisor New York, NY


