Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote AI Philosophy Benchmark Architect

Mercor Inc

New York, NY
  • Remote job

Mercor seeks expert philosophers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous philosophy questions across core domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. This fully remote, asynchronous role requires 10+ hours per week. You will be assigned to Question Authoring or Question Verification, and must provide edits, difficulty ratings, and 1 correct #J-18808-Ljbffr Mercor Inc

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote AI Philosophy Benchmark Architect in New York, NY vacancy
  •  ...author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous MCQs...  ...domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will be assigned one of two... 
    Remote job

    Mercor

    New York, NY
    2 days ago
  •  ...review academic assessment content for an AI research initiative. You will create and...  ...domains and help establish gold-standard benchmarks for AI evaluation. Role focuses on expert...  ...validation, with 10+ hours per week, fully remote, asynchronous collaboration. #J-18808-... 
    Remote job
    10 hours per week

    Mercor

    New York, NY
    2 days ago
  • Mercor is seeking AI experts for a benchmark dataset project evaluating models on visual document understanding and instruction-following in the Education domain. This remote role offers ~15-20 hours per week with flexible scheduling from the US or Canada. Payments are... 
    Remote job
    Flexible hours

    Obsidian

    San Francisco, CA
    4 days ago
  •  ...quality academic assessment content for an AI research initiative. You will craft and...  ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will...  ...Authoring or Question Verification, with remote collaboration and flexible scheduling.... 
    Remote job
    Flexible hours

    Mercor Inc

    New York, NY
    3 days ago
  • Mercor seeks contributors for a benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Architecture domain. This is a fully remote independent contractor role, roughly 15-20 hours per week, available to candidates... 
    Remote job
    Weekly pay
    For contractors
    Flexible hours

    Mercor

    San Francisco, CA
    19 hours ago
  • A leading AI research accelerator is seeking a Biomedical Informatics Subject Matter Expert to design challenging evaluation questions...  ...rubrics to measure outcomes in bioinformatics domains. This remote position requires a Master’s degree or PhD in Biomedical Informatics... 
    Remote work
    Contract work

    Turing

    United States
    19 hours ago
  • A leading AI research accelerator is seeking a highly skilled Biomedical Informatics Subject Matter Expert to develop rigorous evaluation...  ...skills in Python or R. The successful candidate will work remotely on cutting-edge AI projects, with a commitment of at least 30 hours... 
    Remote work
    Contract work

    Turing

    Los Angeles, CA
    3 days ago
  • A top AI research partner is seeking a Biomedical Informatics Subject Matter Expert to develop rigorous evaluation questions for advanced...  ..., and practical experience in the field. This position is fully remote and offers the opportunity to work on exciting AI projects with... 
    Remote work

    Turing

    United States
    3 days ago
  •  ...delivery. The role focuses on authoring and verifying challenging medical questions and establishing rigorous evaluation standards for AI systems. Engagement is fully remote and asynchronous, with a minimum expectation of 10 hours per week. #J-18808-Ljbffr 24-Mag Llc
    Remote job
    Part time
    For contractors
    10 hours per week

    24-Mag Llc

    New York, NY
    4 days ago
  •  ...Astrophysics & Cosmology Expert to design graduate-level problems that test AI systems using real scientific software, including simulations and...  ...and Linux familiarity are essential; 15-20 hours per week work is expected, with remote compute sandbox access. #J-18808-Ljbffr Obsidian
    Remote work

    Obsidian

    San Francisco, CA
    4 days ago
  • YO AI Labs seeks Medical Evaluation Specialists, including medical students, residents, physicians, and biomedical professionals, to...  ...synthesize evidence, and document rationales with citations. This is a remote contractor role emphasizing high-quality medical content. #J-188... 
    Remote job
    For contractors

    YO AI Labs

    Dallas, TX
    1 day ago
  • $110.6k - $178k

    Product Architect - AI & Digital ProductsCome make the world and accelerate...  ...Product team working as a remote employee. We're building a...  ...agents.Experience evaluating, benchmarking, and optimizing Large...  ...Development: Our lifelong learning philosophy means you’ll have access to... 
    Remote work
    Full time
    Local area

    Stanley Black & Decker

    Towson, MD
    1 day ago
  •  ...your goals. EisnerAmper is seeking an AI Platform Architect, Strategy & Enablement who will serve...  ..., deprecated, retired) and evaluation benchmarks across the firm's multi-LLM environment...  ...Conshohocken; Raleigh; Boston; Owings Mills; Dallas; Remote; PrincetonType: Full time
    Remote work
    Full time
    Local area
    Work visa

    EisnerAmper

    Iselin, NJ
    1 day ago
  • YO AI Labs seeks a Product Management Expert on a remote, US-based contractor basis to shape AI-driven product workflows. You will create executive-ready deliverables, evaluate and benchmark PM outputs, and design realistic product scenarios reflecting senior-level thinking... 
    Remote work
    For contractors

    YO AI Labs

    Washington DC
    1 day ago
  • YO AI Labs is seeking Medical Evaluation Specialists to craft and validate challenging...  ...questions and answers for AI evaluation. Remote contractor role focusing on clinical...  ...ensuring accuracy and defensible explanations for AI benchmarks. #J-18808-Ljbffr YO AI Labs
    Remote job
    For contractors

    YO AI Labs

    Palo Alto, CA
    1 day ago
  •  ...author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple...  ..., evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will be assigned one of two... 
    Remote job

    Obsidian

    New York, NY
    2 days ago
  •  ...quality academic assessment content for an AI research initiative. You will write and...  ...quality, and help establish gold-standard benchmarks for advancing AI capabilities. You may...  ...with 10+ hours per week and asynchronous, fully remote work. #J-18808-Ljbffr Mercor Inc
    Remote job
    10 hours per week

    Mercor Inc

    New York, NY
    2 days ago
  •  ...author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple...  ..., evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will be assigned one of two... 
    Remote job

    Mercor

    New York, NY
    4 days ago
  •  ...author and review high-quality academic assessment content for an AI research initiative. You will craft and verify rigorous MCQs...  ...domains, evaluate solution quality, and help establish gold-standard benchmarks for advancing AI capabilities. You will choose between Question... 
    Remote job
    10 hours per week

    Obsidian

    Nashville, TN
    2 days ago
  •  ...high-quality academic assessment content for an AI research initiative. You will write and verify rigorous...  ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will join a fully remote, asynchronous team with a flexible 10+ hours per... 
    Remote job
    10 hours per week
    Flexible hours

    Mercor Inc

    New York, NY
    2 days ago
  •  ...quality academic assessment content for an AI research initiative. You will write and...  ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will...  ...or question verification, with a remote, asynchronous schedule. #J-18808-Ljbffr... 
    Remote job

    Obsidian

    New York, NY
    2 days ago
  •  ...quality academic assessment content for an AI research initiative. You will write and...  ...quality, and help establish gold-standard benchmarks used to advance AI capabilities. You will...  ...Authoring or Question Verification, with remote, asynchronous collaboration and a... 
    Remote job
    10 hours per week

    Obsidian

    New York, NY
    2 days ago
  • Mercor is seeking a researcher to work on a benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Architecture domain. The role is fully remote and can be completed on your own schedule, with ~15-20 hours per week... 
    Remote job
    Weekly pay
    Contract work
    Part time
    For contractors

    Mercor

    New York, NY
    2 days ago
  • $50 - $63 per hour

     ...Develop and validate rigorous academic philosophy assessments that help benchmark AI capabilities. This fully remote role focuses on creating and reviewing challenging multiple-choice questions across Formal Ontology and Knowledge Representation, AI Ethics, Applied Epistemology... 
    Remote work
    Hourly pay
    10 hours per week

    SaidGig

    United States
    16 days ago
  • Mercor is seeking a remote independent contractor to work on a benchmark dataset project evaluating AI models for visual document understanding and instruction-following within the architecture domain. The role offers flexible scheduling and can be completed on your own... 
    Remote job
    Weekly pay
    Part time
    For contractors
    Flexible hours

    Obsidian

    San Francisco, CA
    4 days ago
  •  ...The Principal Architect AI SME in Delivery leads architecture for large Azure programs that...  ...qualification and estimating using playbooks, benchmarks, and templates so the organization...  .... Experience leading hybrid and remote teams and working with multiple stakeholders... 
    Remote work
    Full time
    Temporary work
    Visa sponsorship
    Work visa

    3Cloud

    Remote
    16 days ago
  • Mercor is hiring PhD and Master's scientists to author AI evaluation tasks for a new benchmark in scientific computing. You will create original, executable research problems that today’s frontier models cannot solve, sourced from papers, datasets, or your own scenarios... 
    Part time

    Mercor

    San Diego, CA
    1 day ago
  • Vals AI, Inc. in San Francisco is seeking exceptional researchers and research engineers to design and build the next generation of AI benchmarks that evaluate real-world LLM capabilities. You will lead development of novel benchmarks shaping how foundation models are... 

    Vals AI, Inc.

    San Francisco, CA
    4 days ago
  •  ...design a central evaluation framework and reusable pipelines that span models and teams. You will implement evaluation pipelines, benchmark suites, baselines, and visualization tools to turn results into shared, actionable understanding. Strong SWE and statistics are essential... 

    Causal

    San Francisco, CA
    4 days ago
  • Full Time| Monday-Friday 8:00AM-4:30PM| Weekly Earned Wage Access is an option for this position. 60/40 onsite/remote Job Purpose or Goals: The AI architect defines the technical architecture for enterprise‑level AI systems and their integrations, ensuring they are designed... 
    Remote work
    Full time
    Monday to Friday

    Choctaw Nation

    Durant, OK
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote AI Philosophy Benchmark Architect. Be the first to apply!