Average salary: $45,180 /yearly

More stats
Get new jobs by email
  • $16.6 per hour

    Host/Hostess Position Four Corners is a leading, Chicago-based hospitality group that owns and operates unique establishments, each thoughtfully created to offer an exceptional social experience, creative menus, and superior service. We started with a neighborhood bar...
    Suggested
    Casual work

    Four Corners

    Chicago, IL
    5 days ago
  • $80 - $100 per hour

     ...unable to provide offer letters or employment verification for this role. What You'll Be Doing Design and build the coding benchmarks and evaluation pipelines used to test frontier AI models on real software engineering work: Design coding benchmarks that... 
    Suggested
    Remote job
    Full time
    Contract work
    For contractors

    G2i

    Miami, FL
    1 day ago
  • Dynamo AI is seeking a candidate to lead LLM evaluation and benchmarking in San Francisco, California. You will generate high-quality data and develop innovative methods for assessing the safety and helpfulness of LLMs. The role requires domain knowledge in evaluation... 
    Suggested

    Capitolis

    San Francisco, CA
    4 days ago
  • $150k - $250k

     ...envelope of AI utilization in enterprise. This requires creative researchers who don’t just want to drive incremental improvements on benchmarks or optimize an existing process but instead are looking to creatively redefine how software is used. Our researchers come from... 
    Suggested
    Work at office
    3 days per week

    Distyl

    New York, NY
    5 days ago
  • $60 - $90 per hour

     ...seeks a Bioinformatician for an hourly contract position to analyze sequencing data and test AI models. The role involves designing benchmark tasks, creating workflows from real datasets, and using command-line bioinformatics tools. Candidates must be a graduate student... 
    Suggested
    Remote job
    Hourly pay
    Contract work
    Flexible hours

    Crossing Hurdles

    New York, NY
    4 days ago
  • $206k - $333k

     ...March 2025. Learn more at About this role We're looking for a Principal Engineer to be the technical lead of CoreWeave's Benchmarking & Performance team. You will be responsible for our planet-scale performance data warehouse: Ingesting, storing, transforming and... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    13 days ago
  • LILT (Production) is looking for experienced professionals in Robotics to join a cutting-edge AI benchmarking project. As a Subject Matter Expert, you will create and review high-quality natural sciences scenarios that evaluate AI systems contextually. Your role involves... 
    Suggested

    LILT (Production)

    New York, NY
    1 day ago
  •  ...Job Description Job Description Skills Required: Deep Knowledge in: o CIS Benchmarks and controls. o Relevant OS – Primarily Windows 11/Server 2022, Microsoft 365 Edge/Chrome/Firefox Browsers, Microsoft SQL Server also preferred o Active Directory Group Policy... 
    Suggested
    Work experience placement

    Tier4 Group

    West Virginia
    6 days ago
  • $50 per hour

    A leading AI research accelerator is seeking remote PhD candidates in Mathematics or related fields to design math problems and evaluate AI performance. The role involves collaboration with researchers and offers flexible hours at a pay rate of $50+/hour. Ideal candidates...
    Suggested
    Remote job
    Hourly pay
    Flexible hours

    Turing

    San Francisco, CA
    4 days ago
  •  ...team in Houston. This full-time role involves maintaining the quality of indexes and engaging market participants to enhance GX benchmarks. Candidates should possess a BA/BS, at least three years of experience in crude oil or refined products, and exceptional numerical... 
    Suggested
    Full time

    General Index Limited

    Houston, TX
    4 days ago
  • $50 per hour

    A leading research accelerator for AI labs is seeking skilled PhDs in Mathematics or related fields for a fully remote position. You will help fine-tune AI models by designing advanced math problems and evaluating AI performance. The role offers a pay rate of $50+/hour,...
    Suggested
    Remote job
    Flexible hours

    Turing

    Chicago, IL
    4 days ago
  •  ...Benchmark Dataset Data Specialist is a remote evaluation track for reviewing benchmark dataset data evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured feedback the modeling... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    1 day ago
  •  ...AI Benchmark and Task Designer is a remote review track for evaluating AI outputs across ai benchmark and task designer creative review workflows. Reviewers grade craft, constraints, and tooling adherence; flag production-readiness issues; and document the corrected approach... 
    Suggested
    Remote job
    Hourly pay
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    20 hours ago
  • $37 per hour

    Kforce is immediately adding a full-time AI Systems Benchmarking & Product Engineer in support of our nationally recognized, Consumer and Commercial Electronic R&D client in Fort Collins CO. Summary: Our client is seeking a contractor to support AI performance benchmarking... 
    Suggested
    Full time
    Contract work
    For contractors
    Local area
    Immediate start
    Fort Collins, CO
    14 days ago
  • $50 per hour

    A leading AI research accelerator is seeking PhDs in Mathematics or related fields for remote contract roles. You will work on cutting-edge AI projects, designing math problems and evaluating AI performance. Strong analytical skills and clear communication are essential...
    Suggested
    Remote job
    Contract work
    Flexible hours

    Turing

    Toronto, OH
    4 days ago
  • $182k - $242k

     ...company (Nasdaq: CRWV) in March 2025. Learn more at About this role We're looking for a Senior Engineer for CoreWeave's Benchmarking & Performance team. You will have an integral part in our planet-scale performance data warehouse: Ingesting, storing, transforming... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    21 days ago
  • LILT (Production) is seeking an experienced Subject Matter Expert for a healthcare AI benchmarking project in the United States. You will help design real-world healthcare scenarios and evaluate responses for correctness and quality. Your expertise will ensure cultural... 

    LILT (Production)

    New York, NY
    5 days ago
  • $50 per hour

     ...performance and collaborating with researchers. Responsibilities include developing solutions, evaluating AI outputs, and refining benchmarks. The pay rate is $50+/hour, depending on expertise. This flexibly scheduled position offers the chance to work on cutting-edge AI... 
    Remote job
    Contract work

    Turing

    San Francisco, CA
    5 days ago
  • $50 per hour

     ...Engineering, or related fields. You will design advanced chemistry problems to evaluate AI performance and collaborate with researchers on benchmarks. This role offers flexible hours and a pay rate of over $50/hour. Ideal candidates will have strong reasoning and problem-solving... 
    Remote job
    Hourly pay
    Flexible hours

    Turing

    New York, NY
    5 days ago
  •  ...continuously increasing their personal and technical development. RLE International has an immediate position for a commodity benchmarking analyst who will be located on-site in Allen Park, MI full-time Job Description: This position supports the Global Total... 
    Full time
    Work at office
    Local area
    Immediate start
    Worldwide

    RLE International

    Allen Park, MI
    12 days ago
  • A top AI research partner is seeking a Biomedical Informatics Subject Matter Expert to develop rigorous evaluation questions for advanced AI models. This role involves designing complex bioinformatics questions, ensuring clarity and validity in assessment materials. Ideal...
    Remote job

    Turing

    Denver, CO
    4 days ago
  • A leading AI research accelerator is seeking a highly skilled Biomedical Informatics Subject Matter Expert to develop rigorous evaluation questions for advanced AI models. This role requires deep technical expertise in biomedical data processing and strong coding skills...
    Remote job
    Contract work

    Turing

    Los Angeles, CA
    4 days ago
  • $100k - $130k

     ...Job Description Job Description Client: Benchmark Space Systems Position: Electrical Engineer Reports to: VP of Electric Propulsion Website: Location: Burlington, VT Salary: $100,000 - $130,000/year depending on experience Position Type: Full-time... 
    Full time
    Work at office
    Flexible hours

    Gallagher, Flynn & Company

    Burlington, VT
    28 days ago
  •  ...authority across large-scale infrastructure or real estate portfolios, combining hands-on project delivery with leadership of internal benchmarking and cost intelligence initiatives. You will orchestrate remote delivery teams, leveraging global program data to establish cost... 
    Contract work
    For contractors
    For subcontractor
    Remote work
    Flexible hours

    Turner & Townsend

    Concord, NC
    13 days ago
  • A leading AI research accelerator is seeking a Biomedical Informatics Subject Matter Expert to design challenging evaluation questions that assess advanced AI models. You will develop coding questions and detailed rubrics to measure outcomes in bioinformatics domains. This...
    Remote job
    Contract work

    Turing

    Boston, MA
    4 days ago
  • $177k - $237k

     ...enterprise customers. As a Technical Program Manager, you will lead complex, cross-functional programs across Performance & Benchmarking within our AI/ML Platform Services organization. This team is responsible for ensuring CoreWeave's infrastructure is performant,... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    CoreWeave

    San Francisco, CA
    4 days ago
  •  ...Join a pioneering AI initiative focused on building the next generation of evaluation benchmarks for frontier AI models. We are seeking experienced QA and Test Engineers to ensure every benchmark is reliable, reproducible, and accurately measures real AI capabilities... 
    Full time
    Contract work
    For contractors
    Remote work
    Flexible hours

    Weekday

    Remote
    4 days ago
  • Job Description Job Description Description: The Bartender is responsible for providing the highest quality of service to guests in an efficient and courteous manner. The Restaurant Server will take food and beverage orders, retrieve and serve alcoholic, non-alcoholic...
    Full time
    Temporary work
    Part time
    Local area
    Shift work
    Afternoon shift

    Bentley Legacy Group

    Bozeman, MT
    13 days ago
  • $50 per hour

    A leading research accelerator for AI is seeking remote PhD graduates in Biology, Biotechnology, or Biochemistry to assist in fine-tuning AI models. Responsibilities include designing biology questions to evaluate AI, developing logical solutions, and collaborating with...
    Remote job
    Flexible hours

    Turing

    Los Angeles, CA
    2 days ago
  • $75k - $150k

     ...Job Description Job Description Position: Quantum Benchmarking Initiative Researcher Beware of fraudulent job offers and postings! Technergetics will never extend an offer of employment without a thorough interview process involving face to face interviews... 
    Full time
    Remote work
    Flexible hours

    Technergetics

    Utica, NY
    13 days ago