Cybersecurity Expert for AI Model Evaluation
$65 per hourSaidGig
Cybersecurity professionals apply offensive and defensive security expertise to design domain-specific prompts and evaluate large language model outputs for AI research projects, improving model behavior, safety, and relevance in security-related scenarios.
Key Responsibilities- Develop domain-specific prompts, test cases, and evaluation criteria for LLMs focused on security use cases.
- Evaluate and annotate LLM responses for technical accuracy, safety, and potential misuse.
- Identify unsafe model behaviors, incorrect guidance, and vulnerabilities exposed by model outputs.
- Provide clear, actionable feedback and examples to guide model improvement.
- Learn new skills while contributing across security and adjacent domains as project needs arise.
- At least 3 years of professional experience in security engineering or offensive security in roles such as Security Engineer, Penetration Tester, Red Team Engineer, Malware Analyst, Threat Hunter, Threat Intelligence Analyst, Offensive Security Engineer, Vulnerability Researcher, or Security Consultant.
- Hands-on experience with one or more of the following tools: Burp Suite, Nmap, Wireshark, Metasploit Framework, BloodHound, Impacket, Nessus, or Hashcat.
- Ability to work independently in a remote, asynchronous setting and to document findings clearly in writing.
- Remote, asynchronous work from any location.
- Part-time engagement with flexible hours, and no minimum weekly time commitment.
- The program runs year-round, however placement into specific projects depends on project availability by domain.
Up to $65/hr (depending on the project)
Eligibility- F-1 students with CPT or OPT may be eligible, subject to your school’s rules. Confirm eligibility with your Designated School Official. If your school requires a CPT course, this program may not satisfy that requirement. STEM OPT is not supported.
- Create an account and complete your profile on the program site.
- Complete the required identity verification.
- Apply to or join projects that match your expertise and complete project-specific onboarding.
- Begin contributing to project work and receive payment for your participation.
- ...Prolific is seeking Biology Experts and Life Science Professionals to join an expert network that evaluates and trains AI models. This role involves reviewing AI-generated scientific content for accuracy and validation, requiring candidates with a BS, MS, or PhD in relevant...SuggestedRemote workFlexible hours
- ...Role Overview Drive the creation and evaluation of challenging STEM problems used to fine-tune and benchmark large language models. You will design multi-step physics and math problems... ..., the company accelerates frontier AI research and helps enterprises deploy reliable...SuggestedContract workFor contractorsFreelanceRemote work
$60 - $80 per hour
...building foundational large language models by applying deep insurance domain expertise to create, evaluate, and refine training data. You... ...claims practice. Evaluate AI model outputs against... ...Collaborate with other subject-matter experts to ensure consistency and accuracy...SuggestedHourly payWeekday work- ...improve and validate large language models, by creating realistic retail... ...LLC and placed with a leading AI lab. Key Responsibilities... ...in real retail practice. Evaluate AI model outputs against... ...Collaborate with other subject matter experts to ensure consistency and...SuggestedHourly payWeekday work
$55 per hour
...Biology experts contribute their scientific knowledge to AI research projects by helping improve how large language models understand and explain specialized biological concepts. Role Overview... ...specific prompts for AI systems. Evaluate large language model responses for...SuggestedHourly payPart timeRemote workFlexible hours$60 per hour
...Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...Hourly payRemote workWork from homeFlexible hours$60 per hour
Prolific is seeking Biology Experts and Life Science Professionals to join their Expert Network to evaluate AI-generated science. This role allows you to work from home with... ...pay rate of up to $60 per hour for reviewing model responses, validating technical claims, and critiquing...Remote jobHourly payWork from homeFlexible hours- Prolific is seeking Biology Experts and Life Science Professionals in Jacksonville, Florida, to join our Expert Network for evaluating AI-generated science models. Candidates should hold a BS, MS, or PhD in relevant fields and have experience in research or academia. Responsibilities...Remote jobHourly payFlexible hours
$60 per hour
Prolific, located in Arizona, is seeking Biology Experts and Life Science Professionals to join their Expert Network. This role involves evaluating and training AI models with real scientific expertise. Successful candidates will review AI-generated content for accuracy...Hourly pay$60 per hour
Prolific is seeking Chemistry Experts and Chemical Engineers to join their Expert Network. Participants will evaluate AI-generated chemistry through tasks that assess factual accuracy... ..., enabling cutting-edge advancements in AI models. The position requires a strong educational...Hourly pay- ...Physics Specialist to contribute deep scientific expertise to AI model evaluation. You will craft and assess challenging physics problems,... ...electromagnetism, use adversarial prompting to surface errors, and provide expert critique of AI responses while working with project #J-1880...Remote job
$50 per hour
...step solutions to help fine-tune large language models such as ChatGPT. You will probe model limitations, contribute evaluation benchmarks across physics curricula, and work... .... Projects also provide experience applying AI to improve analytical workflows. About the...Contract workFor contractorsFreelanceRemote work$17 - $54 per hour
...Music & Lyrics Expert - French | Remote AI Model Evaluation is a remote evaluation track for reviewing french generalist evaluation prompts and responses against AuraOne's quality rubric. Reviewers compare paired outputs, label edge cases, and write the kind of structured...Remote jobFor contractors10 hours per week$55 per hour
...Role Overview Biology experts apply their domain knowledge to design domain-specific prompts, evaluate large language model outputs, and guide AI research across biological subfields. This role supports AI model evaluation and improvement by providing specialist review...Part timeRemote workFlexible hours$20 - $40 per hour
...expertise to improve how next-generation AI systems learn and reason by analyzing... ...becomes high-quality training data and evaluations for AI models. The role supports a unique customer project... ..., technical clarifications, and expert-level commentary on complex engineering...Hourly payFor contractorsRemote work$80 per hour
...on retail freight and distribution experience to craft expert-level training content and evaluate AI-generated responses against real-world transportation... ...Provide structured, written feedback that helps improve AI model performance and response quality. Work...Part timeRemote work$80 - $100 per hour
...to help improve next-generation AI systems through practical technical input, evaluations, and high-quality training materials... ...Key Responsibilities Provide expert analysis, feedback, and practical... ...data challenges for AI model development. Document advanced...Hourly payFor contractorsRemote work$50 per hour
...Specialist roles apply advanced chemistry knowledge to create and assess domain-specific prompts and to evaluate large language model responses, helping improve AI performance on chemistry problems and explanations. Key Responsibilities Develop domain-specific...Part timeH1bRemote workVisa sponsorship10 hours per weekFlexible hours$50 - $101 per hour
...knowledge to help train next-generation AI systems. In this remote contractor role supporting... ...programs and nutrition guidance so models learn accurate, practical, safe fitness... ...general wellness guidance. Create and evaluate sample fitness programs and nutrition plans...Hourly payFor contractorsRemote work$20 - $75 per hour
...Role Overview Provide expert telecommunications guidance to help train and evaluate next-generation AI systems. You will convert real-world telecom knowledge into high-quality... ...data, evaluations, and feedback that improve model learning, reasoning, and performance. micro1...Hourly payFor contractorsRemote work- ...definition of excellent enterprise selling for a cutting-edge generative AI team by auditing multi-step sales workflows, producing end-to-end expert examples, and shaping the evaluation standards the model learns from. You will apply deep, practical selling experience to...Hourly payFull timeWork at officeRemote work
$40 per hour
A healthcare technology company in the United States seeks medical experts to evaluate AI chatbots. Responsibilities include presenting healthcare problems to AI and assessing their responses for correctness. Candidates must be fluent in English and possess a current or...Hourly payRemote work$208k - $300k
...Machine Learning Engineer - Model Evaluations, Public Sector The Public Sector ML team at Scale deploys advanced AI systems—including LLMs, agentic models, and multimodal... ...with operations teams and subject matter experts to produce high-quality evaluation datasets...Full time$40 per hour
A biotechnology company is seeking a Biotechnology R&D Scientist to train AI models and evaluate their outputs. This role involves measuring the progress of AI chatbots with complex biology questions and ensuring their performance and correctness. Candidates should have...Hourly payRemote workFlexible hours$224k - $356.5k
...people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts... ...computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting the...Full time$40 per hour
A forward-thinking AI development firm seeks experienced quantitative professionals to evaluate AI-generated work, applying their skills in statistical analysis, predictive modeling, and technical writing. This fully remote opportunity offers a flexible schedule and projects...Hourly payRemote workFlexible hours$40 per hour
A leading AI development company is seeking experienced quantitative professionals to evaluate AI-generated analyses and design quantitative problems for AI training. This fully remote role offers flexibility in project selection and scheduling, with competitive pay starting...Hourly payRemote work$40 per hour
...A leading AI development firm is looking for experienced quantitative professionals to evaluate AI-generated work and design problems for AI training. This fully remote position allows for a flexible schedule, offering competitive pay starting at $40+ per hour. Ideal candidates...Hourly payRemote workFlexible hours- ...legal practice experience to improve how frontier AI models perform real legal work. In this position you will evaluate model outputs, create high-quality instruction... ...please consider the corresponding Legal Domain Expert listings that match those seniority levels and rates...Hourly payFull timeFreelanceInternshipLive inRelocationRelocation package
$40 per hour
A leading AI training company is seeking a Biotechnology R&D Scientist to evaluate and improve AI models by posing complex biological questions. This remote position allows candidates to choose projects and work on their own schedule, with rates starting at $40+ per hour...Hourly payRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Cybersecurity Expert for AI Model Evaluation. Be the first to apply!
- fruit expert United States
- subject matter expert United States
- expert data analyst United States
- guest service support expert United States
- expert systems engineer United States
- technology expert United States
- fulfillment expert United States
- subject matter expert senior United States
- subject matter expert work from home United States
- sql expert United States




