Stop applying to jobs that are already dead.
Every listing verified, aged honestly, expired when filled.

All listings

Weekday AI via Workable

AI Evaluation Specialist

Level not stated $70/hr United States
still open verified 16h ago posted 68d ago checked just now
Apply at apply.workable.com

This is the employer's own posting, not a copy on a job board.

What we know

Is it still open?

Confirmed still open

Last checked 16h ago — checked against the employer's own applicant tracking system, which is the company answering directly.

We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.

Check this listing's status as JSON

How old is it?

Posted 68d ago

The date the source published, not the day we noticed it (2026-07-09). Last seen at its source just now.

Is it remote?

Marked remote on the employer's board

Their board carries a remote setting on this posting — a field they filled in, not wording we read. The location field names somewhere specific, which is usually where the team or the entity sits.

Who may apply?

United States

The description states no restriction of its own. This is the source's own tag.

Pay

$70/hr

Read out of the job description by us, not from a structured field. Shown in the posting's own currency and period; we never convert.

Skills named in the ad

EditorialLLM

Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.

Carried by 1 source

The listing

This role is for one of our clients

Compensation: $70 per hour

Join a cutting-edge AI research initiative focused on improving the quality, accuracy, and reasoning capabilities of next-generation artificial intelligence systems. We are seeking analytical professionals with exceptional critical thinking and communication skills to evaluate AI-generated responses across a variety of topics.

In this role, you will assess AI outputs, identify strengths and weaknesses in reasoning, and provide structured, evidence-based feedback that helps improve model performance. This opportunity is ideal for individuals who enjoy careful analysis, attention to detail, and working independently on intellectually challenging tasks.

This is a fully remote, contract-based opportunity with flexible working hours.

Requirements

Key Responsibilities

Evaluate AI Responses

  • Review AI-generated content for accuracy, logical reasoning, completeness, and clarity.
  • Identify factual errors, reasoning gaps, inconsistencies, and unsupported conclusions.
  • Assess responses using structured evaluation frameworks and detailed quality guidelines.

Provide High-Quality Feedback

  • Write clear, concise, and evidence-based rationales explaining evaluation decisions.
  • Highlight both strengths and areas for improvement in AI-generated outputs.
  • Apply consistent judgment across a wide range of evaluation tasks.

Maintain Evaluation Quality

  • Follow detailed project instructions and standardized assessment criteria.
  • Ensure evaluations are objective, accurate, and reproducible.
  • Complete assignments independently while maintaining high quality standards.

Required Qualifications

  • Bachelor's degree from a globally recognized university (top-ranked institutions preferred).
  • Excellent analytical thinking and problem-solving abilities.
  • Strong written communication skills with the ability to explain complex reasoning clearly and precisely.
  • Exceptional critical reading skills, including the ability to identify:
    • Nuanced arguments
    • Implicit meaning
    • Logical inconsistencies
    • Missing context
    • Weak or unsupported reasoning
  • Strong attention to detail and ability to consistently apply structured evaluation guidelines.
  • Ability to work independently and manage assigned tasks efficiently.
  • Native-level English fluency.

Preferred Qualifications

  • Experience in content evaluation, research, quality assurance, editing, or analytical review.
  • Familiarity with artificial intelligence, large language models, or AI evaluation methodologies.
  • Experience working with structured annotation or assessment frameworks.
  • Ability to produce thoughtful, objective, and well-supported written evaluations under defined quality standards.

Engagement Details

  • Independent contractor engagement.
  • Fully remote with flexible working hours.
  • Work completed on your own schedule.
  • Project duration may be extended, shortened, or concluded based on business needs and performance.
  • Weekly payments processed through supported payment platforms.

Why Join

  • Contribute to the development of next-generation AI technologies.
  • Help improve the reasoning, accuracy, and reliability of advanced AI systems.
  • Work on intellectually engaging projects with real-world impact.
  • Collaborate indirectly with leading AI researchers through high-quality evaluation work.

Equal Opportunity Statement

All qualified applicants will be considered without regard to legally protected characteristics. Reasonable accommodations are available upon request.

Contract Information

  • Independent contractor engagement.
  • Fully remote work completed on your own schedule.
  • Weekly payments are processed based on approved work completed.
  • Work does not involve access to confidential or proprietary information from any employer, client, or institution.
  • Please note that visa sponsorship is not available for this opportunity.
Apply at apply.workable.com