Stop applying to remote jobs that are already dead.
Every listing shows the evidence: when we last checked it, how, and when it was posted and closed.

All listings

Linda Werner & Associates via Greenhouse

Content Specialist III | AI Evaluation & Prompting

United States Level not stated
still open verified 1d ago posted 7d ago seen 1h ago
Apply at job-boards.greenhouse.io

This is the employer's own posting, not a copy on a job board.

What we know

Is it still open?

Confirmed still open

Last checked 1d ago — checked against the employer's own applicant tracking system, which is the company answering directly.

We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.

Check this listing's status as JSON

How old is it?

Posted 7d ago

The date the source published, not the day we noticed it (2026-10-03). Last seen at its source 1h ago.

We have tracked this listing since 9 Oct 2026 (1 days). The employer's own board has carried it every time we have read it, most recently 1 hour ago.

Is it remote?

The listing says yes

The location field doesn't say remote, so our assessment is based on the title or the description. Read the listing before applying.

Who may apply?

United States

The description agrees: it names United States.

What the ad says
…Location : United States (Remote)…

Pay not stated

Similar roles pay $83.8k–162.8k/yr

Middle 50% of 227 listings that do state pay — Content & Media · all levels · United States · USD/year. This employer has published no salary; this is what comparable listings we hold disclose, never converted between currencies or periods. How this is calculated.

Skills named in the ad

EditorialLLM

Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.

Carried by 1 source

The listing

We are seeking an experienced Content Specialist III to help evaluate, refine, and improve advanced AI models and products.

In this role, you will work directly with evolving AI systems, testing how they respond across a wide range of topics and conversation types, identifying where they succeed or fall short, and helping shape the prompts, quality standards, and evaluation frameworks that influence how they communicate and behave.

This is a highly hands on opportunity for someone who brings together strong writing and content expertise, curiosity about AI, thoughtful judgment, and a sharp eye for quality. Your work will directly contribute to improving real world AI product experiences and how these systems perform for users.

What You Will Do

• Test new AI model versions across a variety of topics, use cases, and conversation types

• Evaluate model responses against established rubrics, guidelines, and quality standards

• Identify and document specific examples of successful and unsuccessful model behavior

• Write and refine system prompts that help shape model personality, tone, and behavior

• Assess whether evaluation rubrics and quality standards effectively measure model performance

• Perform quality reviews of both human and agent based evaluations to ensure accuracy and compliance

• Conduct hands on experiments with AI models and products

• Investigate model failures and document findings clearly

• Apply detailed instructions and evaluation criteria consistently across a high volume of work

• Adapt quickly as product priorities, model behavior, and evaluation needs evolve

Top Skills

• AI Model Evaluation

• Prompt Writing and Refinement

• Writing and Editing

• Rubric Based Quality Assessment

• Fact Checking and Analytical Judgment

Qualifications

• 5+ years of experience in writing, editing, journalism, production, linguistics, STEM, coding, policy, or another relevant subject matter field

• 1+ year of hands on AI experience preferred, including prompting, annotation, evaluation, or red teaming

• Bachelor’s degree or equivalent experience

• Deep expertise in at least one subject area with the ability to evaluate content as a subject matter expert

• Strong judgment and attention to detail with the ability to consistently apply detailed rubrics, verify information, and identify quality issues

Preferred Experience

• Experience evaluating generative AI or large language model outputs

• Strong fact checking and research skills

• Ability to verify claims against original sources

• Experience conducting detailed failure investigations

• Clear written communication and documentation skills

• Ability to flag questions and blockers early

• Comfortable shifting priorities as product needs change

This team operates in a fast moving environment where priorities can change from day to day. Successful candidates will be comfortable balancing quality, speed, volume, and independent judgment while following detailed instructions.

You should enjoy getting into the details, identifying subtle differences in AI responses, determining why something works or fails, and translating those observations into clear and actionable feedback.

Location: United States (Remote)

Role type: Contract 6 Month Position

Expected hours: 40 per week

Benefits:

  • Dental insurance
  • Health insurance
  • Health savings account
  • Life insurance
  • Paid time off
  • Retirement plan
  • Vision insurance

Schedule:

  • 8 hour shift
  • Monday to Friday

Application Question(s):

  • Do you or will you in the future require any sponsorship to work in the US?

Language:

  • English  (Required)

 

Apply at job-boards.greenhouse.io