STEM Researchers - Benchmark & First-Author Research
This is the employer's own posting, not a copy on a job board.
What we know
Is it still open?
Confirmed still open
Last checked 22h ago — checked against the employer's own applicant tracking system, which is the company answering directly.
We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.
How old is it?
Posted 10d ago
The date the source published, not the day we noticed it (2026-09-05). Last seen at its source just now.
Is it remote?
Marked remote on the employer's board
Their board carries a remote setting on this posting — a field they filled in, not wording we read. The location field names somewhere specific, which is usually where the team or the entity sits.
Who may apply?
Brazil, India, United Kingdom, Singapore, United States, Australia
The description states no restriction of its own. This is the source's own tag.
Pay
$30–80/hr
Read out of the job description by us, not from a structured field. Shown in the posting's own currency and period; we never convert.
Carried by 1 source
-
workable employer's own board first seen 10d ago · last seen just now
- posted 2026-09-14
The listing
About Gramian
Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.
About the Role
We’re working with a highly specialized AI research lab building new benchmarks for how frontier AI systems perform real scientific research. They are looking for PhD students and postdocs across STEM fields to help reproduce papers, validate AI-generated research, and define what “good research” should look like for AI systems.
We are looking for one Fellow per STEM field to help create a new benchmark for evaluating how well frontier AI models perform real scientific research.
The standout part of this opportunity: the benchmark will be released publicly with a research paper and open evaluation set, and Fellows will be named lead / first authors.
CONTRACT: Research fellowship / contractor
LOCATIONS: Remote, global
COMMITMENT: Flexible hours
COMPENSATION: $30–$80/hour
PROCESS: Application → research review → interview
Fields: Biology, Medicine, Neuroscience, Materials Science, Chemistry, Physics, Mechanical Engineering, Chemical Engineering.
Responsibilities
- Validate AI-generated paper reproductions, simulations, and research results.
- Map key research areas and taxonomy within your field.
- Define what “good research” looks like for AI systems.
- Create benchmark tasks and evaluation standards for frontier models.
- Contribute directly to the benchmark paper and public eval release.
Requirements
- Current PhD student or postdoc in a relevant STEM field.
- Based at a strong research university or institute.
- At least one published paper in your field.
- Able to critically review research and identify methodological or technical errors.
- Currently using AI for Science in your own research.
- Significant hands-on use of tools such as Claude Code, Codex, or similar AI agents.
Benefits
- Be a lead / first author on a public benchmark paper.
- Help create one of the first research benchmarks in your scientific field.
- Work directly with a niche AI research lab and frontier AI researchers.
- Flexible remote work designed around academic commitments.