Stop applying to remote jobs that are already dead.
Every listing shows the evidence: when we last checked it, how, and when it was posted and closed.

All listings

Yonder Media Mobile via Bamboohr

AI Training and Prompt Engineering Specialist

Poland Level not stated
still open verified 11h ago posted 47d ago seen 3h ago
Apply at yomobile.bamboohr.com

This is the employer's own posting, not a copy on a job board.

What we know

Is it still open?

Confirmed still open

Last checked 11h ago — checked against the employer's own applicant tracking system, which is the company answering directly.

We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.

Check this listing's status as JSON

How old is it?

Posted 47d ago

The date the source published, not the day we noticed it (2026-08-18). Last seen at its source 3h ago.

Is it remote?

Marked remote on the employer's board

Their board carries a remote setting on this posting — a field they filled in, not wording we read. The location field names somewhere specific, which is usually where the team or the entity sits.

Who may apply?

Poland

The description states no restriction of its own. This is the source's own tag.

Pay not stated

Similar roles pay PLN 302.4k–399.8k/yr

Middle 50% of 40 listings that do state pay — Engineering · all levels · Poland · PLN/year. This employer has published no salary; this is what comparable listings we hold disclose, never converted between currencies or periods. How this is calculated.

Skills named in the ad

Code ReviewEditorialGitLLMLangChainPython

Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.

Carried by 1 source

The listing

ABOUT THE ROLE 

House of YO is a connectivity and distribution business with compute layered on top. YOlanda is our AI concierge.

On V3 of the YO platform, YOlanda is the interface. Users talk to her instead of tapping through menus. She answers questions about plans, data, top-ups and YOYO$ balances, and she takes the action on the user’s account. When she is wrong, the user is stuck and the product has failed.

Her knowledge is the YO ecosystem and nothing else. When a question falls outside it she hands off to YOnC, our compute cooperative, which routes it to a specialist model. Knowing where that line sits, and holding it, is central to the job.

She works to one golden rule: within five responses she has put a product or service in front of the user. Not five responses to be helpful. Five to arrive somewhere real.

This role makes her right. You write the prompts that govern how she thinks. You build the evaluations that prove she works. You write the training data that fixes her when she does not. You ship the code that runs all of it.

YOlanda runs on our own LangChain and LangGraph build. You will open pull requests against that graph alongside the two engineers who own it.

You will not manage anyone. You will be in the model every day.

Who this is for

You break things on purpose and then explain exactly why they broke. You can read a bad answer and say in one sentence what is wrong with it. You would rather run fifty tests than win one argument about which prompt is better. You have opinions about tone and you can defend them. You write well enough that people notice.

You are relentless. The first version is never the one you ship.


Location:
Remote, global
Work format:
Remote, full time

WHAT YOU'LL DO

  • Own YOlanda's system prompt — persona, tone, safety rules, what she refuses — and the prompts behind every user action: plan changes, top-ups, balance queries, offers, support.
  • Own the boundary: decide what's inside YOlanda's frame of reference and what hands off to YOnC, and make the handoff seamless — same conversation, same voice.
  • Build the few-shot library and manage context, so she never opens on a blank prompt or repeats a question she already answered.
  • Build the evaluation set: every task gets test cases, a pass condition and a score you can track release over release. Nothing ships if the score went down.
  • Rank and critique outputs (RLHF), rewrite weak answers into reference/fine-tuning data, and fact-check plan prices, balances and YOYO$ maths — the hallucination target on money is zero.
  • Red team her: adversarial prompts, prompt injection, jailbreaks and abuse of the top-up and rewards flows.
  • Write Python to run and score evaluations at volume, work in the API (streaming, function calling, structured output), and keep prompts versioned in git.


WHAT YOU NEED

  • 2+ years building with large language models in production. Using a chat window every day is not this.
  • Python you can ship: scripts, API clients, data handling, tests.
  • LangChain and LangGraph — you can read a graph you didn't write and change it without breaking it.
  • Retrieval in production: chunking, embeddings, reranking, grounding, and knowing when retrieval is the wrong tool.
  • Exceptional written English, and bilingual Spanish and English to the same standard — YOlanda has one voice that has to survive the crossing.
  • Method: change one variable, test it, record it, move on.

Nice to have: fine-tuning (SFT, DPO, LoRA); evaluation frameworks (promptfoo, Braintrust, LangSmith, DeepEval); mobile, telecoms or fintech; editorial or linguistics training.

WHAT WE OFFER

  • Top-notch products disrupting the world of the traditional media-services industry
  • A brilliant, highly collaborative team with a shared drive to achieve goals
  • Flexible working hours
  • Corporate equipment for work
  • Competitive salary
  • Real opportunity for personal and professional growth
Apply at yomobile.bamboohr.com