AI QA Engineer
This is the employer's own posting, not a copy on a job board.
What we know
Is it still open?
Confirmed still open
Last checked 1d ago — checked against the employer's own applicant tracking system, which is the company answering directly.
We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.
How old is it?
Posted 7d ago
The date the source published, not the day we noticed it (2026-09-07). Last seen at its source just now.
Is it remote?
The listing says yes
The location field doesn't say remote, so our assessment is based on the title or the description. Read the listing before applying.
Who may apply?
Europe, Ukraine
The description states no restriction of its own. This is the source's own tag.
Pay not stated
Similar roles pay €75k–114.5k/yr
Middle 50% of 80 listings that do state pay — Engineering · all levels · Europe · EUR/year. This employer has published no salary; this is what comparable listings we hold disclose, never converted between currencies or periods. How this is calculated.
Skills named in the ad
Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.
Carried by 1 source
-
greenhouse employer's own board first seen 7d ago · last seen just now
The listing
Our value to you:
- Flexible hours and remote-first mode,
- Competitive compensation,
- Complete Hardware/Software setup – anything you need for work,
- Open-door culture, transparent communication, and top management at a handshake distance,
- Health insurance, vacation, sick leaves, holidays, paid maternity/paternity leave,
- Access to our learning & development center: workshops, webinars, training platform, and edutainment events,
- Virtual team buildings and social activities.
If you feel like you’re the perfect match for this role, drop us your CV!
There are no limitations, no barriers when the right people are on your way — apply for the vacancy and succeed with us!
Innovecs is an equal opportunity employer. All hiring decisions are based on professional qualifications, skills, and experience. We are committed to a transparent, merit-based recruitment process that prevents discrimination and ensures equal opportunities for all candidates. Reasonable accommodations are available upon request throughout the recruitment process to support accessibility and inclusion.
Project description
Our client is building a next-generation Warehouse Management System (WMS) that combines traditional supply chain execution with advanced Artificial Intelligence and Agentic AI capabilities. The platform is evolving into an intelligent operational assistant capable of decision support, warehouse optimisation, exception management, inventory analysis, workforce productivity recommendations, and conversational interactions.
The solution leverages LLMs, AI agents, and enterprise integrations to automate warehouse operations and improve decision-making across fulfillment, inventory, labor management, and logistics processes.
As part of this transformation, quality assurance plays a critical role in validating both traditional enterprise software functionality and AI-driven behaviours. The project requires advanced testing capabilities spanning deterministic business workflows, complex system integrations, AI agents, LLM-powered features, and autonomous decision-making processes
Technology Stack
AI & Agentic Technologies: LLMs, AI Agents and Agent Workflows, RAG, MCP integrations;
Agent orchestration frameworks: LangGraph, CrewAI, AutoGen;
Prompt engineering and evaluation;
Vector databases;
AI Evaluation & Quality Frameworks: DeepEval, LangSmith, Braintrust, etc.
Test Automation & Quality Engineering: Playwright, Selenium, Cypress, PyTest, JUnit, Postman, Rest Assured;
Integrations: REST APIs, Event-driven integrations, Databases (SQL/NoSQL), Enterprise messaging platforms, Warehouse system integrations (ERP/TMS/WCS integrations);
CI/CD & DevOps: GitHub Actions / Azure DevOps, Jenkins, Docker
AI Observability & Monitoring: MLflow, Prometheus/Grafana
Cloud Platforms: Azure (preferred), AWS (nice to have)
Position requirements
* 5+ years of experience in Quality Assurance and Test Automation.
* 2+ years of experience testing AI-enabled applications, LLM solutions, or agentic systems.
* Strong Python/ Java/ Javascript skills.
* Experience with Playwright, Selenium, Cypress, or similar automation frameworks.
* Solid understanding of software testing methodologies and test design techniques.
* Experience with API, integration, and database testing.
* Experience with CI/CD pipelines and DevOps practices.
* Knowledge of AI evaluation frameworks.
* Understanding of AI-specific failure modes: hallucinations, prompt injection, jailbreak attacks, model drift, bias and fairness issues, RAG failures.
* Experience working in cloud-native environments (Azure or AWS).
* Understanding of MCP architecture and agent-to-tool integrations.
* Experience with LangSmith, Braintrust, MLflow, Arize, or Evidently AI.
* Familiarity with EU AI Act and AI governance frameworks.
**Strong plus:**
* Background in Supply Chain, Logistics, WMS, TMS, OMS, or ERP platforms.
* ISTQB CT-AI certification or equivalent AI testing credentials.
* Contributions to AI testing frameworks or open-source projects.
* Leadership experience driving AI quality strategy across multiple teams.
* Knowledge of warehouse operations and fulfillment processes.
Position responsibilities
**Focus:** application-layer LLM QA, and enterprise/AISDLC quality engineering.
**AI Quality Engineering:**
* Design and execute quality strategies for AI-powered warehouse assistants, copilots, and autonomous agents.
* Build evaluation frameworks for LLM-based warehouse workflows and agentic systems.
* Define and maintain golden datasets and benchmark scenarios for warehouse operations.
* Validate AI-generated recommendations, decisions, and workflow actions.
* Perform hallucination, factuality, grounding, and response-quality testing.
* Validate RAG implementations and retrieval accuracy.
* Design agent trajectory tests for multi-step autonomous workflows.
* Conduct adversarial testing, prompt injection testing, and jailbreak assessments.
* Monitor model drift, output quality degradation, and AI performance trends.
**Test Automation & Enterprise Quality:**
* Create and maintain scalable automation frameworks for functional, API, integration, and end-to-end testing.
* Design CI/CD quality gates for AI and non-AI releases.
* Validate enterprise integrations across WMS, ERP, TMS, OMS, WCS, and third-party platforms.
* Implement automated regression suites for business-critical warehouse processes.
* Ensure release readiness through risk-based testing and quality metrics.
* Perform root-cause analysis of complex system-level defects.
* Drive quality best practices across engineering teams.
**Governance & Compliance**
* Define AI quality standards and testing methodologies.
* Support compliance with EU AI Act, ISO 42001, and responsible AI principles.
* Maintain audit-ready quality evidence and evaluation reports.
* Establish measurable KPIs for AI reliability, safety, and business accuracy.
* Contribute to AI quality engineering standards and Centers of Excellence.
**Cross-Functional Collaboration**
* Partner with AI Product Engineer, Solution Architect, and UX Lead.
* Participate in AI solution design reviews.
* Manage warehouse business requirements and maintain validation strategies.