← Усі вакансії

AQA Engineer — AI Evaluation & Automation

Windmill Digital, Ukraine
Формат:
повний remote
Джерело:
jobs.dou.ua
Відгукнутись на вакансію →

We’re looking for a quality-focused engineer who combines a solid manual testing foundation with hands-on automation experience and a growing interest in AI evaluation. You care deeply about how systems behave, you’re not satisfied with surface-level testing, and you’re excited about what modern QA looks like in an AI-driven world.

What you’ll do

Design and execute evaluation strategies for AI-powered features, including LLM-based flows, recommendation engines, and decision logic.

Validate AI outputs for correctness, consistency, bias, hallucination risk, and edge cases — defining clear qualitative and quantitative acceptance criteria.

Perform deep manual, exploratory, and scenario-based testing across the product, from AI-driven flows to core application functionality.

Build and maintain automated UI and API test suites, and keep them running reliably in CI/CD.

Build and maintain automated evaluation and regression pipelines for AI-enabled systems.

Identify non-determinism and reliability risks specific to AI systems and propose practical mitigations.

Partner closely with engineers and product teams to improve testability, observability, and overall release quality.

Performance, load, and security testing.

What we’re looking for

Sharp analytical thinking and attention to detail — you notice what others miss.

4+ years of hands-on manual testing experience, including exploratory testing, test planning, and writing acceptance criteria for complex or enterprise systems.

4 +years of test automation experience using frameworks such as Playwright, Cypress, or Selenium — covering UI and/or API layers.

Confident writing and maintaining test code in TypeScript/JavaScript, Python, Java.

Demonstrated ability to own a full test strategy, not just execute test cases.

Strong understanding of AI system behavior, including non-determinism, prompt sensitivity, and the unique challenges of validating model outputs.

Reads backend code comfortably to debug test failures and understand system behavior.

Nice to have

Experience with AI evaluation frameworks, prompt testing, or model validation techniques.

Familiarity with agentic QA concepts and AI-assisted testing workflows.

Hands-on SDET experience — from framework design to internal tooling.

Exposure to observability tooling and how it supports quality in production systems.

What We Offer

Competitive compensation and benefits package.

Remote-first work with a flexible schedule.

Opportunities for professional growth, training, and certifications.

A dynamic environment where innovation, security, operational excellence, and cutting-edge ML technologies are highly valued.

Projects that span cutting-edge AI, UX, fintech, and healthtech domains.

Схожі вакансії

З блогу Trackr

Усі статті →

Знайдено через trackr.help/jobs · Канал: @trackrhelp · Бот для персональних сповіщень: @trackrhelpBot