AI-First Senior QA Engineer (Manual)
Ukraine, Kyiv · On-Site · Full-Time
Company: Orbox.ai — AI-Native PropTech Operating System for Real Estate
Location: Ukraine, Kyiv on-site, office in the city centre
Employment: Full-time, permanent
Seniority: Senior
Reports to: Technical Lead / Head of Delivery & Product
Works with: Product Analyst, Product Designer, Engineering team, Fleet of Agents
Mobilization reservation: Full reservation (deferment) from mobilization for male employees
PropTech experience: Not essential; very good if you have direct or indirect experience
English: Advanced (B2+)
About
We are building an AI-native operating system for real estate. The product is a PropTech platform where LLMs and AI agents are not bolted on top — they are the architecture. Owners, operators and guests interact with the AI system through natural language. AI autonomously generates compliance documents, routes staff, answers guest queries from a RAG-indexed knowledge base in real time and orchestrates financial split-payments.
But the product is only half of it.
We are equally an AI-first delivery organization. The way we discover, define and ship is as differentiated as what we ship. Specs, stories and designs are machine-readable artifacts our coding and review agents consume directly — and quality is verified by a human who knows exactly where agents fail, where specs lie and where the product breaks in real hands.
The role
This is the role that decides whether what we ship actually works for humans and for the agents that built it.
You will report directly to the Head of Delivery & Product, with short decision loops to founders and direct access to engineering.
This is a hands-on, manual-first QA role at the core of the delivery machine. Most of our code is written by AI agents against machine-readable specs. That makes human quality judgment more valuable, not less: agents pass their own checks and still ship broken flows, wrong edge-case behaviour and UX that technically matches the spec but fails the user. You are the last honest reader of the product before it reaches staging review and then customers on prod.
You will own manual testing end to end: exploratory testing, spec-based test design, UAT coordination, regression passes and release sign-off. You will test not just deterministic UI flows, but AI-driven behaviour agent responses, RAG answers, document generation where “expected result” is a judgment call, not an assert.
You will work side by side with the AI-First Product Analyst (who writes the specs you test against) and engineering (whose agents implement them). You will use our shared Claude Team / Cowork with premium seating along with the company's AI tools daily, deeply and natively — not as a novelty, but as the default way you work.
We operate in a fast-moving startup environment where speed and precision both matter. You must be opinionated about which QA tasks AI does well today (test-case generation, coverage analysis, log triage), which it does not (exploratory judgment, UX intuition, risk feel) — and how to combine both deliberately.
What you will own & do
AI-First Way of Working
• Daily, deep, native use of AI tools for test-case generation from specs, coverage gap analysis, bug report drafting, log and session triage, regression checklist upkeep.
• Treat AI tools the way an engineer treats their IDE: opinionated, tuned, integrated into every step of your workflow.
• Contribute to internal evals from the QA side: define what “good” looks like for AI-generated code, AI-generated test artifacts and our user-facing AI features.
Manual Testing & Test Design
• Own manual testing end to end: exploratory sessions, spec & story-based test design, edge cases, negative paths, cross-role and multi-tenant scenarios.
• Convert machine-readable specs & stories with acceptance criteria into structured, traceable test cases and checklists — every requi


