AI training simulatorsTwo prototypes · demos are live

Train the decision,
not the slide deck.

Role-play training where the trainee actually decides under pressure, with a budget, a clock and consequences. Every choice is scored against your playbook, not against a quiz key, and the trainer gets a readiness report instead of a completion checkbox.

2 live demosopen in a browser, nothing to install
15–30 minone scenario run, start to debrief
282 optionsauthored decisions behind three scenarios
Repeatablesame scenario, measurable second attempt
scenario
● SCORED LIVEEvery option moves the competency bars on the right. There is no single correct button — there are trade-offs your playbook has an opinion about.
Monday 07:45 · Barcelona
decision 1 of 2budget €5,000

Sarah Chen · VIP

BCN → LHR · VY8101 cancelled

Critical board meeting at 14:00 in London. Cannot be postponed.

Pick an option to see how it scores.

scorecard
Prioritisation40
Response speed35
Budgeting45
decisions made0 / 2
off-playbook0
readinessnot started

A two-decision miniature of the real engine. The shipped simulators run 15–30 minutes with a live clock, arriving customers and a full debrief.

Two prototypes you can open right now.

Same engine underneath, different playbook on top. Both are prototypes: they run, they score, they have not yet been through a year of corporate rollout.

Crisis Command

Prototype

Crisis response training for aviation and travel companies. A disruption hits, travellers queue up, and the trainee has to rebook, communicate and compensate inside a fixed budget while the clock runs.

  • Three scenarios: flight cancellation (15 min), storm across Northern Europe (25 min, three waves), French ATC strike (30 min, four waves)
  • 25 travellers and 282 authored decision options — every traveller has a name, a company, a reason for the trip and a personality that decides how long they wait before escalating
  • Four levers per traveller: transport, communication, hotel, compensation — including EU261 compensation, so the trainee practises the regulation rather than reading it
  • A response-time SLA per tier — five minutes for a VIP, ten for business, fifteen for standard — and missing it cuts the satisfaction score by a third no matter how good the rebooking was
  • Full debrief afterwards: decision timeline, per-traveller breakdown, and skill gaps grouped into prioritisation, response speed and budgeting
aviationtravelcrisis opsEN
Open the demo →

Delivery Ace

Prototype

Onboarding for delivery couriers. Learning material first, then the same material as situations: documents, vehicle check, equipment, the app, the delivery zone. An AI instructor reacts to every answer.

  • Six modules planned, module one built end to end with seven scenes
  • Scored on five competencies: safety, service, efficiency, professionalism, stress resistance
  • Readiness verdict from the number of off-playbook choices, plus time on task
  • Ships in en, ru, bg and uk via the URL; a client build is authored in the client's language
logisticsonboardingfront-lineEN · RU · BG · UK
Open the demo →

Neither demo is your business. The point of the call is to find the twenty minutes of your work where a wrong decision is expensive, and build that.

Talk it through

Where a simulator beats a course.

Anywhere the job is judgement under pressure and the cost of learning it live is a lost customer, a fine, or an injury.

!

Crisis and disruption response

Cancellations, outages, recalls, evacuations. The team practises the first ninety minutes, when the decisions are worst and the pressure is highest, before it happens for real.

in your crisis playbook → out scored runs per person
☎

Difficult customer situations

The complaint that escalates, the refund that is not owed, the angry caller. Trainees try the wrong phrasing in a simulator instead of on your actual customer.

in real transcripts → out branching scenarios
→

Onboarding to a front-line role

Courier, driver, cashier, warehouse operator, reception. New hires reach a known standard without blocking an experienced colleague for three weeks.

in SOPs and manuals → out module per role
§

Compliance and safety drills

Data protection at the front desk, lockout-tagout, permit-to-work, anti-money-laundering checks. Every run is logged, so "we trained them" has a record behind it.

in regulation + internal rules → out auditable run log
€

Sales and negotiation

Discount pressure, procurement games, the objection nobody has a good answer to. Scored on outcome and on margin given away, not on enthusiasm.

in your pricing rules → out deal-level scoring
◎

Assessment before and after hiring

The same scenario given to candidates and to the existing team shows where the gap actually is. Useful in hiring, useful in deciding what to train next.

in one scenario → out comparable scores across people

Not sure which one is yours? Bring the incident that made you think about training in the first place.

Book a call

What makes it a simulator and not a quiz.

E-learning asks whether the trainee remembers the rule. A simulator asks what they do when two rules collide, the budget is short and the clock is running. Those are different skills, and only one of them shows up on a bad day.

  • 1
    Decisions branch, they do not just get marked. What the trainee chose for the first traveller changes the budget, the clock and the queue for the next one. The message they send is written for the booking they actually made, so choosing a worse flight and an empathetic phone call produces different words on screen than the same call after a good one.
  • 2
    Scoring follows your playbook. Competency axes, weights and the "best" option come from your rules, not from a generic model opinion. When your policy changes, the scenario changes with it.
  • 3
    Constraints are real, and time is one of them. A fixed budget, a departure time, a queue that grows while the trainee deliberates, and a response deadline per customer tier that silently discounts the score when missed. Pressure is the thing being trained; without it the exercise is a reading comprehension test.
  • 4
    The debrief explains, then lets you retry. Every choice comes back with why it scored that way and what the playbook would have done. The second run through the same scenario is the measurement that matters.
LLM-authored scenariosReactdeterministic scoringruns in a browserLMS embed via iframeany language
Cohort debrief · one scenarioillustrative
TraineeLargest skill gapAvg CSATSeverity
A. Ivanova—88none
M. PetrovBudgeting81opportunity
S. DimitrovResponse speed64improve
K. GeorgievaResponse speed47critical
Two of four lost points to the same thing: they solved the problem well and told the customer too late. That is not four conversations to have — it is one line missing from the induction, and the debrief is what makes it visible.

How a simulator for your team gets built.

The engine exists. What takes the time is turning your actual rules into situations where the wrong choice is tempting.

01 · week 1

Playbook intake

What good looks like, in your words. Policies, SOPs, the incidents that went badly, and what an experienced person would have done instead.

02 · weeks 2–3

Scenario authoring

Situations, branches, options and scoring drafted with a language model, then reviewed line by line by your subject expert. Nothing ships that they have not signed off.

03 · week 4

Pilot cohort

A small group runs it while we watch. Options nobody picks get cut, options everyone picks get harder, wording that confused people gets fixed.

04 · week 5 →

Rollout

Embedded in your LMS or on its own link, with the trainer report. New scenarios added as your rules change.

Where we are honest before you ask.

Two working prototypes and no corporate rollout yet. Here is exactly what that means for you.

These are prototypes, not a platform.

Both demos run end to end and score honestly, and Crisis Command already has a full debrief with skill-gap analysis behind it. What does not exist yet is the boring half: no admin panel, no single sign-on, no LMS connector, no multi-cohort reporting. Those are scoped per client rather than waiting on a shelf, and they are real work, not a checkbox.

The content is the work, not the engine.

A generic simulator teaches nothing. Yours has to encode your rules, your edge cases and your idea of a good outcome. That is why the build starts with your playbook and ends with your expert signing off every branch.

The score is only as good as the playbook.

The simulator measures agreement with the standard you gave it. If your standard is wrong or unwritten, the tool will faithfully train the wrong thing. Sometimes the useful outcome of week one is discovering the playbook does not exist.

It is not a certification.

A good run is evidence that someone handled a simulated situation well. It is not a licence, an accreditation, or a legal defence on its own. Where a regulator requires certified training, this sits alongside it, not instead of it.

Asked on every first call

Why not just use an off-the-shelf e-learning platform?+

Use one, for the parts that are genuinely "read this and confirm". A simulator earns its cost only where the job is judgement under pressure and a wrong call is expensive. If your training problem is knowledge transfer, a course is cheaper and we will say so.

Does the AI improvise, or is it scripted?+

Both, deliberately. Scenarios, branches and scoring are authored and reviewed, so the simulator is deterministic and defensible. The language model is used to write the material fast and, where it helps, to play the other person in free conversation. What gets scored is never left to improvisation.

Can it run inside our LMS?+

As a standalone link today, and as an embedded frame in Moodle, WebTutor or an equivalent once built — the simulators are plain web applications, so embedding is the easy part. Single sign-on and pushing results back into your LMS are scoped per client and are not built yet.

What languages?+

Whatever your team works in. Crisis Command is in English. Delivery Ace ships en, ru, bg and uk. Scenario text is authored per client, so the language is a choice, not a limit.

What does it cost?+

A first scenario set is a fixed-scope project, priced after the playbook intake, because the honest estimate depends on how many branches your rules actually need. After that, adding scenarios is incremental. We would rather quote one narrow module you will use than a six-module programme you will not.

Which decision costs you most when it goes wrong?

Thirty minutes. Open a demo together, then find the twenty minutes of your team's work that would be worth simulating, and be honest about whether it is worth building.

Book a call with Artem