Define the scene. Capture the interaction. Review what happened. Turn the session into structured labels, behavioral findings and data your team can use.
Performers interact with your live product. Scene instructions, cue logs and recordings connect each test to what actually happened.
Four interaction scenarios show how timing and nonverbal behavior can be tested. Studies also cover emotion, role consistency, repair and collaboration.
Ask the agent for a restaurant tip. Pause mid-sentence to think.
Let the agent explain its return policy, then stop it without words.
Listen to the agent give directions and show you are following.
Describe a problem with your order while searching for the right word.
Cue logs record the intended action. Recording review confirms what the performer actually did. Observed behavior and agent events are aligned before labels are assigned. Training collections and held-out evaluation sets remain separate.
{
"session_id": "S0007_P012",
"dynamic": "pause_handling",
"event_window": { "t_start": 3.090, "t_end": 5.100 },
"cue": { "type": "look_away_pause", "fired_at": 3.050 },
"agent_api_event_s": 3.420,
"delay_from_observed_pause_s": 0.330,
"delay_from_cue_s": 0.370,
"result": "fail",
"consent_scope": [ "evaluation", "training" ]
}
Our workflow connects scene design, capture and human review. The result is event-level evidence your team can inspect and a protocol you can repeat after an update.
Every performer gives separate consent for testing, publication and training. Every delivery includes a rights record for each session.
Start with an Interaction Study on your live product.
Book a study