
We put AI into real conversations with performers, test how it responds, and turn those interactions into data your team can use.
AI that sees, hears and responds to people face to face now has a name: Human Interaction Models. They still miss the cues people catch without thinking.
A correct answer can still arrive at the wrong moment. Research benchmarks expose gaps in timing and nonverbal understanding; live studies examine how those behaviors appear in your own product.
Test how your AI responds when people hesitate, disagree, change their tone or expect it to stay in character. We design scenes around your product and evaluate the behaviors that matter to your users.
Each study focuses on agreed scenarios, participants and behavioral criteria. Receive recorded interactions, labeled evidence and findings tied to your product.
Our longer-term work extends into how robots behave around people: when to approach, how much space to leave, when to yield and how to respond to human signals.
Scope an interaction studyHuman Interaction Models bring voice, vision and behavior into real-time conversation. Our longer-term direction extends that work into how robots approach, respond and share space with people. Physical systems need their own data and validation.

We test and train the AI faces people already interact with: support agents, interviewers, tutors and assistants.

Adapt scene capture with robotics partners to study human signals from the robot’s perspective.

Performers work in the room with your robot. Its motion data joins the same session file, ready for robot training.

The behavioral test every humanoid passes before it works around people.
Data vendors sell hours of human work. We sell instrumented interactions: the system knows exactly what the person did and when.
The performer learns the situation and goal, plus exact lines when a scene needs them.
They meet your live AI. The app cues the key moment at an exact time.
Both sides and your AI's own events land on one clock.
Recording review verifies observed behavior. Labels follow agreed criteria, with retests after your next release.
Four ways to work with us, all built on live sessions between real people and your AI.
Performers test your live product on agreed scenes built around your use case.
Comparative studies with consistent conditions and documented criteria.
The same scenes on every build, with regressions flagged.
Labeled human-to-AI sessions built around what your model gets wrong.

Does it wait when a candidate freezes, stop when they cut in, and stay steady when they get defensive?
We design those moments, run them with performers on the live coach, document what breaks, and retest after every update.

Remote sessions from your phone. Most scenes give you a situation and a goal, not a script. You choose what your recordings can be used for, and you're paid for every accepted session.
Tell us what you're building. We'll scope a study around the behaviors that matter for your release.
Book a study