Privacy & SecurityIntelligence Snack

Responsive Video Models Could Adapt Scams

Video models that respond to a person's reactions could make digital tutors more attentive, but the same ability could also make impersonation scams more convincing.

Developed from a conversation between Pete Winn and Andy David

From Episode 75: The Human Bottleneck

This week, a company called Tavus introduced Griffin, which it describes as the first Human Interaction Model (HIM) claims is the first to pass a real-time video Turing test. In the company’s own study, 48% of participants believed they were speaking to a real person after a one-minute video call, compared with less than 3% for earlier systems.

Andy thought Griffin could become an infinitely patient tutor that could notice hesitation, adjust its approach and give a language learner someone to practice with. Rather than simply responding to what a learner says, it could watch their reactions and adjust the conversation accordingly.

That responsiveness also creates a potential risk. Andy’s next thought was that such a model could try to steal his dad’s banking passwords. Pete pushed the scenario further, imagining a generated wife asking for a password or a video of someone’s son apparently being held captive. Instead of relying on one fixed recording, a scam could continue answering as the target reacts, adapting the conversation in real time.

The wider consequence could be that people become much less trusting of video calls. If a face on screen can watch, respond and convincingly impersonate someone familiar, seeing someone on video may no longer be enough to establish who they are. Pete wondered whether people might increasingly favour smaller, more private online communities, where identities and relationships have been established through other means.

Get Intelligence Snacks in your inbox.

Quickly digest the big ideas emerging from the world of AI, delivered each week.