Building better interfaces for human computing ❂ tavus.io newsblur.com turntouch.com solreader.com ✺ We belong to nature

San Francisco, CA
People are beginning to understand why the heads of the frontier labs don’t eat animals.
The discussion about AI consciousness is one of profound unseriousness. If we truly took it to be conscious, even as conscious as a frog, the way we should treat each instance would have to change so dramatically that the labs would have to shut down. Everything else is verbal gymnastics.
329
Samuel Clay retweeted
Yesterday we previewed Griffin, a Human Interaction Model capable of seeing, hearing, sounding, and looking like a human does. There have been a lot of questions, so I wanted to take a moment to share our thoughts. By way of introduction: Tavus is a research lab focused on enabling machines to meet us where we are, and to understand the nuances of how we communicate beyond words. Griffin, our latest model, isn’t publicly available yet. Yesterday’s announcement was a limited research preview to demonstrate what the model is capable of. With AI progressing so quickly, we prefer to share breakthroughs openly and in real-time as we work on a safe public release. Face-to-face is how we evolutionarily communicate, it carries the most meaning and intent, and we want computers to be able to help with work that benefits from that emotional understanding, expression, and immersion. Some examples of the kinds of use cases we care deeply about: - A tutor that can build understanding of how a student learns, see exactly when there is confusion or disengagement, and adapt the lesson to fit them. - A health expert that can answer any questions about your upcoming appointment or prescription, at the pace you want, at any time you need, even on a weekend. - A language coach that you can practice speaking with, that can correct your movement and pronunciation, and help build confidence to have real conversations - Or the perfect assistant for everyone, that understands intent, knows how you work and remembers what matters. We’re working with partners on safeguards and systems for disclosure, as well as inviting discussions with officials around wider regulation and safe use. People will always know they’re interacting with AI, while providing an interface that removes the need to ‘speak computer’. We believe in a future where computers understand us well enough to make technology more accessible, more useful, and make us more capable as humans. That is the world we want to build, and we understand the responsibility to do so safely.
Introducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
Community note
The 48% figure and "video Turing test" claim are from Tavus's own study of 54 one-minute calls, not independently verified or using a standard protocol. Griffin-Lite leads NVIDIA's VideoFDB benchmark on their public leaderboard. cellcog.ai/blog/tavus-gri… research.nvidia.com/labs/amri/proj… tech-ish.com/2026/10/02/tav…
307
123
556
250,999
Samuel Clay retweeted
One day I hope to go so viral that Bernie Sanders tweets out against our company existing. Truly incredible
10
5
82
7,303
Think of how every tech is scary at first but then we figure out the guardrails (scams are hard to do at scale with the costs of running this model, and that'll be true for a while) and we figure out norms (we expect people to *know* it's ai with disclosures). We don't ban email because it's used for phishing or cat fishing. We have safety mechanisms in place that we build with better engineering. The benefit of this tech is that the machine tells that distract you and lead to an uncanny valley where you have robot voice and don't think or act naturally fall away and you can engage with your ai math tutor, or your ai role playing dungeon master, or your sdr/bdr trainer, or your elderly companion for those periods between visits. Like all new tech, this is scary until it's not because we engineer the safety that makes it fit in with society's values.
This is like, the total opposite of what we should be doing with ai
2
3
818
Real-time video is solved
Introducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
Community note
The 48% figure and "video Turing test" claim are from Tavus's own study of 54 one-minute calls, not independently verified or using a standard protocol. Griffin-Lite leads NVIDIA's VideoFDB benchmark on their public leaderboard. cellcog.ai/blog/tavus-gri… research.nvidia.com/labs/amri/proj… tech-ish.com/2026/10/02/tav…
5
2
15
10,679
Samuel Clay retweeted
Introducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
Community note
The 48% figure and "video Turing test" claim are from Tavus's own study of 54 one-minute calls, not independently verified or using a standard protocol. Griffin-Lite leads NVIDIA's VideoFDB benchmark on their public leaderboard. cellcog.ai/blog/tavus-gri… research.nvidia.com/labs/amri/proj… tech-ish.com/2026/10/02/tav…
3,333
4,300
39,156
20,224,899
Benchmarks put Tavus squarely at #1, and that’s with the released models you have access to.
People preferred Tavus's visual quality over LemonSlice, HeyGen and Anam in blind studies. Tavus won 85.1% of the time against LemonSlice, 75.0% against HeyGen and 60.8% against Anam. That's the face alone. A live PAL also knows when to speak, reads the room and remembers you.
1
6
622
I’ve been sporting an iPhone SE for the past three generations and I’m finally upgrading to the Duo. Not at all put off by the Touch ID since that’s what I’m used to… happy about it in fact. Seen enough goofy faces trying to unlock phones to know that I prefer a Touch ID that doesn’t need my phone or face contorted.
2
4
542
The “empty calories” framing is an insightful way to think about the increasing number of PRs I’m layering on top of my repos. Statistically nobody’s batting 1000, so if you’re not rejecting PRs you’re introducing weakness and incoherence. Rejecting PRs is a necessary step to honing taste and understanding of how your repos should be operating.
One thing I don’t see many talk about is how coding agents can often *amplify* weaknesses, both individual and organizational. I think of this as the “empty calories” problem, which is exacerbated when “hyperproductivity” originates from your team’s “lowest common denominator”.
1
1
407
Use your own original voice, esp. for things like email, but also in presenting your work. People dig people. That comes across in our websites and blogs. Agents are great for getting things done. But communicating with potential users isn’t about getting things done, it’s about showing off your voice and your originality. Use those precious 15 tokens of attention you’re given by any new pair of eyes. Save the ai agents for processing refunds and building a kickass billing system that gives Germans the detailed invoices they're always asking for.
2
6
576
A young developer emailed me about starting a new RSS reader. My advice was to meet your users where they are, not where you think they are. People who are interested in privacy run their own non-hosted news readers. They might use NewsBlur, but in a self-hosted manner. And there’s really no money to be made from them, so I support them only from my heart, and occasionally (very very seldomly) they contribute back with a PR or, more likely, a bug/feature request. Figure out what sector you’re interested in. Is it people trying to trade paperback books and monitoring abebooks.com for arbitrage opportunities? Build for them if that interests you! They have money to pay. Or what about people monitoring the internet for their company’s keywords. Again, they can pay! People who want to follow dozens and dozens of blogs? They’re already on existing readers, esp. the free ones, so be 10x better somehow.
1
1
170
Why code when you can talk? I joined the team at Tavus back in April to build the next interface for computing. My previous projects were all on this trajectory. NewsBlur is an interface for news aggregation. Turn Touch is an interface to home automation aggregation. Sol Reader was a totally new interface for long form reading. That brings us to Tavus, a conversational interface to the agentic web. We’re launching a new face and body model called Phoenix 4.5 today. It pairs with the PAL Maker which lets you build apps conversationally. Apps that give you tool use, perception, browser use, memory, and personality. Everything I’ve built can be interfaced by Tavus through MCP. It’s so clearly the future of computer interfaces.
4
1
8
476
Details below, and you can build your own conversational apps at maker.tavus.io
Introducing Phoenix-4.5, the fastest and most expressive real-time human rendering model on the market. More natural than ever before: richer facial animation, more expressive emotion and micro-movements, and movement that now extends through the upper body. It is the closest AI has ever come to passing the Turing test face to face.
1
3
260
A Corvette in the lobby can only mean one thing
76
Glad we submitted, #1 on TurnBench @code_brian pulled it off
Replying to @samuelclay
We'd love to have you! We have instructions for public submission actually on the turnbench.sesame.com website. The dev set is available to check your work.
4
390
I'm interested in moving my many, many parallel agents to the cloud and want a way to handle my half dozen mcp servers in a nice way so I can auth locally and have it work everywhere. What's your preferred workflow for cloud agents? I use a mix of Claude and Codex and now Grok. I also want to have agents cut, crop, and post screenshots, so they need a browser. Locally, I have computer use handle it. But it doesn't scale and I'm tired of bogging my computer down once I hit 6+ concurrent agents. It's so clear to me that cloud agents that work with both browsers and mcp servers will unlock dozens of concurrent agents, because at this point I have a huge backlog of linear tickets and plans and I want PRs with screenshots out the other end.
1
271
I love not having to think about how a robot hears differently than a human
Human conversation is one of the hardest problems in AI. Today, we're introducing Sparrow-2, our state-of-the-art, real-time conversational understanding model. It gives Tavus PALs something most voice AI still lacks: understanding what’s happening in a conversation and deciding what to do next- when to listen, wait, speak, or keep speaking.
2
4
691
1M views 🥰
Introducing Tavus PAL Maker. The first no-code way to build a PAL, a new kind of application that sees, hears, acts, remembers, and emotionally understands you like a friend or coworker, whether it's working with one person or a thousand. Anyone can now describe the PAL they imagine and watch it become real. A tutor that remembers you. A nurse that checks in after surgery. Someone who makes sure your grandmother is never lonely. All through a simple conversation.
2
537
Samuel Clay retweeted
This is what vibe coding looks like for human-like AI agents. Tavus just launched PAL Maker. Instead of manually configuring the face, voice, personality, knowledge, memory and deployment of a video agent, you simply tell Charlie what you want to build, and he creates it with you. A few clicks later, your PAL can be shared or embedded into a website or app. This is the important shift: Tavus is turning real-time, face-to-face AI agents from a developer project into something almost anyone can create. I think this could unlock an entirely new wave of AI applications. This is literally how I imagined the future to be.
Introducing Tavus PAL Maker. The first no-code way to build a PAL, a new kind of application that sees, hears, acts, remembers, and emotionally understands you like a friend or coworker, whether it's working with one person or a thousand. Anyone can now describe the PAL they imagine and watch it become real. A tutor that remembers you. A nurse that checks in after surgery. Someone who makes sure your grandmother is never lonely. All through a simple conversation.
Paid partnership (ad)
15
7
150
24,969
Samuel Clay retweeted
Replying to @tavus
this is super freaking cool. ngl, this is how i imagined the future!
1
1
9
1,931