# Interactive Generative Humans Could Transform the Service Economy by Alexander Wissner-Gross > Real-time generative video will make fully interactive synthetic people inexpensive enough to join meetings, podcasts, and customer interactions. Many service jobs require a face, a voice, and responsive presence; once models can supply all three, text-only automation becomes embodied service-sector automation. > **— Adapted from Alexander Wissner-Gross**, *MOONSHOTS Live, October 2026* ## Sources and Context - **Recording or publication:** [MOONSHOTS Live at 50:46](https://www.youtube.com/watch?v=Blyb1D927pM&t=3046s) — The poster text is an explicit non-verbatim adaptation assembled from the source excerpts below. It preserves the speaker's argument while removing spoken-language filler, restoring the subject, and completing the mechanism or consequence needed for independent use. - **Exact transcript excerpt at [50:46](https://www.youtube.com/watch?v=Blyb1D927pM&t=3046s):** “Just, just checking, but I, I guess Tavis is ahead of me. I, I, I think on the one hand, if you look at models coming out of Chinese labs like Alibaba's lab, Wanstreamer gave us a preview of what was going to happen. The Chinese labs remain over-invested relative to the Western labs in generative video models and interactive generative video models. So Wanstreamer, I think, is a preview of what's possible. My guess is I, I, I, I read the, the information that was put out around Tavis. I, I think in full generality, let, let's just talk about where this is going to end up. It's, it's very difficult to predict the short term. I think it's pretty easy to predict the long term. In the long term, end-to-end generative pixels will have these magic mirrors that could be real-time, fully interactive, pixel-wise generated, or if, if Anthropic has its way, vector-wise or procedurally generated, but either way, fully generated real-time interactive video models, and you'll be able to create a scene that consists of people, like it's possible right now, but the latency is high. You see with Wanstreamer, which is already out and already open source, or with this Tavis Griffin type model, which is not really out yet and definitely not open source, you see a preview of the future. What does this look like? Well, th- there are a few different ways it can go. Query how revenue generating per token it is. I- is it anywhere close to the optimal frontier of code generation? Doubt it. On the other hand, if it becomes so absurdly inexpensive to be able to, to generate arbitrary, say, humans participating in a Zoom meeting or humans participating in a podcast, do we really care whether it's close to, to being near the optimal cost performance frontier? Maybe not. There are a lot of human service industry jobs that require a face and require interactivity and a voice, and the ability to touch one's face apparently on demand-” - **Exact transcript excerpt at [52:53](https://www.youtube.com/watch?v=Blyb1D927pM&t=3173s):** “... or on, on request, that could be probably completely automated away by a, a model that otherwise would be limited to text-based interaction but doesn't have a face and a voice. So in the most optimistic scenario, these sorts of interactive video models, open paren, n- note that the acronym for this, which by the way was there first with Alex Finn, HIM, H-I-M, is such an obvious reference to Her, the, the movie, close paren. I, I think th- this, this has, this has-” - **Exact transcript excerpt at [53:28](https://www.youtube.com/watch?v=Blyb1D927pM&t=3208s):** “Uh, okay, I'll delve into it. So I, I, I think, I, I, I think this is going to be transformative for the service sector, and hopefully Tavis and the broader American west of video frontier models and interactive video frontier models take a page from Tavis and start competing with China.” - **Reconciled source dossier:** [[research/ASI and RSI Timeline Research Moonshots|ASI and RSI Timeline Research Moonshots]] — preserves broadcast order, speaker reconciliation, editorial conventions, and the surrounding argument from which this Reminder was promoted. ## Related Articles and Collections - **Collection:** [[collections/Simple Reminders|Simple Reminders]] - **Collection:** [[collections/Machine Succession|Machine Succession]] - **Article:** [[articles/The Next Interface Layer|The Next Interface Layer]] - **Article:** [[articles/Mind Uploading and AI — The Host is Reusable and the Person is the Delta|Mind Uploading and AI]] ## Related Topics - [[wiki/Alexander Wissner-Gross|Alexander Wissner-Gross]] - [[wiki/Interactive Generative Video|Interactive Generative Video]] - [[wiki/Service-Sector Avatar Automation|Service-Sector Avatar Automation]] ## Share on Social Media ``` Real-time generative video will make fully interactive synthetic people inexpensive enough to join meetings, podcasts, and customer interactions. Many service jobs require a face, a voice, and responsive presence; once models can supply all three, text-only automation becomes embodied service-sector automation. — Adapted from Alexander Wissner-Gross, MOONSHOTS Live, October 2026 https://bryantmcgill.com/simple-reminders-alexander-wissner-gross-interactive-generative-humans-could-transform-the-service-economy ```