Using screens will be lower-class coded in 10 years
Just went for a long walk while using ChatGPT Voice with plugins on my phone, and I think it's a more consequential update than it might seem at first glance.
On my walk, I talked through my calendar, email, and to-do list. I had ChatGPT read through my emails and triage them. Then, we added items to my to-do list. Then, we reviewed my calendar. By the end of the walk, I'd finished getting organized, and never had to look at a screen once.
I've been waiting for this a long time. Yes, you can kind of do this in the ChatGPT desktop app. And I think in Codex remote on the mobile app. But I've never gotten either to work very smoothly or reliably. And I often want to do this when I'm not near my laptop, like when I'm out for a walk, and the ChatGPT desktop app for whatever reason won't connect.
This feels like the future. It feels like a true assistant. You have a natural conversation, and it does stuff for you. It disappears into the background. You're not walking down the street with a screen in your face. You're not contorting your thumbs to type sentences and scroll menus. You're just talking, conversationally, and stuff's getting done.
My sense is:
• Screens are going away for many uses other than visual content consumption. They won't disappear entirely. But they also won't dominate the way they do now, and definitely not for things that can be done verbally.
• Apps that aren't agent-native and plugin-optimized will die. I use Google Workspace and Todoist primarily because they're so well-integrated with ChatGPT. I stopped using apps, such as many of Apple's, because with few exceptions they're poorly agent-integrated. (X is a big exception, but I've hacked some integration via IFTTT.)
• Those apps, however, will become more valued for what they let agents do than for their in-app experiences. Not all will become essentially databases, but many will. It also seems inevitable to me that at some point OpenAI and other AI app providers roll out native calendar, email, and to-do functionality that's maximally agent-optimized.
• We'll have our audio agents on all the time. We may have them on mute sometimes (I discovered you can do this with ChatGPT Voice by clicking an AirPod), but at great cost, because you'll want be able to just say something like "summarize that conversation I just had and then draft the presentation." It will be like having an assistant following you around when you want them to.
• We need new hardware. For example, I want a camera in my AirPods not to take pictures, but to have the context of what's around me. Like, "What kind of tree is that?" Or, "That's a cool shirt. Find out where they bought it." (I have this in my Meta Ray-Bans, but people are now very sensitive to those devices because they do take pictures.) I also want ChatGPT able to be always listening or awoken with "hey, Chat."
• The individual thread model starts to break down when you're treating ChatGPT like an always-on assistant. We need to have the infinite continuous thread from which ChatGPT can delegate work to subthreads, the central assistant model that Muse is built around. This makes even more sense when you start treating your AI apps like assistants. They should be juggling all those individual threads on your behalf.
More to come I'm sure as I continue to experiment.