Design
Vision Pro: A Motion Designer's First Take on Spatial Computing
On Monday at WWDC, Apple announced the Vision Pro: a $3,499 headset, shipping sometime next year, that the keynote pointedly refused to call a headset. The framing was "spatial computing," the demo was apps floating in your living room, and the internet immediately split into two camps: people declaring the future arrived and people making jokes about the external battery pack and that slightly uncanny video-call persona. Both camps are having fun. Neither is asking the question that matters for people who make brand and motion work for a living: what is our craft when the canvas is a room?
I have not touched the device; almost nobody outside Cupertino has. These are first-take notes from the keynote and the developer sessions, dated accordingly, and I reserve the right to be wrong by autumn.
The interface language is telling
Look past the hardware and study the visionOS design sessions, because they reveal Apple's actual bet. Interfaces are made of translucent materials that sit in your real space, respond to real light, and are navigated by eyes and fingertips. There is no controller to point. Attention itself is the cursor.
For motion design, this is a bigger shift than the resolution numbers. Our discipline has spent decades choreographing rectangles: things enter frame, exit frame, transition between screens. Spatial interfaces have no frame. Elements have position, depth, scale relative to a human body, and physical plausibility. Easing curves start needing to obey something like real inertia, because motion that ignores physics reads as broken far faster when it happens in your room rather than on a distant screen. The nearest existing craft is not app design. It is stage design, exhibition design, and honestly a fair amount of what our 3D team already knows.
What brands should actually do right now
Here is where I depart from this week's hot takes: almost nothing, and certainly nothing expensive. The device costs $3,499 and ships next year in one country. The audience for brand experiences on it in 2024 will be developers, journalists, and the professionally curious. Any client asking about a "Vision Pro strategy" this month, and one already has, is really asking the AI-strategy question in new packaging, and it deserves Priya's same answer: what business problem, for which customers, provable how?
What is worth doing is cheap and structural:
- Audit your brand for dimensionality. Does your identity system exist only as flat lockups, or do your materials, motion rules, and type have a considered answer to depth? That question pays off in 3D and motion work today regardless of what any headset sells.
- Let your motion and 3D people play. The simulator is free with the SDK. A skunkworks afternoon per month is the right size of bet.
- Watch the interaction patterns, not the device. Eye-plus-hand input, spatial audio as interface, attention-driven UI: if any of this works, it leaks into every platform, the way phone gestures colonized the desktop.
The honest uncertainty
It is entirely possible this becomes the iPhone arc: derided at announcement, category-defining within five years. It is also possible it becomes a beautiful, expensive developer kit for a future that arrives on some cheaper descendant, or does not arrive. The price, the battery, the sheer social oddness of wearing a computer on your face at breakfast: these are real headwinds, and the keynote's most careful choice was showing people using it alone.
Our position as a studio: genuinely curious, financially unmoved. We will build spatial craft the way we built motion craft, through small experiments ahead of client demand, so that if the demand comes we are ready and if it does not we have deepened our 3D practice anyway. That is the nice thing about betting on craft instead of devices. The bet pays either way.
Building something this could apply to?
We take on a small number of flagship projects each quarter.
Start a project