NO. 01 — manifesto

We are a
Spatial AI lab.

We build foundation models that understand the 3D world — perceive, reason, simulate, and generate within it. One representation, many tasks.

FIG. 01.Aspatial primitive
NO. 02 — what we do

Image models produce images.Video models produce motion.Spatial models produce worlds.

From models to worlds
The shift

Spatial AI is the idea that a model carries a coherent 3D representation of what it sees — the way humans do without thinking. Objects exist in space. Surfaces have geometry. Occlusion is real. The world stays consistent when you move.

Why it matters

Reconstruction, understanding, editing, and generation become different ways of interrogating the same world model. One foundation, many tasks — a platform shift for the whole field.

No. 04 — ManifestoFig. 04. B

The world is not flat,
AI shouldn't be either.

No. 05.S.01

CV is fragmented.

Computer vision — how AI perceives the world — is still many narrow models for many narrow tasks. CV today sits where NLP did seven years ago: lots of pipelines, no shared foundation.

No. 05.S.02

3D is foundational.

3D is not a feature you bolt on. It is the substrate — under perception, under planning, under interaction.

No. 05.S.03

The 3D gap.

Text, images, and video each have foundation models. 3D does not — yet. Closing that gap is the open problem we are building for.