Can it compose a space?
Scale, layout, light and materials, and whether the scene still holds together as you move around it.
We give AI models spatial design tasks and report what they can and can’t do, so creators know what to expect before they build with one.
We open everything a model builds and report back on three questions.
Scale, layout, light and materials, and whether the scene still holds together as you move around it.
Which objects a visitor can use, whether each action can be undone, and whether the scene resets cleanly.
What runs in an ordinary browser, what needs Safari on Apple Vision Pro, and what has only been checked in simulation.
A model built this furnished garden in code. Switch the lanterns, open a journal or serve tea, then reset the scene and judge how well it did.
Open Living Pavilion
OpenAI · GPT-6 Astra · Reasoning: high
A rendered still. Open the artifact to explore its browser interactions.
Spatial Web Evals puts each model’s work side by side, labelled with provider, model and reasoning level, so you can compare them before choosing one for your own project.
Explore Spatial Web EvalsGive each model the same kind of spatial design task, with clear goals and limits.
The model produces a working experience you can open, explore and test.
We record what worked, what broke, and what still needs checking on a device.
Every result links to the code, media and notes behind it. Use them to set expectations, pick a model, or write a better brief.
Build an artifact