Open Spatial Models

SpatialAxiom

An Open Spatial Intelligence Model for General Spatial Reasoning

RELEASED

SpatialAxiom-9B

Built on Qwen3.5-9B · compact dense model

RELEASED

SpatialAxiom-35B-A3B

Built on Qwen3.5-35B-A3B · MoE with 3B active params

Overview

Overview

SpatialAxiom is a spatial intelligence model built on the Qwen3.5 family with large-scale spatial supervision spanning indoor scenes, egocentric views, and multi-camera settings, delivering strong general spatial reasoning without altering the base architecture.

Leading Spatial Reasoning

On average, our models surpass proprietary models and larger open-source alternatives, with leading results on VSI-Bench, MMSI-Bench, MindCube, ViewSpatial, and EmbSpatial.

Spatial Data-Centric Training Recipe

A systematic taxonomy of spatial tasks, balanced task distribution, and data synthesis to raise data quality. SpatialAxiom is trained purely with full-parameter SFT and serves as a clean starting point for downstream fine-tuning or RL.

Qwen3.5 VLM Backbone

Inherits the Qwen3.5 vision-language model architecture, preserving a general-purpose multimodal design without task-specific architectural modifications.

Open-Weight Release

SpatialAxiom-9B and SpatialAxiom-35B-A3B are publicly released on Hugging Face and ModelScope, compatible with transformers and vLLM out of the box.

Evaluation

Evaluation Results

Measured on 8 spatial reasoning benchmarks.

Average Score
Over 8 spatial reasoning benchmarks
Capability Profile
Per-benchmark scores across models

Per-benchmark comparison

Score on each of the 8 spatial benchmarks

Detailed ranking table

Click a column header to sort