A Very Big Video Reasoning Suite

We bet on a future that video reasoning is the next fundamental intelligence paradigm, after language reasoning, where spatiotemporal embodied world experiences could be more naturally captured.

Data Engines

View All
circle_central_dot
GitHub
Knowledge out-of-domain testset
A row of dots is shown. Circle the dot that is in the middle by count (the one with an equal number of dots on each side).
First Frame
Last Frame
shape_color_then_move
GitHub
Abstraction out-of-domain testset
The scene shows a sequential analogy A→B→C :: D→?→? with three shapes per row. On the top row, a color_145 T_shape becomes color_197, then moves down. On the bottom row, apply the same two-step transformation to the rectangle: first change color from color_145 to color_197, then move down. Show both transformations sequentially in the video.
First Frame
Last Frame
find_keys_and_open_doors
GitHub
Spatiality training set
In the maze, the agent is the green circle. First move the agent to collect the key (diamond shape), then move the agent to the door (hollow rectangle). Use the shortest path for each movement. Show the complete movement step by step.
First Frame
Last Frame
object_packing
GitHub
Transformation training set
The scene shows objects on the left side and a container on the right side. Place the objects into the container one by one in the color order: orange - brown. Each object must be placed individually in the exact order specified, and all objects must end up inside the container.
First Frame
Last Frame
draw_midpoint_perpendicular_line
GitHub
Perception out-of-domain testset
Draw a vertical red line through the middle point that is perpendicular to the horizontal parallel lines. The line should extend from the upper parallel line to the lower parallel line.
First Frame
Last Frame

Inference Results

View Full Bench
Glass Refraction - Samples
00
01
02
03
04
Task Domains 1/5
Glass Refraction
Knowledge in-domain testset
Next Figure (Small-Large Alt)
Abstraction out-of-domain testset
Grid Shortest Path
Spatiality in-domain testset
Track Object Movement
Transformation in-domain testset
Locate Intersection
Perception out-of-domain testset
Prompt
Loading...
Ground Truth
First
First Frame
Final
Final Frame
Model Outputs
1/
VBVR-Wan2.2
VBVR-Wan2.2
CogVideoX 1.5
Kling 2.6
LTX-2
Runway Gen-4
Sora 2
Veo 3
Wan 2.2 I2V
Hunyuan I2V
Seedance 2.0

Leaderboard

Modality
Split
Type
Category
2026-04-28