A Very Big Video Reasoning Suite
We bet on a future that video reasoning is the next fundamental intelligence paradigm, after language reasoning, where spatiotemporal embodied world experiences could be more naturally captured.
circle_central_dot
GitHub
Prompt
A row of dots is shown. Circle the dot that is in the middle by count (the one with an equal number of dots on each side).
First Frame
Last Frame
Video
shape_color_then_move
GitHub
Prompt
The scene shows a sequential analogy A→B→C :: D→?→? with three shapes per row. On the top row, a color_145 T_shape becomes color_197, then moves down. On the bottom row, apply the same two-step transformation to the rectangle: first change color from color_145 to color_197, then move down. Show both transformations sequentially in the video.
First Frame
Last Frame
Video
find_keys_and_open_doors
GitHub
Prompt
In the maze, the agent is the green circle. First move the agent to collect the key (diamond shape), then move the agent to the door (hollow rectangle). Use the shortest path for each movement. Show the complete movement step by step.
First Frame
Last Frame
Video
object_packing
GitHub
Prompt
The scene shows objects on the left side and a container on the right side. Place the objects into the container one by one in the color order: orange - brown. Each object must be placed individually in the exact order specified, and all objects must end up inside the container.
First Frame
Last Frame
Video
draw_midpoint_perpendicular_line
GitHub
Prompt
Draw a vertical red line through the middle point that is perpendicular to the horizontal parallel lines. The line should extend from the upper parallel line to the lower parallel line.
First Frame
Last Frame
Video
Glass Refraction - Samples
00
01
02
03
04
Prompt
Loading...
Ground Truth
First
Final
Model Outputs
1/
VBVR-Wan2.2
VBVR-Wan2.2
CogVideoX 1.5
Kling 2.6
LTX-2
Runway Gen-4
Sora 2
Veo 3
Wan 2.2 I2V
Hunyuan I2V
Seedance 2.0
Leaderboard
Modality
Split
Type
Category