Skip to content
Novus Examples
mp417.8 KB

Depth Input — Three-Layer Parallax Scene

Three planes translating at 0.6, 2.4 and 6.0 pixels per frame — a 1 : 4 : 10 velocity ratio that defines their relative depth by motion alone. There are no pictorial depth cues to fall back on: no perspective convergence, no shading, no familiar object sizes. A monocular depth model that has learned pictorial cues rather than motion parallax will do poorly here, which is precisely what makes the fixture informative.

Browser-playable test video · 2 seconds

Specifications

Resolution
640x360
Fps
24
Layer Velocities
0.6, 2.4, 6.0 px/frame (far, mid, near)
Velocity Ratio
1 : 4 : 10
Role
depth-estimation input
Reference
vid-depth-parallax-gt
Crf
19
Base Plate
parallax

Testing contract

Expected to pass
Scenario
Estimate depth for the clip and check the ordering and ratio of the three planes against the ground truth.
Expected result
The three planes translate at 0.6, 2.4, 6.0 px/frame (far, mid, near), a velocity ratio of 1 : 4 : 10, at 640x360 24 fps; the reference is vid-depth-parallax-gt. Parallax is the only depth cue present - there is no perspective, occlusion gradient or focus falloff - so a model that returns a plausible depth map here is genuinely using motion.

What is a .mp4 file?

MP4 (MPEG-4 Part 14) is the dominant container for digital video, holding video, audio, subtitle, and metadata tracks in a tree of typed boxes. `ftyp` declares the brand, `moov` carries the sample tables that make seeking possible, and `mdat` holds the media; when an encoder writes `moov` last, playback cannot begin until the file has fully downloaded, which a faststart remux fixes. It derives from Apple's QuickTime format, generalized by ISO into the ISOBMFF base that MOV, 3GP, and HEIF share.

How to use this file

Use an example MP4 to test box parsing and track detection, seeking, range-request streaming, and transcode or thumbnail pipelines: verifying that a file with a trailing `moov` box is still handled, and that codec support is checked per track rather than inferred from the extension.

How to use this file for testing

“Depth Input — Three-Layer Parallax Scene” is a deterministic Novus Examples fixture for Depth from video, Frame interpolation, Video QA. Layered scenes whose planes translate at documented, fixed ratios, so relative depth is defined by measurable motion rather than inferred from pictorial cues, giving depth-from-video output an objective reference.

Documented properties for this file: depth-estimation input · 24 fps. Compare results against paired or grouped companions on this page when present (clean↔damaged, searchable↔scanned, or format twins) so scores stay reproducible across runs.

Download the file once, keep the path stable in CI or local scripts, and treat the spec table as the contract: dimensions, seeds, field lists, and roles are intentional. Corrupt or invalid samples are labelled as such, expect parsers to fail loudly rather than silently accept them.

Media fixtures are short and synthetic by design. Prefer waveform or transcript ground truth in the same group when measuring ASR, trim, upscale, or sync tools; do not assume broadcast-quality masters.

Code examples

<video controls preload="metadata" width="640" src="parallax-scene.mp4"></video>

Generated by generation/video_ai_suites.py. Free for any use, no attribution required, license.