LAYERED VIDEO + JSON DATASET

JSON Layered Video Compositions Dataset

Access a large-scale library of more than 700,000 layered video compositions representing structured visual scenes rather than flattened video outputs. Each composition typically contains 2–4 independently rendered visual layers, including base video, HUD graphics, overlays, atmospheric effects, flares, particles, bokeh, graphs, and other composited elements.

Every composition includes structured JSON describing its visual layers, composition structure, asset relationships, positioning, and rendering information. With approximately 2,917 hours of content, the collection provides structured scene-level data for visual generation, compositing, multimodal reasoning, scene understanding, retrieval, and model development. Structured composition data of this kind complements the rendered generative AI training data available across the wider library.

700,000+
Compositions
~2,917
Total Hours
~15s
Average Duration
2–4
Typical Layers
JSON
Structured Metadata
None
Audio

Dataset Preview

Representative previews showing multi-layer compositions assembled from base footage, interface graphics, overlays, particles, atmospheric effects, flares, bokeh, graphs, and other visual elements.

The rendered examples show how independently addressable layers combine into the final composition, providing a visual reference for layer relationships and compositing structure.

Key Highlights

  • Separate visual elements include base video, HUD graphics, overlays, particles, flares, bokeh, and atmospheric effects
  • Preview renders and production metadata associated with compositions
  • Non-destructive scene structure preserved instead of flattened video alone

Metadata Fields

composition_id Identifier for the individual layered composition
layers Structured representation of the visual layers contained within the composition
composition_structure Information describing how individual layers are organized within the scene
asset_relationships Relationships between the assets used to build the composition
positioning Positioning information associated with visual elements
rendering_information Rendering information associated with the composition and its layers
preview_reference Reference to the corresponding rendered composition preview
production_metadata Production information associated with the composition

Technical Specifications

Dataset Type

Layered Video Composition Dataset

Library Size

700,000+ compositions

Total Duration

Approximately 2,917 hours

Average Composition Duration

Approximately 15 seconds

Typical Layer Count

2–4 visual layers per composition

Metadata Format

Structured JSON

Composition Structure

Multiple independently rendered visual elements assembled into structured scenes

Layer Types

Base video, HUD graphics, overlays, atmospheric effects, flares, particles, bokeh, graphs, and other composited assets

Preview Assets

Rendered composition previews

Asset Relationships

Preserved within structured composition metadata

Positioning Information

Available within composition JSON

Rendering Information

Available within composition JSON

Audio

None

Licensing, Documentation, and Delivery

Wavebreak Media can provide applicable licensing terms, dataset documentation, and structured JSON metadata information for the layered video composition content, with dataset delivery organized around the intended model-development, research, retrieval, or visual-generation workflow. See our Dataset Licensing & Compliance and Dataset Delivery & Security pages for details on rights documentation, packaging, and secure transfer.

Need the Full Dataset?

Request access to the JSON Layered Video Compositions Dataset or discuss composition structures, layer types, metadata, licensing, and technical delivery requirements with Wavebreak Media.

Selected Partners

Selected Wavebreak Media partners