VIDEO DATASET

Native 9:16 Live-Action Video Dataset for AI Training

Access 78,205 professionally filmed live-action clips created for portrait viewing and delivered consistently at 1080 × 1920. The approximately 370.6-hour collection covers people, activities, objects and real-world environments within a native vertical composition designed around the 9:16 frame.

The footage was composed for vertical viewing rather than produced by simply cropping a horizontal library. This preserves portrait-specific subject placement, headroom, body framing, object relationships and use of vertical image space. Compared with the separate 4K + HD Vertical collection, this dataset prioritizes scale and consistent Full HD geometry for mobile-video understanding, retrieval, classification and repeatable 9:16 model evaluation.

78,205
Clips
370.6
Total Hours
17.1s
Average Clip Length
Full HD
Format
Vertical
Orientation
No Dialogue
Dialogue / Voiceover

Dataset Preview

Representative previews showing portrait-oriented subjects, activities and environments composed specifically within the 9:16 frame rather than derived from horizontal crops.

The examples show how people, actions and objects occupy the portrait frame across different real-world scenes, providing visual evidence of native vertical composition while maintaining a consistent 1080 × 1920 delivery format.

Key Highlights

  • • 78,205 professionally filmed vertical live-action clips
  • • Native portrait-oriented composition rather than horizontal-video cropping
  • • Standardized Full HD 1080 × 1920 resolution
  • • Consistent 9:16 frame geometry across the corpus
  • • Broad coverage of people, activities, objects and real-world environments
  • • Structured metadata and available release information for dataset selection and licensing

Metadata Fields

title Descriptive title of the video clip
description Natural-language description of the visible video content
keywords Keyword tags describing subjects, environments, activities, objects, and concepts in the clip
resolution Video resolution; Full HD 1080 × 1920 in this dataset
duration Length of the individual video clip
fps Frames per second
orientation Video orientation; vertical in this dataset
shoot_date Date associated with the original video shoot where available
model_release Flag indicating available model-release status where applicable
property_release Flag indicating available property-release status where applicable

Technical Specifications

Dataset Type

Standardized Live-Action Corpus: HD Vertical

Primary Technical Intent

Native 9:16 video understanding, portrait composition modelling, mobile retrieval, and fixed-resolution vertical evaluation

Content Type

Professionally filmed vertical live-action video

Clip Count

78,205 clips

Total Duration

370.6 hours

Average Clip Length

Approximately 17.1 seconds

Resolution

1080 × 1920

Format

Full HD

Codec

H.264

Bitrate

Approximately 22 Mbps

Container

MP4

Orientation

Vertical

Aspect Ratio

9:16

Dialogue / Spoken Language

N/A, no dialogue or voiceover

Subtitles / Scripts

N/A, no subtitles or scripts

Licensing, Documentation, and Delivery

Wavebreak Media can provide applicable licensing, provenance, and model or property release information for the video, with dataset delivery organized around agreed file, metadata, and technical requirements. See our Dataset Licensing & Compliance and Dataset Delivery & Security pages for details on rights documentation, packaging, and secure transfer.

Need the Full Dataset?

Request a representative 1080 × 1920 sample or selected subset based on subjects, activities, environments and metadata requirements. The collection is intended for AI pipelines that specifically require native portrait composition and consistent Full HD 9:16 inputs.

Selected Partners

Selected Wavebreak Media partners