VIDEO DATASET

HD Vertical AI-Generated Video Dataset for AI Training

Access 147,007 AI-generated clips delivered exclusively in vertical Full HD 1080 × 1920. The approximately 220.3-hour corpus is built for mobile video generation, portrait framing, short-form creative outputs, vertical retrieval, and evaluation of models that must preserve subjects and action inside a 9:16 composition.

Metadata can distinguish portrait-native generation from adapted or reframed output where known, while separately recording the generation workflow, provenance, synthetic-content flag, generation date, model or production source where permitted, prompt availability, reference-input status, human review, and known limitations. Portrait-native material can also be licensed as one component of a broader generative AI training data requirement.

147,007
Clips
220.3
Total Hours
5.4s
Average Clip Length
Full HD
Format
Vertical
Orientation
AI-Generated
Content Type

Dataset Preview

Representative 9:16 AI-generated previews showing portrait framing, mobile composition, subject placement, motion, scene variation, and synthetic artifacts.

The examples illustrate generation behaviour in vertical compositions, including subject framing, edge detail, movement, and artifact patterns relevant to short-form outputs.

Key Highlights

  • Portrait 9:16 geometry for mobile, social, and short-form model workflows
  • Portrait-native versus adapted or reframed status can be recorded where known
  • Explicit synthetic-content flag and vertical-format validation
  • Generation workflow, date, provenance, and permitted source information can be included
  • Prompt and reference-input availability recorded as asset-level statuses
  • Human-review and limitation fields can capture framing, edge, text, anatomy, and temporal issues

Metadata Fields

clip_id Unique identifier linking the vertical file to its manifest record
title Descriptive title of the AI-generated video clip
description Natural-language description of the visible generated content
keywords Keyword tags describing subjects, environments, objects, visual concepts, and scenes in the clip
resolution Video resolution; Full HD 1080 × 1920 in this dataset
duration Length of the individual video clip
fps Frames per second
orientation Video orientation; vertical in this dataset
vertical_source_status Portrait-native generation, reframed, cropped, extended, or unknown status where production history is retained
generation_workflow Text-to-video, image-to-video, video-to-video, or multi-stage portrait-production workflow where retained
provenance Available source, ownership, generation, reframing, ingestion, and transformation history
synthetic_content Explicit boolean or controlled-value flag identifying the clip as AI-generated
generation_date Generation or production date where retained
model_or_production_source Generating model, platform, studio, or production source where known and permitted for disclosure
prompt_availability Availability level for the prompt: full, partial, summary, unavailable, or restricted
reference_input_status Whether image, video, or other reference inputs were used and whether they can be supplied
human_review Recorded review of vertical composition, technical conformity, visible artifacts, labeling, or policy requirements
known_limitations Known issues involving portrait framing, crop safety, edge artifacts, faces, anatomy, text rendering, motion, or temporal stability

Technical Specifications

Dataset Type

Portrait Synthetic Video: HD Vertical

Primary Technical Intent

Mobile generation, 9:16 composition, short-form outputs, and vertical-model evaluation

Clip Count

147,007 clips

Total Duration

220.3 hours

Total Duration in Minutes

Approximately 13,217.4 minutes

Average Clip Length

Approximately 5.4 seconds

Resolution

1080 × 1920

Format

Full HD

Codec

H.264

Bitrate

Approximately 22 Mbps

Container

MP4

Orientation

Vertical

Aspect Ratio

9:16

Portrait-Origin Status

Native vertical, reframed, cropped, extended, or unknown where this can be established

Synthetic-Content Label

Explicit asset-level flag

Generation Documentation

Workflow, date, provenance, and source fields supplied where retained and permitted

Prompt & Reference Inputs

Availability and restriction status recorded separately

Vertical QA

Human-review state and known framing, edge, text, anatomy, or temporal limitations supplied where recorded

Licensing, Documentation, and Delivery

Wavebreak Media can provide applicable licensing, generation provenance, and portrait-source disclosure status information for the vertical AI-generated video content, with dataset delivery organized around agreed file, metadata, and technical requirements. See our Dataset Licensing & Compliance and Dataset Delivery & Security pages for details on rights documentation, packaging, and secure transfer.

Need the Full Dataset?

Request the vertical synthetic video corpus with the required 9:16 selection, portrait-origin status, generation metadata, prompt and reference availability, human-review state, and limitation reporting.

Selected Partners

Selected Wavebreak Media partners