Large-Scale HD AI-Generated Video Dataset for AI Training
Access 162,810 AI-generated horizontal video clips delivered through a single Full HD 1920 × 1080 profile. With approximately 242.5 hours of synthetic content, the collection provides a large standardized corpus for synthetic-video detection, representation learning, retrieval, classification and model evaluation.
The dataset is designed for workloads where training volume and consistent input geometry are more important than resolution diversity. Unlike the separate 4K + HD AI-generated collection, this corpus removes mixed-resolution variation and provides a substantially larger pool of standardized synthetic examples for training, validation and held-out evaluation. Asset-level records can also document synthetic status, generation workflow, provenance, generation date, prompt availability, reference-input status, human review and known generation limitations where retained.
Dataset Preview
Representative previews from the standardized horizontal HD AI-generated corpus, including varied synthetic subjects, environments, actions, visual styles, and generation artifacts.
The examples illustrate content and artifact variation within a consistent HD format, making it easier to compare generated scenes without resolution differences.
Key Highlights
- • 162,810 AI-generated clips for large training, validation and evaluation splits
- • Single 1920 × 1080 horizontal delivery profile across the corpus
- • Explicit synthetic-content status for positive-class construction
- • Consistent frame geometry for controlled training and model comparison
- • Generation workflow, provenance, prompt and reference-input status available where retained
- • Human-review state and known generation limitations can support QA and error analysis
Metadata Fields
Technical Specifications
Dataset Type
Standardized Synthetic Video: HD Horizontal
Primary Technical Intent
Scalable classification, retrieval, synthetic-video detection, and efficient model training
Clip Count
162,810 clips
Total Duration
242.5 hours
Total Duration in Minutes
Approximately 14,548.1 minutes
Average Clip Length
Approximately 5.4 seconds
Resolution
1920 × 1080
Format
Full HD
Codec
H.264
Bitrate
Approximately 22 Mbps
Container
MP4
Orientation
Horizontal
Aspect Ratio
16:9
Format Consistency
Single 1920 × 1080 delivery profile for predictable decode and batching
Synthetic-Content Label
Explicit asset-level flag suitable for classification targets
Generation Provenance
Workflow, source, and date fields supplied where retained and permitted
Prompt & Reference Status
Availability values recorded without implying that source inputs are always deliverable
Human QA
Review status and known limitation notes supplied where recorded
Licensing, Documentation, and Delivery
Wavebreak Media can provide applicable licensing, generation provenance, and disclosure status information for the standardized HD AI-generated video content, with dataset delivery organized around agreed file, metadata, and technical requirements. See our Dataset Licensing & Compliance and Dataset Delivery & Security pages for details on rights documentation, packaging, and secure transfer.
Need the Full Dataset?
Request a representative Full HD synthetic-video sample or a larger subset based on training volume, metadata requirements, generation provenance, prompt status, review state and known limitation categories.
Related Datasets
4K & HD AI-Generated Video Dataset for AI Training
Mixed-resolution horizontal AI-generated video in 4K/UHD and HD for high-resolution evaluation, synthetic-content detection, and resolution-robust pipelines.
View Dataset Video Datasets for AI TrainingHD Vertical AI-Generated Video Dataset for AI Training
A 1080 × 1920 AI-generated video corpus for mobile generation, portrait composition, short-form outputs, and vertical-model evaluation.
View Dataset Video Datasets for AI TrainingHD Animation & Motion Graphics Dataset for AI Training
A standardized 1920 × 1080 horizontal animation corpus for scalable decoding, classification, retrieval, and efficient motion-model training.
View Dataset
