HD AI-Generated Video Dataset for AI Training
Access 162,810 horizontal AI-generated clips delivered in a consistent Full HD 1920 × 1080 format. With approximately 242.5 hours of synthetic video, the corpus is designed for high-volume classification, retrieval, synthetic-video detection, embedding development, and efficient training pipelines that benefit from standardized decoding.
The metadata can pair uniform technical delivery with asset-level generation records: workflow, provenance, synthetic-content flag, generation date, model or production source where disclosure is permitted, prompt availability, reference-input status, human review, and known limitations.
Dataset Preview
Representative previews from the standardized horizontal HD corpus, including varied synthetic subjects, environments, actions, visual styles, and generation artifacts.
Preview assets shown here represent content available within the dataset.
Key Highlights
- • 162,810 synthetic clips in one horizontal Full HD format
- • Approximately 242.5 hours with an average clip length of approximately 5.4 seconds
- • Consistent 1920 × 1080, 16:9 delivery for predictable decoding and batching
- • Large corpus suited to class-balanced sampling, retrieval indexing, and representation learning
- • Explicit synthetic-content flag for positive-class construction
- • Generation workflow, date, provenance, and permitted production-source status can be included
- • Prompt and reference-input availability recorded independently from the media file
- • Human-review state and known generation limitations can support QA and error analysis
Metadata Fields
Example Metadata Record
Illustrative delivery record. Unknown and restricted provenance fields remain explicitly labeled rather than inferred.
clip_id: aihd_000001
title: AI-generated urban movement scene
description: Standardized Full HD synthetic clip showing people moving through a generated city environment
keywords: AI-generated video, synthetic video, Full HD, horizontal, city, people, movement
resolution: 1920x1080
format_conformance: Validated for HD horizontal delivery
duration: 00:00:05
fps: Available in source metadata
orientation: Horizontal
generation_workflow: Generation and HD packaging stages recorded where retained
provenance: Source and transformation history supplied to the available level
synthetic_content: true
generation_date: Available where retained
model_or_production_source: Disclosed where known and permitted
prompt_availability: full / partial / summary / unavailable / restricted
reference_input_status: none / image / video / unavailable / restricted
human_review: Review type and status available where recorded
known_limitations: Asset-level QA observations where identified
Technical Specifications
Dataset Type
Standardized Synthetic Video: HD Horizontal
Primary Technical Intent
Scalable classification, retrieval, synthetic-video detection, and efficient model training
Clip Count
162,810 clips
Total Duration
242.5 hours
Total Duration in Minutes
Approximately 14,548.1 minutes
Average Clip Length
Approximately 5.4 seconds
Resolution
1920 × 1080
Format
Full HD
Codec
H.264
Bitrate
Approximately 22 Mbps
Container
MP4
Orientation
Horizontal
Aspect Ratio
16:9
Format Consistency
Single 1920 × 1080 delivery profile for predictable decode and batching
Synthetic-Content Label
Explicit asset-level flag suitable for classification targets
Generation Provenance
Workflow, source, and date fields supplied where retained and permitted
Prompt & Reference Status
Availability values recorded without implying that source inputs are always deliverable
Human QA
Review status and known limitation notes supplied where recorded
Licensing, Documentation, and Delivery
Wavebreak Media can provide applicable licensing, generation provenance, and disclosure status information for the standardized HD AI-generated video content, with dataset delivery organized around agreed file, metadata, and technical requirements. See our Dataset Licensing & Compliance page for more detail.
Need the Full Dataset?
Request the standardized HD synthetic video corpus with the required sample size, class schema, provenance fields, prompt and reference status, review state, and delivery manifest.

