2.2K+ Hours English and Japanese Vertical Storyline Video Dataset

Generative AI & Computer Vision

Tags and Keywords

Storyline

Vertical

Video

English

Japanese

Multilingual

Ai

Annotation

2.2K+ Hours English and Japanese Vertical Storyline Video Dataset Dataset on Opendatabay data marketplace

£24,500

About

2.2K+ Hours English & Japanese Vertical Storyline Video Dataset

Description

The 2.2K+ Hours English & Japanese Vertical Storyline Video Dataset is a large-scale collection of high-quality vertical videos featuring scripted and natural storyline-based content in both English and Japanese. Designed for AI and machine learning applications, this dataset supports the development of advanced computer vision, multimodal AI, video understanding, action recognition, scene segmentation, video captioning, content moderation, recommendation systems, and generative AI models.
The dataset captures diverse real-world scenarios, environments, subjects, and storytelling styles, making it suitable for research, commercial AI development, and multimedia analytics. With more than 2,200 hours of multilingual video content, it provides extensive coverage for training robust and scalable AI systems
Note: Pricing varies depending on several factors, including the total video hours, video quality and resolution, metadata availability, number of attributes, annotation requirements, and customization needs. The final price will be determined based on the specific dataset requirements.

Data Product Features

The dataset may include the following metadata fields (availability depends on the licensed version):
FeatureDescription
LanguageLanguage spoken in the video (English or Japanese).
Video FileVertical video in portrait orientation.
DurationLength of the video in seconds or minutes.
ResolutionVideo resolution (e.g., 720×1280, 1080×1920).
Frame RateFrames per second (FPS).
Storyline CategoryType of storyline or narrative represented.
EnvironmentIndoor, outdoor, studio, public places, workplace, etc.
ParticipantsNumber of visible individuals.
ActivityPrimary action or event occurring in the video.
Camera MotionStatic, handheld, panning, tracking, etc.
Lighting ConditionNatural, artificial, daylight, low-light, mixed lighting.

Distribution

The dataset is provided as an organized collection of vertical storyline video files.
  • Format: MP4 videos
  • Data Volume:
    • Video Duration: 2.2K+ Hours
    • Languages: English and Japanese
    • Orientation: Vertical (Portrait)

Usage

This data product is ideal for a variety of applications:
  • Video Understanding: Train AI models to understand scenes, activities, and narratives.
  • Multimodal AI: Develop models combining video, speech, and text.
  • Computer Vision: Support action recognition, object detection, and scene analysis.
  • Video Captioning: Generate automated descriptions for video content.
  • Recommendation Systems: Improve personalized video recommendations.
  • Content Moderation: Detect inappropriate or policy-violating content.
  • Generative AI: Train video generation and multimodal foundation models.
  • Academic Research: Support studies in multimedia analytics and artificial intelligence.

Coverage

The dataset provides broad multilingual coverage suitable for global AI applications.
  • Geographic Coverage: Global
  • Time Range: Multiple collection periods
  • Languages: English and Japanese

License

CC BY 4.0 (Creative Commons Attribution 4.0 International)

AI Training Rights

InfoBay.AI ensures that all datasets are sourced, curated, and managed with proper ownership verification, licensing documentation, and data provenance records. We hold the necessary rights to license and sublicense the datasets we provide through formal agreements with our data vendors, which grant us the required permissions for commercial licensing and AI training use cases. To ensure transparency and compliance, we maintain relevant documentation and have previously shared redacted agreements for selected datasets as evidence of our data rights and licensing authority.

Data Dictionary

Column NameData TypeDescriptionPossible Values/Notes
LanguageStringLanguage spoken in the videoEnglish, Japanese
Video_FileStringVideo file name or pathMP4
DurationFloatVideo durationSeconds or minutes
ResolutionStringVideo resolution720×1280, 1080×1920, etc.
Frame_RateIntegerFrames per second24, 25, 30, 60, etc.
Storyline_CategoryStringStory or narrative categoryVaries by dataset
EnvironmentStringRecording environmentIndoor, Outdoor, Studio, Public, Residential
ParticipantsIntegerNumber of visible individualsPositive integer
ActivityStringPrimary action in the sceneWalking, Talking, Shopping, Working, etc.
LightingStringLighting conditionNatural, Artificial, Mixed, Low-light

Considerations


This dataset is provided for research and educational purposes only. It contains only sample data.

Listing Stats

VIEWS

1

DELIVERY

CUSTOM, S3

LISTED

24/07/2026

UPDATED

05/08/2026

REGION

GLOBAL

Universal Data Trust Rating UDTRTRUST

5 / 5

Loading...

£24,500

Download Dataset in VIDEO Format