2.2K+ Hours English and Japanese Vertical Storyline Video Dataset
Generative AI & Computer Vision
Tags and Keywords

£24,500
About
2.2K+ Hours English & Japanese Vertical Storyline Video Dataset
Description
The 2.2K+ Hours English & Japanese Vertical Storyline Video Dataset is a large-scale collection of high-quality vertical videos featuring scripted and natural storyline-based content in both English and Japanese. Designed for AI and machine learning applications, this dataset supports the development of advanced computer vision, multimodal AI, video understanding, action recognition, scene segmentation, video captioning, content moderation, recommendation systems, and generative AI models.
The dataset captures diverse real-world scenarios, environments, subjects, and storytelling styles, making it suitable for research, commercial AI development, and multimedia analytics. With more than 2,200 hours of multilingual video content, it provides extensive coverage for training robust and scalable AI systems
Note: Pricing varies depending on several factors, including the total video hours, video quality and resolution, metadata availability, number of attributes, annotation requirements, and customization needs. The final price will be determined based on the specific dataset requirements.
Data Product Features
The dataset may include the following metadata fields (availability depends on the licensed version):
| Feature | Description |
|---|---|
| Language | Language spoken in the video (English or Japanese). |
| Video File | Vertical video in portrait orientation. |
| Duration | Length of the video in seconds or minutes. |
| Resolution | Video resolution (e.g., 720×1280, 1080×1920). |
| Frame Rate | Frames per second (FPS). |
| Storyline Category | Type of storyline or narrative represented. |
| Environment | Indoor, outdoor, studio, public places, workplace, etc. |
| Participants | Number of visible individuals. |
| Activity | Primary action or event occurring in the video. |
| Camera Motion | Static, handheld, panning, tracking, etc. |
| Lighting Condition | Natural, artificial, daylight, low-light, mixed lighting. |
Distribution
The dataset is provided as an organized collection of vertical storyline video files.
-
Format: MP4 videos
-
Data Volume:
- Video Duration: 2.2K+ Hours
- Languages: English and Japanese
- Orientation: Vertical (Portrait)
Usage
This data product is ideal for a variety of applications:
- Video Understanding: Train AI models to understand scenes, activities, and narratives.
- Multimodal AI: Develop models combining video, speech, and text.
- Computer Vision: Support action recognition, object detection, and scene analysis.
- Video Captioning: Generate automated descriptions for video content.
- Recommendation Systems: Improve personalized video recommendations.
- Content Moderation: Detect inappropriate or policy-violating content.
- Generative AI: Train video generation and multimodal foundation models.
- Academic Research: Support studies in multimedia analytics and artificial intelligence.
Coverage
The dataset provides broad multilingual coverage suitable for global AI applications.
- Geographic Coverage: Global
- Time Range: Multiple collection periods
- Languages: English and Japanese
License
CC BY 4.0 (Creative Commons Attribution 4.0 International)
AI Training Rights
InfoBay.AI ensures that all datasets are sourced, curated, and managed with proper ownership verification, licensing documentation, and data provenance records. We hold the necessary rights to license and sublicense the datasets we provide through formal agreements with our data vendors, which grant us the required permissions for commercial licensing and AI training use cases. To ensure transparency and compliance, we maintain relevant documentation and have previously shared redacted agreements for selected datasets as evidence of our data rights and licensing authority.
Data Dictionary
| Column Name | Data Type | Description | Possible Values/Notes |
|---|---|---|---|
| Language | String | Language spoken in the video | English, Japanese |
| Video_File | String | Video file name or path | MP4 |
| Duration | Float | Video duration | Seconds or minutes |
| Resolution | String | Video resolution | 720×1280, 1080×1920, etc. |
| Frame_Rate | Integer | Frames per second | 24, 25, 30, 60, etc. |
| Storyline_Category | String | Story or narrative category | Varies by dataset |
| Environment | String | Recording environment | Indoor, Outdoor, Studio, Public, Residential |
| Participants | Integer | Number of visible individuals | Positive integer |
| Activity | String | Primary action in the scene | Walking, Talking, Shopping, Working, etc. |
| Lighting | String | Lighting condition | Natural, Artificial, Mixed, Low-light |
Considerations
This dataset is provided for research and educational purposes only. It contains only sample data.
Loading...
£24,500
Download Dataset in VIDEO Format
Recommended Datasets
Loading recommendations...
