4K+ Hours Marathi Single-Channel Audio Podcast

Audio, Speech & Acoustic Datasets

Tags and Keywords

Marathi

Podcast

Dataset,

Speech

Audio

Corpus,

Asr

Nlp

4K+ Hours Marathi Single-Channel Audio Podcast Dataset on Opendatabay data marketplace

£31,300

About

Marathi Single-Channel Podcast Audio Dataset

Description

The Marathi Single-Channel Podcast Audio Dataset is a high-quality collection of 4,162 hours of Marathi podcast recordings designed for training, fine-tuning, and evaluating speech and language AI systems. The dataset contains single-channel audio featuring natural Marathi speech captured from podcast recordings, making it suitable for real-world speech processing applications.
This dataset is ideal for Automatic Speech Recognition (ASR), Natural Language Processing (NLP), speech analytics, speaker recognition, speaker diarization, audio transcription, language modeling, conversational AI, voice AI, and machine learning.
Note: Pricing varies depending on several factors, including the language, total audio hours, metadata availability, number of attributes, annotation requirements, and customization needs. The final price will be determined based on the specific dataset requirements.

Data Product Features

The dataset includes the following key components:
FeatureDescription
Audio FormatHigh-quality WAV, OBB, MP3 audio files
Channel TypeSingle-Channel
LanguageMarathi
Speech TypeNatural podcast conversations and discussions
Audio QualityClean recordings with natural speaking styles
AI ReadySuitable for speech and language AI model development

Distribution

The dataset is distributed in an organized directory structure for easy integration into AI and machine learning workflows.
  -Format: WAV, MP3, OBB
  -Structure: Audio files organized by podcast recordings

Data Volume

-Records: Podcast recordings
-Volume: 4,162 hours
-Channel Configuration: Single-Channel 

Usage

This data product is suitable for a wide range of AI and speech technology applications.
-Automatic Speech Recognition (ASR): Train and evaluate speech recognition systems.
-Speech-to-Text: Develop transcription models.
-Conversational AI: Build intelligent virtual assistants and customer support systems.
-Speech Analytics: Analyze customer interactions and speech patterns.
-Machine Learning: Train and benchmark speech-based AI models.
-Natural Language Processing (NLP): Support spoken language understanding and downstream NLP tasks.
-Voice Technology: Develop and evaluate voice-enabled applications.
-Academic Research: Conduct research in speech processing, AI, and computational linguistics.
-Speaker Recognition: Build speaker identification and verification systems.
-Speaker Diarization: Separate multiple speakers in conversations.

Coverage

The dataset covers Marathi-language call center recordings.
-Geographic Coverage: Marathi-speaking regions
-Time Range: Not specified
-Demographics: Adult speakers with diverse accents, genders, speaking styles, and podcast topics

License

CC BY 4.0 (Creative Commons Attribution 4.0 International)

AI Training Rights


InfoBay.AI ensures that all datasets are sourced, curated, and managed with proper ownership verification, licensing documentation, and data provenance records. We hold the necessary rights to license and sublicense the datasets we provide through formal agreements with our data vendors, which grant us the required permissions for commercial licensing and AI training use cases. To ensure transparency and compliance, we maintain relevant documentation and have previously shared redacted agreements for selected datasets as evidence of our data rights and licensing authority.

Data Dictionary

Column NameData TypeDescriptionPossible Values/Notes
audio_fileStringName or path of the audio fileWAV, OBB, MP3 filename
languageStringLanguage of the recordingMarathi
channel_typeStringAudio channel configurationSingle-Channel
speaker_countIntegerNumber of speakers in the recording1 or more
duration_secondsFloatAudio duration in secondsPositive numeric value
sample_rateIntegerAudio sampling ratee.g., 16000, 22050, 44100, 48000 Hz
bit_depthIntegerAudio bit depthTypically 16 or 24-bit
audio_formatStringAudio file formatWAV, OBB, MP3

Considerations


This dataset is provided for research and educational purposes only. It contains only sample data.

Listing Stats

VIEWS

3

DELIVERY

CUSTOM, S3

LISTED

15/07/2026

UPDATED

24/07/2026

REGION

GLOBAL

Universal Data Trust Rating UDTRTRUST

5 / 5

Loading...

£31,300

Download Dataset in AUDIO Format