The Abuse Project Audio Dataset (TAPAD). Think MNIST for audio profanity.
-
Updated
Mar 25, 2020 - Python
The Abuse Project Audio Dataset (TAPAD). Think MNIST for audio profanity.
[AAAI 2023] AVCAffe: A Large Scale Audio-Visual Dataset of Cognitive Load and Affect for Remote Work
ParquetToHuggingFace processes raw audio data, converts it into Parquet files, and uploads them to Hugging Face. The README explains how to set up the environment, configure paths, and run the scripts to generate and upload the data.
A utility for wrapping the Free Spoken Digit Dataset into PyTorch-ready data set splits.
A comprehensive voice persona dataset for character consistency in voice synthesis, generated using advanced audio-language model Qwen2-Audio-7B, with a GPU-optimized pipeline
Source code for baseline obtenience
Fine tuning Whisper-Small LLM for Hinglish Audio dataset
HSK-graded Chinese sentence dataset: sentences + pinyin + English + per-word glosses + official grammar-point tags + normal/slow audio. Verified zero out-of-level words and wordlist coverage per level (HSK 3.0, GF0025-2021).
CNN Based Audio and Image Captcha Breaker Project
These are different files I created to do different tasks when I was working on creating ASR model for mTEDx dataset.
Download and convert the original MTG Jamendo dataset to opus
For converting audio datasets from one format into another.
Audio watermarking support using WavMark (https://github.com/wavmark/wavmark).
Scalable creation of audio datasets for generating stem continuations of music files
Synthetic doctor-patient consultation dataset generation pipeline with structured clinical outputs and optional full-consultation audio.
Curated Indian-English + Hindi single-speaker speech corpus for expressive TTS (~62 min). Human-verified transcripts and emotion tags, built with Sarvam ASR, diarization and LLM.
Visualization plugins for the audio-dataset-converter library.
Phonemization plugin based on the phonemizer library.
Audio watermarking support using WavMark (https://github.com/facebookresearch/audioseal).
Add a description, image, and links to the audio-dataset topic page so that developers can more easily learn about it.
To associate your repository with the audio-dataset topic, visit your repo's landing page and select "manage topics."