datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Audio-Video-Engineering-Agentic-Tasks-1M
Audio/Video Engineering Agentic Tasks (1M)
Abstract
A highly specialized dataset comprising 1,029,459 in-context troubleshooting prompts and execution commands built for the deepest levels of media production. Unlike standard datasets that simulate clean, theoretical instructions, this matrix captures the chaotic, highly-detailed, and conversational reality of professional audio engineers, composers, and video editors mid-session. It is engineered to train multimodal AI… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Audio-Video-Engineering-Agentic-Tasks-1M.VCDB-Core-AudioVideo
VCDB Core Audio-Video Retrieval
This repository packages synchronized video and extracted audio from the
528-video core set of VCDB as a symmetric video+audio-to-video+audio
retrieval task for MTEB/MOEB. The separate 100,000-video background collection
is not included.
Terms and provenance
The source dataset is provided by Fudan University for research purposes
only. The source authors and Fudan University make no warranties about the
dataset, including… See the full description on the dataset page: https://huggingface.co/datasets/pranitchawla/VCDB-Core-AudioVideo.shkolkovo-bobr.video-webinars-audio
shkolkovo-bobr.video-webinars-audio
Dataset of audio of ≈2573 webinars from bobr.video with text transcription made with whisper and VAD. Webinars are parts of free online school exams training courses made by Shkolkovo.
Language: Russian, includes some webinars on English
Dataset structure:
mp3 files in format ID.mp3, where ID is webinar ID. You can check original webinar with url like bobr.video/watch/ID. Some webinars may contain multiple speakers and music.
txt file in format… See the full description on the dataset page: https://huggingface.co/datasets/ZeroAgency/shkolkovo-bobr.video-webinars-audio.cat_video_audiotendances-audio-video-barometre
Tendances audio-vidéo - Baromètre
[!NOTE]
Ce jeu de données Hugging Face est vide. Cette carte sert seulement à référencer le jeu de données Tendances audio-vidéo - Baromètre qui est disponible à l'adresse https://www.data.gouv.fr/datasets/6836e0dc1baaf48fbb8b1851
Description
L’Arcom publie les données du volet quantitatif de son Baromètre Tendances audio-vidéo. L’étude est conduite auprès d’un échantillon représentatif de Français âgés de 15 ans et plus.
Elle vise à… See the full description on the dataset page: https://huggingface.co/datasets/french-open-data/tendances-audio-video-barometre.Video-Audio-MMEyt-video-audio-setHDTF_audio_videoDitto_videos_hdtf_400_audio_60s_chunks_all_preprocessvideo-gen7-audiosample-video-audio
