Model Hub

Browse PQC-verified AI models, datasets, and tools

M
MIT/ast-finetuned-audioset-10-10-0.4593 HF Unverified

Audio-ClassificationTransformersPyTorchSafetensorsAudio-Spectrogram-Transformer MEDIUM
F
fishaudio/s2-pro HF Unverified

Text-To-SpeechSafetensorsFish_qwen3_omniInstruction-FollowingMultilingualJw HIGH
F
FunAudioLLM/Fun-CosyVoice3-0.5B-2512 HF Unverified

Text-To-SpeechONNXSafetensors HIGH
F
facebook/audiobox-aesthetics HF Unverified

Audio-ClassificationSafetensorsModel_hub_mixinPytorch_model_hub_mixin MEDIUM
B
bosonai/higgs-audio-v2-generation-3B-base HF PQC Verified

Text-To-SpeechTransformersSafetensorsHiggs_audio_v2Text-To-Audio HIGH
aline-gassenn/MedDialog-Audio HF Unverified

MedDialogue-Audio English Medical Dialogue Corpus for Speech Recognition Research. This repository contains MedDialogue-Audio, an English audio corpus designed for research in Automatic Speech Recognition (ASR) in the healthcare domain. The dataset was published in the proceedings of the 7th SBBD Dataset Showcase Workshop, and is available online at the following link: https://sol.sbc.org.br/index.php/dsw/article/view/37199 Dataset Description MedDialogue-Audio is… See the full description on the dataset page: https://huggingface.co/datasets/aline-gassenn/MedDialog-Audio.

Task_categories:automatic-Speech-RecognitionLanguage:enSize_categories:100K<n<1MDoi:10.57967/hf/5889Medical
HKUSTAudio/Audio-FLAN-Dataset HF Unverified

Audio-FLAN Dataset (Paper) (the FULL audio files and jsonl files are still updating) An Instruction-Tuning Dataset for Unified Audio Understanding and Generation Across Speech, Music, and Sound. 1. Dataset Structure The Audio-FLAN-Dataset has the following directory structure: Audio-FLAN-Dataset/ ├── audio_files/ │ ├── audio/ │ │ └── 177_TAU_Urban_Acoustic_Scenes_2022/ │ │ └── 179_Audioset_for_Audio_Inpainting/ │ │ └── ... │ ├── music/ │ │ └──… See the full description on the dataset page: https://huggingface.co/datasets/HKUSTAudio/Audio-FLAN-Dataset.

Task_categories:text-To-SpeechTask_categories:text-To-AudioTask_categories:automatic-Speech-RecognitionLanguage:enLanguage:zhSize_categories:10M<n<100M
Showing 7 of 7 items (page 1 of 1)
Prev Next