Model Hub

Browse PQC-verified AI models, datasets, and tools

A
alefiury/wav2vec2-large-xlsr-53-gender-recognition-librispeech HF Unverified

Audio-ClassificationTransformersPyTorchSafetensorsWav2vec2Generated_from_trainer HIGH
artur-muratov/multilingual-speech-commands-15lang HF Unverified

Multilingual Speech Commands Dataset (15 Languages, Augmented) This dataset contains augmented speech command samples in 15 languages, derived from multiple public datasets. Only commands that overlap with the Google Speech Commands (GSC) vocabulary are included, making the dataset suitable for multilingual keyword spotting tasks aligned with GSC-style classification. Audio samples have been augmented using standard audio techniques to improve model robustness (e.g., time-shifting… See the full description on the dataset page: https://huggingface.co/datasets/artur-muratov/multilingual-speech-commands-15lang.

Language:enLanguage:ruLanguage:kkLanguage:ttLanguage:arLanguage:tr
S
speechbrain/emotion-recognition-wav2vec2-IEMOCAP HF Unverified

Audio-ClassificationSpeechbrainEmotionRecognitionWav2vec2PyTorch MEDIUM
X
xbgoose/hubert-large-speech-emotion-recognition-russian-dusha-finetuned HF Unverified

Audio-ClassificationTransformersPyTorchSafetensorsHubertSER HIGH
M
microsoft/speecht5_tts HF Unverified

Text-To-SpeechTransformersPyTorchSpeecht5Text-To-AudioAudio MEDIUM
S
speechbrain/lang-id-voxlingua107-ecapa HF Unverified

Audio-ClassificationSpeechbrainEmbeddingsLanguageIdentificationPyTorch MEDIUM
ATH-MaaS/Marco_Longspeech HF Unverified

Marco-LongSpeech Dataset Marco-LongSpeech is a multi-task long speech understanding dataset containing 8 different speech understanding tasks designed to benchmark Large Language Models on lengthy audio inputs. 📊 Dataset Statistics Task Statistics Task Train Val Test Total Unique Audios ASR 71,275 15,273 15,274 101,822 101,822 Temporal_Relative_QA 5,886 1,261 1,262 8,409 8,409 summary 4,366 935 937 6,238 6,238… See the full description on the dataset page: https://huggingface.co/datasets/ATH-MaaS/Marco_Longspeech.

Task_categories:automatic-Speech-RecognitionTask_categories:audio-ClassificationTask_categories:text-GenerationLanguage:enLanguage:zhSize_categories:10K<n<100K
japanese-asr/whisper_transcriptions.reazon_speech_all HF Unverified

Size_categories:10M<n<100MFormat:parquetModality:audioModality:textLibrary:datasetsLibrary:dask
Showing 8 of 8 items (page 1 of 1)
Prev Next