Model Hub

Browse PQC-verified AI models, datasets, and tools

S
SWivid/F5-TTS HF PQC Verified

Text-To-SpeechF5-Tts HIGH
A
amunchet/rorshark-vit-base HF Unverified

Image-ClassificationTransformersTensorboardSafetensorsVitVision MEDIUM
N
nvidia/mit-b2 HF Unverified

Image-ClassificationTransformersPyTorchTfSegformerVision MEDIUM
mvp-lab/LLaVA-OneVision-1.5-Mid-Training-85M HF Unverified

🚀 LLaVA-One-Vision-1.5-Mid-Training-85M Dataset is being uploaded 🚀 Upload Status All Completed: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics 📜 Cite If you find LLaVA-One-Vision-1.5-Mid-Training-85M useful in your research, please consider to cite the following related papers: @misc{an2025llavaonevision15fullyopenframework, title={LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training}… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-1.5-Mid-Training-85M.

Size_categories:10M<n<100MFormat:parquetModality:imageModality:textLibrary:datasetsLibrary:dask
C
CompVis/stable-diffusion-v1-4 HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-DiffusersDiffusers:StableDiffusionPipeline CRITICAL
T
timm/vit_tiny_patch16_224.augreg_in21k_ft_in1k HF Unverified

Image-ClassificationTimmPyTorchSafetensorsTransformers MEDIUM
T
timm/vit_base_patch8_224.augreg2_in21k_ft_in1k HF Unverified

Image-ClassificationTimmPyTorchSafetensorsTransformers MEDIUM
P
pnnbao-ump/VieNeu-TTS-v3-Turbo HF Unverified

Text-To-SpeechONNXSafetensorsVieneu_v3_turboVoice-CloningCode-Switching HIGH
T
timm/vit_small_patch16_224.augreg_in21k_ft_in1k HF Unverified

Image-ClassificationTimmPyTorchSafetensorsTransformers MEDIUM
A
apple/mobilevit-small HF PQC Verified

Image-ClassificationTransformersPyTorchTfCoremlMobilevit MEDIUM
T
timm/vit_base_patch16_224.augreg2_in21k_ft_in1k HF Unverified

Image-ClassificationTimmPyTorchSafetensorsTransformers MEDIUM
B
buildborderless/CommunityForensics-DeepfakeDet-ViT HF Unverified

Image-ClassificationTransformersSafetensorsVitTimmDetection MEDIUM
M
microsoft/VibeVoice-1.5B HF Unverified

Text-To-SpeechTransformersSafetensorsVibevoiceText GenerationPodcast HIGH
anon8231489123/ShareGPT_Vicuna_unfiltered HF Unverified

Further cleaning done. Please look through the dataset and ensure that I didn't miss anything. Update: Confirmed working method for training the model: https://huggingface.co/AlekseyKorshuk/vicuna-7b/discussions/4#64346c08ef6d5abefe42c12c Two choices: Removes instances of "I'm sorry, but": https://huggingface.co/datasets/anon8231489123/ShareGPT_Vicuna_unfiltered/blob/main/ShareGPT_V3_unfiltered_cleaned_split_no_imsorry.json Has instances of "I'm sorry, but":… See the full description on the dataset page: https://huggingface.co/datasets/anon8231489123/ShareGPT_Vicuna_unfiltered.

Language:en
nvidia/SAGE-10k HF Unverified

SAGE-10k SAGE-10k is a large-scale interactive indoor scene dataset featuring realistic layouts, generated by the agentic-driven pipeline introduced in "SAGE: Scalable Agentic 3D Scene Generation for Embodied AI". The dataset contains 10,000 diverse scenes spanning 50 room types and styles, along with 565K uniquely generated 3D objects. 🔑 Key Features SAGE-10k integrates a wide variety of scenes, and particularly, preserves small items for… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/SAGE-10k.

Task_categories:text-To-3dLanguage:enSize_categories:10K<n<100KScene-GenerationInteractive-ScenesEmbodied-AI
M
MCG-NJU/videomae-base HF Unverified

Video-ClassificationTransformersPyTorchSafetensorsVideomaePretraining MEDIUM
N
nvidia/segformer-b0-finetuned-ade-512-512 HF Unverified

Image-SegmentationTransformersPyTorchTfSafetensorsSegformer MEDIUM
M
microsoft/VibeVoice-Realtime-0.5B HF PQC Verified

Text-To-SpeechTransformersSafetensorsVibevoice_streamingRealtime TTSStreaming text input HIGH
stanford-vision-lab/gpic HF Unverified

GPIC: A Giant Permissive Image Corpus for Visual Generation Keshigeyan&nbsp;Chandrasegaran*1,&nbsp; Kyle&nbsp;Sargent*1,&nbsp; Suchir&nbsp;Agarwal1,&nbsp; Michael&nbsp;Jang1,&nbsp; Michael&nbsp;Poli1,2,&nbsp; Juan&nbsp;Carlos&nbsp;Niebles1,4,&nbsp; Justin&nbsp;Johnson3,&nbsp; Jiajun&nbsp;Wu1,&nbsp; Li&nbsp;Fei-Fei1 1&nbsp;Stanford University&nbsp;&nbsp; 2&nbsp;Radical Numerics&nbsp;&nbsp; 3&nbsp;University of Michigan&nbsp;&nbsp; 4&nbsp;Salesforce… See the full description on the dataset page: https://huggingface.co/datasets/stanford-vision-lab/gpic.

Language:en
N
nvidia/segformer-b2-finetuned-ade-512-512 HF Unverified

Image-SegmentationTransformersPyTorchTfSegformerVision MEDIUM
Showing 20 of 83 items (page 2 of 5)