Model Hub

Browse PQC-verified AI models, datasets, and tools

D
depth-anything/Depth-Anything-V2-Large HF Unverified

Depth-EstimationDepth-Anything-V2DepthRelative depthEnglish HIGH
J
John6666/prefect-illustrious-xl-v3-sdxl HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlAnime HIGH
D
distilbert/distilbert-base-cased-distilled-squad HF Unverified

Question AnsweringTransformersPyTorchTfRustSafetensors HIGH
N
nvidia/segformer-b1-finetuned-ade-512-512 HF Unverified

Image-SegmentationTransformersPyTorchTfSegformerVision MEDIUM
J
John6666/amanatsu-illustrious-v11-sdxl HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlAnime HIGH
P
PekingU/rtdetr_r50vd HF Unverified

Object-DetectionTransformersSafetensorsRt_detrVisionEnglish MEDIUM
F
facebook/vjepa2-vitl-fpc64-256 HF Unverified

Video-ClassificationTransformersSafetensorsVjepa2Feature ExtractionVideo HIGH
K
krea/Krea-2-Turbo HF Unverified

Text-to-ImageDiffusersSafetensorsBase_model:krea/Krea-2-RawBase_model:finetune:krea/Krea-2-RawDiffusers:Krea2Pipeline CRITICAL
J
John6666/obsession-illustriousxl-v10-sdxl HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlAnime HIGH
N
nvidia/LocateAnything-3B HF Unverified

Image-Text-to-TextTransformersSafetensorsLocateanythingImage-Feature-ExtractionNvidia HIGH
F
facebook/mask2former-swin-tiny-coco-instance HF Unverified

Image-SegmentationTransformersPyTorchSafetensorsMask2formerVision MEDIUM
fixie-ai/common_voice_17_0 HF Unverified

Size_categories:10M<n<100MFormat:parquetModality:audioModality:textLibrary:datasetsLibrary:dask
K
keremberke/yolov8m-table-extraction HF Unverified

Object-DetectionUltralyticsTensorboardV8UltralyticsplusYolov8 MEDIUM
facebook/anli HF Unverified

Dataset Card for "anli" Dataset Summary The Adversarial Natural Language Inference (ANLI) is a new large-scale NLI benchmark dataset, The dataset is collected via an iterative, adversarial human-and-model-in-the-loop procedure. ANLI is much more difficult than its predecessors including SNLI and MNLI. It contains three rounds. Each round has train/dev/test splits. Supported Tasks and Leaderboards More Information Needed Languages English… See the full description on the dataset page: https://huggingface.co/datasets/facebook/anli.

Task_categories:text-ClassificationTask_ids:natural-Language-InferenceTask_ids:multi-Input-Text-ClassificationAnnotations_creators:crowdsourcedAnnotations_creators:machine-GeneratedLanguage_creators:found
V
valhalla/distilbart-mnli-12-1 HF Unverified

Zero-Shot ClassificationTransformersPyTorchJAXBartText Classification HIGH
W
Wan-AI/Wan2.2-TI2V-5B-Diffusers HF Unverified

Text-To-VideoDiffusersSafetensorsDiffusers:WanPipelineEnglishChinese HIGH
isaacus/open-australian-legal-corpus HF Unverified

Open Australian Legal Corpus ‍⚖️ The Open Australian Legal Corpus by Isaacus, a foundational legal AI research company, is the first and only multijurisdictional open corpus of Australian legislative and judicial documents. Comprised of 229,122 texts totalling over 60 million lines and 1.4 billion tokens, the Corpus includes every in force statute and regulation in the Commonwealth, New South Wales, Queensland, Western Australia, South Australia, Tasmania and Norfolk Island, in… See the full description on the dataset page: https://huggingface.co/datasets/isaacus/open-australian-legal-corpus.

Task_categories:text-GenerationTask_categories:fill-MaskTask_categories:text-RetrievalTask_ids:language-ModelingTask_ids:masked-Language-ModelingTask_ids:document-Retrieval
HPLT/HPLT2.0_cleaned HF Unverified

NB: HPLT2.0 is now superseded by a newer release: HPLT3.0 We recommed switching to v3.0, unless you have a compelling reason to stay on 2.0. This is a large-scale collection of web-crawled documents in 191 world languages, produced by the HPLT project. The source of the data is mostly Internet Archive with some additions from Common Crawl. For a detailed description of the dataset, please refer to our website and our pre-print. The Cleaned variant of HPLT Datasets v2.0 This is… See the full description on the dataset page: https://huggingface.co/datasets/HPLT/HPLT2.0_cleaned.

Task_categories:fill-MaskTask_categories:text-GenerationTask_ids:language-ModelingMultilinguality:multilingualLanguage:aceLanguage:af
Helsinki-NLP/fineweb-edu-translated HF PQC Verified

Helsinki-NLP/fineweb-edu-translated fineweb-edu-tanslated is a collection of automatically translated documents from fineweb-edu. Translations are based on OPUS-MT and HPLT-MT models. The data covers 36,704,000 documents with over 28 billion space-searated tokens of English data translated into 36 languages. The total data set is incudes of over 960 billion tokens and the translated documents are aligned across all languages. More information about how the data has been produced can… See the full description on the dataset page: https://huggingface.co/datasets/Helsinki-NLP/fineweb-edu-translated.

Task_categories:translationTask_categories:text-GenerationLanguage:bosLanguage:bulLanguage:catLanguage:ces
nebius/SWE-rebench-V2-PRs HF Unverified

SWE-rebench-V2-PRs Dataset Summary SWE-rebench-V2-PRs is a large-scale dataset of real-world GitHub pull requests collected across multiple programming languages, intended for training and evaluating code-generation and software-engineering agents. The dataset contains 126,300 samples covering Go, Python, JavaScript, TypeScript, Rust, Java, C, C++, Julia, Elixir, Kotlin, PHP, Scala, Clojure, Dart, OCaml, and other languages. For log parser functions, base Dockerfiles, and… See the full description on the dataset page: https://huggingface.co/datasets/nebius/SWE-rebench-V2-PRs.

Task_categories:text-GenerationLanguage:enSize_categories:100K<n<1MFormat:parquetModality:textLibrary:datasets
Showing 20 of 898 items (page 29 of 45)