Model Hub

Browse PQC-verified AI models, datasets, and tools

CohereLabs/xP3x HF Unverified

Dataset Card for xP3x Dataset Summary xP3x (Crosslingual Public Pool of Prompts eXtended) is a collection of prompts & datasets across 277 languages & 16 NLP tasks. It contains all of xP3 + much more! It is used for training future contenders of mT0 & BLOOMZ at project Aya @Cohere Labs 🧡 Creation: The dataset can be recreated using instructions available here together with the file in this repository named xp3x_create.py. We provide this version to save processing… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/xP3x.

Task_categories:otherAnnotations_creators:expert-GeneratedAnnotations_creators:crowdsourcedMultilinguality:multilingualLanguage:afLanguage:ar
HuggingFaceFW/finephrase HF PQC Verified

Dataset Card for HuggingFaceFW/finephrase Dataset Summary Synthetic data generated by DataTrove: Model: HuggingFaceTB/SmolLM2-1.7B-Instruct (main) Source dataset: HuggingFaceFW/fineweb-edu, config sample-350BT, split train Generation config: temperature=1.0, top_p=1.0, top_k=50, max_tokens=2048, model_max_context=8192 Speculative decoding: {"method":"suffix","num_speculative_tokens":32} System prompt: None Input column: text Prompt families: faq prompt Rewrite the… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceFW/finephrase.

Task_categories:text-GenerationTask_ids:language-ModelingAnnotations_creators:machine-GeneratedLanguage_creators:foundSource_datasets:HuggingFaceFW/fineweb-Edu/sample-350BTLanguage:en
M
mattmdjaga/segformer_b2_clothes HF Unverified

Image-SegmentationTransformersPyTorchONNXSafetensorsSegformer MEDIUM
C
cagliostrolab/animagine-xl-4.0 HF PQC Verified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlBase_model:stabilityai/stable-Diffusion-Xl-Base-1.0 CRITICAL
Kazimir-ai/text-to-image-prompts HF Unverified

The dataset of the most popular text-to-image prompts. Dataset Details Dataset Description Curated by: kazimir.ai Funded by [optional]: [More Information Needed] Shared by [optional]: https://kazimir.ai License: apache-2.0 Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Free to use. Dataset Structure CSV file… See the full description on the dataset page: https://huggingface.co/datasets/Kazimir-ai/text-to-image-prompts.

Language:enSize_categories:10K<n<100KFormat:csvModality:textLibrary:datasetsLibrary:pandas
mvp-lab/LLaVA-OneVision-2-Data HF Unverified

LLaVA-OneVision-2-Data Training data for the LLaVA-OneVision-2 multimodal model family, covering large-scale video and spatial reasoning corpora used in mid-training. Dataset Composition Subset Format Description mid_training_video/60s_rest/ WebDataset (.tar) 10,809 shards of ~60s video clips mid_training_video/caption_v0/split_30s.jsonl JSONL Captions for 30-second video clips mid_training_video/caption_v0/split_60s.jsonl JSONL Captions for… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-2-Data.

Task_categories:video-Text-To-TextTask_categories:visual-Question-AnsweringTask_categories:image-Text-To-TextLanguage:enSize_categories:n<1KFormat:parquet
L
Lykon/DreamShaper HF Unverified

Text-to-ImageDiffusersStable-DiffusionStable-Diffusion-DiffusersArtArtistic CRITICAL
J
jameslahm/yolov10s HF Unverified

Object-DetectionYolov10SafetensorsComputer-VisionPytorch_model_hub_mixin MEDIUM
K
krea/Krea-2-Turbo HF Unverified

Text-to-ImageDiffusersSafetensorsBase_model:krea/Krea-2-RawBase_model:finetune:krea/Krea-2-RawDiffusers:Krea2Pipeline CRITICAL
T
tencent/HunyuanImage-3.0 HF Unverified

Text-to-ImageTransformersSafetensorsHunyuan_image_3_moeText GenerationCustom_code CRITICAL
B
black-forest-labs/FLUX.2-dev HF PQC Verified

Image-To-ImageDiffusersSafetensorsImage GenerationImage-EditingFlux CRITICAL
B
ByteDance/SDXL-Lightning HF Unverified

Text-to-ImageDiffusersStable-Diffusion CRITICAL
P
PramaLLC/BEN2 HF Unverified

Image-SegmentationBen2ONNXSafetensorsBEN2Background-Remove HIGH
J
John6666/nova-furry-xl-il-v120-sdxl HF Unverified

Text-to-ImageDiffusersSafetensorsStable-DiffusionStable-Diffusion-XlNot-For-All-Audiences HIGH
F
facebook/vjepa2-vitg-fpc64-256 HF Unverified

Video-ClassificationTransformersSafetensorsVjepa2Feature ExtractionVideo HIGH
rajpurkar/squad HF Unverified

Dataset Card for SQuAD Dataset Summary Stanford Question Answering Dataset (SQuAD) is a reading comprehension dataset, consisting of questions posed by crowdworkers on a set of Wikipedia articles, where the answer to every question is a segment of text, or span, from the corresponding reading passage, or the question might be unanswerable. SQuAD 1.1 contains 100,000+ question-answer pairs on 500+ articles. Supported Tasks and Leaderboards Question Answering.… See the full description on the dataset page: https://huggingface.co/datasets/rajpurkar/squad.

Task_categories:question-AnsweringTask_ids:extractive-QaAnnotations_creators:crowdsourcedLanguage_creators:crowdsourcedLanguage_creators:foundMultilinguality:monolingual
M
myshell-ai/MeloTTS-Japanese HF Unverified

Text-To-SpeechTransformersKorean MEDIUM
C
cross-encoder/nli-deberta-v3-large HF Unverified

Zero-Shot ClassificationSentence-TransformersPyTorchONNXSafetensorsDeberta-V2 HIGH
openai/openai_humaneval HF Unverified

Dataset Card for OpenAI HumanEval Dataset Summary The HumanEval dataset released by OpenAI includes 164 programming problems with a function sig- nature, docstring, body, and several unit tests. They were handwritten to ensure not to be included in the training set of code generation models. Supported Tasks and Leaderboards Languages The programming problems are written in Python and contain English natural text in comments and docstrings.… See the full description on the dataset page: https://huggingface.co/datasets/openai/openai_humaneval.

Annotations_creators:expert-GeneratedLanguage_creators:expert-GeneratedMultilinguality:monolingualSource_datasets:originalLanguage:enSize_categories:n<1K
jsulz/FIFA23 HF Unverified

About this dataset Context The datasets provided include the players data for the Career Mode from FIFA 15 to FIFA 23. The data allows multiple comparisons for the same players across the last 9 versions of the video game. Some ideas of possible analysis: Historical comparison between Messi and Ronaldo (what skill attributes changed the most during time - compared to real-life stats); Ideal budget to create a competitive team (at the level of top n teams in Europe) and… See the full description on the dataset page: https://huggingface.co/datasets/jsulz/FIFA23.

Task_categories:tabular-ClassificationTask_categories:tabular-RegressionLanguage:enSize_categories:10M<n<100MModality:tabularTabular
Showing 20 of 738 items (page 22 of 37)