Model Hub
Browse PQC-verified AI models, datasets, and tools
DeepSeek's reasoning model. Chain-of-thought reasoning with 671B MoE architecture, rivaling frontier closed models.
Subset of LAION-5B filtered for aesthetic quality. 600M image-text pairs scored >5.0 by aesthetic predictor. Standard for image generation training.
FineWeb Tokenized > 4 trillion tokens of the pre-tokenized data the š web has to offer What is it? This is a pre-tokenized version of the HuggingFaceFW/fineweb dataset (currently in-progress, tokenization of the ~15 trillion tokens corpus is ongoing). The data is being pre-processed and tokenized using the AnisoleAI BPE tokenizer (52,022 vocabulary size) and packed into compact uint16 Parquet shards. By distributing the pre-tokenized corpus, we eliminate⦠See the full description on the dataset page: https://huggingface.co/datasets/anisoleai/fineweb-tokenized.
Meta's Llama 3.1 8B parameter instruction-tuned model. Optimized for dialogue and instruction following with 128K context.
Distilled Whisper model. 809M parameters ā 6x faster than large-v3 with minimal quality loss.
OpenAI's first open-source model release. 20B parameter GPT architecture trained on diverse web data.