Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
nm-testing
's Collections
Models in CI
KV Cache Quantization
FP8-Block Quantized Models
LLM Compressor testing
Speculators testing
Sparse-Llama-3.1-8B-2of4
SparseGPT LLMs
FP8 Models
FP8-Block Quantized Models
updated
1 day ago
Collection of State-of-the-art FP8 Block Quantized Models
Upvote
-
Sort: Collection
RedHatAI/Qwen3-8B-FP8-block
Text Generation
•
8B
•
Updated
Dec 31, 2025
•
111
RedHatAI/Qwen3-32B-FP8-block
Text Generation
•
33B
•
Updated
Oct 24, 2025
•
1.09k
RedHatAI/Qwen3-14B-FP8-block
Text Generation
•
15B
•
Updated
Oct 24, 2025
•
67
RedHatAI/Llama-3.1-8B-Instruct-FP8-block
Text Generation
•
8B
•
Updated
Oct 29, 2025
•
1.3k
nm-testing/Qwen3-VL-235B-A22B-Instruct-FP8-BLOCK
Text Generation
•
Updated
Oct 27, 2025
RedHatAI/Llama-3.3-70B-Instruct-FP8-block
Text Generation
•
71B
•
Updated
Oct 24, 2025
•
344
nm-testing/Qwen3-30B-A3B-FP8-block
Text Generation
•
3B
•
Updated
Oct 27, 2025
•
4.19k
Upvote
-
Sort: Collection
Share collection
View history
Collection guide
Browse collections