Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
26.2
TFLOPS
ddh0
ddh0
93
107
1080
Follow
ubergarm's profile picture
TheBloke's profile picture
KatyTheCutie's profile picture
171 followers
·
103 following
ddh0
AI & ML interests
None yet
Recent Activity
reacted
to
eaddario
's
post
with ❤️
about 12 hours ago
Experimental global target bits‑per‑weight quantization of **XHToken/Spark-X2.5-1.7B** and **XHToken/Spark-X2.5-4B**. Unlike standard llama.cpp quantization that rely on fixed type heuristics (e.g., Q4_K_M), the Target BPW approach automatically optimizes per-tensor precision where it matters the most, and produces high quality models that meet a precise global size target. Key Advantages: - VRAM Maximization: Can generate high quality models sized exactly to fit hardware constraints (e.g., fitting the model into exactly 24GB VRAM). - Data-Driven Precision: Quantization mix is determined by actual weight error sensitivity rather than hardcoded rules, often yielding better PPL/KLD size trade-offs. Full benchmarks (PPL, KLD, ARC, GPQA, MMLU, etc.) and methodology in the model's card. https://huggingface.co/eaddario/Spark-X2.5-1.7B-GGUF https://huggingface.co/eaddario/Spark-X2.5-4B-GGUF
reacted
to
eaddario
's
post
with 🔥
about 12 hours ago
Experimental global target bits‑per‑weight quantization of **XHToken/Spark-X2.5-1.7B** and **XHToken/Spark-X2.5-4B**. Unlike standard llama.cpp quantization that rely on fixed type heuristics (e.g., Q4_K_M), the Target BPW approach automatically optimizes per-tensor precision where it matters the most, and produces high quality models that meet a precise global size target. Key Advantages: - VRAM Maximization: Can generate high quality models sized exactly to fit hardware constraints (e.g., fitting the model into exactly 24GB VRAM). - Data-Driven Precision: Quantization mix is determined by actual weight error sensitivity rather than hardcoded rules, often yielding better PPL/KLD size trade-offs. Full benchmarks (PPL, KLD, ARC, GPQA, MMLU, etc.) and methodology in the model's card. https://huggingface.co/eaddario/Spark-X2.5-1.7B-GGUF https://huggingface.co/eaddario/Spark-X2.5-4B-GGUF
updated
a model
about 18 hours ago
ddh0/imatrices
View all activity
Organizations
models
90
Sort: Recently updated
ddh0/DeepSeek-V4-Flash-Vision-Exp-Uncensored-GGUF
284B
•
Updated
about 16 hours ago
•
310
ddh0/imatrices
14.2M
•
Updated
about 18 hours ago
•
74
•
1
ddh0/GLM-5.3-Flash-GGUF
321B
•
Updated
4 days ago
•
1.47k
•
2
ddh0/DeepSeek-V4-Flash-Vision-Exp-GGUF
284B
•
Updated
13 days ago
•
609
•
1
ddh0/Qwen3.8-27B-GGUF
27B
•
Updated
Aug 15
•
338
ddh0/data
Updated
Aug 14
ddh0/Muse-Glimmer-30B-GGUF
28B
•
Updated
Aug 10
•
792
•
1
ddh0/DeepSeek-V4-Flash-GGUF
284B
•
Updated
Aug 7
•
613
•
7
ddh0/DeepSeek-V4-Flash-0731-GGUF
284B
•
Updated
Aug 3
•
92
•
3
ddh0/MiniMax-M2.5-GGUF
229B
•
Updated
Jun 23
•
18
View 90 models
datasets
0
None public yet