Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
antirez
/
deepseek-v4-gguf
like
412
Text Generation
GGUF
English
quantized
deepseek
deepseek-v4
deepseek-v4-flash
Mixture of Experts
mixture-of-experts
2-bit
4-bit precision
iq2_xxs
q2_k
q4_k
ds4
apple-silicon
metal
License:
mit
Model card
Files
Files and versions
xet
Community
20
Copy to bucket
new
TemporalMesh Transformer: 29.4 PPL at 48% compute — beats Mamba, new open-source architecture
#12
by
vigneshwar234
- opened
Jun 7
Discussion
vigneshwar234
Jun 7
This comment has been hidden (marked as Spam)
Edit
Preview
Upload images, audio, and videos by dragging in the text input, pasting, or
clicking here
.
Tap or paste here to upload images
Comment
·
Sign up
or
log in
to comment