Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
36.9
TFLOPS
ManniX
PRO
ManniX-ITA
35
6
28
Follow
JohnRoger's profile picture
leiqiang's profile picture
bujchen's profile picture
98 followers
·
20 following
https://github.com/mann1x
mann1x
AI & ML interests
None yet
Recent Activity
posted
an
update
1 day ago
🚀 Qwen3.6-27B-A3B-CoderX — the long-horizon sibling to A3B-Coder. Same 256→184 expert budget (~35B→27B, A3B active), different selection: our saliency map picks the keep-set, a REAP-style per-layer floor (p=24) protects the tail, and the 72 evicted experts per layer are folded DERN-style into the survivors instead of discarded. No fine-tuning, no distillation. 📊 Q6_K + imatrix, llama.cpp b9700, greedy, one pinned geometry per bench, same host — CoderX / A3B-Coder / unpruned 256e: ⚡ LiveCodeBench v6 (77q, 24k think) — 72.73 / 61.04 / 61.04 → +11.7pp over both ✅ HumanEval+ (164) — 96.95 / 95.12 / 93.90 → best of the three 🤝 MultiPL-E-100 (rs+java+js) — 88.67 / 89.00 / 91.00 ⚠️ Read that last row honestly: a same-basis repeat of MultiPL-E moved 1.0pp on batch-scheduling nondeterminism alone. The 0.33pp CoderX↔Coder gap is INSIDE that band — a tie. The 2.33pp gap to the base is outside it and real. CoderX takes Rust (0.85 vs 0.81), gives up JS (0.92 vs 0.96). 🎯 Ships top-8, and that was measured, not assumed: MBPP-full 78.4 / 79.0 at top-8 vs 73.2 / 73.0 at top-10. Opposite call from A3B-Coder, which bakes top-10. 🧠 It thinks long — LCB median completion ~15.8k tokens vs ~2.2k for Coder. The length is where the win comes from; give it context headroom rather than clamping it. 🔬 Not measured yet: the canonical 9-bench. GPQA / MATH-500 / IFEval are deliberately NOT quoted — treat the non-code profile as unknown. Coder remains the one with a published 9-bench table. 📦 bf16 safetensors (text-only) · 19 GGUF tiers, EVERY K/I-quant imatrix-built and verified by reading quantize.imatrix.* back out of each uploaded file · Ollama 39 tags (19 text + 19 vision-<tier> + :latest). MTP in every tier — draft_num_predict 3 gives 190→252 tok/s (+33%) on an RTX 5080. 🔗 https://huggingface.co/ManniX-ITA/Qwen3.6-27B-A3B-CoderX 🔗 https://huggingface.co/ManniX-ITA/Qwen3.6-27B-A3B-CoderX-MTP-GGUF 🔗 https://ollama.com/mannix/qwen3.6-27b-a3b-coderx
updated
a collection
1 day ago
Qwen-3.6
updated
a collection
1 day ago
Qwen-3.6
View all activity
Organizations
None yet
ManniX-ITA
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
ManniX-ITA/Qwen3.6-27B-A3B-Coder
3 days ago
Frequency-rnorm correlations
6
#2 opened 18 days ago by
sfrav
I really like what I'm seeing here
6
#1 opened about 1 month ago by
sometimesanotion
New activity in
LiquidAI/LFM2.5-2.6B
16 days ago
Speculative Decoding
🔥
1
4
#6 opened 17 days ago by
Harpere
New activity in
ManniX-ITA/gemma-4-A4B-98e-v7-coder-it-GGUF
24 days ago
Gets stuck in loops using llama.cpp
21
#1 opened 3 months ago by
DPS-900
New activity in
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4-MTP-GGUF
26 days ago
My life-saver: ManniX-ITA/Qwen3.6-27B-Omnimerge-v4-MTP-GGUF.
1
#2 opened 27 days ago by
VictorDerbobo
commented
a paper
26 days ago
Group Entropy-Controlled Policy Optimization
Paper
•
2607.16850
•
Published
Jul 18
•
29
•
3
New activity in
google/gemma-4-12B-it
30 days ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
8
#38 opened 2 months ago by
ManniX-ITA
New activity in
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4-MTP-GGUF
about 2 months ago
My experience
👍
2
10
#1 opened 3 months ago by
tooltd
New activity in
ManniX-ITA/gemma-4-A4B-98e-v7-coder-it-GGUF
2 months ago
Repeated pauses in OpenCode
11
#2 opened 2 months ago by
mohkamfer
New activity in
google/gemma-4-12B-it
2 months ago
gemma-4-12B-it: deterministic "thought\n thought\n …" degenerate loop on long agent prompts (~60% reproducer, 4-bit, repros across temperatures)
9
#41 opened 2 months ago by
Raullen
New activity in
bartowski/google_gemma-4-26B-A4B-it-GGUF
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
#6 opened 2 months ago by
ManniX-ITA
New activity in
unsloth/gemma-4-26B-A4B-it-GGUF
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
#43 opened 2 months ago by
ManniX-ITA
New activity in
unsloth/gemma-4-26B-A4B-it
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
#2 opened 2 months ago by
ManniX-ITA
New activity in
google/gemma-4-E2B-it
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
#36 opened 2 months ago by
ManniX-ITA
New activity in
google/gemma-4-E4B-it
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
#37 opened 2 months ago by
ManniX-ITA
New activity in
google/gemma-4-31B-it
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
#119 opened 2 months ago by
ManniX-ITA
New activity in
google/gemma-4-26B-A4B-it
2 months ago
Chat template may re-inject prior-turn reasoning during multi-turn tool use → repetition loops
2
#48 opened 2 months ago by
ManniX-ITA
New activity in
ManniX-ITA/gemma-4-A4B-98e-v6-coder-it-GGUF
3 months ago
Are there any plans for the qwen3.6-35B model?
3
#1 opened 3 months ago by
jian2023
New activity in
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4-GGUF
3 months ago
MTP version
3
#1 opened 3 months ago by
DzmitryTheOtherOne
New activity in
ManniX-ITA/Qwen3.6-27B-Omnimerge-v4
3 months ago
can give a mlx 4bit version?
2
#1 opened 3 months ago by
zwqjoy
Load more