Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
36.9
TFLOPS
ManniX
PRO
ManniX-ITA
35
6
28
Follow
Tetramatr1x's profile picture
ZeraOS's profile picture
DzmitryTheOtherOne's profile picture
97 followers
Β·
20 following
https://github.com/mann1x
mann1x
AI & ML interests
None yet
Recent Activity
posted
an
update
about 17 hours ago
π Qwen3.6-27B-A3B-CoderX β the long-horizon sibling to A3B-Coder. Same 256β184 expert budget (~35Bβ27B, A3B active), different selection: our saliency map picks the keep-set, a REAP-style per-layer floor (p=24) protects the tail, and the 72 evicted experts per layer are folded DERN-style into the survivors instead of discarded. No fine-tuning, no distillation. π Q6_K + imatrix, llama.cpp b9700, greedy, one pinned geometry per bench, same host β CoderX / A3B-Coder / unpruned 256e: β‘ LiveCodeBench v6 (77q, 24k think) β 72.73 / 61.04 / 61.04 β +11.7pp over both β HumanEval+ (164) β 96.95 / 95.12 / 93.90 β best of the three π€ MultiPL-E-100 (rs+java+js) β 88.67 / 89.00 / 91.00 β οΈ Read that last row honestly: a same-basis repeat of MultiPL-E moved 1.0pp on batch-scheduling nondeterminism alone. The 0.33pp CoderXβCoder gap is INSIDE that band β a tie. The 2.33pp gap to the base is outside it and real. CoderX takes Rust (0.85 vs 0.81), gives up JS (0.92 vs 0.96). π― Ships top-8, and that was measured, not assumed: MBPP-full 78.4 / 79.0 at top-8 vs 73.2 / 73.0 at top-10. Opposite call from A3B-Coder, which bakes top-10. π§ It thinks long β LCB median completion ~15.8k tokens vs ~2.2k for Coder. The length is where the win comes from; give it context headroom rather than clamping it. π¬ Not measured yet: the canonical 9-bench. GPQA / MATH-500 / IFEval are deliberately NOT quoted β treat the non-code profile as unknown. Coder remains the one with a published 9-bench table. π¦ bf16 safetensors (text-only) Β· 19 GGUF tiers, EVERY K/I-quant imatrix-built and verified by reading quantize.imatrix.* back out of each uploaded file Β· Ollama 39 tags (19 text + 19 vision-<tier> + :latest). MTP in every tier β draft_num_predict 3 gives 190β252 tok/s (+33%) on an RTX 5080. π https://huggingface.co/ManniX-ITA/Qwen3.6-27B-A3B-CoderX π https://huggingface.co/ManniX-ITA/Qwen3.6-27B-A3B-CoderX-MTP-GGUF π https://ollama.com/mannix/qwen3.6-27b-a3b-coderx
updated
a collection
about 17 hours ago
Qwen-3.6
updated
a collection
about 17 hours ago
Qwen-3.6
View all activity
Organizations
None yet
ManniX-ITA
's datasets
1
Sort:Β Recently updated
ManniX-ITA/osync-code
Viewer
β’
Updated
Jan 12
β’
1
β’
34