Unfolding the Leech Lattice: Fused Multi-Shell Decoding and VRAM Layouts for 2-Bit LLM Weights Paper • 2609.02652 • Published 6 days ago
view article Article I put a 24 dimensional lattice in a CUDA kernel to run Qwen3-4B in 2.6 GB Pier-Jean • 9 days ago
view article Article Docling Studio — Open-Source Visual Inspection for Docling Pipelines Pier-Jean • Apr 5
ibm-granite/granite-docling-258M Image-Text-to-Text • 0.3B • Updated Sep 23, 2025 • 210k • 1.26k
docling-project/SmolDocling-256M-preview Image-Text-to-Text • 0.3B • Updated Sep 17, 2025 • 25.6k • 1.62k