Malandrino PRO
Pier-Jean
AI & ML interests
Document parsing
Quantization
Recent Activity
authored a paper 5 days ago
Unfolding the Leech Lattice: Fused Multi-Shell Decoding and VRAM Layouts for 2-Bit LLM Weights published an article 10 days ago
I put a 24 dimensional lattice in a CUDA kernel to run Qwen3-4B in 2.6 GB updated a model 21 days ago
Pier-Jean/Qwen3-4B-LLVQ-2bit