ZipTok3D: High-Fidelity 3D Tokenization with Compact Token Prefixes Paper • 2609.01740 • Published 19 days ago • 28
FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation Paper • 2605.09430 • Published May 10 • 2
Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion Paper • 2608.19567 • Published Aug 20 • 33
PolicyTrim: Boosting Intrinsic Policy Efficiency of Vision-Language-Action Models Paper • 2606.22540 • Published Jun 21 • 5
Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization Paper • 2605.15980 • Published May 15 • 35
Less Detail, Better Answers: Degradation-Driven Prompting for VQA Paper • 2604.04838 • Published Apr 6 • 11
Less is More: Improving LLM Reasoning with Minimal Test-Time Intervention Paper • 2510.13940 • Published Oct 15, 2025 • 7
Emerging Properties in Unified Multimodal Pretraining Paper • 2505.14683 • Published May 20, 2025 • 136
Neighboring Autoregressive Modeling for Efficient Visual Generation Paper • 2503.10696 • Published Mar 12, 2025 • 8
NAR Collection Neighboring Autoregressive Modeling for Efficient Visual Generation • 10 items • Updated Mar 17, 2025 • 3
ZipVL: Efficient Large Vision-Language Models with Dynamic Token Sparsification and KV Cache Compression Paper • 2410.08584 • Published Oct 11, 2024 • 12