VoiceMem: Streaming Dual-Brain Memory for Real-Time Interaction Paper • 2608.26005 • Published 12 days ago • 177
Self-OPD: On-Policy Distillation for Flow Matching Models without Teacher Paper • 2608.26872 • Published 11 days ago • 82
JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution Paper • 2608.25593 • Published 12 days ago • 117
view article Article Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers tomaarsen • 12 days ago • 120
WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report Paper • 2608.24053 • Published 13 days ago • 70
QUASAR: Lowering the Loss Floor of Quantization-Aware Training with Loss-Aware Reconstruction Paper • 2608.13966 • Published 24 days ago • 3
Apodex 1.1: Scaling Agentic Intelligence for Complex Work Paper • 2608.23283 • Published 14 days ago • 205
view article Article Measuring benchmark optimization in speech recognition +5 tlebryk02, bezzam, aliceebaird, dayllon, jpc, jens-hume-ai, tzirakis • 17 days ago • 65
view article Article Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers +1 tomaarsen, NohTow, raphaelsty • 20 days ago • 106
view article Article State of Open Models: Summer 2026 Observations +1 AdinaY, multimodalart, irenesolaiman • 24 days ago • 186
view article Article LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge LiquidAI • 25 days ago • 50
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • 28 days ago • 110
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • 28 days ago • 40
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published Jul 29 • 140
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published Jul 30 • 311