SAKI: Maximal-Coupling-Routed Teacher Supervision for On-Policy Distillation Paper • 2609.36601 • Published 6 days ago • 94
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 12 days ago • 23
APM-Bench: Benchmarking Cross-session Persistent Memory for Egocentric Streaming Video Assistants Paper • 2609.37559 • Published 6 days ago • 46
Omni-IO Skills: Harnessing Your Agent Omni-Native Paper • 2609.31847 • Published 10 days ago • 292
YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality Paper • 2609.33757 • Published 8 days ago • 244
CoWindow Attention: Full Causal Coverage Is a Collective Property Paper • 2609.32704 • Published 9 days ago • 67
Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching Paper • 2608.09444 • Published 10 days ago • 13
DataoceanAI/Chinese_Female_Speech_Synthesis_Corpus_Live_Streaming_for_Sales Updated Jan 10, 2025 • 69 • 8
PierrunoYT/higgs-audio-v2-generation-3B-base Text-to-Speech • 6B • Updated Jul 28, 2025 • 51 • 4
ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds Paper • 2609.30199 • Published 11 days ago • 29
VladS159/common_voice_16_1_romanian_speech_synthesis Viewer • Updated Feb 26, 2024 • 39.1k • 109 • 6
Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 12 days ago • 56
szhengac25/higgs-audio-v2-generation-3B-base Text-to-Speech • 6B • Updated Aug 29, 2025 • 57 • 6
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 14 days ago • 103