AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model Paper • 2604.19747 • Published Apr 21 • 41
GenConViT: Deepfake Video Detection Using Generative Convolutional Vision Transformer Paper • 2307.07036 • Published Jul 13, 2023 • 1
Veritas: Generalizable Deepfake Detection via Pattern-Aware Reasoning Paper • 2508.21048 • Published Feb 27 • 1
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published 16 days ago • 280
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 20 days ago • 763
Gated Recurrent Transformers: Expressive Depth through Recurrent Modulation Paper • 2608.15062 • Published 4 days ago • 10
Task-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection Paper • 2608.20169 • Published 6 days ago • 11
UniSpace: Unified Visual Representation and Scalable Multimodal Modeling Paper • 2608.08676 • Published 21 days ago • 14
V-Rubrics: Visual Faithfulness via Rubric-Based Reinforcement Learning Paper • 2608.25580 • Published 4 days ago • 15
Agent-G^2: Gaussian Guidance for Agentic Reinforcement Learning Paper • 2608.23318 • Published 6 days ago • 25
Block3D: Efficient Text-to-3D Generation via Block-Wise Diffusion Paper • 2608.19567 • Published 10 days ago • 30
CyberFactory: Scaling Cyber Security Capabilities with Instances from the Wild Paper • 2608.23181 • Published 6 days ago • 33
KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation Paper • 2607.14202 • Published Jul 15 • 43
AsySplat: Efficient Asymmetric 3D Gaussian Splatting for Long-Sequence Scene Modeling Paper • 2607.10995 • Published Jul 13 • 10
TimeChat-Captioner: Scripting Multi-Scene Videos with Time-Aware and Structural Audio-Visual Captions Paper • 2602.08711 • Published Feb 9 • 29
SemanticMoments: Training-Free Motion Similarity via Third Moment Features Paper • 2602.09146 • Published Feb 9 • 22