FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations Paper • 2609.20817 • Published 3 days ago • 17
TIPSv2 Collection TIPSv2 foundational vision-language models. Webpage: https://gdm-tipsv2.github.io/ • 9 items • Updated Jul 21 • 44
view article Article Gotchas in Tokenizer Behavior Every Developer Should Know qgallouedec • Apr 18, 2025 • 72
view article Article cocogold: training Marigold for text-grounded segmentation pcuenq • Jul 8, 2025 • 30
Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion Paper • 2412.13389 • Published Dec 18, 2024 • 7