FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations Paper • 2609.20817 • Published 15 days ago • 40
TIPSv2 Collection TIPSv2 foundational vision-language models. Webpage: https://gdm-tipsv2.github.io/ • 9 items • Updated Jul 21 • 45
view article Article Gotchas in Tokenizer Behavior Every Developer Should Know qgallouedec • Apr 18, 2025 • 72
view article Article cocogold: training Marigold for text-grounded segmentation pcuenq • Jul 8, 2025 • 30
Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion Paper • 2412.13389 • Published Dec 18, 2024 • 7