iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning Paper • 2509.19552 • Published Dec 5, 2025
CAViAR: A Causal Video Dataset for Fine-Grained Accident Reasoning in Real-World Scenarios Paper • 2608.19380 • Published 26 days ago
AutoScape: Geometry-Consistent Long-Horizon Scene Generation Paper • 2510.20726 • Published Oct 23, 2025
Running on Zero MCP Featured 3.48k Wan2.2 14B Fast 🎥 3.48k generate a video from an image with a text prompt
Learning Semantic Segmentation from Multiple Datasets with Label Shifts Paper • 2202.14030 • Published Feb 28, 2022
Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations Paper • 2508.02047 • Published Aug 4, 2025 • 1
MM-TTA: Multi-Modal Test-Time Adaptation for 3D Semantic Segmentation Paper • 2204.12667 • Published Apr 27, 2022
Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations Paper • 2508.02047 • Published Aug 4, 2025
AIDE: An Automatic Data Engine for Object Detection in Autonomous Driving Paper • 2403.17373 • Published Mar 26, 2024
sparshgarg57/swin-tiny-patch4-window7-224-finetuned-birdclef Image Classification • 27.7M • Updated Mar 21, 2025 • 6
sparshgarg57/swin-tiny-patch4-window7-224-finetuned-birdclef Image Classification • 27.7M • Updated Mar 21, 2025 • 6