TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 8 days ago • 139
A Frozen 12B Beats Frontier Models on Verified Work: 100% Accuracy, 0 Tokens, Bit-Exact, Forever Paper • 2607.23806 • Published 11 days ago • 7
text to image Collection ai models that create images from text prompts • 7 items • Updated 1 day ago • 3
MiniCPM-V 4.6 Collection A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone • 11 items • Updated May 24 • 12