AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks Paper • 2609.38288 • Published 9 days ago • 140
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 10 days ago • 167
m3hrdadfi/wav2vec2-xlsr-greek-speech-emotion-recognition Automatic Speech Recognition • Updated Jul 6, 2021 • 57 • 14
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 15 days ago • 24
The Past Frames the Future: Memory for Autoregressive Video Generation Paper • 2609.28466 • Published 15 days ago • 64
speechbrain/emotion-recognition-wav2vec2-IEMOCAP Audio Classification • Updated Jul 23, 2024 • 48.1k • 197
Geometric and Semantic Coupling for Interaction Understanding in 3D Scenes Paper • 2609.25247 • Published 17 days ago • 10
From Pattern Recognizers to Personalized Companions: A Survey of Large Language Models in Mental Health Paper • 2609.25186 • Published 17 days ago • 32
xbgoose/hubert-large-speech-emotion-recognition-russian-dusha-finetuned Audio Classification • 0.3B • Updated Apr 26, 2024 • 236k • 20
WorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory Paper • 2609.24984 • Published 17 days ago • 157