Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper • 2609.00111 • Published 6 days ago • 371
Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition Paper • 2601.16211 • Published Jul 2 • 54
Running Agents 31 Physical AI Bench Leaderboard 🤖 31 Benchmark for Physical AI generation and understanding
naver-hyperclovax/HyperCLOVAX-SEED-Vision-Instruct-3B Text Generation • 4B • Updated Sep 16, 2025 • 11.5k • 221
stabilityai/stable-diffusion-xl-base-1.0 Text-to-Image • 3B • Updated Oct 30, 2023 • 1.79M • • 8.11k