view article Article LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge LiquidAI • 4 days ago • 40
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • 7 days ago • 96
view article Article Making Knowledge Distillation Cheap Enough to Run at Scale MultiverseComputingCAI • 6 days ago • 30
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 19 days ago • 139
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents Paper • 2607.28227 • Published 18 days ago • 303
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Paper • 2607.24904 • Published 21 days ago • 37
view article Article Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community ResterChed • 30 days ago • 195