Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 2 days ago • 158 • 2
Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space Paper • 2608.29188 • Published 7 days ago • 8 • 2
Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments Paper • 2609.04148 • Published 2 days ago • 223 • 1
PACE: Towards Surfacing Hidden Conflicts in User Requests Paper • 2609.03293 • Published 2 days ago • 22 • 2
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published 4 days ago • 107 • 2
Compile by Training: Turning Natural-Language Specifications into Local Neural Functions Paper • 2609.04199 • Published 2 days ago • 271 • 3
Let Confidence Change, Not the Prediction: Prediction-Preserving Repair for Post-hoc Calibration Paper • 2609.01072 • Published 3 days ago • 13 • 2
The Missing Temporal Link: Temporal Context Routing for Script-Driven Audio-Video Generation Paper • 2609.02367 • Published 3 days ago • 33 • 2
QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation Paper • 2608.29253 • Published 7 days ago • 13 • 5
CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation Paper • 2609.04083 • Published 2 days ago • 24 • 2
DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training Paper • 2609.04094 • Published 2 days ago • 23 • 2
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 2 days ago • 204 • 3
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 9 days ago • 149 • 2
Beyond Retrieval: Progressive Latent Memory Evolution for Streaming Video Understanding Paper • 2609.04131 • Published 2 days ago • 29 • 2
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 2 days ago • 73 • 2
A Common Measure of Communication for Speech Brain-Computer Interfaces Paper • 2609.02887 • Published 3 days ago • 8 • 2
Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States Paper • 2609.04196 • Published 2 days ago • 65 • 3
Select, Compress, Reinvest: A Controlled Study of Visual-Token Allocation in Long-Video MLLMs Paper • 2609.03820 • Published 2 days ago • 15 • 3