FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving Paper • 2608.19758 • Published 3 days ago • 16
AVA-Encoder: Towards Agent-Native Video Representation Learning Paper • 2608.12313 • Published 12 days ago • 40
ASI-Bench: At the Dawn of Artificial Superintelligence Paper • 2608.17271 • Published 6 days ago • 60
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published 7 days ago • 148
MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling Paper • 2608.14783 • Published 10 days ago • 19
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published 12 days ago • 30
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 14 days ago • 338
Skaling: Chinchilla's Exponents Meet Kaplan's Coupling Paper • 2608.07222 • Published 17 days ago • 10
Sol-Attn: Accelerating Video Generation Inference via On-the-Fly Attention Sparsification Paper • 2607.24027 • Published 28 days ago • 37