LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published 7 days ago • 195
SeKV: Resolution-Adaptive KV Cache with Hierarchical Semantic Memory for Long-Context LLM Inference Paper • 2606.31145 • Published 23 days ago • 11
timaeus/rl-lm-pythia160m-sentiment-pos-beta0-grpo-nostd-gs4-tp1-tk0-pt80000-steerDotIncL1c512s16-seed21 Updated 20 days ago • 2 • 1
DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization Paper • 2605.31455 • Published May 29 • 6
Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality? Paper • 2605.22109 • Published May 21 • 171