CalVerT: Augmenting Agents with Calibrated Verifier Telemetry Improves Action and Learning in Knowledge-Intensive Tasks Paper • 2606.21777 • Published Jun 19 • 5
No Hidden Prompts Needed! You Can Game AI Peer Review with Presentation-Only Revisions Paper • 2606.13044 • Published Jun 11 • 11
ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining Paper • 2603.28737 • Published Mar 30
TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning Paper • 2603.12529 • Published Mar 13 • 19
Beyond Test-Time Training: Learning to Reason via Hardware-Efficient Optimal Control Paper • 2603.09221 • Published Mar 10 • 1
nabla-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space Paper • 2603.04948 • Published Mar 5 • 2
EntRGi: Entropy Aware Reward Guidance for Diffusion Language Models Paper • 2602.05000 • Published Feb 4 • 2
view post Post 1655 Are you familiar with reverse residual connections or looping in language models?Excited to share my Looped-GPT blog post and codebase 🚀https://github.com/sanyalsunny111/Looped-GPTTL;DR: looping during pre-training improves generalization.Plot shows GPT2 LMs pre-trained with 15.73B OWT tokensP.S. This is my first post here — I have ~4 followers and zero expectations for reach 😄 See translation 3 replies · 🧠 6 6 👍 3 3 + Reply
Mitigating Catastrophic Forgetting in Mathematical Reasoning Finetuning through Mixed Training Paper • 2512.13706 • Published Dec 5, 2025 • 1