SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 8 days ago • 97
ResearchMath-14K: Scaling Research-Level Mathematics via Agents Paper • 2605.28003 • Published May 27 • 50
How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data Paper • 2604.14164 • Published Mar 23 • 35
Running 247 MedGemma - Radiology Explainer Demo 🩺 247 Radiology Image & Report Explainer Demo. Built with MedGemma
Running on CPU Upgrade Agents 10.1k Kolors Virtual Try-On 👕 10.1k Generate virtual try‑on images of a person wearing a chosen garment
Running on Zero MCP Featured 2.03k Stable Video Diffusion 1.1 📺 2.03k Generate a short video from a single image