심은서
juliann2026
AI & ML interests
None yet
Recent Activity
upvoted a paper 3 days ago
Dynamic Important Example Mining for Reinforcement Finetuning upvoted a paper 25 days ago
DAPD: Dual-Anchored Policy Distillation upvoted a paper 25 days ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement LearningOrganizations
None yet