심은서
juliann2026
AI & ML interests
None yet
Recent Activity
upvoted a paper 4 days ago
Dynamic Important Example Mining for Reinforcement Finetuning upvoted a paper 25 days ago
DAPD: Dual-Anchored Policy Distillation upvoted a paper 26 days ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement LearningOrganizations
None yet