ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning Paper • 2603.05863 • Published 18 days ago • 5
ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning Paper • 2603.05863 • Published 18 days ago • 5
nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-FP8 Text Generation • 32B • Updated 9 days ago • 1.45M • • 324
naver-hyperclovax/HyperCLOVAX-SEED-Think-32B Text Generation • 33B • Updated Jan 6 • 34.5k • 396