arxiv:2606.02404
Seungone Kim PRO
seungone
AI & ML interests
Large Language Models, LLM-as-a-Judge, Reward Model Overoptimization, Personalized Alignment
Recent Activity
upvoted a paper about 2 hours ago
SPADE: Self-Play in Adaptive Synthetic Executable Environments upvoted an article 20 days ago
Bringing Fusion Down to Earth: ML for Stellarator Optimization upvoted a paper about 1 month ago
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL