Seong Joon Oh
coallaoh
AI & ML interests
Scalable Trustworthy AI
https://scalabletrustworthyai.github.io/
Recent Activity
upvoted a paper about 15 hours ago
Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR