lv Solex
Solex0259
ยท
AI & ML interests
Robust Machine Learning
Recent Activity
upvoted a paper 7 minutes ago
CAST: Game Solvers as Turn-Level Teachers for LLM Agents upvoted a paper 9 minutes ago
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization upvoted a paper 9 months ago
A Theoretical Study on Bridging Internal Probability and
Self-Consistency for LLM ReasoningOrganizations
None yet