xiaotong
xtongji
AI & ML interests
None yet
Recent Activity
upvoted a paper about 2 hours ago
Does This Action Still Explain the Task? Reverse Scoring for Diffusion Language Model Agents upvoted a paper about 2 hours ago
Composable Decoding on the Probability Simplex: Theory and Implementation upvoted a paper about 2 hours ago
The Weakest Link: Distilling LLM Reasoning with Worst-Case Constrained Reinforcement LearningOrganizations
None yet