PsychoXiong
PsychoO
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
TROJail: Trajectory-Level Optimization for Multi-Turn Large Language Model Jailbreaks with Process Rewards upvoted a paper about 2 months ago
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security upvoted a paper 10 months ago
Quantile Advantage Estimation for Entropy-Safe ReasoningOrganizations
None yet