andresnowak/qwen38-27b-math-grpo-baseline-drgrpo-step100 Reinforcement Learning • 27B • Updated 7 days ago • 23
andresnowak/qwen38-27b-math-grpo-monitor-drgrpo-step100 Reinforcement Learning • 27B • Updated 7 days ago • 36
andresnowak/qwen38-27b-math-grpo-baseline-drgrpo-step100 Reinforcement Learning • 27B • Updated 7 days ago • 23
andresnowak/qwen38-27b-math-grpo-monitor-drgrpo-step100 Reinforcement Learning • 27B • Updated 7 days ago • 36
XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B Image-Text-to-Text • 9B • Updated 9 days ago • 12.1k • 580
Running Reproduction: GCIB: Graph Contrastive Information Bottleneck for Multi-Behavior Recommendation 🎯 Explore project logs, code, and traces in an interactive web logbook
Running Reproduction: GCIB: Graph Contrastive Information Bottleneck for Multi-Behavior Recommendation 🎯 Explore project logs, code, and traces in an interactive web logbook
Running Reproduction: CSG: Cognitive Structure Generation for Intelligent Education 🎯 Explore code, traces, and workspace in an interactive logbook
Running Reproduction: CSG: Cognitive Structure Generation for Intelligent Education 🎯 Explore code, traces, and workspace in an interactive logbook