Adapters and trajectory artifacts for matched-compute PRM guidance and ORM reranking in discrete diffusion reasoning.
🔄 In a Training Loop
Yan Zhan
YanZhanPKU
·
AI & ML interests
LLM
Recent Activity
liked a dataset about 9 hours ago
YanZhanPKU/dLLM-PRM-Gap-Datasets liked a model about 9 hours ago
YanZhanPKU/dLLM-PRM-Gap-c3a-causal-lasttoken-dream7b-s42 liked a model about 9 hours ago
YanZhanPKU/dLLM-PRM-Gap-c3a-causal-lasttoken-dream7b-s43Organizations
None yet