Run 3: Qwen2.5-0.5B GRPO 500 steps | reward 0.96→1.75 | conflict 0.375→0.1875 | dashboard chart | agents wired 0a0a4fc smiit commited on Apr 25
feat: improve UI visualization and training evaluation with smooth transitions and before/after metrics c334ad7 chathurvitha commited on Apr 25
Fix: task difficulty counts, dashboard legend modal, training metrics, reward curve, /metrics/history endpoint cfa1b2b smiit commited on Apr 25
Setup project environment, fixed Python version issues, added venv, installed dependencies, and configured server execution ca7e89f chathurvitha commited on Apr 21
feat: implement dashboard UI and integrate uv dependency management 4a8a672 Vizxal commited on Mar 30
feat: LLM agent, Leaflet map, batch step, reward fix, inference all tasks 6025d24 smiit commited on Mar 29