Run 3: Qwen2.5-0.5B GRPO 500 steps | reward 0.96→1.75 | conflict 0.375→0.1875 | dashboard chart | agents wired 0a0a4fc smiit commited on Apr 25
remove filter: let LLM produce natural conflicts for LLM vs LoRA comparison be71a02 smiit commited on Apr 25
fix: enforce sensor ownership server-side, drop cross-agent proposals 4101a93 smiit commited on Apr 25
fix: restrict LLM proposals to agent-type sensors, prevent S1 conflicts 2dfcf50 smiit commited on Apr 25
Fix: task difficulty counts, dashboard legend modal, training metrics, reward curve, /metrics/history endpoint cfa1b2b smiit commited on Apr 25
fix: implement Phase 1 core fixes (capability matrix, agent wiring, multi-agent inference, robust observation handling) f6ad663 chathurvitha commited on Apr 22
feat: integrate trained LoRA adapters into inference and server, enabling real multi-agent learning deployment c47b19a chathurvitha commited on Apr 22
Fix all audit issues: conflict schema, curriculum wiring, density_factor, eval path, dead files removed a7cb63d smiit commited on Apr 22
feat: activate full multi-agent pipeline with negotiation, conflict handling, and GRPO training 1c97666 chathurvitha commited on Apr 22
feat: activate full multi-agent pipeline with negotiation, conflict handling, and GRPO training 949a94e chathurvitha commited on Apr 22
Fix LLM multi-agent proposals: agent-type sensor filtering, prompt improvements, debug logging daaa149 smiit commited on Apr 22
Setup project environment, fixed Python version issues, added venv, installed dependencies, and configured server execution ca7e89f chathurvitha commited on Apr 21
feat: implement dashboard UI and integrate uv dependency management 4a8a672 Vizxal commited on Mar 30
feat: LLM agent, Leaflet map, batch step, reward fix, inference all tasks 6025d24 smiit commited on Mar 29