AryaX / server.py

Commit History

Keep mx_obs after episode done, prevent reset_multi first error
53a8778

smiit commited on

Fix API_BASE_URL to correct HF inference /v1 endpoint
892a583

smiit commited on

Fix HF Inference API URL to include model path /v1
6f64e17

smiit commited on

Fix API_BASE_URL - remove duplicate /v1
4588d74

smiit commited on

Switch to HF Inference API + zephyr-7b (no provider needed)
0445d40

smiit commited on

Switch to Phi-3.5-mini-instruct (free HF router)
9cbe0b9

smiit commited on

Switch to Llama-3.2-3B-Instruct (supported by HF router)
ccce53a

smiit commited on

Add timeout+stderr logging to LLM calls, reduce max_tokens to 64
3e1d372

smiit commited on

Log to stderr + gunicorn capture-output for HF diagnosis
f9c6cbc

smiit commited on

Cache bust + debug env print for HF token diagnosis
b23f827

smiit commited on

Fix checkpoint path + default model to Qwen2.5-0.5B for HF API
fc9640e

smiit commited on

Run 3: Qwen2.5-0.5B GRPO 500 steps | reward 0.96→1.75 | conflict 0.375→0.1875 | dashboard chart | agents wired
0a0a4fc

smiit commited on

add /game route for game mode
0729b9c

smiit commited on

fix: clean up auto_multi, set mx_obs=None on done
1acf997

smiit commited on

remove filter: let LLM produce natural conflicts for LLM vs LoRA comparison
be71a02

smiit commited on

debug max_steps + safety counter in runAllMulti
29f3ccc

smiit commited on

debug: log filtered proposals before step_multiagent
d7e15ea

smiit commited on

fix: enforce sensor ownership server-side, drop cross-agent proposals
4101a93

smiit commited on

fix: restrict LLM proposals to agent-type sensors, prevent S1 conflicts
2dfcf50

smiit commited on

add lora checkpoints and training updates
099deaf

smiit commited on

fix: rewards, conflict detection, training comparison panel
f114961

smiit commited on

Fix: task difficulty counts, dashboard legend modal, training metrics, reward curve, /metrics/history endpoint
cfa1b2b

smiit commited on

fix: implement Phase 1 core fixes (capability matrix, agent wiring, multi-agent inference, robust observation handling)
f6ad663

chathurvitha commited on

Merge branch 'main' into chathu01
8be5cf6
unverified

chathurvitha commited on

feat: integrate trained LoRA adapters into inference and server, enabling real multi-agent learning deployment
c47b19a

chathurvitha commited on

Fix all audit issues: conflict schema, curriculum wiring, density_factor, eval path, dead files removed
a7cb63d

smiit commited on

Merge pull request #8 from smritis21/chathu01
c67e73a
unverified

chathurvitha commited on

feat: activate full multi-agent pipeline with negotiation, conflict handling, and GRPO training
1c97666

chathurvitha commited on

feat: activate full multi-agent pipeline with negotiation, conflict handling, and GRPO training
949a94e

chathurvitha commited on

Fix LLM multi-agent proposals: agent-type sensor filtering, prompt improvements, debug logging
daaa149

smiit commited on

Setup project environment, fixed Python version issues, added venv, installed dependencies, and configured server execution
ca7e89f

chathurvitha commited on

error fixed
7470a2a

Vizxal commited on

submission ready
4835868

Vizxal commited on

feat: implement dashboard UI and integrate uv dependency management
4a8a672

Vizxal commited on

feat: serve dashboard at / as well as /ui
7654594

smiit commited on

feat: LLM agent, Leaflet map, batch step, reward fix, inference all tasks
6025d24

smiit commited on

improve: better prompt + fixed grader normalization, score 0.77
b360381

smiit commited on

Round 1 submission — SentinelEnv OpenEnv
dd82d4f

smiit commited on