Add model card

#1
by nielsr HF Staff - opened
Files changed (1) hide show
  1. README.md +10 -0
README.md ADDED
@@ -0,0 +1,10 @@
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: transformers
3
+ pipeline_tag: text-classification
4
+ ---
5
+
6
+ This repository contains the judge model described in [Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory](https://huggingface.co/papers/2605.03228).
7
+
8
+ ShadowMem is a defensive framework that maintains a dedicated, safety-focused agentic memory — inspired by the "shadow stack" abstraction in systems security — to distill and retain safety-critical context across an agent's full execution trajectory, and uses this shadow memory to proactively assess the risk of pending actions prior to their execution. This model is a fine-tuned Qwen3-based language model used as the safety judge that produces those risk assessments.
9
+
10
+ Code: https://github.com/ZJUWYH/ShadowMem