ChipHolmes/Code_Vulnerability_Security_DPO-archive Viewer • Updated about 19 hours ago • 4.66k • 9 • 1
ChipHolmes/All-CVE-Records-Training-Dataset-archive Viewer • Updated about 19 hours ago • 297k • 10 • 1
mradermacher/GPT-OSS-Cybersecurity-20B-Merged-i1-GGUF Text Generation • 21B • Updated Dec 7, 2025 • 1.43k • 9
jason-oneal/mitre-stix-cve-exploitdb-dataset-alpaca-chatml-harmony Viewer • Updated May 2 • 1.93M • 156 • 9
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 7 days ago • 27
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 6 days ago • 144
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources Paper • 2606.29538 • Published 13 days ago • 141
Long-Horizon-Terminal-Bench: Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading Paper • 2607.08964 • Published 20 days ago • 76