Alexander Reinthal
reinthal
AI & ML interests
Technical AI safety
Jailbreaking, CyberSecurity Red-teaming with Agents, AI Control
Recent Activity
updated a dataset 23 days ago
reinthal/chinese-llm-sensitive-questions published a dataset 23 days ago
reinthal/chinese-llm-sensitive-questions updated a model 2 months ago
reinthal/qwen3.5-9b-truthmix-v3-lora