-
CyberSecEvalTest
📈73Evaluate LLMs' cybersecurity risks and capabilities
-
meta-llama/Llama-Guard-3-8B
Text Generation • 8B • Updated • 45.6k • • 329 -
meta-llama/Prompt-Guard-86M
Text Classification • 0.3B • Updated • 4.07M • • 413 -
protectai/deberta-v3-base-prompt-injection-v2
Text Classification • 0.2B • Updated • 849k • • 117
🤝 Open to Collab
Shyam Sunder Kumar
theainerd
AI & ML interests
Natural Language Processing
Recent Activity
liked a Space 4 days ago
huggingface/whos-shipping-open-source-ai liked a model 4 days ago
Qwen/Qwen-Image-2.1 liked a dataset 9 days ago
nasa-ibm-ai4science/Surya-bench-solarwindOrganizations
Agents
-
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
Search-o1: Agentic Search-Enhanced Large Reasoning Models
Paper • 2501.05366 • Published • 106 -
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 107 -
Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Paper • 2501.10893 • Published • 26
Large Language Models Utils
Utils useful for LLM
- RunningAgents115
Predict Memory
🧮115Estimate model memory usage and see detailed plots
- Running on CPU UpgradeAgentsFeatured1.02k
Model Memory Utility
🚀1.02kCalculate GPU memory needed for training Hugging Face models
- RunningAgents81
Transformers Timeline
🤗81Interactive timeline to explore the 🤗Transformers models
- Running on CPU UpgradeFeatured3.31k
The Smol Training Playbook
📚3.31kThe secrets to building world-class LLMs
Reasoning
-
Training Large Language Models to Reason in a Continuous Latent Space
Paper • 2412.06769 • Published • 93 -
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Paper • 2408.03314 • Published • 68 -
Evolving Deeper LLM Thinking
Paper • 2501.09891 • Published • 116 -
Kimi k1.5: Scaling Reinforcement Learning with LLMs
Paper • 2501.12599 • Published • 130
Safety & Security
- Running73
CyberSecEvalTest
📈73Evaluate LLMs' cybersecurity risks and capabilities
-
meta-llama/Llama-Guard-3-8B
Text Generation • 8B • Updated • 45.6k • • 329 -
meta-llama/Prompt-Guard-86M
Text Classification • 0.3B • Updated • 4.07M • • 413 -
protectai/deberta-v3-base-prompt-injection-v2
Text Classification • 0.2B • Updated • 849k • • 117
Large Language Models Utils
Utils useful for LLM
- RunningAgents115
Predict Memory
🧮115Estimate model memory usage and see detailed plots
- Running on CPU UpgradeAgentsFeatured1.02k
Model Memory Utility
🚀1.02kCalculate GPU memory needed for training Hugging Face models
- RunningAgents81
Transformers Timeline
🤗81Interactive timeline to explore the 🤗Transformers models
- Running on CPU UpgradeFeatured3.31k
The Smol Training Playbook
📚3.31kThe secrets to building world-class LLMs
Agents
-
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
Search-o1: Agentic Search-Enhanced Large Reasoning Models
Paper • 2501.05366 • Published • 106 -
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 107 -
Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Paper • 2501.10893 • Published • 26
Reasoning
-
Training Large Language Models to Reason in a Continuous Latent Space
Paper • 2412.06769 • Published • 93 -
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Paper • 2408.03314 • Published • 68 -
Evolving Deeper LLM Thinking
Paper • 2501.09891 • Published • 116 -
Kimi k1.5: Scaling Reinforcement Learning with LLMs
Paper • 2501.12599 • Published • 130