-
CyberSecEvalTest
📈73Evaluate LLMs' cybersecurity risks and capabilities
-
meta-llama/Llama-Guard-3-8B
Text Generation • 8B • Updated • 56.1k • • 324 -
meta-llama/Prompt-Guard-86M
Text Classification • 0.3B • Updated • 4.5M • • 403 -
protectai/deberta-v3-base-prompt-injection-v2
Text Classification • 0.2B • Updated • 866k • • 117
🤝 Open to Collab
Shyam Sunder Kumar
theainerd
AI & ML interests
Natural Language Processing
Recent Activity
liked a dataset 3 days ago
nasa-ibm-ai4science/Surya-bench-solarwind liked a model 3 days ago
sulabhkatiyar/indian-ne-multilingual-tts reacted to Banaxi-Tech's post with ❤️ 6 days ago
AGI has arrived.
Just gotta wait for the GLM distill.Organizations
Agents
-
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
Search-o1: Agentic Search-Enhanced Large Reasoning Models
Paper • 2501.05366 • Published • 106 -
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 107 -
Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Paper • 2501.10893 • Published • 26
Large Language Models Utils
Utils useful for LLM
- RunningAgents115
Predict Memory
🧮115Estimate model memory usage and see detailed plots
- Running on CPU UpgradeAgentsFeatured1.01k
Model Memory Utility
🚀1.01kCalculate GPU memory needed for training Hugging Face models
- RunningAgents81
Transformers Timeline
🤗81Interactive timeline to explore the 🤗Transformers models
- Running on CPU UpgradeFeatured3.3k
The Smol Training Playbook
📚3.3kThe secrets to building world-class LLMs
Reasoning
-
Training Large Language Models to Reason in a Continuous Latent Space
Paper • 2412.06769 • Published • 93 -
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Paper • 2408.03314 • Published • 68 -
Evolving Deeper LLM Thinking
Paper • 2501.09891 • Published • 116 -
Kimi k1.5: Scaling Reinforcement Learning with LLMs
Paper • 2501.12599 • Published • 130
Safety & Security
- Running73
CyberSecEvalTest
📈73Evaluate LLMs' cybersecurity risks and capabilities
-
meta-llama/Llama-Guard-3-8B
Text Generation • 8B • Updated • 56.1k • • 324 -
meta-llama/Prompt-Guard-86M
Text Classification • 0.3B • Updated • 4.5M • • 403 -
protectai/deberta-v3-base-prompt-injection-v2
Text Classification • 0.2B • Updated • 866k • • 117
Large Language Models Utils
Utils useful for LLM
- RunningAgents115
Predict Memory
🧮115Estimate model memory usage and see detailed plots
- Running on CPU UpgradeAgentsFeatured1.01k
Model Memory Utility
🚀1.01kCalculate GPU memory needed for training Hugging Face models
- RunningAgents81
Transformers Timeline
🤗81Interactive timeline to explore the 🤗Transformers models
- Running on CPU UpgradeFeatured3.3k
The Smol Training Playbook
📚3.3kThe secrets to building world-class LLMs
Agents
-
Agent Laboratory: Using LLM Agents as Research Assistants
Paper • 2501.04227 • Published • 95 -
Search-o1: Agentic Search-Enhanced Large Reasoning Models
Paper • 2501.05366 • Published • 106 -
Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Paper • 2501.11425 • Published • 107 -
Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments
Paper • 2501.10893 • Published • 26
Reasoning
-
Training Large Language Models to Reason in a Continuous Latent Space
Paper • 2412.06769 • Published • 93 -
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Paper • 2408.03314 • Published • 68 -
Evolving Deeper LLM Thinking
Paper • 2501.09891 • Published • 116 -
Kimi k1.5: Scaling Reinforcement Learning with LLMs
Paper • 2501.12599 • Published • 130