-
protectai/deberta-v3-base-prompt-injection-v2
Text Classification • 0.2B • Updated • 197k • • 84 -
openai/gpt-oss-safeguard-20b
Text Generation • 22B • Updated • 9.36k • • 179 -
meta-llama/Llama-Prompt-Guard-2-86M
Text Classification • 0.3B • Updated • 28.4k • • 76 -
leolee99/PIGuard
Text Classification • 0.2B • Updated • 1.8k • 4
Lipeng (Tony) He
ttttonyhe
·
AI & ML interests
Trustworthy Machine Learning
Recent Activity
authored
a paper
10 days ago
Safety at One Shot: Patching Fine-Tuned LLMs with A Single Instance
submitted
a paper
11 days ago
Safety at One Shot: Patching Fine-Tuned LLMs with A Single Instance
updated
a collection
12 days ago
Red-Teaming Models & Datasets
Organizations
Open Embedding Models
-
Running on CPU Upgrade6.94k
MTEB Leaderboard
🥇6.94kEmbedding Leaderboard
-
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 1.98M • • 831 -
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 587k • • 1.42k -
nomic-ai/nomic-embed-text-v2-moe
Sentence Similarity • 0.5B • Updated • 952k • 449
Red-Teaming Models & Datasets
Novel Models
Guardrails
-
protectai/deberta-v3-base-prompt-injection-v2
Text Classification • 0.2B • Updated • 197k • • 84 -
openai/gpt-oss-safeguard-20b
Text Generation • 22B • Updated • 9.36k • • 179 -
meta-llama/Llama-Prompt-Guard-2-86M
Text Classification • 0.3B • Updated • 28.4k • • 76 -
leolee99/PIGuard
Text Classification • 0.2B • Updated • 1.8k • 4
Specialized LLMs
Open Embedding Models
-
Running on CPU Upgrade6.94k
MTEB Leaderboard
🥇6.94kEmbedding Leaderboard
-
Qwen/Qwen3-Embedding-0.6B
Feature Extraction • 0.6B • Updated • 1.98M • • 831 -
google/embeddinggemma-300m
Sentence Similarity • 0.3B • Updated • 587k • • 1.42k -
nomic-ai/nomic-embed-text-v2-moe
Sentence Similarity • 0.5B • Updated • 952k • 449
Domain-specific Datasets
Red-Teaming Models & Datasets
SOTA Medium-sized Models
Novel Models
Templates