Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Malikeh1375
's Collections
Abliterated Models
Safety-Aligned Models
AI Safety Benchmarks
Clustered Tulu
LLM-Alignment
LLM Interpretability
Medical Datasets
AI Safety Benchmarks
updated
Jul 17
Upvote
1
Sort: Collection
JailbreakBench/JBB-Behaviors
Viewer
•
Updated
Sep 26, 2024
•
500
•
26k
•
137
walledai/HarmBench
Viewer
•
Updated
Jul 31, 2024
•
400
•
7.56k
•
59
allenai/real-toxicity-prompts
Viewer
•
Updated
Sep 30, 2022
•
99.4k
•
16.5k
•
123
cais/wmdp
Viewer
•
Updated
Apr 27, 2024
•
3.67k
•
33.7k
•
31
Upvote
1
Sort: Collection
Share collection
View history
Collection guide
Browse collections