Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

OpenSafetyLab

non-profit
https://open-trust-lab.vercel.app
Activity Feed Request to join this org

AI & ML interests

None defined yet.

Recent Activity

WRHC  authored a paper about 3 hours ago
SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models
WRHC  authored a paper about 3 hours ago
EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models
WRHC  authored a paper about 3 hours ago
ELV-Halluc: Benchmarking Semantic Aggregation Hallucinations in Long Video Understanding
View all activity

RW's profile picture XuHao Hu's profile picture Bowen Dong's profile picture Lijun Li's profile picture
Organization Card
Community About org cards

Edit this README.md markdown file to author your organization card.

spaces 2

Sleeping
Agents
8

Salad Bench Leaderboard

🏢

Display benchmark results for models across different taxonomies

Mar 25, 2024

models 3

OpenSafetyLab/MD-Judge-v0_2-internlm2_7b

Text Generation • 8B • Updated Mar 8, 2025 • 556 • 17

OpenSafetyLab/ImageGuard

Image-to-Text • Updated Jan 19, 2025 • 6

OpenSafetyLab/MD-Judge-v0.1

Text Generation • 7B • Updated May 20, 2024 • 1.42k • 19

datasets 3

OpenSafetyLab/Salad-Data

Viewer • Updated Jul 23 • 30.7k • 1.5k • 33

OpenSafetyLab/t2i_safety_dataset

Updated Aug 5, 2025 • 499 • 3

OpenSafetyLab/t2isafety_evaluation

Preview • Updated Feb 10, 2025 • 85
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs