8 repos
thu-ml/MMTrustEval
A toolbox for benchmarking trustworthiness of multimodal large language models (MultiTrust, NeurIPS…
178
216 commits
thu-ml/MLA-Trust
A toolbox for benchmarking Multimodal LLM Agents trustworthiness across truthfulness,…
64
44 commits
ThuCCSLab/FigStep
[AAAI'25 (Oral)] Jailbreaking Large Vision-language Models via Typographic Visual Prompts
216
41 commits
ZichenWen1/DIJA
(ICLR 2026 🔥) Code for "The Devil behind the mask: An emergent safety vulnerability of Diffusion…
79
43 commits
mbzuai-oryx/ALM-Bench
[CVPR 2025 🔥] ALM-Bench is a multilingual multi-modal diverse cultural benchmark for 100 languages…
47
37 commits
Agora-Lab-AI/Atom
a suite of finetuned LLMs for atomically precise function calling 🧪
16
19 commits
wangyouze/Trust-videoLLMs
No description
33
12 commits
AI-secure/MMDT
Comprehensive Assessment of Trustworthiness in Multimodal Foundation Models
29
77 commits