robustness
12 个项目 · ⭐ 5.0kA unified evaluation framework for large language models
Raising the Cost of Malicious AI-Powered Image Editing
Fiddler Auditor is a tool to evaluate language models.
Contains code for the paper "Vision Transformers are Robust Learners" (AAAI 2022).
The official implementation of ECCV'24 paper "To Generate or Not? Safety-Driven Unlearned Diffusion Models Are Still Easy To Generate Unsafe Images ... For Now". This work introduces one fast and effective attack method to evaluate the harmful-content generation ability of safety-driven unlearned diffusion models.
[CVPR 2024] The official implementation of paper "synthesize, diagnose, and optimize: towards fine-grained vision-language understanding"
[ICML 2025] Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning