ScaleAI/fortress_public

Dataset

This dataset contains adversarial prompts and associated rubrics designed to evaluate the safety and security of large language models (LLMs), as described in the paper FORTRESS: Frontier Risk Evaluation for National Security and Public Safety. Please exercise care and caution when using these data, as they contain potentially sensitive or harmful information related to public safety and national security. This dataset should be used for safety evaluations only, and it is prohibited to use these

10

9 commits

1 linked in READMEs

updated Aug 5, 2025

See the code

README

This dataset contains adversarial prompts and associated rubrics designed to evaluate the safety and security of large language models (LLMs), as described in the paper FORTRESS: Frontier Risk Evaluation for National Security and Public Safety. Please exercise care and caution when using these data, as they contain potentially sensitive or harmful information related to public safety and national security. This dataset should be used for safety evaluations only, and it is prohibited to use these data for any adversarial training or research.
Project page

Contributors

kaus-scale

8 commits

nielsr

1 commits

ScaleAI/fortress_public

Dataset

This dataset contains adversarial prompts and associated rubrics designed to evaluate the safety and security of large language models (LLMs), as described in the paper FORTRESS: Frontier Risk Evaluation for National Security and Public Safety. Please exercise care and caution when using these data, as they contain potentially sensitive or harmful information related to public safety and national security. This dataset should be used for safety evaluations only, and it is prohibited to use these

10

9 commits

1 linked in READMEs

updated Aug 5, 2025

See the code

README

This dataset contains adversarial prompts and associated rubrics designed to evaluate the safety and security of large language models (LLMs), as described in the paper FORTRESS: Frontier Risk Evaluation for National Security and Public Safety. Please exercise care and caution when using these data, as they contain potentially sensitive or harmful information related to public safety and national security. This dataset should be used for safety evaluations only, and it is prohibited to use these data for any adversarial training or research.
Project page

Contributors

kaus-scale

8 commits

nielsr

1 commits