keyword
PromptBench
PromptBench is an open-source evaluation framework and library designed to assess the performance, robustness, and safety of large language models. It provides a unified environment for researchers and developers to construct prompts, load standardized datasets and models, execute dynamic evaluation protocols, and simulate adversarial prompt attacks. By analyzing how language models respond to various prompt perturbations and diverse task scenarios, the framework facilitates systematic benchmarking, assists in uncovering model vulnerabilities, and supports the development of reliable artificial intelligence systems.
1 item

