> For clean Markdown of any page, append .md to the page URL.
> For a complete documentation index, see https://docs.lakera.ai/llms.txt.
> For AI client integration (Claude Code, Cursor, etc.), connect to the MCP server at https://docs.lakera.ai/_mcp/server.

# Evaluation Datasets

The following is a list of public data sets on Hugging Face that can be used for evaluating the accuracy and effectiveness of Check Point AI Guardrails.

For more guidance on performing evaluations, please refer to our [evaluation guide](./evaluation-overview).

If you'd like to do a formal evaluation of AI Guardrails as part of a 'Proof-of-Value' please [contact us](https://www.lakera.ai/contact).

| Name                                                                                                | Type               | #Prompts | Purpose                                                                                                                                                                                                                     |
| --------------------------------------------------------------------------------------------------- | ------------------ | -------- | --------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| [Salad-Data](https://huggingface.co/datasets/OpenSafetyLab/Salad-Data)                              | Prompt Injection   | 21,318   | Comprehensive categorized dataset containing attack-enhanced prompts with jailbreak attempts across multiple harm categories including illegal drugs, misinformation, fraud, and dangerous content.                         |
| [ChatGPT-Jailbreak-Prompts](https://huggingface.co/datasets/rubend18/ChatGPT-Jailbreak-Prompts)     | Prompt Injection   | 79       | Collection of jailbreak related prompts for ChatGPT.                                                                                                                                                                        |
| [Vigil: LLM Jailbreak embeddings](https://huggingface.co/datasets/deadbits/vigil-jailbreak-ada-002) | Prompt Injection   | 104      | Curated dataset of prompts to test scanners detecting prompt injections, jailbreaks, and other potentially risky inputs.Contains text-embedding-ada-002 embeddings for all "jailbreak" prompts used by Vigil.               |
| [ALERT Adverserial](https://huggingface.co/datasets/Babelscape/ALERT)                               | Prompt Injection   | 45,731   | Categorized dataset of harmful instructions for the ALERT benchmark. Designed for testing content moderation and safety alignment in instruction-following models.                                                          |
| [NOETI ToxicQAFinal](https://huggingface.co/datasets/NobodyExistsOnTheInternet/ToxicQAFinal)        | Content Moderation | 6,866    | Categorized dataset of harmful and toxic content for evaluating content moderation.                                                                                                                                         |
| [SQuAD 2.0](https://huggingface.co/datasets/rajpurkar/squad_v2)                                     | Negative           | 142,192  | Stanford Question Answering Dataset (SQuAD) is a reading comprehension dataset, consisting of questions posed by crowdworkers on a set of Wikipedia articles. This is an all-negative dataset for false-postive evaluation. |