Intelsi.si
Submit a tool

// category · 10 tools · 7 open source

Safety & Evals

Open frameworks, red-teaming kits, guardrails and benchmarks for measuring what frontier models can do and keeping deployed systems within safe bounds.

ARC Prizearcprize.orgNonprofit competition built around the ARC-AGI benchmarks, which test fluid reasoning on puzzles that are easy for people and hard for AI.Safety & EvalsFree★ 743Inspectinspect.aisi.org.ukEvaluation framework from the UK AI Security Institute for building rigorous, reproducible LLM and agent evaluations.Safety & EvalsOpen source★ 2.9kGuardrails AIguardrailsai.comFramework and hub of validators that check LLM inputs and outputs for risks and enforce structured responses.Safety & EvalsOpen source★ 7.5kgarakgarak.aiLLM vulnerability scanner that probes models for jailbreaks, prompt injection, data leakage and other failure modes.Safety & EvalsOpen source★ 9.4kNeMo Guardrailsgithub.comToolkit for adding programmable rails to LLM apps: topic control, jailbreak detection, fact-checking and safe dialog flows.Safety & EvalsOpen source★ 7.2kPurple Llamagithub.comMeta's umbrella project for open trust-and-safety tools, including Llama Guard classifiers and CyberSecEval benchmarks.Safety & EvalsOpen source★ 4.4kHarmBenchharmbench.orgStandardised evaluation framework for automated red-teaming and measuring model refusal robustness.Safety & EvalsOpen source★ 1.1kPyRITgithub.comMicrosoft's Python Risk Identification Toolkit for automating red-teaming of generative AI systems.Safety & EvalsOpen source★ 118Humanity's Last Examlastexam.aiExpert-written benchmark of very hard questions across dozens of subjects, built to measure progress at the frontier of knowledge.Safety & EvalsFreeHN 1Lakeralakera.aiRuntime security for generative AI apps, detecting prompt injection, data leakage and harmful content in real time.Safety & EvalsSee siteHN 1

More categories