Alice raises $140M for frontier AI safety and security [Read the news]
See All Resources
benchmark

The LLM Safety Review: Benchmarks & Analysis

Get the benchmark

Aug 1, 2023

As GenAI tools and the LLMs behind them impact the daily lives of billions, this report examines whether these technologies can be trusted to keep users safe.

What you’ll learn:

  • How LLMs respond to risky prompts from bad actors and vulnerable users
  • Where current models show safety strengths and weaknesses
  • Actionable steps to improve LLM safety and reduce harmful outcomes

Overview

In this first independent benchmarking report on the LLM safety landscape, ActiveFence’s subject-matter experts put leading models to the test. More than 20,000 prompts were used to analyze how six LLMs respond across seven major languages and four high-risk abuse areas: child exploitation, hate speech, self-harm, and misinformation. The report provides comparative insight into each model’s relative safety strengths and weaknesses, helping teams understand where gaps exist and where additional resources may be required.

What’s new from Alice

Not Your Typical Red Teaming: How we Go Above and Beyond

blog
Sep 3, 2026
,
 
Sep 3, 2026
 -
7
 min read
Sep 3, 2026
 -
7
 min watch
September 3, 2026

Red teaming usually tests if your agent can be broken. Alice also tests over-refusal, how often it wrongly blocks a real customer, the silent failure that costs you business.

Learn More

5 Ways Your Third-Party CX Agent Gets Broken

whitepaper
Jul 31, 2026
,
 
Jul 31, 2026
 -
This is some text inside of a div block.
 min read
Jul 31, 2026
 -
This is some text inside of a div block.
 min watch
July 31, 2026

Third-party CX agents create hidden liability. Learn the 5 attack patterns vendors miss and how WonderSuite closes the gap.

Learn More

Secure the keys to GenAI wonderland?

Get a demo
Guardrails
Red-Team Lab