Alice raises $140M for frontier AI safety and security [Read the news]
Back
Company News

ActiveFence Introduces: The Dawn of Explainable AI

Jul 24, 2024
Learn more about ActiveFence Solutions in Generative AI Safety
Learn More

NEW YORK, July 23, 2024 — ActiveFence, a leading technology solution for Trust and Safety intelligence, management, and content moderation, is proud to announce the launch of AI Explainability, a groundbreaking feature of its ActiveScore AI models. Explainability opens the “black box” of AI models, offering unprecedented transparency and insight into AI decision-making processes.

Explainability addresses a crucial need in the market by providing a detailed breakdown of why content, such as images or videos, is classified as violative. For example, if an image is flagged for promoting terror, Explainability will indicate the signals in the image—like the existence of logos and flags, or the presence of known terrorists—that contributed to this detection.

With Explainability, ActiveFence continues its mission to create safer and more compliant online environments. By unveiling how models decide on content violations, Explainability enables moderators to make more informed decisions by exposing the components that contribute to the assessment of risk. With Explainability, moderators make better decisions, thereby improving user trust and increasing user retention and usage. Furthermore, Explainability aids in the review process for content appeals by exposing the reason an item was flagged, ensuring compliance with online safety regulations like the EU’s Digital Services Act (DSA). 

Iftach Orr, Co-founder and CTO at ActiveFence:“Explainability is a game-changer in the field of AI moderation, We are excited to provide our clients with a level of transparency and understanding that has never been seen before. By revealing the inner workings of our AI models, we empower moderators to make more accurate and fair decisions, ultimately creating a safer online space for all users. “

For more information on how to safeguard online platforms and users against online harm, visit our website at alice.io

About ActiveFence:
ActiveFence is the leading Trust and Safety provider for online platforms, protecting over three billion users daily from malicious behavior and content. Trust and Safety teams of all sizes rely on ActiveFence to keep their users safe from the widest spectrum of online harms, including child abuse, disinformation, hate speech, terror, fraud, and more. We offer a full stack of capabilities with our deep intelligence research, AI-driven harmful content detection and moderation platform. ActiveFence protects platforms globally, in over 100 languages, letting people interact and thrive safely online.

Share

Alice Data Advantage

Alice is the world’s largest collector and manager of adversarial intelligence data. Our data is the cornerstone for protecting platform, tech, and users online.

 Learn More >

What’s new from Alice

Alice Raises $140M to Make Sure AI Does Exactly What It's Supposed to Do

blog
Aug 25, 2026
,
 
Aug 25, 2026
 -
4
 min read
Aug 25, 2026
 -
4
 min watch
August 25, 2026

Alice raised $140M led by Apax Digital, bringing total funding to $280M. The AI trust, safety, and security company works with 8 of the 10 top AI labs.

Learn More

5 Ways Your Third-Party CX Agent Gets Broken

whitepaper
Jul 31, 2026
,
 
Jul 31, 2026
 -
This is some text inside of a div block.
 min read
Jul 31, 2026
 -
This is some text inside of a div block.
 min watch
July 31, 2026

Third-party CX agents create hidden liability. Learn the 5 attack patterns vendors miss and how WonderSuite closes the gap.

Learn More