AI Briefing
KO

ServiceNow-AI Unveils AprielGuard, a Safety Model for Agents

·2025.12.23 23:07

Key point

AprielGuard, an 8B-parameter guardrail model designed for the security and safety of agentic workflows, has been released.

1 / 2

Details

As LLMs evolve beyond simple text assistants into Agentic systems capable of tool use and reasoning, new threats are emerging that existing single-message-based safety classifiers struggle to handle.

AprielGuard is a 8B-parameter safety and security guardrail model designed to respond to these changes.

Key Features and Characteristics:

  • Broad Detection Coverage: Detects 16 safety risk categories (toxicity, misinformation, illegal activity, etc.) as well as various adversarial attacks including prompt injection, jailbreaks, memory poisoning, and tool manipulation.
  • Optimized for Agentic Workflows: Supports complex agentic environments including multi-turn conversations, long context, tool calls, and reasoning traces, not just simple prompts.
  • Two Operating Modes: Provides both a Reasoning mode for cases requiring explainable classification and a Non-reasoning mode for low-latency production environments.

This model offers an integrated solution for managing the complex security threats arising in the modern LLM agent ecosystem.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.