LATIDIA · Ciberseguridad
Persona Guardrail: un marco de defensa de nivel de producción para sistemas genéticos
arXiv: 2610.03434v1Tipo de anuncio: nuevo Resumen: Los agentes basados en modelos de lenguaje grande se implementan cada vez más para realizar tareas específicas del dominio al interactuar con el conocimiento, las herramientas y los servicios externos de la empresa. Existin
WhatsApp ↗Telegram ↗
La noticia
arXiv:2610.03434v1 Announce Type: new Abstract: Large language model-based agents are increasingly deployed to perform domain-specific tasks by interacting with enterprise knowledge, tools, and external services. Existing runtime guardrails primarily target prompt injection and other attack-specific behaviors under a black-box threat model, but provide limited guarantees that agents operate within their intended functionality. As a result, production agents remain vulnerable to malicious requests and out-of-domain queries that existing defenses often fail to distinguish. We present Persona Guardrail, a production-grade runtime defense framework that enforces explicit functional boundaries for customer-facing agentic AI systems through synchronous input and output validation driven by semantic allowlist and blocklist specifications. We also introduce PAGE (Persona-Aware Guardrail