Cumulative Risk in LLM Dialogues: Safety Goes Stateful
Today's guardrails evaluate each prompt-response pair in isolation, overlooking risks that emerge only over multi-turn dialogues. A new framework tracks semantic drift and information accumulation, with direct implications for on-premise deployment a...