HE-Guardrail: A Homomorphic Guardrail Against Jailbreak Attacks for Encrypted Large Language Model Inference

Researchers propose HE-Guardrail, a framework for protecting against jailbreak attacks on encrypted large language model inference. The framework evaluates guardrail mechanisms over encrypted data, allowing servers to control the return of model responses without revealing the input or output.

RSS Score 0 9/21/2026, 4:00:00 AM Original Source
Save an API key to vote.