Tag: inference time
Safety-Aware Decoding for LLMs: Inference-Time Guardrails Explained
Explore how safety-aware decoding and inference-time guardrails secure LLMs without retraining. Learn about SafeDecoding, SSD, and ShieldHead, plus their impact on latency and robustness.