inference 2 Enabling Encoder–Decoder Models on Intel XPU in SGLang: A KV-Cache Story Aug 30, 2026 FP8 Quantization and Beyond: From First Principles to Intel Neural Compressor Mar 25, 2026