Date, time, and room will be added once confirmed.
This workshop explains how SGLang uses KV cache from prefix reuse to hierarchical caching. It covers RadixAttention and Unified RadixAttention for cache reuse, HiCache for moving cache across storage tiers, and the Unified Memory Pool for hybrid-attention models.