Xiang Li

2 articles on SOTA Papers

DeepSeek Cuts KV Cache Footprint Fourfold

A causal encoder-decoder MoE pairs cross-layer sparse-attention reuse with FP4 caching to reduce global context memory to 890 bytes per token.

Sep 19, 20264 min2609.19969

XPolicyLab Cuts Robot Policy Integration From NM to N Plus M

A shared policy contract and isolated client/server runtime let 42 robot policies connect to simulations and real-robot evaluation without pairwise adapters.

Aug 18, 20264 min2608.09892