Min Zhang

2 articles on SOTA Papers

DeepSeek Cuts KV Cache Footprint Fourfold

A causal encoder-decoder MoE pairs cross-layer sparse-attention reuse with FP4 caching to reduce global context memory to 890 bytes per token.

Sep 19, 20264 min2609.19969

Pressure Reveals a Fragile Bright Magnetic Exciton

Hydrostatic compression extinguishes NiPS3 photoluminescence by 1.5 GPa without magnetic, structural, or electronic reconstruction.

Aug 11, 20264 min2607.27695