J. Yang

2 articles on SOTA Papers

DeepSeek Cuts KV Cache Footprint Fourfold

A causal encoder-decoder MoE pairs cross-layer sparse-attention reuse with FP4 caching to reduce global context memory to 890 bytes per token.

Sep 19, 20264 min2609.19969

XENONnT Extends Light Dark Matter Limits With S2-Only Data

A 7.83 tonne-year ionization-only analysis models low-energy backgrounds in four observables and excludes spin-independent scattering above 6.0×10⁻⁴⁵ cm² at 5 GeV/c².

Aug 3, 20264 min2601.11296