Congrats to
@deepseek_ai on releasing DeepSeek-V4.1-Flash!
> 552B backbone, with a causal encoder-decoder activating just 8B params at prefill, 16B at decode
> 196B Engram memory accessed through sparse lookups
> ~8x less persistent KV than V4-Flash through bounded replay