Register and share your invite link to earn from video plays and referrals.

alphaXiv
@askalphaxiv
High fidelity research. Dm for promo
Joined November 2023
72 Following    50.7K Followers
OCR but with a working memory...?! This paper, Unlimited OCR, replaces decoder full self-attention with Reference Sliding Window Attention, giving the model a working memory where each token attends to the fixed visual and prompt references plus only the most recent output tokens, making decode KV cache constant instead of linear in generation length. Combined with DeepEncoder’s 16x visual compression, this working-memory-style attention enables one-shot multi-page OCR with stable memory and latency, while improving OmniDocBench performance over DeepSeek OCR.
Show more