RTPrune: Reading-Twice Inspired Token Pruning for Efficient DeepSeek-OCR Inference
DGX agentarXiv:2605.00392v1 Announce Type: new Abstract: DeepSeek-OCR leverages visual-text compression to reduce long-text processing costs and accelerate inference, yet visual tokens remain prone to redundan