PaddlePaddle Introduces Unlimited-OCR, a New OCR Model Designed for Large Document Processing.
PaddlePaddle Introduces Unlimited-OCR, a New OCR Model Designed for Large Document Processing.
PaddlePaddle has announced the Unlimited-OCR model, capable of processing large documents without speed loss.
According to the developers, the model can process hundreds of pages in a single pass without noticeable speed loss. This has been made possible by the R-SWA (Reference Sliding Window Attention) mechanism, which maintains a constant size of the KV cache during decoding.
In the OmniDocBench benchmark, the model scored 93%, outperforming DeepSeek-OCR by 6%.
Why it matters
AnalysisThis new model can significantly simplify the processing of large documents, marking an important step in the development of OCR technologies. Its high performance and accuracy open up new possibilities for document processing automation.
Discuss in community
Share your questions and insights with developers