What infiniflow/ragflow shipped
Written by FoxPlug from public releases; not affiliated with InfiniFlow. An automatic summary of the public release, pull request and commit data of github.com/infiniflow/ragflow. InfiniFlow did not write it and does not use or endorse FoxPlug. Every line links to the public change it describes.
Week of September 21, 2026
What shipped
- Image chunk previews now stream directly to storage instead of batching in memory, reducing peak heap usage for large PDFs. Pull request #20165
- Image and table chunk cropping now runs in parallel across chunks instead of serially, improving chunker performance. Pull request #20183
- New model provider instances now auto-merge newly discovered models by default while preserving user-removed models. Pull request #20177
- Tool calling is now enabled by default for new provider model instances. Pull request #20195
- Re-parsing documents with existing chunks now shows a confirmation dialog before proceeding. Pull request #20210
- Ingestion log modal now polls for messages during parsing and stops at terminal states. Pull request #20206
- Shared chats and embedded bots can now access document images via beta tokens. Pull request #20185
- Rerank models now take effect during agent retrieval. Pull request #20197
- Disabled documents are now excluded from Wiki product datasets and graph entities. Pull request #20154
- Removed unused tenant_llm table from database. Pull request #20204
Why it matters
This week focused on performance improvements for document parsing with optimized image handling, better defaults for model provider configuration, and fixes for chat and retrieval features. Several refactoring changes cleaned up dead code in the agent execution path while fixes addressed dataset parsing, reranking, and document management issues.
Changelog entry
- feat(chunker): stream cropped previews to storage instead of batching in memory Pull request #20165
- perf(chunker): parallelize cropImageChunks across chunks Pull request #20183
- feat(setting): auto-merge newly listed models into draft instance Pull request #20177
- feat(setting): default tool calling on for new provider instances Pull request #20195
- fix(dataset): confirm before re-parse Go documents with chunks Pull request #20210
- fix(dataset): poll ingestion log while parsing and stop after terminal Pull request #20206
- fix(router): accept beta tokens on document image routes for shared chats Pull request #20185
- fix: rerank not working in agent retrieval Pull request #20197
- fix: exclude disabled documents from Wiki products Pull request #20154
- fix: remove table tenant_llm Pull request #20204
- fix(setting): suppress PaddleOCR form autofill and default URL to first option Pull request #20189
- fix: eval error Pull request #20158
ragflow shipped parallel image cropping for faster chunking, streaming previews to reduce memory usage, auto-merge for model discovery, and fixes for reranking and document parsing.
ragflow this week improved document parsing performance with parallel image chunk cropping and streaming preview uploads that reduce peak memory. Model provider setup now auto-merges newly discovered models while preserving removals, and tool calling defaults to enabled. Fixes address agent reranking, document re-parsing confirmations, ingestion log polling, and shared chat image access.