Tencent has released WeVisDoc, an end-to-end document parser designed for page images. This model, fine-tuned from Qwen3-VL-2B-Instruct and Qwen3-VL-4B-Instruct, converts page images into structured Markdown, including LaTeX formulas and HTML tables. WeVisDoc-4B has demonstrated top performance on benchmarks like OmniDocBench and PureDocBench, outperforming other end-to-end parsers in multiple settings. AI
IMPACT This tool could improve document processing and data extraction from image-based documents.
RANK_REASON This is a specific product release from a major tech company, but not a frontier model release.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →