Tencent releases WeVisDoc-4B, a document parser that leads OmniDocBench v1.6
Original titletencent/WeVisDoc-4B
AISummary
Tencent's WeVisDoc-4B, fine-tuned from Qwen3-VL-4B-Instruct, converts page images into structured Markdown with LaTeX formulas and HTML tables.
It scores 95.38 Overall on OmniDocBench v1.6 and a mean Overall of 75.54 across three PureDocBench tracks, ranking first among compared end-to-end parsers in all four reported settings.
The model is available on Hugging Face and runs through vLLM, which requires version 0.11.1 or later.
Source: Tencent · new models on Hugging Face · huggingface.coPublished · added here