Tencent Releases WeVisDoc-2B and WeVisDoc-4B Document Parsing Models on Hugging Face
Original titletencent/WeVisDoc-2B
AISummary
Tencent's WeVisDoc-4B, fine-tuned from Qwen3-VL-4B-Instruct, scores 95.38 Overall on OmniDocBench v1.6 and 75.54 mean Overall across three PureDocBench tracks.
The end-to-end parser converts page images into structured Markdown with LaTeX formulas and HTML tables, and the 2B variant is also available.
The repository provides vLLM serving scripts with a 32768-token default context and a Python client for batch processing.
Source: Tencent · new models on Hugging Face · huggingface.coPublished · added here