Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Yuliang
PRO
Leon5201314
6
12
Follow
SanchezBoyzLLC's profile picture
webxos's profile picture
RightToken's profile picture
3 followers
·
7 following
https://www.yuliang-liu.com/
Yuliang-Liu
AI & ML interests
Document-native intelligence.
Recent Activity
liked
a model
12 days ago
wpj20000/DeltaV-2B
upvoted
a
paper
12 days ago
DeltaV: Thinking with Visual State Updates in Unified Large Multimodal Models
replied
to
their
post
14 days ago
0.7B MonkeyOCRv2 Outperforms Larger Models on 17-Language Document Parsing MonkeyOCRv2-B-Parsing reaches 83.3 on MDPBench, a multilingual benchmark covering digital-born and photographed documents across 17 languages. Results among evaluated open-source models: • MonkeyOCRv2-B, 0.7B: 83.3 • dots.mocr, 3B: 80.5 • HunyuanOCR-1.5, 1B: 76.8 • PaddleOCR-VL-1.6, 0.9B: 75.0 • MinerU2.5-Pro, 1.2B: 71.0 The central idea is simple: before asking an LLM to reason over a document, the vision encoder must preserve every character stroke, digit, punctuation mark, and layout cue. MonkeyOCRv2 is pretrained on 113M document images across 17 languages using joint image-to-text generation and pixel-level reconstruction. Models: https://huggingface.co/collections/zenosai/monkeyocrv2 Paper: https://huggingface.co/papers/2607.11562 GitHub: https://github.com/Yuliang-Liu/MonkeyOCRv2 Code and model weights are available under Apache-2.0. We welcome tests on difficult multilingual, photographed, and visually ambiguous documents—especially failure cases.
View all activity
Organizations
Leon5201314
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
12 days ago
wpj20000/DeltaV-2B
2B
•
Updated
26 days ago
•
52
•
4
liked
a Space
14 days ago
Running
5
MDPBench Leaderboard
📄
5
Explore multilingual document parsing benchmark rankings
liked
a dataset
15 days ago
zenosai/MonkeyDocv2
Updated
15 days ago
•
88
•
2
liked
a model
15 days ago
zenosai/MonkeyOCRv2-S
Image Feature Extraction
•
28.5M
•
Updated
8 days ago
•
10.9k
•
12
liked
a model
17 days ago
zenosai/MonkeyOCRv2-B-Parsing
Image-Text-to-Text
•
0.9B
•
Updated
8 days ago
•
1.44k
•
12
liked
2 datasets
4 months ago
Delores-Lin/MDPBench
Benchmark
•
Updated
16 days ago
•
1.53k
•
21
ling99/OCRBench_v2
Viewer
•
Updated
Feb 24, 2025
•
10k
•
3.13k
•
20
liked
a dataset
over 1 year ago
lmms-lab/OCRBench-v2
Viewer
•
Updated
Feb 9, 2025
•
10k
•
1.69k
•
12
liked
a model
over 1 year ago
zwzhang/Accountable-Textual-Visual-Chat
Updated
Jun 17, 2023
•
2
liked
a model
almost 2 years ago
echo840/Monkey
Text Generation
•
Updated
Apr 7, 2024
•
358
•
33
liked
a dataset
about 2 years ago
ByteDance/MTVQA
Viewer
•
Updated
May 30, 2024
•
8.79k
•
487
•
42
liked
a Space
over 2 years ago
Running
Agents
224
Ocrbench Leaderboard
🏆
224
View OCR model leaderboard rankings