Dolphin

by bytedance · Developer Tools

The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

New0 ratings9,047 starsActive
Developer ToolsDeveloper ToolPythondocument-analysis
GitHub⚑ Report
Dolphin — image 1
Dolphin — image 2
Dolphin — image 3
Dolphin — image 4
1 / 4

About this project

Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting Dolphin-v2 is an enhanced universal document parsing model that substantially improves upon the original Dolphin. It seamlessly handles any document type—whether digital-born or photographed—through a document-type-aware two-stage architecture with scalable anchor prompting. 📑 Overview Document image parsing is challenging due to diverse document types and complexly intertwined elements such as text paragraphs, figures, formulas, tables, and code blocks. Dolphin-v2 addresses these challenges through a document-type-aware two-stage approach: 1. 🔍 Stage 1: Document type classification (digital vs. photographed) + layout…

Technologies

Pythondocument-analysisocrlayout-analysis

Project health

Low recent activity
Last update5 months ago
Contributors3
Latest release
Open issues & PRs77
LicenseOther
On GitHubsince 2025

GitHub

9,047
stars
775
forks
3
contributors
77
open issues & PRs
Python
language
5 months ago
last commit
View on GitHub ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review Dolphin.

Built by

bytedance
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain bytedance/Dolphin? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like