DeepSeek-VL2

by deepseek-ai · Developer Tools

DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

New0 ratings5,372 starsActive
Developer ToolsDeveloper ToolPythonCSS
GitHub⚑ Report
DeepSeek-VL2 — image 1
DeepSeek-VL2 — image 2
1 / 2

About this project

📥 Model Download ⚡ Quick Start 📜 License 📖 Citation 📄 Paper Link 📄 Arxiv Paper Link 👁️ Demo 1. Introduction Introducing DeepSeek-VL2, an advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly improves upon its predecessor, DeepSeek-VL. DeepSeek-VL2 demonstrates superior capabilities across various tasks, including but not limited to visual question answering, optical character recognition, document/table/chart understanding, and visual grounding. Our model series is composed of three variants: DeepSeek-VL2-Tiny, DeepSeek-VL2-Small and DeepSeek-VL2, with 1.0B, 2.8B and 4.5B activated parameters respectively. DeepSeek-VL2 achieves…

Technologies

JavaScriptPythonMakefileCSS

Project health

Inactive for over a year
Last updatelast year
Contributors7
Latest release
Open issues & PRs120
LicenseMIT
On GitHubsince 2024

GitHub

5,372
stars
1,811
forks
7
contributors
120
open issues & PRs
Python
language
last year
last commit
View on GitHub ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review DeepSeek-VL2.

Built by

deepseek-ai
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain deepseek-ai/DeepSeek-VL2? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like