🚀NEW LABGetting Started with Claude AgentsStart lab
Architecture · Multimodal

DeepSeek-VL2

First page
DeepSeek-VL2
Paper summary

a new series of vision-language models featuring dynamic tiling for high-resolution images and efficient MoE architecture, achieving competitive performance across visual tasks; achieves competitive or state-of-the-art performance with similar or fewer activated parameters compared to existing open-source dense and MoE-based models.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack