往下拉回到首頁
DeepSeek released 'Thinking-with-Visual-Primitives' framework

DeepSeek released 'Thinking-with-Visual-Primitives' framework

DeepSeek released 'Thinking-with-Visual-Primitives' framework

DeepSeek, in collaboration with Peking University and Tsinghua University, has released the paper "Thinking with Visual Primitives" along with its open-source repository, introducing a new multimodal reasoning framework. The core approach of this framework is to elevate spatial tokens—specifically coordinate points and bounding boxes—into the "minimal units of thought" within the model's chain-of-thought. These are directly interleaved during the reasoning process, enabling the model to "point"