<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>视觉推理 on GPT资讯  --  知识铺</title>
    <link>https://index.zshipu.com/gpt/tags/%E8%A7%86%E8%A7%89%E6%8E%A8%E7%90%86/</link>
    <description>Recent content in 视觉推理 on GPT资讯  --  知识铺</description>
    <generator>Hugo -- gohugo.io</generator>
    <language>zh-CN</language>
    <lastBuildDate>Tue, 01 Sep 2026 00:00:00 +0000</lastBuildDate>
    <atom:link href="https://index.zshipu.com/gpt/tags/%E8%A7%86%E8%A7%89%E6%8E%A8%E7%90%86/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>DeepSeek 305B 多模态模型不做看图说话 只专注视觉推理</title>
      <link>https://index.zshipu.com/gpt/post/20260901/DeepSeek-305B-%E5%A4%9A%E6%A8%A1%E6%80%81%E6%A8%A1%E5%9E%8B%E4%B8%8D%E5%81%9A%E7%9C%8B%E5%9B%BE%E8%AF%B4%E8%AF%9D-%E5%8F%AA%E4%B8%93%E6%B3%A8%E8%A7%86%E8%A7%89%E6%8E%A8%E7%90%86/</link>
      <pubDate>Tue, 01 Sep 2026 00:00:00 +0000</pubDate>
      <guid>https://index.zshipu.com/gpt/post/20260901/DeepSeek-305B-%E5%A4%9A%E6%A8%A1%E6%80%81%E6%A8%A1%E5%9E%8B%E4%B8%8D%E5%81%9A%E7%9C%8B%E5%9B%BE%E8%AF%B4%E8%AF%9D-%E5%8F%AA%E4%B8%93%E6%B3%A8%E8%A7%86%E8%A7%89%E6%8E%A8%E7%90%86/</guid>
      <description>DeepSeek 8 月 31 日上线的 DeepSeek-V4-Flash-Vision-Exp 是 305B 参数的 V4 家族首款多模态模型，却明确不是用来做图像描述的。这款实验模型在原有语言底座上接入视觉编码器和 Aligner 后获得图像理解能力，其对多模态的定义与主流做法形成直接反差。 DeepSeek-V4-Flash-Vision-Exp 的出现让多模态这个概念在中文开源社区里有了新含义。过去几年，多数多模态大模型把重点放在描述图像</description>
    </item>
  </channel>
</rss>
