<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>本地部署 on 知识铺的博客</title>
    <link>https://index.zshipu.com/ai001/tags/%E6%9C%AC%E5%9C%B0%E9%83%A8%E7%BD%B2/</link>
    <description>Recent content in 本地部署 on 知识铺的博客</description>
    <generator>Hugo -- gohugo.io</generator>
    <language>zh-CN</language>
    <lastBuildDate>Fri, 04 Sep 2026 00:00:00 +0000</lastBuildDate>
    <atom:link href="https://index.zshipu.com/ai001/tags/%E6%9C%AC%E5%9C%B0%E9%83%A8%E7%BD%B2/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>AMD Threadripper Halo Station 工作站支持超 1T 参数模型本地运行</title>
      <link>https://index.zshipu.com/ai001/post/20260904/AMD-Threadripper-Halo-Station-%E5%B7%A5%E4%BD%9C%E7%AB%99%E6%94%AF%E6%8C%81%E8%B6%85-1T-%E5%8F%82%E6%95%B0%E6%A8%A1%E5%9E%8B%E6%9C%AC%E5%9C%B0%E8%BF%90%E8%A1%8C/</link>
      <pubDate>Fri, 04 Sep 2026 00:00:00 +0000</pubDate>
      <guid>https://index.zshipu.com/ai001/post/20260904/AMD-Threadripper-Halo-Station-%E5%B7%A5%E4%BD%9C%E7%AB%99%E6%94%AF%E6%8C%81%E8%B6%85-1T-%E5%8F%82%E6%95%B0%E6%A8%A1%E5%9E%8B%E6%9C%AC%E5%9C%B0%E8%BF%90%E8%A1%8C/</guid>
      <description>AMD 在 IFA 2026 上公布的 Threadripper Halo Station 工作站，基于 96 核心锐龙 Threadripper PRO 处理器，搭配两块 Instinct MI350P 显卡和 2TB 系统内存，支持参数超 1T 的大型模型本地运行。未来 4 卡配置可将 HBM3E 内存总容量提升至 576GB，所有核心算力组件均采用液冷散热。这把‘个人超级计算机’直接把万亿参数模型推理拉到桌面。 96 核 Threadripper PRO 处理器构成本地算力底座</description>
    </item>
    <item>
      <title>AMD NPU 在 Linux 上跑 Whisper 与 LLM 的实际效果</title>
      <link>https://index.zshipu.com/ai001/post/20260903/AMD-NPU-%E5%9C%A8-Linux-%E4%B8%8A%E8%B7%91-Whisper-%E4%B8%8E-LLM-%E7%9A%84%E5%AE%9E%E9%99%85%E6%95%88%E6%9E%9C/</link>
      <pubDate>Thu, 03 Sep 2026 00:00:00 +0000</pubDate>
      <guid>https://index.zshipu.com/ai001/post/20260903/AMD-NPU-%E5%9C%A8-Linux-%E4%B8%8A%E8%B7%91-Whisper-%E4%B8%8E-LLM-%E7%9A%84%E5%AE%9E%E9%99%85%E6%95%88%E6%9E%9C/</guid>
      <description>AMD NPU 在 Linux 上跑 Whisper 与 LLM 的实际效果 在 MSI Stealth A16 AI+ 笔记本搭载的 Ryzen AI 9 365 处理器上，XDNA2 NPU 成功驱动了 OpenAI 的 whisper-large-v3-turbo 模型进行语音转录。整个过程没有调用 CPU 或 GPU，实时因子 RTF 达到约 0.18，一段 30 秒的音频在 5.2 秒内完成转录。相比 CPU 执行相同任务，NPU 的能耗大约只有前者的十分之一。同时，同一块 NPU 还能运行大</description>
    </item>
  </channel>
</rss>
