<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>NVIDIA Blackwell on 知识铺的博客</title>
    <link>https://index.zshipu.com/ai002/tags/NVIDIA-Blackwell/</link>
    <description>Recent content in NVIDIA Blackwell on 知识铺的博客</description>
    <generator>Hugo -- gohugo.io</generator>
    <language>zh-CN</language>
    <lastBuildDate>Wed, 09 Sep 2026 00:00:00 +0000</lastBuildDate>
    <atom:link href="https://index.zshipu.com/ai002/tags/NVIDIA-Blackwell/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>AWS SageMaker AI 上小 LLM 推理基准测试：G7 与 G5、G6 实例对比</title>
      <link>https://index.zshipu.com/ai002/post/20260909/AWS-SageMaker-AI-%E4%B8%8A%E5%B0%8F-LLM-%E6%8E%A8%E7%90%86%E5%9F%BA%E5%87%86%E6%B5%8B%E8%AF%95G7-%E4%B8%8E-G5G6-%E5%AE%9E%E4%BE%8B%E5%AF%B9%E6%AF%94/</link>
      <pubDate>Wed, 09 Sep 2026 00:00:00 +0000</pubDate>
      <guid>https://index.zshipu.com/ai002/post/20260909/AWS-SageMaker-AI-%E4%B8%8A%E5%B0%8F-LLM-%E6%8E%A8%E7%90%86%E5%9F%BA%E5%87%86%E6%B5%8B%E8%AF%95G7-%E4%B8%8E-G5G6-%E5%AE%9E%E4%BE%8B%E5%AF%B9%E6%AF%94/</guid>
      <description>两个 30B MoE 模型的基准测试设置 AWS 在 SageMaker AI 上针对两个 30B Mixture-of-Experts 模型展开基准测试。这两个模型分别是 Qwen3-Coder-30B 和 NVIDIA Nemotron-3-Nano-30B。测试覆盖 G5、G6、G6e 以及 G7 GPU 实例，核心指标包括吞吐量、延迟和每 token 成本。 测试直接在 Amazon SageMaker AI 环境中运行，目的是让开发者了解不同 GPU 实例在实际小 LLM 推理场景下</description>
    </item>
  </channel>
</rss>
