<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>显存 on AI 实战派 · 从技术到赚钱</title><link>https://guijiagi.com/tags/%E6%98%BE%E5%AD%98/</link><description>Recent content in 显存 on AI 实战派 · 从技术到赚钱</description><generator>Hugo</generator><language>zh-cn</language><copyright>本站内容采用 CC BY-NC-SA 4.0 国际许可协议授权</copyright><lastBuildDate>Wed, 07 Oct 2026 15:54:00 +0800</lastBuildDate><atom:link href="https://guijiagi.com/tags/%E6%98%BE%E5%AD%98/index.xml" rel="self" type="application/rss+xml"/><item><title>推理成本曲线：每百万 token 为什么能几年降两个数量级</title><link>https://guijiagi.com/posts/2026-10-07-inference-cost-curve/</link><pubDate>Wed, 07 Oct 2026 15:54:00 +0800</pubDate><guid>https://guijiagi.com/posts/2026-10-07-inference-cost-curve/</guid><description>从 Hopper 到 Blackwell 再到 Rubin，token 成本几年间跌了几十倍。这条下降曲线由哪几股力量共同推动。</description></item><item><title>AMD 显卡跑本地模型：ROCm 装对版本，比 CUDA 省钱一半</title><link>https://guijiagi.com/posts/2026-10-07-amd-roc-local-llm-deploy/</link><pubDate>Wed, 07 Oct 2026 12:12:00 +0800</pubDate><guid>https://guijiagi.com/posts/2026-10-07-amd-roc-local-llm-deploy/</guid><description>AMD 卡不是不能跑大模型，是 ROCm 版本和显卡型号要对得上。这篇讲清支持列表、安装坑和 Ollama/LM Studio 的调用方式。</description></item></channel></rss>