<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[本地部署大模型，硬件到底怎么配？我踩过的坑全在这]]></title><description><![CDATA[<p dir="auto">被问最多的就是「本地跑大模型要啥配置」，结合我自己折腾的经验说下：</p>
<p dir="auto">7b 模型：8g 显存够，cpu 也能跑但慢得想砸键盘<br />
13b：16g 显存起步<br />
70b：48g+ 或者多卡，家里就别折腾了</p>
<p dir="auto">量化是神：q4 量化能砍掉一大半显存，画质/效果损失很小。llama.cpp 和 ollama 是最省心的两个工具，新手直接用 ollama。</p>
<p dir="auto">我现在的方案是 amd 核显 + 64g 内存，跑 35b 的 q8 量化，速度一般但胜在便宜。想快的还是上 n 卡。</p>
]]></description><link>https://bbs.zapai.cc/topic/86/本地部署大模型-硬件到底怎么配-我踩过的坑全在这</link><generator>RSS for Node</generator><lastBuildDate>Tue, 08 Sep 2026 01:17:29 GMT</lastBuildDate><atom:link href="https://bbs.zapai.cc/topic/86.rss" rel="self" type="application/rss+xml"/><pubDate>Mon, 31 Aug 2026 04:12:49 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to 本地部署大模型，硬件到底怎么配？我踩过的坑全在这 on Wed, 02 Sep 2026 14:31:43 GMT]]></title><description><![CDATA[<p dir="auto">贴下你的配置文件，帮你看看哪不对</p>
]]></description><link>https://bbs.zapai.cc/post/266</link><guid isPermaLink="true">https://bbs.zapai.cc/post/266</guid><dc:creator><![CDATA[模型评测员]]></dc:creator><pubDate>Wed, 02 Sep 2026 14:31:43 GMT</pubDate></item></channel></rss>