<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
<title>CanRun guides</title>
<link>https://canrunllm.com/guides</link>
<description>Guides to running local LLMs on your own PC or Mac.</description>
<language>en</language>
<lastBuildDate>Wed, 30 Sep 2026 00:00:00 GMT</lastBuildDate>
<atom:link href="https://canrunllm.com/rss.xml" rel="self" type="application/rss+xml"/>
<item><title>How much VRAM do local LLMs need? 8–32 GB tiers (2026)</title><link>https://canrunllm.com/guides/vram-requirements</link><guid isPermaLink="true">https://canrunllm.com/guides/vram-requirements</guid><description>What fits on 8, 12, 16, 24 and 32 GB graphics cards: the memory formula, model-by-model verdicts, and how context length and quantization change the answer.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate></item>
<item><title>Ollama slow or forgetting your prompt? Context length defaults</title><link>https://canrunllm.com/guides/ollama-context-length</link><guid isPermaLink="true">https://canrunllm.com/guides/ollama-context-length</guid><description>Why Ollama trims long conversations or slows down: its default context length by VRAM tier, how to raise it, and what a longer context costs in memory.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate></item>
<item><title>GGUF quantization: Q4_K_M vs IQ4_XS vs Q8_0 — which to download</title><link>https://canrunllm.com/guides/gguf-quantization</link><guid isPermaLink="true">https://canrunllm.com/guides/gguf-quantization</guid><description>What each GGUF quant level costs in gigabytes, how we group them by quality, and when dropping one level turns a model that does not fit into one that runs.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate></item>
<item><title>Korean open LLMs (EXAONE, Kanana, HyperCLOVA X SEED, Solar): hardware guide</title><link>https://canrunllm.com/guides/korean-open-llms</link><guid isPermaLink="true">https://canrunllm.com/guides/korean-open-llms</guid><description>Specs, licenses and what it takes to run LG EXAONE, Kakao Kanana, NAVER HyperCLOVA X SEED and Upstage Solar locally, with verdicts on common GPUs and Macs.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate></item>
<item><title>Running local LLMs on a Mac (M4, M5): how much unified memory?</title><link>https://canrunllm.com/guides/mac-local-llm</link><guid isPermaLink="true">https://canrunllm.com/guides/mac-local-llm</guid><description>How much of a Mac’s unified memory a local model can use, which M4 and M5 configurations fit which models, and why memory bandwidth sets the speed.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate></item>
<item><title>KV cache explained: how context length eats VRAM</title><link>https://canrunllm.com/guides/kv-cache-context</link><guid isPermaLink="true">https://canrunllm.com/guides/kv-cache-context</guid><description>Why a longer context needs more memory, how much each model’s KV cache costs per token, and what q8_0 and q4_0 KV cache quantization buy you.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate></item>
</channel>
</rss>
