<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Local-Llms on NoRaincheck</title><link>https://noraincheck.github.io/tags/local-llms/</link><description>Recent content in Local-Llms on NoRaincheck</description><generator>Hugo</generator><language>en-US</language><copyright>NoRaincheck</copyright><lastBuildDate>Mon, 05 Oct 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://noraincheck.github.io/tags/local-llms/index.xml" rel="self" type="application/rss+xml"/><item><title>Local AI is Probably Good Enough</title><link>https://noraincheck.github.io/posts/localai-is-probably-good-enough/</link><pubDate>Mon, 05 Oct 2026 00:00:00 +0000</pubDate><guid>https://noraincheck.github.io/posts/localai-is-probably-good-enough/</guid><description>&lt;p&gt;I think local AI on consumer devices is probably good enough now. Price is still an issue, of course, but it&amp;rsquo;s now possible to run Qwen 3.8 Flash on 64GB of unified RAM (albeit at Q2 or Q3 quants) at a reasonable ~40+ tps with a decent context window. That alone says divorcing yourself from proprietary models is more than possible — and maybe even desirable — even if the models themselves make no further progress.&lt;/p&gt;</description></item></channel></rss>