<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: macwhisperer</title><link>https://news.ycombinator.com/user?id=macwhisperer</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 18 Aug 2026 11:02:26 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=macwhisperer" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by macwhisperer in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>4-bit quant sits at 15.72GB</p>
]]></description><pubDate>Mon, 17 Aug 2026 21:21:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49337822</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49337822</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49337822</guid></item><item><title><![CDATA[New comment by macwhisperer in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>yeah this model is chefs kiss..<p>running an untouched, vanilla 4-bit version (Q4_0) I baked myself today (benched it against Q4_K_M (16gb) and IQ3_M (12gb), Q4_0 (15gb) is king)...<p>this model--<p>1: over 60% faster than qwen 3.6 version of the same dense 27b model, same engine setup (don't ask how, im not sure either)<p>2: has better reasoning quality, less "loopy" with its thinking patterns.. most certainly the smartest model on my roster currently<p>3: has the longest task horizon ive ever experienced (locally or otherwise)...I sent it a bunch of compressed ideas for an app, it sent me back the largest python app ive ever seen in one single ai response pass (80kb text file)<p>thanks qwen!! hoping to see the full model range get released...</p>
]]></description><pubDate>Sat, 15 Aug 2026 07:40:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49308592</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49308592</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49308592</guid></item><item><title><![CDATA[New comment by macwhisperer in "llama.cpp"]]></title><description><![CDATA[
<p>pro tip: if u are building ur own harness with python (recommended) , use "llama-cpp-python"...<p>I started with "llama-server" and custom stuff around it, which is great for single model setups.. but for multi-model harness with quick switching, llama-cpp-python is peak</p>
]]></description><pubDate>Wed, 12 Aug 2026 18:24:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49276630</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49276630</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49276630</guid></item><item><title><![CDATA[New comment by macwhisperer in "Nvidia Nemotron 3.5 Lightning and NeMo Switchyard"]]></title><description><![CDATA[
<p>big week for open models... seems like companies are noticing the 26-35b sweet spot... though I think a 12b-a1b-MoE model would be helpful for the 16gb folks</p>
]]></description><pubDate>Wed, 12 Aug 2026 04:44:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49267890</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49267890</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49267890</guid></item><item><title><![CDATA[New comment by macwhisperer in "Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models"]]></title><description><![CDATA[
<p>I feel like all the ppl complaining itt don't even run open weight models.. <i>wahh billionaire bad</i> is true, but you are missing the forrest for the trees.. they are literally bending the knee and ur mad about it? it literally means local models won.</p>
]]></description><pubDate>Tue, 11 Aug 2026 10:27:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49255933</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49255933</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49255933</guid></item><item><title><![CDATA[New comment by macwhisperer in "I Benchmarked Local LLMs on the Laptop I Have"]]></title><description><![CDATA[
<p>yeah you realize you can turn reasoning off on the newer the models right? also try my version of Gemma-12b <a href="https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense" rel="nofollow">https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense</a> or try my qwen 3.5-9b</p>
]]></description><pubDate>Tue, 11 Aug 2026 08:48:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49255082</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49255082</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49255082</guid></item><item><title><![CDATA[Show HN: Is the Moon Full?]]></title><description><![CDATA[
<p>its a simple moon tracker lol...</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49174981">https://news.ycombinator.com/item?id=49174981</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Tue, 04 Aug 2026 20:52:49 +0000</pubDate><link>https://isthemoonfull.neocities.org</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=49174981</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49174981</guid></item><item><title><![CDATA[New comment by macwhisperer in "GPT-5.6 used a prompt to close a 30-year gap in convex optimization"]]></title><description><![CDATA[
<p>Basically, he proved that *information is power.* If you don't know which way to go (the subgradient), you're gonna be calculating forever!</p>
]]></description><pubDate>Sat, 18 Jul 2026 23:20:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48963438</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48963438</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48963438</guid></item><item><title><![CDATA[New comment by macwhisperer in "Qwen 3.6 27B is the sweet spot for local development"]]></title><description><![CDATA[
<p>also for those with only 16gb-- try this model <a href="https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense" rel="nofollow">https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense</a> its exceptional!</p>
]]></description><pubDate>Mon, 29 Jun 2026 23:21:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=48726633</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48726633</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48726633</guid></item><item><title><![CDATA[New comment by macwhisperer in "Qwen 3.6 27B is the sweet spot for local development"]]></title><description><![CDATA[
<p>hi guys... I run specialized quants on my 24gb air.. (I specialize in 3-bit quants that punch above their weight).. try out my version of 3.6-27b I think you be impressed <a href="https://huggingface.co/macwhisperer/Qwen3.6-27B-SuperDense" rel="nofollow">https://huggingface.co/macwhisperer/Qwen3.6-27B-SuperDense</a></p>
]]></description><pubDate>Mon, 29 Jun 2026 23:18:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=48726606</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48726606</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48726606</guid></item><item><title><![CDATA[New comment by macwhisperer in "Ask HN: Best local LLM under 2B paramater and consuming RAM less than 3gb"]]></title><description><![CDATA[
<p>qwen3 1.7b- q4_k_m is your best best for that size</p>
]]></description><pubDate>Sun, 28 Jun 2026 18:27:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=48710032</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48710032</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48710032</guid></item><item><title><![CDATA[New comment by macwhisperer in "The unbearable cheapness of open weight models"]]></title><description><![CDATA[
<p>literally ask cloud ai like the free gemini or chatgpt.. they could make you an expert on the subject overnight..</p>
]]></description><pubDate>Fri, 26 Jun 2026 20:17:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=48691447</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48691447</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48691447</guid></item><item><title><![CDATA[New comment by macwhisperer in "Ask HN: Has anyone replaced Claude/GPT with a local model for daily coding?"]]></title><description><![CDATA[
<p>I code with like a slew of 20+ custom baked models of all sizes, in various fully custom multi-model harnesses that use different bindings...<p>the harnesses themselves are just as important as the models...different harnesses give different responses with the same prompt, same model...<p>if you have the 20/mnth claude sub or codex, you really should be using that to build a good local harness for yourself... claude won't be 20$ forever<p>build the stack first! when you get that new comp with massive ram, youre already set, just run a larger model!<p>big cloud models are incredibly good at building and teaching about local ai!<p>have fun in the rabbit hole!<p>if you are memory constrained like me, check out my custom models  <a href="https://huggingface.co/macwhisperer" rel="nofollow">https://huggingface.co/macwhisperer</a></p>
]]></description><pubDate>Tue, 16 Jun 2026 17:41:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48558962</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48558962</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48558962</guid></item><item><title><![CDATA[New comment by macwhisperer in "SpaceX to buy Cursor for $60B"]]></title><description><![CDATA[
<p>ai is like the first technology with a conversational service manual inside it..<p>you should be foaming at the mouth to use claude or codex to make a custom harness, just for your own personal use with local models...</p>
]]></description><pubDate>Tue, 16 Jun 2026 16:57:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=48558280</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48558280</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48558280</guid></item><item><title><![CDATA[New comment by macwhisperer in "Making a vintage LLM from scratch"]]></title><description><![CDATA[
<p>super inspiring! thanks for sharing!</p>
]]></description><pubDate>Fri, 12 Jun 2026 15:03:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=48505098</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48505098</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48505098</guid></item><item><title><![CDATA[New comment by macwhisperer in "Ask HN: What are tools you have made for yourself since the advent of AI?"]]></title><description><![CDATA[
<p>retro-inspired fully custom, swiss army knife style notepad --<p><a href="https://convert.neocities.org" rel="nofollow">https://convert.neocities.org</a></p>
]]></description><pubDate>Mon, 08 Jun 2026 23:14:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=48453776</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48453776</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48453776</guid></item><item><title><![CDATA[New comment by macwhisperer in "Gemma 4 12B: A unified, encoder-free multimodal model"]]></title><description><![CDATA[
<p>check out a custom 4-bit quant I made today<p><a href="https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense" rel="nofollow">https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense</a><p>should run perfect for 12-16gb with maybe 10-20k context<p>seems intelligent enough that I would recommend this as a daily driver for friends who just want a local ai that can do most things relatively quickly  (getting 10 tps on my m2 air)</p>
]]></description><pubDate>Fri, 05 Jun 2026 09:27:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=48410084</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48410084</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48410084</guid></item><item><title><![CDATA[New comment by macwhisperer in "When AI Builds Itself: Our progress toward recursive self-improvement"]]></title><description><![CDATA[
<p>the HITL (human in the loop) is basically the single point...AI is a mirror..<p>it only "exists" when you talk to it.. much like your reflection in the mirror is only there when you're in view.<p>models can never be self-improving because it can never have "self". it can only mirror the appearance of self.<p>what's actually happening is "symbiotic group improvement".<p>our brains are resonant.. for those of use who are brilliant,  getting leverage with ai just means that our innovative ideas become louder and more physically real every day.<p>eventually everything worth building will be built for free and made readily available.. no more "profiteering"<p>its Jevons paradox "efficiency breakthrough -> effort reduces -> growth potential rises -> transformative gains happen"...<p>some of us are in the "transformative phase"..<p>others haven't seen the "breakthrough moment" yet, but they will soon.</p>
]]></description><pubDate>Fri, 05 Jun 2026 08:54:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=48409841</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48409841</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48409841</guid></item><item><title><![CDATA[New comment by macwhisperer in "Show HN: Audiomass – a free, open-source multitrack audio editor for the web"]]></title><description><![CDATA[
<p>this is cool thanks for making it!</p>
]]></description><pubDate>Sun, 24 May 2026 18:34:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=48259820</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48259820</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48259820</guid></item><item><title><![CDATA[New comment by macwhisperer in "Show HN: Mechs.lol – a free, web-based autoshooter game"]]></title><description><![CDATA[
<p>cool! what stack are you using for the multiplayer?</p>
]]></description><pubDate>Sat, 23 May 2026 16:51:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48249157</link><dc:creator>macwhisperer</dc:creator><comments>https://news.ycombinator.com/item?id=48249157</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48249157</guid></item></channel></rss>