<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: XCSme</title><link>https://news.ycombinator.com/user?id=XCSme</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 17 Aug 2026 08:27:42 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=XCSme" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>Also, note that the Luna was run and costs were calculated through OpenRouter, via API. With a ChatGPT subscription it's likely even cheaper.<p>Also, Qwen 3.7 27B is actually Terra level.</p>
]]></description><pubDate>Mon, 17 Aug 2026 07:14:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49327413</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49327413</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49327413</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>I removed the extra links to sources for the electricity prices, but that's the average cost in EU, where I live.<p><a href="https://ec.europa.eu/eurostat/web/products-eurostat-news/w/ddn-20260505-1" rel="nofollow">https://ec.europa.eu/eurostat/web/products-eurostat-news/w/d...</a></p>
]]></description><pubDate>Mon, 17 Aug 2026 07:12:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49327411</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49327411</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49327411</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>My comparison of its reasoning  efforts[0] seems to show that it only really supports 3 modes: none, low, xhigh.<p>Low and medium are basically the same.<p>Also, the electricity it costs to run on a 3090 is not negligible, so that it's cheaper to use Luna high via API than Qwen 3.8 27b locally, hardware costs excluding.<p>[0]: <a href="https://aibenchy.com/compare/qwen-qwen3-8-27b-high/qwen-qwen3-8-27b-medium/qwen-qwen3-8-27b-low/qwen-qwen3-8-27b-none/" rel="nofollow">https://aibenchy.com/compare/qwen-qwen3-8-27b-high/qwen-qwen...</a></p>
]]></description><pubDate>Mon, 17 Aug 2026 06:58:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49327331</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49327331</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49327331</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>It's an amazing model, it's GLM-5.2 level[0], running locally...<p>I tested it on my 3090, took like 8 hours to benchmark it and my room became a furnace (35+ deg outside temp), but it's really good.<p>Now, in theory, you can talk directly to your computer and tell it what to do, and it does everything locally.<p>[0]: <a href="https://aibenchy.com/compare/qwen-qwen3-8-27b-medium/z-ai-glm-5-2-high/" rel="nofollow">https://aibenchy.com/compare/qwen-qwen3-8-27b-medium/z-ai-gl...</a></p>
]]></description><pubDate>Sat, 15 Aug 2026 12:30:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49310050</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49310050</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49310050</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>Msi afterburner, you go around 900mv curve editor, raise it up to normal clock frequency, and it uses like 260w instead 300w for same performance.<p>There's some YouTube guides for it.<p>I also undervolted my new 5070ti, same tdp, around 260w instead of 300w and like 8% better performance.</p>
]]></description><pubDate>Fri, 14 Aug 2026 23:38:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49305908</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49305908</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49305908</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>5900x, 3090 24gb (slightly undervolted), 128gb ddr4, running via Ollama.<p>I am benchmarking it now locally, will put the results and speed/tps on aibenchy.com</p>
]]></description><pubDate>Fri, 14 Aug 2026 19:31:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49303533</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49303533</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49303533</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>Oh, ok, that's like the average tps for most AI providers</p>
]]></description><pubDate>Fri, 14 Aug 2026 19:10:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49303283</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49303283</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49303283</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>Why slow? I see ~50tps on a single 3090</p>
]]></description><pubDate>Fri, 14 Aug 2026 18:59:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49303150</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49303150</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49303150</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>With default config via Ollama and 65k context I get 50tps on   a 3090.</p>
]]></description><pubDate>Fri, 14 Aug 2026 18:59:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49303135</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49303135</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49303135</guid></item><item><title><![CDATA[New comment by XCSme in "Gemini 3.7 Flash"]]></title><description><![CDATA[
<p>They both have horizontal scroll on mobile...</p>
]]></description><pubDate>Thu, 13 Aug 2026 18:56:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49290416</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49290416</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49290416</guid></item><item><title><![CDATA[New comment by XCSme in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>So where does the guarantee stop? At the GPU driver level? Firmware level?<p>What if the GPU has a custom bios flash that somehow logs the unencrypted prompts?</p>
]]></description><pubDate>Thu, 13 Aug 2026 10:49:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49284079</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49284079</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49284079</guid></item><item><title><![CDATA[New comment by XCSme in "Pixel Watch 5"]]></title><description><![CDATA[
<p>Why do I feel like there haven't been any hardware releases since the Macbook Neo? Market is so slow, probably due to the DRAM shortages.<p>No new GPUs, CPUs, technologies, etc...</p>
]]></description><pubDate>Thu, 13 Aug 2026 10:23:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49283893</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49283893</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49283893</guid></item><item><title><![CDATA[New comment by XCSme in "Grok 4.6"]]></title><description><![CDATA[
<p>Yeah, Sol models have higher in/out cost but are incredibly token efficient.<p>Also, in those tests Sol Low did better, but you can also compare the price vs Sol High, then it's getting a bit closer.<p>So Grok 4.6 is still not the best choice when paying API rates, but they are improving fast.<p>Also, the more important difference is that sol is a lot faster.</p>
]]></description><pubDate>Thu, 13 Aug 2026 10:12:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49283800</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49283800</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49283800</guid></item><item><title><![CDATA[New comment by XCSme in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>Open source doesn't matter if someone else is running it, right? They can change it?<p>As long as the prompt is not encrypted at some point, and I don't think LLMs can run on encrypted prompts, then it can be read.</p>
]]></description><pubDate>Thu, 13 Aug 2026 03:30:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49281472</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49281472</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49281472</guid></item><item><title><![CDATA[New comment by XCSme in "Nvidia Nemotron 3.5 Lightning and NeMo Switchyard"]]></title><description><![CDATA[
<p>This is not a solution for now, just a moonshot solution for maybe 5-10 years from now.<p>There's no immediate fix for the RAM supply issue, but as long as it's manufacturing issue, not a raw materials issue, time always solves it.</p>
]]></description><pubDate>Thu, 13 Aug 2026 01:19:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280738</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49280738</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280738</guid></item><item><title><![CDATA[New comment by XCSme in "Grok 4.6"]]></title><description><![CDATA[
<p>Grok 4.6 vs Sol 5.6 vs Opus 5:<p><a href="https://aibenchy.com/compare/openai-gpt-5-6-sol-low/x-ai-grok-4-6-high/anthropic-claude-opus-5-high" rel="nofollow">https://aibenchy.com/compare/openai-gpt-5-6-sol-low/x-ai-gro...</a></p>
]]></description><pubDate>Thu, 13 Aug 2026 00:12:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280266</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49280266</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280266</guid></item><item><title><![CDATA[New comment by XCSme in "Grok 4.6"]]></title><description><![CDATA[
<p>Does really well and ~2x cheaper than Qwen3.8 2.4T, they have same pricing but grok is around 2x more token efficient:<p><a href="https://aibenchy.com/compare/qwen-qwen3-8-2-4t-a95b-low/x-ai-grok-4-6-high/" rel="nofollow">https://aibenchy.com/compare/qwen-qwen3-8-2-4t-a95b-low/x-ai...</a></p>
]]></description><pubDate>Thu, 13 Aug 2026 00:10:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280248</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49280248</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280248</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen3.8-2.4T"]]></title><description><![CDATA[
<p>Interestingly, the high variant does a lot worse and failed to generate a valid SVG (and the low variant use more tokens than the high one, so maybe their reasoning efforts are not working properly).<p>The solar system animation is also the coolest looking I've seen, unfortunately the animation doesn't work:<p><a href="https://aibenchy.com/compare/qwen-qwen3-8-2-4t-a95b-low/qwen-qwen3-8-2-4t-a95b-high/x-ai-grok-4-6-high/?showcase=solar-system-css#showcase=c098d6ee18b748dd" rel="nofollow">https://aibenchy.com/compare/qwen-qwen3-8-2-4t-a95b-low/qwen...</a></p>
]]></description><pubDate>Thu, 13 Aug 2026 00:07:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280229</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49280229</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280229</guid></item><item><title><![CDATA[New comment by XCSme in "Qwen3.8-2.4T"]]></title><description><![CDATA[
<p>That's a really cool hamster [0], unfortunately it's really expensive now, 2x more expensive than Grok 4.6[1].<p>[0]: <a href="https://aibenchy.com/compare/x-ai-grok-4-6-high/bytedance-seed-seed-2-1-turbo-low/qwen-qwen3-8-2-4t-a95b-low/bytedance-seed-seed-2-0-code-low/#showcase=9f4b661ee684c1c2" rel="nofollow">https://aibenchy.com/compare/x-ai-grok-4-6-high/bytedance-se...</a><p>[1]: <a href="https://aibenchy.com/compare/x-ai-grok-4-6-high/bytedance-seed-seed-2-1-turbo-low/qwen-qwen3-8-2-4t-a95b-low/bytedance-seed-seed-2-0-code-low/" rel="nofollow">https://aibenchy.com/compare/x-ai-grok-4-6-high/bytedance-se...</a></p>
]]></description><pubDate>Thu, 13 Aug 2026 00:02:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280190</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49280190</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280190</guid></item><item><title><![CDATA[New comment by XCSme in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>I don't understand how this works?<p>Is it another proxy on top? What stops the provider from reading/storing the prompts at the LLM execution level?</p>
]]></description><pubDate>Wed, 12 Aug 2026 23:38:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280002</link><dc:creator>XCSme</dc:creator><comments>https://news.ycombinator.com/item?id=49280002</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280002</guid></item></channel></rss>