<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: scrlk</title><link>https://news.ycombinator.com/user?id=scrlk</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 17 Aug 2026 04:39:42 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=scrlk" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by scrlk in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>Unsloth publishes KL divergence numbers which measures how much the quantised probability distribution changes vs unquantised: <a href="https://unsloth.ai/docs/models/qwen3.8#quantization-analysis">https://unsloth.ai/docs/models/qwen3.8#quantization-analysis</a><p>It's a bit bare at the moment, I assume they are going to add further detail later (eg comparison to other quants), similar to their other releases.</p>
]]></description><pubDate>Fri, 14 Aug 2026 15:48:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49300405</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49300405</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49300405</guid></item><item><title><![CDATA[New comment by scrlk in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>Beats Opus 4.7 Max (w/ Claude Code) on DeepSWE (42.2 vs 40). Looks like Qwen's 27B models continue to pack some punch.<p>Unsloth's GGUF quants are up: <a href="https://huggingface.co/unsloth/Qwen3.8-27B-GGUF" rel="nofollow">https://huggingface.co/unsloth/Qwen3.8-27B-GGUF</a></p>
]]></description><pubDate>Fri, 14 Aug 2026 15:11:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49299805</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49299805</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49299805</guid></item><item><title><![CDATA[New comment by scrlk in "Show HN: C# Game Engine with its own scripting language and IDE"]]></title><description><![CDATA[
<p>Reminds me of the Tomorrow Corporation (developer of World of Goo) tech demo: <a href="https://www.youtube.com/watch?v=72y2EC5fkcE" rel="nofollow">https://www.youtube.com/watch?v=72y2EC5fkcE</a><p>Their entire stack is custom (language, compiler, IDE, build system, engine). Has neat features like time-travelling debugging and instant hot-reloading for code and assets.</p>
]]></description><pubDate>Fri, 14 Aug 2026 10:29:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49296853</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49296853</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49296853</guid></item><item><title><![CDATA[New comment by scrlk in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>With Opus 4.7, there's a 10 point improvement in the AA coding agent index when you swap out Claude Code for OpenCode: <a href="https://artificialanalysis.ai/agents/coding-agents#harness-comparison" rel="nofollow">https://artificialanalysis.ai/agents/coding-agents#harness-c...</a><p>To put that in perspective, the difference between GPT-5.6 Sol Max and 5.6 Luna Max is 8 points. That's a lot of extra performance that you can get for free just by using the best harness.</p>
]]></description><pubDate>Wed, 12 Aug 2026 23:10:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49279825</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49279825</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49279825</guid></item><item><title><![CDATA[New comment by scrlk in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>What harness are you using? DS V4 is harness sensitive.</p>
]]></description><pubDate>Wed, 12 Aug 2026 18:36:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49276768</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49276768</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49276768</guid></item><item><title><![CDATA[New comment by scrlk in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>Benefits of having a well performing hedge fund funding DeepSeek.<p>IIRC, Demis attempted to start a fund inside DeepMind but it was killed off. In an alternative world where he manages to pull that off, perhaps DeepMind would still be independent with Demis at the helm.</p>
]]></description><pubDate>Wed, 12 Aug 2026 17:35:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49275991</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49275991</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49275991</guid></item><item><title><![CDATA[New comment by scrlk in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>Benchmarks:<p><pre><code>    | Benchmark                | DS-V4-Pro | DS-V4-Flash | DS-V4-Pro | DS-V4-Flash | GLM-5.2   | Kimi-K3   | Opus-4.8  | Fable 5       |
    |                          | 0813      | 0731        | Preview   | Preview     |           |           |           | (w/ fallback) |
    |--------------------------|-----------|-------------|-----------|-------------|-----------|-----------|-----------|---------------|
    | HLE (wo/w tools)         | 42.7/60.0 | 37.8/51.5   | 37.7/48.2 | 34.8/45.1   | 40.5/54.7 | 43.5/56.0 | 49.8/57.9 | 53.3/63.0     |
    | Terminal Bench 2.1       | 87.9      | 82.7        | 72.1      | 61.8        | 81.0      | 88.3      | 85.0      | 88.0          |
    | NL2Repo                  | 61.5      | 54.2        | 38.5      | 39.4        | 48.9      | -         | 69.7      | -             |
    | Cybergym                 | 83.3      | 76.7        | 52.7      | 38.7        | -         | 80.0      | 78.3      | 83.1          |
    | DeepSWE                  | 62.7      | 54.4        | 12.8      | 7.3         | 46.2      | 67.5      | 58.0      | 70.0          |
    | Toolathlon-Verified      | 74.1      | 70.3        | 55.9      | 49.7        | 59.9      | 76.5      | 76.2      | 77.9          |
    | Agents' Last Exam        | 25.7      | 25.2        | 16.5      | 15.8        | 23.8      | 27.6      | 25.7      | -             |
    | AutomationBench (Public) | 31.8      | 25.1        | 12.8      | 10.8        | 12.9      | 30.8      | 27.2      | 29.1          |
    | DSBench-FullStack        | 71.1      | 68.7        | 41.8      | 37.0        | 61.8      | 73.7      | 71.6      | 77.2          |
    | DSBench-Hard             | 67.2      | 59.6        | 31.1      | 25.8        | 54.5      | 63.0      | 71.7      | 68.3          |
</code></pre>
Source: <a href="https://reddit.com/r/LocalLLaMA/comments/1vmi0fg/deepseek_v4pro0813_benchmarks/" rel="nofollow">https://reddit.com/r/LocalLLaMA/comments/1vmi0fg/deepseek_v4...</a></p>
]]></description><pubDate>Wed, 12 Aug 2026 16:42:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49275180</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49275180</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49275180</guid></item><item><title><![CDATA[New comment by scrlk in "Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows"]]></title><description><![CDATA[
<p>Will be interesting to see how Qwen3.8 27B compares against this once it releases this week. Seems like dense 30B is back in fashion?<p>EDIT: An open weight version of Muse Spark 1.2 is going to be released as well:<p><a href="https://x.com/alexandr_wang/status/2086756152034066792" rel="nofollow">https://x.com/alexandr_wang/status/2086756152034066792</a><p><a href="https://xcancel.com/alexandr_wang/status/2086756152034066792" rel="nofollow">https://xcancel.com/alexandr_wang/status/2086756152034066792</a></p>
]]></description><pubDate>Mon, 10 Aug 2026 10:50:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49241998</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49241998</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49241998</guid></item><item><title><![CDATA[New comment by scrlk in "The Alpha 21264 CPU: NT's Greatest RISC (1998)"]]></title><description><![CDATA[
<p>IIRC, the Sunway CPUs used in Chinese supercomputers have an Alpha-like design.</p>
]]></description><pubDate>Sun, 09 Aug 2026 15:17:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49232215</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49232215</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49232215</guid></item><item><title><![CDATA[New comment by scrlk in "Qwen3.8 Max now ranked as the best overall model by agentic index"]]></title><description><![CDATA[
<p>Different benchmarks:<p>> Artificial Analysis Agentic Index: Represents the weighted average of agentic capabilities benchmarks in the Artificial Analysis Intelligence Index (GDPval-AA v2, Tau³-Banking)<p>> Artificial Analysis Coding Agent Index v1.3 incorporates 3 benchmarks: DeepSWE, Terminal-Bench v2, and SWE-Atlas-QnA<p>Qwen3.8 Max is 55.4 on the Agentic Index but hasn't been tested for the Coding Agent Index.</p>
]]></description><pubDate>Thu, 06 Aug 2026 19:19:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49201068</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49201068</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49201068</guid></item><item><title><![CDATA[New comment by scrlk in "Discovery Loop"]]></title><description><![CDATA[
<p>"You're absolutely right! I shouldn't have pushed the anti-mass spectrometer to 105% power, causing a resonance cascade. This was a major oversight on my part."</p>
]]></description><pubDate>Wed, 05 Aug 2026 17:52:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49186386</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49186386</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49186386</guid></item><item><title><![CDATA[New comment by scrlk in "Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs"]]></title><description><![CDATA[
<p>What happened to the "strike team" that Sergey Brin was leading to try and improve Gemini's coding performance?</p>
]]></description><pubDate>Wed, 05 Aug 2026 16:33:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49185180</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49185180</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49185180</guid></item><item><title><![CDATA[New comment by scrlk in "Civilian plane crash in New Mexico tied to military GPS blocking"]]></title><description><![CDATA[
<p>To add to this, there's a good lecture called <i>Children of the Magenta Line</i> that discusses automation dependency: <a href="https://www.youtube.com/watch?v=5ESJH1NLMLs" rel="nofollow">https://www.youtube.com/watch?v=5ESJH1NLMLs</a></p>
]]></description><pubDate>Wed, 05 Aug 2026 12:18:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49181796</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49181796</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49181796</guid></item><item><title><![CDATA[New comment by scrlk in "Smaller, faster, safer: running Kimi and GLM at scale"]]></title><description><![CDATA[
<p>IMO, yes. For example, Qwen 3.x is insensitive to weight and KV cache quantisation, whereas Gemma 4 is more sensitive: <a href="https://localbench.substack.com/p/kv-cache-quantization-benchmark" rel="nofollow">https://localbench.substack.com/p/kv-cache-quantization-benc...</a></p>
]]></description><pubDate>Tue, 04 Aug 2026 09:26:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49166167</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49166167</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49166167</guid></item><item><title><![CDATA[New comment by scrlk in "Smaller, faster, safer: running Kimi and GLM at scale"]]></title><description><![CDATA[
<p>KL divergence is your friend when it comes to evaluating the effects of quantisation: <a href="https://en.wikipedia.org/wiki/Kullback%E2%80%93Leibler_divergence" rel="nofollow">https://en.wikipedia.org/wiki/Kullback%E2%80%93Leibler_diver...</a></p>
]]></description><pubDate>Mon, 03 Aug 2026 22:17:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49162200</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49162200</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49162200</guid></item><item><title><![CDATA[New comment by scrlk in "Smaller, faster, safer: running Kimi and GLM at scale"]]></title><description><![CDATA[
<p>Nice to see a provider being transparent about KV cache quantisation. I've been suspecting that some providers do this silently whilst heavily promoting their unquantised weights, even though KV quantisation can degrade quality more than weight quantisation.<p>However, I wish their testing were more detailed. Firstly, some model families are more sensitive to KV quantisation than others (only Kimi K2.6 was tested). Secondly, the evaluation suite they use to claim that  FP8 KV quantisation is indistinguishable is noticeably lacking coding benchmarks; in long-running tasks, minor tool call errors compound over time.</p>
]]></description><pubDate>Mon, 03 Aug 2026 19:18:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49160216</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49160216</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49160216</guid></item><item><title><![CDATA[New comment by scrlk in "Situational Awareness down 67% in July in AI stock rout"]]></title><description><![CDATA[
<p>> Aschenbrenner party blamed short sellers who targeted the firm’s positions for exacerbating the fund’s losses, the letter said. The letter compared Situational’s experience to a bank run.<p>4 years ago, it was SBF blaming Changpeng Zhao for shorting FTT and triggering a  run on FTX.<p>Now another EA has followed the path of making a lot of money relatively quickly and losing it just as fast, using the exact same arguments for why it happened.</p>
]]></description><pubDate>Fri, 31 Jul 2026 14:13:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49123393</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49123393</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49123393</guid></item><item><title><![CDATA[New comment by scrlk in "Situational Awareness down 67% in July in AI stock rout"]]></title><description><![CDATA[
<p><a href="https://archive.ph/PCjtG" rel="nofollow">https://archive.ph/PCjtG</a></p>
]]></description><pubDate>Fri, 31 Jul 2026 14:10:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49123365</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49123365</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49123365</guid></item><item><title><![CDATA[New comment by scrlk in "IMAX vs. IMAX 70mm: The difference between these two cinema formats"]]></title><description><![CDATA[
<p>1.43 IMAX isn't exclusive to 15/70mm, it can also be screened with IMAX's dual laser projection system.</p>
]]></description><pubDate>Fri, 31 Jul 2026 13:28:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49122893</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49122893</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49122893</guid></item><item><title><![CDATA[Leopold Aschenbrenner unwinds all public stock positions after steep losses]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.cnbc.com/2026/07/30/leopold-aschenbrenners-hedge-fund-is-facing-steep-ai-losses.html">https://www.cnbc.com/2026/07/30/leopold-aschenbrenners-hedge-fund-is-facing-steep-ai-losses.html</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49110556">https://news.ycombinator.com/item?id=49110556</a></p>
<p>Points: 12</p>
<p># Comments: 1</p>
]]></description><pubDate>Thu, 30 Jul 2026 14:29:57 +0000</pubDate><link>https://www.cnbc.com/2026/07/30/leopold-aschenbrenners-hedge-fund-is-facing-steep-ai-losses.html</link><dc:creator>scrlk</dc:creator><comments>https://news.ycombinator.com/item?id=49110556</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49110556</guid></item></channel></rss>