<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: nabakin</title><link>https://news.ycombinator.com/user?id=nabakin</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 10 Sep 2026 17:05:12 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=nabakin" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by nabakin in "Jamesob's guide to running SOTA LLMs locally"]]></title><description><![CDATA[
<p>Are you running qwen3.6-27b on one 3090 with your KV cache at q4? Ime there is significant long-context recall accuracy degradation at that precision. I prefer putting the KV cache at q8 and working with the 120k context</p>
]]></description><pubDate>Fri, 03 Jul 2026 19:52:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=48779207</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=48779207</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48779207</guid></item><item><title><![CDATA[New comment by nabakin in "U.S. DOJ demands Apple and Google unmask over 100k users of car-tinkering app"]]></title><description><![CDATA[
<p>Then you can sandbox</p>
]]></description><pubDate>Fri, 15 May 2026 21:29:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=48154149</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=48154149</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48154149</guid></item><item><title><![CDATA[New comment by nabakin in "DeepSeek v4"]]></title><description><![CDATA[
<p>I think they were mistaken or maybe they were just referring to inference because I don't see anyone making that claim and it would be quite the news.</p>
]]></description><pubDate>Sun, 26 Apr 2026 12:39:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=47909849</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47909849</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47909849</guid></item><item><title><![CDATA[New comment by nabakin in "DeepSeek v4"]]></title><description><![CDATA[
<p>Probably because you said you used DeepSeek. People don't want to see AI in the comments and don't trust AI responses.</p>
]]></description><pubDate>Sat, 25 Apr 2026 00:38:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=47897500</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47897500</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47897500</guid></item><item><title><![CDATA[New comment by nabakin in "DeepSeek v4"]]></title><description><![CDATA[
<p>Yes, that's the footnote from citation [5].</p>
]]></description><pubDate>Fri, 24 Apr 2026 18:01:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=47893719</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47893719</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47893719</guid></item><item><title><![CDATA[New comment by nabakin in "DeepSeek v4"]]></title><description><![CDATA[
<p>> Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips.<p>That is a huge claim to make with no evidence.<p>I researched what you said, and I have found no statement to that effect in their paper[0], on huggingface[1], twitter[2], WeChat[3], or in their news release[4].<p>They only mention as a footnote in only the Chinese version of their news release that they plan to reduce inference costs with the Ascend 950 supernode when it releases[5]. The only mention of Huawei in their paper is that they validated a technique to lower interconnect bandwidth on Ascend NPUs and Nvidia GPUs[6].<p>[0] <a href="https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main/DeepSeek_V4.pdf" rel="nofollow">https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main...</a><p>[1] <a href="https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro" rel="nofollow">https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro</a><p>[2] <a href="https://xcancel.com/deepseek_ai/status/2047516922263285776" rel="nofollow">https://xcancel.com/deepseek_ai/status/2047516922263285776</a><p>[3] <a href="https://mp.weixin.qq.com/s/8bxXqS2R8Fx5-1TLDBiEDg" rel="nofollow">https://mp.weixin.qq.com/s/8bxXqS2R8Fx5-1TLDBiEDg</a><p>[4] <a href="https://api-docs.deepseek.com/news/news260424" rel="nofollow">https://api-docs.deepseek.com/news/news260424</a><p>[5] <a href="https://api-docs.deepseek.com/zh-cn/img/v4-price.png" rel="nofollow">https://api-docs.deepseek.com/zh-cn/img/v4-price.png</a><p>[6] Page 16</p>
]]></description><pubDate>Fri, 24 Apr 2026 16:55:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=47892830</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47892830</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47892830</guid></item><item><title><![CDATA[New comment by nabakin in "NIST scientists create 'any wavelength' lasers"]]></title><description><![CDATA[
<p>> When it comes to information transfer and processing, light can do things that electricity can’t. Photons — particles of light — are far zippier than electrons at working their way through circuits.<p>Electrons themselves don't move at the speed of light, but information transfer (i.e. communication) via electrons does happen close to the speed of light.<p>A subtle, but important, distinction that's often misunderstood and means computational performance gains would probably come from bandwidth, not latency.</p>
]]></description><pubDate>Sun, 19 Apr 2026 00:14:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=47820684</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47820684</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47820684</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>Dunno but there's a PR for it. Probably also more performant than Modular.</p>
]]></description><pubDate>Thu, 02 Apr 2026 20:55:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=47620051</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47620051</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47620051</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>If OP meant they have the fastest implementation of Gemma 4 on Blackwell at the moment, I guess that is technically true. I doubt that will hold up when TensorRT-LLM finishes their implementation though.</p>
]]></description><pubDate>Thu, 02 Apr 2026 19:34:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=47619130</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47619130</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47619130</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>I know Arc AGI 2 has a private test set and they have a good amount of results[0] but it's not a conventional benchmark.<p>Looking around, SWE Rebench seems to have decent protection against training data leaks[1]. Kagi has one that is fully private[2]. One on HuggingFace that claims to be fully private[3]. SimpleBench[4]. HLE has a private test set apparently[5]. LiveBench[6]. Scale has some private benchmarks but not a lot of models tested[7]. vals.ai[8]. FrontierMath[9]. Terminal Bench Pro[10]. AA-Omniscience[11].<p>So I guess we do have some decent private benchmarks out there.<p>[0] <a href="https://arcprize.org/leaderboard">https://arcprize.org/leaderboard</a><p>[1] <a href="https://swe-rebench.com/about" rel="nofollow">https://swe-rebench.com/about</a><p>[2] <a href="https://help.kagi.com/kagi/ai/llm-benchmark.html" rel="nofollow">https://help.kagi.com/kagi/ai/llm-benchmark.html</a><p>[3] <a href="https://huggingface.co/spaces/DontPlanToEnd/UGI-Leaderboard" rel="nofollow">https://huggingface.co/spaces/DontPlanToEnd/UGI-Leaderboard</a><p>[4] <a href="https://simple-bench.com/" rel="nofollow">https://simple-bench.com/</a><p>[5] <a href="https://agi.safe.ai/" rel="nofollow">https://agi.safe.ai/</a><p>[6] <a href="https://livebench.ai/" rel="nofollow">https://livebench.ai/</a><p>[7] <a href="https://labs.scale.com/leaderboard" rel="nofollow">https://labs.scale.com/leaderboard</a><p>[8] <a href="https://www.vals.ai/about" rel="nofollow">https://www.vals.ai/about</a><p>[9] <a href="https://epoch.ai/frontiermath/" rel="nofollow">https://epoch.ai/frontiermath/</a><p>[10] <a href="https://github.com/alibaba/terminal-bench-pro" rel="nofollow">https://github.com/alibaba/terminal-bench-pro</a><p>[11] <a href="https://artificialanalysis.ai/articles/aa-omniscience-knowledge-hallucination-benchmark" rel="nofollow">https://artificialanalysis.ai/articles/aa-omniscience-knowle...</a></p>
]]></description><pubDate>Thu, 02 Apr 2026 19:30:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=47619074</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47619074</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47619074</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>It's easy to game and human evaluation data has its trade-offs, but it's way easier to fake public benchmark results. I wish we had a source of high quality private benchmark results across a vast number of models like Lmarena. Having high quality human evaluation data would be a plus too.</p>
]]></description><pubDate>Thu, 02 Apr 2026 18:00:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=47617896</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47617896</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47617896</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>Faster than TensorRT-LLM on Blackwell? Or do you not consider TensorRT-LLM open source because some dependencies are closed source?</p>
]]></description><pubDate>Thu, 02 Apr 2026 17:43:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=47617646</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47617646</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47617646</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>It's referring to the Lmsys Leaderboard/Lmarena/Arena.ai[0]. It's very well-known in the LLM community for being one of the few sources of human evaluation data.<p>[0] <a href="https://arena.ai/leaderboard/chat" rel="nofollow">https://arena.ai/leaderboard/chat</a></p>
]]></description><pubDate>Thu, 02 Apr 2026 17:17:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=47617258</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47617258</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47617258</guid></item><item><title><![CDATA[New comment by nabakin in "Google releases Gemma 4 open models"]]></title><description><![CDATA[
<p>Public benchmarks can be trivially faked. Lmarena is a bit harder to fake and is human-evaluated.<p>I agree it's misleading for them to hyper-focus on one metric, but public benchmarks are far from the only thing that matters. I place more weight on Lmarena scores and private benchmarks.</p>
]]></description><pubDate>Thu, 02 Apr 2026 17:00:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=47617034</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47617034</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47617034</guid></item><item><title><![CDATA[New comment by nabakin in "Big data on the cheapest MacBook"]]></title><description><![CDATA[
<p>And nowadays we have Debian running in a VM on Android [1]<p>[1] <a href="https://www.zdnet.com/article/how-to-use-the-new-linux-terminal-on-android/" rel="nofollow">https://www.zdnet.com/article/how-to-use-the-new-linux-termi...</a></p>
]]></description><pubDate>Thu, 12 Mar 2026 16:16:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=47353103</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47353103</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47353103</guid></item><item><title><![CDATA[New comment by nabakin in "MacBook Pro with M5 Pro and M5 Max"]]></title><description><![CDATA[
<p>I would consider it reasonable if this was 4x TTFT and Throughput, but it seems like it's only for TTFT.</p>
]]></description><pubDate>Tue, 03 Mar 2026 16:49:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=47235140</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=47235140</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47235140</guid></item><item><title><![CDATA[New comment by nabakin in "Tesla kills Autopilot, locks lane-keeping behind $99/month fee"]]></title><description><![CDATA[
<p>And right to repair</p>
]]></description><pubDate>Fri, 23 Jan 2026 20:11:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=46737241</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=46737241</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46737241</guid></item><item><title><![CDATA[New comment by nabakin in "JetBlue flight averts mid-air collision with US Air Force jet"]]></title><description><![CDATA[
<p>Ty this is great</p>
]]></description><pubDate>Tue, 16 Dec 2025 18:02:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=46291868</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=46291868</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46291868</guid></item><item><title><![CDATA[New comment by nabakin in "JetBlue flight averts mid-air collision with US Air Force jet"]]></title><description><![CDATA[
<p>TIL Europe still has some presence in the Americas. Thought all of that was gone with the Monroe Doctrine</p>
]]></description><pubDate>Tue, 16 Dec 2025 02:39:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=46284104</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=46284104</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46284104</guid></item><item><title><![CDATA[New comment by nabakin in "France threatens GrapheneOS with arrests / server seizure for refusing backdoors"]]></title><description><![CDATA[
<p>Fyi it doesn't look like this post is listed on the frontpage anymore, even with the points it has. Not sure if it's intentional</p>
]]></description><pubDate>Mon, 24 Nov 2025 22:11:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=46039997</link><dc:creator>nabakin</dc:creator><comments>https://news.ycombinator.com/item?id=46039997</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46039997</guid></item></channel></rss>