<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: npodbielski</title><link>https://news.ycombinator.com/user?id=npodbielski</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 02 Oct 2026 10:14:14 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=npodbielski" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by npodbielski in "Framework Desktop with 192GB RAM"]]></title><description><![CDATA[
<p>Yeah, I did the same. Now I am happily running Flash Next.</p>
]]></description><pubDate>Wed, 30 Sep 2026 17:44:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49912075</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49912075</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49912075</guid></item><item><title><![CDATA[New comment by npodbielski in "Framework Desktop with 192GB RAM"]]></title><description><![CDATA[
<p>For that price? It is better to buy DGX Spark or Strix Halo 50% more RAM is not worth about 5% more performance for 3 times the price.</p>
]]></description><pubDate>Wed, 30 Sep 2026 16:07:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49910752</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49910752</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49910752</guid></item><item><title><![CDATA[New comment by npodbielski in "Transformers Explained Visually"]]></title><description><![CDATA[
<p>I have no idea why but I was thinking about Optimus Prime and Megatron when I was clicking the link. I was a bit disappointed.<p>That is good too though.</p>
]]></description><pubDate>Tue, 22 Sep 2026 07:21:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49797687</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49797687</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49797687</guid></item><item><title><![CDATA[New comment by npodbielski in "Shapelearn Qwen 3.8 27B (13.1 GB VRAM)"]]></title><description><![CDATA[
<p>When I changed the number of draft tokens to 3 in both, it helped and they Draft is actually performing a bit better:<p>- draft: 67.17<p>- MTP:   64.18<p>Why they used those examples? Seems strange.</p>
]]></description><pubDate>Fri, 18 Sep 2026 14:39:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49755070</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49755070</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49755070</guid></item><item><title><![CDATA[New comment by npodbielski in "Shapelearn Qwen 3.8 27B (13.1 GB VRAM)"]]></title><description><![CDATA[
<p>Which was not he point because I was testing their solution for MPT and it was just funny addition. But of course in internet you always will find some 'well akchually' person straight from the meme.</p>
]]></description><pubDate>Fri, 18 Sep 2026 14:14:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49754739</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49754739</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49754739</guid></item><item><title><![CDATA[New comment by npodbielski in "Shapelearn Qwen 3.8 27B (13.1 GB VRAM)"]]></title><description><![CDATA[
<p>Well I tested it on 7900XTX with the same prompts and their draft model gave me about 30t/s. Their own snippet of code with regular MTP model gave me 60t/s.<p>Also model with their draft answered incorrectly. 
With MTP it answered correctly.<p>Question was: "Does MikroTik CRS312-4C+8XG-RM have combo ports?". The answer is Yes.</p>
]]></description><pubDate>Fri, 18 Sep 2026 13:30:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49754145</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49754145</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49754145</guid></item><item><title><![CDATA[New comment by npodbielski in "Qwen 3.8 Omni Flash"]]></title><description><![CDATA[
<p>I am using Flash Next for few weeks and it is very capable model. I just wish there would a way to have faster prefill because reloading longer sections of session sometimes can take even 2h. I stopped using Qwen 3.8 27B completely on my Strix Halo.</p>
]]></description><pubDate>Fri, 18 Sep 2026 12:51:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49753700</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49753700</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49753700</guid></item><item><title><![CDATA[New comment by npodbielski in "Qwen 3.8 Omni Flash"]]></title><description><![CDATA[
<p>Yes, for fun I tried OVH AI Endpoint and they do not have cache read at all. They bill you every time you send a prompt regardless if you are hit cache or not. One agent session was like 80M input and 300K output and I paid 30$ for that. Or rather I interrupted it and let my local Qwen finish it because cost was getting radicoulous.</p>
]]></description><pubDate>Fri, 18 Sep 2026 12:39:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49753573</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49753573</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49753573</guid></item><item><title><![CDATA[New comment by npodbielski in "Registration without a phone number on Signal will use zero-knowledge proofs"]]></title><description><![CDATA[
<p>what would be the usecase? sharing account with kids? it is easier to just install something else and register on throwaway email?</p>
]]></description><pubDate>Mon, 14 Sep 2026 13:17:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49696348</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49696348</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49696348</guid></item><item><title><![CDATA[New comment by npodbielski in "DeepSeek v4.1 Flash"]]></title><description><![CDATA[
<p>Or two gorgon halos?</p>
]]></description><pubDate>Thu, 10 Sep 2026 08:52:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49640528</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49640528</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49640528</guid></item><item><title><![CDATA[New comment by npodbielski in "TALA Is Open-Source"]]></title><description><![CDATA[
<p>Looks like really great tool to generate some graphs and diagrams for static file blogs.</p>
]]></description><pubDate>Tue, 08 Sep 2026 12:37:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49609518</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49609518</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49609518</guid></item><item><title><![CDATA[New comment by npodbielski in "What We Tell AI"]]></title><description><![CDATA[
<p>Wow. I just read couple, but... That seems terrible! People do that?<p>On the other hand my own father believes now he is an alien from outer space because someone generated stupid youtube video...</p>
]]></description><pubDate>Sun, 30 Aug 2026 18:12:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49501230</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49501230</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49501230</guid></item><item><title><![CDATA[New comment by npodbielski in "Anthropic's best AI model struggles to attract users as cheaper tools thrive"]]></title><description><![CDATA[
<p>I am running it on 32GB and I did not saw model loosing it context even after 4-5 compactions in pi. I am running sessions for few days sometimes. I think it looped once, but loop police extension stopped it. The only problem I have know is how pi compaction works, which is forcing full prefill which takes time and it is erroring a lot. I wrote my own compaction that should remove full prefil but it does not work. But this is the only problem with this setup and it is more problem with pi then the model. I much more prefer it to use Qwen then paid models: Claude forces me to do reauth every other day and codex models either are too costly or not capable enough.</p>
]]></description><pubDate>Tue, 25 Aug 2026 17:52:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49437901</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49437901</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49437901</guid></item><item><title><![CDATA[New comment by npodbielski in "Migrating a Synology NAS to a UniFi UNAS Pro 8 with Robocopy, SMB Multichannel"]]></title><description><![CDATA[
<p>I bought several cameras and they unvr early this year. It works great but I have no AI features because you have to buy they hardware just to accept some license (!). I mean... OK but no.<p>If they would allow me to send them email saying: "I hereby promise I wont record my neighbours having little tete-a-tete."<p>Sure I understand liability. But requiring consumer to buy another product to fully utilize they other product? Sorry but no. They have really cool hardware and nice UI but policy like that... no.</p>
]]></description><pubDate>Mon, 24 Aug 2026 18:52:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49424256</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49424256</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49424256</guid></item><item><title><![CDATA[New comment by npodbielski in "Everything I own, owned"]]></title><description><![CDATA[
<p>What about Raspberry Pi is not open?</p>
]]></description><pubDate>Mon, 24 Aug 2026 05:03:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49415338</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49415338</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49415338</guid></item><item><title><![CDATA[New comment by npodbielski in "DiffusionGemma Technical Report"]]></title><description><![CDATA[
<p>As you said: everything works on llama.cpp
Why it does not work on vllm? Of course you can say that it is AMD fault but there was an issue of abysmal performance of models on Strix Halo, that is open for half a year (<a href="https://github.com/vllm-project/vllm/issues/34579#issuecomment-5129108179" rel="nofollow">https://github.com/vllm-project/vllm/issues/34579#issuecomme...</a>) and nothing is happening there. They do not care about those use cases. Seems like they are going with bit players that will run vllm inside datacenters racks. Hobbyists does not matter.</p>
]]></description><pubDate>Fri, 21 Aug 2026 08:52:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49385512</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49385512</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49385512</guid></item><item><title><![CDATA[New comment by npodbielski in "DiffusionGemma Technical Report"]]></title><description><![CDATA[
<p>And it fails on rocm of course. This engine is such a hassle on AMD.</p>
]]></description><pubDate>Thu, 20 Aug 2026 20:35:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49379911</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49379911</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49379911</guid></item><item><title><![CDATA[New comment by npodbielski in "DiffusionGemma Technical Report"]]></title><description><![CDATA[
<p>Anybody was able to run this model in a server?</p>
]]></description><pubDate>Thu, 20 Aug 2026 19:00:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49378722</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49378722</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49378722</guid></item><item><title><![CDATA[New comment by npodbielski in "Composable Tests"]]></title><description><![CDATA[
<p>If it is not hard, with an experience of this guy, he could came up with a better example? It is your own words so I am sure you will agree? He did alright job so he could exercise gray matter a bit and came up with some API for saving customer data and then test it his way? I mean... This would be more interesting and "real world" so I am sure it would be worth to write longer blog post about it!</p>
]]></description><pubDate>Tue, 18 Aug 2026 19:45:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49351564</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49351564</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49351564</guid></item><item><title><![CDATA[New comment by npodbielski in "NeoBrowser: An MCP server that drives real Chrome with your logged-in sessions"]]></title><description><![CDATA[
<p>It spits out password in logs according to gif. No thank you.</p>
]]></description><pubDate>Tue, 18 Aug 2026 17:41:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49349437</link><dc:creator>npodbielski</dc:creator><comments>https://news.ycombinator.com/item?id=49349437</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49349437</guid></item></channel></rss>