<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: disiplus</title><link>https://news.ycombinator.com/user?id=disiplus</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 13 Aug 2026 21:58:40 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=disiplus" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by disiplus in "ChatGPT Desktop (Codex Desktop) for Linux"]]></title><description><![CDATA[
<p>You don't care what were you trying to build but for a quick me alone Linux app I used Flutter.
Honestly it's my go-to when I want to have a quick native app on any platform but dont want to bundle electron.</p>
]]></description><pubDate>Thu, 13 Aug 2026 13:25:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49285550</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=49285550</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49285550</guid></item><item><title><![CDATA[New comment by disiplus in "France to ban unsolicited telemarketing calls"]]></title><description><![CDATA[
<p>It's not that they don't care, it's just that the protocol behind it was never designed for it.<p>Compared for example to email where when you want to send email to somebody, you discover its MX settings on the DNS level and then send it directly and the receiver can verify are you allowed to send those emails. With phone calls it's basically a network of interconnects so I'm sending the call directly to my provider but that provider to reach the destination has multiple steps and they have to trust each other with the numbers signalling. Compare it to a regular snail mail where everybody can write anything as a sender.<p>What most of the providers are switching to now is for enduser the case is if the phone number that you want to signal is not registered with us, you are not allowed to signal it. Which is fine and covers most of the cases but imagine that you are a business and have multiple providers And you have boughta number block with one provider and another one is offering you better minute price for outgoing calls. Or you want to have a forwarding to your cell phone number. Or you want to signal your number when you're calling prospects from CRM. The regulation usuall differs from country to country, but most allow you to signal any number that you own with any provider that you use. And this is correct, but the problem is verification.</p>
]]></description><pubDate>Tue, 11 Aug 2026 13:02:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49257636</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=49257636</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49257636</guid></item><item><title><![CDATA[New comment by disiplus in "Kimi Work"]]></title><description><![CDATA[
<p>They all are pretty similar. And if it was copying anything, I would say more inspiration from claude desktop then codex. The Zcode is more like codex.<p><a href="https://drive.google.com/file/d/1JFfgfMO0nO7HR0WHwEqEEjXIIQjzzJ55/view?usp=sharing" rel="nofollow">https://drive.google.com/file/d/1JFfgfMO0nO7HR0WHwEqEEjXIIQj...</a></p>
]]></description><pubDate>Tue, 21 Jul 2026 11:18:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=48990761</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48990761</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48990761</guid></item><item><title><![CDATA[New comment by disiplus in "Annoying and alarming things about OpenCode"]]></title><description><![CDATA[
<p>I found this part funny.<p>> People familiar with OpenCode internals (if you are on the OpenCode dev team I assume this doesn’t include you) might have objected to my python3 example above.</p>
]]></description><pubDate>Mon, 20 Jul 2026 16:04:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=48980743</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48980743</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48980743</guid></item><item><title><![CDATA[New comment by disiplus in "Annoying and alarming things about OpenCode"]]></title><description><![CDATA[
<p>Cache Misses are pretty bad, i have a locally running deepseek v4 flash, i have tuned it now to have 1100-1300 prefill. Its not great but properly useable. Imagine having a session with already 100k and half of it has to be prefilled it would be waiting minutes with worse numbers. And if you are paying by the token, for hosted models, you are wasting money.</p>
]]></description><pubDate>Mon, 20 Jul 2026 14:09:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=48979115</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48979115</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48979115</guid></item><item><title><![CDATA[New comment by disiplus in "I think Anthropic and OpenAI have found product-market fit"]]></title><description><![CDATA[
<p>Which model are you running ?</p>
]]></description><pubDate>Sun, 31 May 2026 16:30:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=48347061</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48347061</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48347061</guid></item><item><title><![CDATA[New comment by disiplus in "The solution might be cancelling my AI subscription"]]></title><description><![CDATA[
<p>Diagnosed with ADHD, ultimately does not change anything for me even through i had the same idea as you. Reason is that i can now start even more stuff in parallel. And some part of them get finished more before i can just prompt more when in focus, but instead of finishing i add more features.</p>
]]></description><pubDate>Sun, 31 May 2026 16:24:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=48347000</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48347000</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48347000</guid></item><item><title><![CDATA[New comment by disiplus in "I think Anthropic and OpenAI have found product-market fit"]]></title><description><![CDATA[
<p>same, but you need more then 100k of hw to run something like kimi k2.6 for a bigger team. on the other hand there is a ds4 flash that you can run on a macbook with 128gb ram. an that one is perfectly usable for a lot of tasks.<p><a href="https://github.com/antirez/ds4" rel="nofollow">https://github.com/antirez/ds4</a></p>
]]></description><pubDate>Wed, 27 May 2026 20:31:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=48300199</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48300199</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48300199</guid></item><item><title><![CDATA[New comment by disiplus in "Agents can now create Cloudflare accounts, buy domains, and deploy"]]></title><description><![CDATA[
<p>The problem is not website, the problem is discovery and discovery is on Instagram, TikTok, and social networks. You don't have any incentive to build a website for a regular audience. What you might do is build an audience on a social network and then try to move them to a website.<p>But at that point you're big enough to build it properly.</p>
]]></description><pubDate>Wed, 06 May 2026 10:13:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=48034481</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48034481</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48034481</guid></item><item><title><![CDATA[New comment by disiplus in "Accelerating Gemma 4: faster inference with multi-token prediction drafters"]]></title><description><![CDATA[
<p>depends, a super small one finetuned to do function calling instead sending it to big model and waiting, instead, you ask for a revenue in last month, i do a small llm function call -> show results. some bigger ones, analysis, summary, classification.
what is great with smaller ones, and im looking at 2b, 4b is you can get a huge throughput with just vllm and a couple of consumer gpus.
what i usually do is basically distillation of a big one onto smaller one.</p>
]]></description><pubDate>Tue, 05 May 2026 17:31:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=48025727</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48025727</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48025727</guid></item><item><title><![CDATA[New comment by disiplus in "Accelerating Gemma 4: faster inference with multi-token prediction drafters"]]></title><description><![CDATA[
<p>i dont know what are you talking about, i replaced an older gpt4o with a finetuned qwen. there is a huge amount of "AI, that can be done with those models, or partly by those models." Huge amount of people would not notice the difference. And if you prepare the context correctly, even bigger slice of people would not notice.</p>
]]></description><pubDate>Tue, 05 May 2026 17:02:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48025270</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48025270</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48025270</guid></item><item><title><![CDATA[New comment by disiplus in "Accelerating Gemma 4: faster inference with multi-token prediction drafters"]]></title><description><![CDATA[
<p>nice, will run it later agains qwen3.6 27b, the speed was one of the reasons why in was running qwen and not gemma. the difference was big, there is some magic that happpens when you have more then 100tps.</p>
]]></description><pubDate>Tue, 05 May 2026 16:54:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=48025139</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=48025139</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48025139</guid></item><item><title><![CDATA[New comment by disiplus in "DeepSeek v4"]]></title><description><![CDATA[
<p>Depends how many users you have and what is "production grade" for you but like 500k gets you a 8x B200 machine.</p>
]]></description><pubDate>Fri, 24 Apr 2026 05:06:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=47885778</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47885778</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47885778</guid></item><item><title><![CDATA[New comment by disiplus in "Kimi K2.6: Advancing open-source coding"]]></title><description><![CDATA[
<p>was part of the beta, its properly good model, in some sense i forgot that im not on opus or gpt. opus is still better. gpt is the one struggling for me. it has some niche in backend work but you can get the same with opus with skills, its lacking in almost all others.</p>
]]></description><pubDate>Mon, 20 Apr 2026 20:40:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=47840225</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47840225</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47840225</guid></item><item><title><![CDATA[New comment by disiplus in "ChatGPT Pro now starts at $100/month"]]></title><description><![CDATA[
<p>It looks like its called prolite.<p><a href="https://snipboard.io/jmGKfM.jpg" rel="nofollow">https://snipboard.io/jmGKfM.jpg</a></p>
]]></description><pubDate>Thu, 09 Apr 2026 18:42:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=47707850</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47707850</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47707850</guid></item><item><title><![CDATA[New comment by disiplus in "I've sold out"]]></title><description><![CDATA[
<p>yet</p>
]]></description><pubDate>Wed, 08 Apr 2026 10:53:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=47688390</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47688390</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47688390</guid></item><item><title><![CDATA[New comment by disiplus in "GLM-5.1: Towards Long-Horizon Tasks"]]></title><description><![CDATA[
<p>i have glm and kimi. kimi was in most of the cases better and my replacement for claude when i run out of tokens. Now im finding myself using glm more then kimi. Its funny that glm vs kimi, is like codex vs claude. Where glm and codex are better for backend and kimi and claude more for frontend.<p>as kimi did a huge amount of claude distilation it seems to be somewhat based in data<p><a href="https://www.anthropic.com/news/detecting-and-preventing-distillation-attacks" rel="nofollow">https://www.anthropic.com/news/detecting-and-preventing-dist...</a></p>
]]></description><pubDate>Tue, 07 Apr 2026 19:18:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=47680049</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47680049</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47680049</guid></item><item><title><![CDATA[New comment by disiplus in "GLM-5.1: Towards Long-Horizon Tasks"]]></title><description><![CDATA[
<p>Yeah it seems they did not align it to much, at least for now. Yesterday it helped me bypass the bot detection on a local marketplace. that i wanted to scrap some listing for my personal alerting system. Al the others failed but glm5.1 found a set of parameters and tweaks how to make my browser in container not be detected.</p>
]]></description><pubDate>Tue, 07 Apr 2026 19:13:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=47680000</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47680000</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47680000</guid></item><item><title><![CDATA[New comment by disiplus in "GLM-5.1: Towards Long-Horizon Tasks"]]></title><description><![CDATA[
<p>basically my expirience as well. Sometimes it can break past 100k and be ok, but mostly it breaks down.</p>
]]></description><pubDate>Tue, 07 Apr 2026 19:08:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=47679931</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47679931</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47679931</guid></item><item><title><![CDATA[New comment by disiplus in "GLM-5.1: Towards Long-Horizon Tasks"]]></title><description><![CDATA[
<p>When it works and its not slow it can impress. Like yesterday it solved something that kimi k2.5 could not. and kimi was best open source model for me. But it still slow sometimes. I have z.ai and kimi subscription when i run out of tokens for claude (max) and codex(plus).<p>i have a feeling its nearing opus 4.5 level if they could fix it getting crazy after like 100k tokens.</p>
]]></description><pubDate>Tue, 07 Apr 2026 18:15:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=47679203</link><dc:creator>disiplus</dc:creator><comments>https://news.ycombinator.com/item?id=47679203</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47679203</guid></item></channel></rss>