<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: benjiro29</title><link>https://news.ycombinator.com/user?id=benjiro29</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 25 Aug 2026 04:40:45 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=benjiro29" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by benjiro29 in "Anthropic's best AI model struggles to attract users as cheaper tools thrive"]]></title><description><![CDATA[
<p>*This is likely because of your thinking level. The difference between max and ultracode is primarily that the latter is max with a bunch of agents.*<p>I do not know why people even use Max or Ultra levels of thinking effort. Most of the time, i found its better to just run any of the frontier level models, with high or medium. Their capability beyond those levels is often diminishing returns, in exchange for double or quadrupling the costs (or usage).<p>And the bonus is that often at those levels, they tend to spin up less sub-agents that just eat away at tokens/usage, like its water in the desert.<p>A Max or Ultra is something that really only belongs if medium or high can not solve a very nasty bug or issue. Even in planning its often too much.</p>
]]></description><pubDate>Mon, 24 Aug 2026 11:13:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49418110</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49418110</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49418110</guid></item><item><title><![CDATA[New comment by benjiro29 in "Anthropic's best AI model struggles to attract users as cheaper tools thrive"]]></title><description><![CDATA[
<p>It did not exactly help that we saw traffic to DS (over OpenCode) jump from around 1.6T tokens per day, to over 14T token in a matter of days. Nobody has the compute to deal with such increases.<p>This keeps happening with every good new model release. People jumping from one to another, and as prices get lower, they start using the models even more.<p>People make not like to hear it but prices and usage limits are ways to shape traffic. The third option is the nuclear one like Kimi did, by just stopping to sell subscription at all. But that is something that DeepSeek can not do as all they offer is API.<p>Even OpenAI despite having the most compute is not immune to client influx = capacity issues. As people found their usage dropping, despite the push to the easier to run Luna models.<p>Reality is, that compute can not keep up with demand, especially when models get more capable and cheaper. What trigger people being using them more, what trigger compute crisis's.<p>This constant up and down cycle is going to keep happening for a long time, as this new market grows and eventually, somewhere in the future stabilizes. But yea, that is still going to be a few more years for sure.</p>
]]></description><pubDate>Mon, 24 Aug 2026 11:02:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49418018</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49418018</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49418018</guid></item><item><title><![CDATA[New comment by benjiro29 in "Anthropic's best AI model struggles to attract users as cheaper tools thrive"]]></title><description><![CDATA[
<p>I spend way too much time in all the LLM related subs, to the point that i consider it unhealthy (inc claude/anthropic subs).<p>Its in my opinion not wide spread at all and as today is literally the first time i ever hear anybody mention this.</p>
]]></description><pubDate>Mon, 24 Aug 2026 10:53:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49417955</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49417955</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49417955</guid></item><item><title><![CDATA[New comment by benjiro29 in "Anthropic appears to be A/B testing reduced effort levels in Claude Code"]]></title><description><![CDATA[
<p>* Anthropic's Cyber Verification Program // Codex + gotTAC approved*<p>Meanwhile the Chinese models are "go ham dude"...<p>If it was not for capacity issues, Chinese models have a higher change to just dominate.<p>> $2t company by the way<p>It used to be that OpenAI and Anthropic had such a moat around them, that such a valuation was worth it. But these days, its gross overvalued (like so many).<p>The more stuff is being pulled like cyber verifications, downgrading effort levels, downgrading usage (OpenAI), the more people move to those Open Weight Chinese models.<p>A fun recent event ...  <a href="https://opencode.ai/data/">https://opencode.ai/data/</a><p>When DeepSeek Flash 0731 came out and provided a massive jump in cheap inference capability. It resulted in a 10x increased OpenCode token usage.<p>It took a 2.5x to 5.0x price increase AND a reduction by 4x usage (later to 2x) usage, and several cheaper models + a free model, to push the traffic down.<p>Traffic towards open weight models is increasing, even if providers can not keep up with the influx of new customers. This is not something you want to see as two companies, trying to go for IPOs.<p>So the idea of stonewalling cyber capabilities, when the rest of the world is just doing whatever with open weight models, on their own hardware even! This entire strategy from Anthropic never made any sense.</p>
]]></description><pubDate>Sat, 22 Aug 2026 18:21:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49402311</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49402311</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49402311</guid></item><item><title><![CDATA[New comment by benjiro29 in "Canada will match US tariffs 'dollar for dollar' as trade talks break down"]]></title><description><![CDATA[
<p>*Nobody is willing to help with Iran even on symbolic level*<p>Most people do not realize how much of a shift that is.<p>In the past there was always a ton of EU countries helping the US out, in whatever crap the US started. And a lot of European soldiers died. Those soldiers did not help Europe but the US, because of being allies.<p>Even when some European countries got called names for not helping in Iraq because they called out that the whole Weapons Of Mass destruction was a load of bull** . It did not stop others from helping the US.<p>Yet, now, what we see as "help" is so minimal, and even countries like Spain not just refusing but active blocking. It shows how much relationships have deteriorated.<p>Trump thinks of Europa like a vassal state, and every time we say no or delay, it drives him nuts. Or our best "yes", if often a "well, we need to have agreement in all the countries, see you in a few years". Do not think its coincidence, its done deliberately.<p>In the past when the US said something, beyond maybe France, most countries quickly followed. This is gone ... We know the world has changed, and the old order is finished.</p>
]]></description><pubDate>Sat, 22 Aug 2026 16:33:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49401346</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49401346</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49401346</guid></item><item><title><![CDATA[New comment by benjiro29 in "Ox Alpha"]]></title><description><![CDATA[
<p>I am guessing its GLM 5.3 Air + Vision. A smaller then 250b model.</p>
]]></description><pubDate>Fri, 21 Aug 2026 11:36:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49386585</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49386585</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49386585</guid></item><item><title><![CDATA[New comment by benjiro29 in "DeepSeek-V4-Flash-Vision-Exp Is Now Live on the DeepSeek API Platform"]]></title><description><![CDATA[
<p>Benchmark results:<p><a href="https://x.com/deepseek_ai/status/2090730032574631962" rel="nofollow">https://x.com/deepseek_ai/status/2090730032574631962</a><p>Vision + increase in benchmark results.</p>
]]></description><pubDate>Fri, 21 Aug 2026 09:56:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49385934</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49385934</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49385934</guid></item><item><title><![CDATA[New comment by benjiro29 in "Claude Code May–August 2026 weekly limits promotion"]]></title><description><![CDATA[
<p>Only on openrouter for some reason. Not OpenAI Subs/API.</p>
]]></description><pubDate>Tue, 18 Aug 2026 20:14:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49351991</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49351991</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49351991</guid></item><item><title><![CDATA[New comment by benjiro29 in "Claude Code May–August 2026 weekly limits promotion"]]></title><description><![CDATA[
<p>Seeing how many people with Max accounts on Codex are complaining, your in for a rude awakening. The days that Codex was the undisputed usage king, seem to have been reversed. Its been most most noticed after they did like 20 resets in a month and half, ...</p>
]]></description><pubDate>Tue, 18 Aug 2026 20:10:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49351943</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49351943</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49351943</guid></item><item><title><![CDATA[New comment by benjiro29 in "Coin-sized device can hack a Boeing 737"]]></title><description><![CDATA[
<p>Fixed ...</p>
]]></description><pubDate>Sat, 15 Aug 2026 12:23:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49310021</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49310021</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49310021</guid></item><item><title><![CDATA[New comment by benjiro29 in "Coin-sized device can hack a Boeing 737"]]></title><description><![CDATA[
<p>> <i>That trust was what led to this incident where someone just walked into the airplane dressed as a maintainence worker and nobody stopped him.</i><p>O, it can be even worse, when you realize its a 30 year old problem ...<p><i>During the 1998 television show Schalkse Ruiters, presenters Bart De Pauw and Tom Lenaerts dressed in fake pilot uniforms. They bypassed security at Brussels Airport (Zaventem), entered a Boeing cockpit, and left undetected to expose safety flaws</i><p>The TV show got cancelled not long after this incident because of pollical backlash.<p>The show had a reputation of finding security flaws (like being able to transfer money from people bank accounts) and other issues, but the airport one was the end of the show.<p>People loved the show, as it forced companies, ... make changes to their processes that they normally never did. They even did follow-up episodes to see if the companies actually made changes.<p>Bit of public shaming to fix security issues... worked great. Until the show was cancelled. The number one rated show of the country ... Yea, there was absolute no correlation between its cancellation and the airport incident. Really ;)</p>
]]></description><pubDate>Sat, 15 Aug 2026 11:51:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49309829</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49309829</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49309829</guid></item><item><title><![CDATA[New comment by benjiro29 in "DeepSeek API Pricing Update"]]></title><description><![CDATA[
<p>The problem with DS Flash/Pro is that they are extreme reasoning heavy and step heavy. Step = cache hit. Reasoning = output hit. So the impact on those price increases will be felt much stronger.<p>I think that Flash is still a usable model but Pro is DOA... Even before the price difference between Flash and Pro, vs the intelligence / problem solving / tool calling did not make sense. But now that gap has widen even more. And there are just too many competitors models now close to that Pro price range.<p>Especially when we compare that competitive models offer subscription services that easily cut down the token price by 1:10. That makes Pro especially a bad value.<p>We shall see what the 3th party market is going to do, but i suspect that prices will be increased. If the argument was that DeepSeek increases price as they lack capacity, a company with access to billions, other 3th party providers that need to rent and have less optimized infrastructures will increase prices. Especially if they get hit hard with people moving around.<p>Its like we always see the same issue with popular models.<p>* GLM 5.2 is good, capacity issues, API price up, subscription heavy nerfs.
* Kimi K3 is good, capacity issues, API price up, subscription heavy nerfs.
* DeepSeek V4 GA is good,  capacity issues, API price up
* OpenAI GLM 5m, 10m active users. Subscription usage is sneakily tightened more and more.
* Anthropic Opus too popular, ...<p>That is the main issue. The AI users are people who actively easily move between companies. Pushing peak loads to each unprepared company, releasing load on the "less desired". And round we go ...</p>
]]></description><pubDate>Thu, 13 Aug 2026 15:03:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49287098</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49287098</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49287098</guid></item><item><title><![CDATA[New comment by benjiro29 in "Grok 4.6"]]></title><description><![CDATA[
<p>Strange that i do not experience this. Its been great in my experience. But that may simple be because i switched from typing most of my prompts. To just dictating my prompts in a long and convoluted way and letting the LLM extra the information.<p>It allows for much more context that flow with your thoughts. Where as when you type, you tend to shorten you thinking process trying to get the bulleting points in, but that often ignores smaller things. And then you think "i can add this later", but that never happens because rabbit chasing the LLM.<p>So far all the suggestion that Opus 5.0 offered me, always aligned with what i wanted. Its not just Opus that i noticed this with.</p>
]]></description><pubDate>Thu, 13 Aug 2026 11:00:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49284176</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49284176</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49284176</guid></item><item><title><![CDATA[New comment by benjiro29 in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p><i>I thought it was impossible to downvote posts?</i><p>User Posts can be downvoted but you need over 500 karma to have access to the downvote button. A Submission can not be downvoted.</p>
]]></description><pubDate>Wed, 12 Aug 2026 19:48:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49277652</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49277652</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49277652</guid></item><item><title><![CDATA[New comment by benjiro29 in "OpenAI’s head of ethics leaves less than a year after joining"]]></title><description><![CDATA[
<p>Its funny because "Shall we get rid of the FAA and the FDA then?" ... matches exactly with the current situation of the FDA and Points 3 > 5 ... The systematic nerfing of the FDA and other organizations their power.</p>
]]></description><pubDate>Wed, 12 Aug 2026 07:22:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49268900</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49268900</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49268900</guid></item><item><title><![CDATA[New comment by benjiro29 in "2027 memory capacity is reportedly sold out"]]></title><description><![CDATA[
<p>HBM requires stacking the chips. So they need to shave the layers, glue, stack more, shave again. They also require a substrate what is even more wafers. The issue is that a error in the stack means a lot of losses.<p>In order to get high bandwidth, you want memory as close as possible to the GPU. The more trace lane length = signal loss, bandwidth loss.. HBM is compact, and so you can stack 24GB modules, 8 around a GPU die.<p>If you tried to do that with normal memory, you need like 64 modules. So a a TON of traces more that all need to be equal length, and because so many = far away from the GPU = less bandwidth.<p>The issue is like stated above, its a process that waste a ton of wafers. Wafers that can make easily 3x more normal memory.<p>Intel with "Crescent Island" is trying to make a 160GB card using LPDDR5x memory but the bandwidth is only ~700GB/s.</p>
]]></description><pubDate>Sat, 08 Aug 2026 00:24:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49217733</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49217733</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49217733</guid></item><item><title><![CDATA[New comment by benjiro29 in "US strikes $1.2B deal to pay German firm to halt offshore wind projects"]]></title><description><![CDATA[
<p>Ironically, we are also moving to more capable / faster models that use less power. DeepSeek V4 Flash 0731 is so extreme capable and comparability to a lot of models cheap to run.<p>We are seeing stuff like AMD buying Taalas, with their Llama 3 8B on a chip, being able to push 17.000 tokens / second.<p><a href="https://chatjimmy.ai/" rel="nofollow">https://chatjimmy.ai/</a><p><i>Generated in 0,033s • 14.212 tok/s</i> ...<p>People need to understand, that even AI companies do not like to build datacenters and the high energy needs. It cost them money and they are also looking at ways to develop better hardware / solutions that gives them more inference for cheaper (what means less power draw / heat generated).</p>
]]></description><pubDate>Fri, 07 Aug 2026 12:32:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49209414</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49209414</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49209414</guid></item><item><title><![CDATA[New comment by benjiro29 in "Qwen3.8 Max now ranked as the best overall model by agentic index"]]></title><description><![CDATA[
<p><i>Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. </i><p>What cost the most in API. Input, Cached Input, or Output. There you have your answer.<p>Unfortunately, we have moved so much of the actual intelligence of models towards reasoning, what results in some models getting good scores, but this is because they are dumping a insane amount of reasoning tokens at the problem.<p>So a mid priced model, with heavy reasoning output, cost the same as a expensive model, with medium reasoning output.<p>Before the GPT Luna price drop of 80%, you actually had the same price if you used Luna High and Sol Low. With the difference that Sol Low was insane fast, and often way better code.<p><a href="https://deepswe.datacurve.ai/">https://deepswe.datacurve.ai/</a><p>Do not look at the top score but more what is on the horizontal axis as you go down. Sol Medium is frankly, was the best performance for dollar, until that Luna price drop. I will even argue that despite the higher price, Sol Medium is still way better despite Luna Max being cheaper.  Or Opus Low, one of the better values also.<p>What do you notice? Is that those models all have a high intelligence start point for their low setting. So that means they do not rely as much on output tokens aka thinking.</p>
]]></description><pubDate>Thu, 06 Aug 2026 20:49:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49202341</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49202341</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49202341</guid></item><item><title><![CDATA[New comment by benjiro29 in "ROI of $100 Claude Code Subscription"]]></title><description><![CDATA[
<p>A very interesting cost analyze of using AI and without by a software engineering.<p>Interesting quote by OP on reddit:<p>> <i>So I wouldn't say "AI saved me $200k". What actually happened is that without it I'd have quit around month nine, like every other side project in my graveyard.</i><p><a href="https://www.reddit.com/r/vibecoding/comments/1vgl5as/the_roi_of_my_100_claude_code_subscription/" rel="nofollow">https://www.reddit.com/r/vibecoding/comments/1vgl5as/the_roi...</a></p>
]]></description><pubDate>Wed, 05 Aug 2026 22:56:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49190183</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49190183</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49190183</guid></item><item><title><![CDATA[ROI of $100 Claude Code Subscription]]></title><description><![CDATA[
<p>Article URL: <a href="https://medium.com/@nenadmitt/roi-of-my-100-claude-code-subscription-028937c9f77e">https://medium.com/@nenadmitt/roi-of-my-100-claude-code-subscription-028937c9f77e</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49190182">https://news.ycombinator.com/item?id=49190182</a></p>
<p>Points: 2</p>
<p># Comments: 1</p>
]]></description><pubDate>Wed, 05 Aug 2026 22:56:40 +0000</pubDate><link>https://medium.com/@nenadmitt/roi-of-my-100-claude-code-subscription-028937c9f77e</link><dc:creator>benjiro29</dc:creator><comments>https://news.ycombinator.com/item?id=49190182</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49190182</guid></item></channel></rss>