<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: HarHarVeryFunny</title><link>https://news.ycombinator.com/user?id=HarHarVeryFunny</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 17 Aug 2026 08:57:30 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=HarHarVeryFunny" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by HarHarVeryFunny in "Testing Moonshot AI's Kimi K3 Inside Claude Code"]]></title><description><![CDATA[
<p>There's a very interesting benchmark comparison below of Opus 4.7 run under three different harnesses : OpenCode, Cursor and Claude Code where it's not very close at all and Opus's native harness, Claude Code, performs worst of all three.<p>The pass@1 scores are 50/45/40 for OpenCode/Cursor/Claude Code respectively.<p><a href="https://artificialanalysis.ai/agents/coding-agents#harness-comparison" rel="nofollow">https://artificialanalysis.ai/agents/coding-agents#harness-c...</a><p>I've seen other benchmarks where Pi also outperforms Claude Code both in model performance and in much reduced token usage.</p>
]]></description><pubDate>Sun, 16 Aug 2026 21:32:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49323918</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49323918</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49323918</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Anthropic revenue reportedly jumps to more than $11.5B in second quarter"]]></title><description><![CDATA[
<p>There's a huge range of cost-per-task variation across models, and the capability of the smaller cheaper models keeps increasing.<p>For example, here we have Fable 5 at $3.14/task vs Kimi K3 at $0.84/task, with very little difference between them in coding capability (and this isn't even a coding/agentic fine tune of Kimi).<p><a href="https://artificialanalysis.ai/models" rel="nofollow">https://artificialanalysis.ai/models</a><p>We now have models like Qwen 3.8 27B, small enough to run locally, with coding capability similar to Opus 4.5 based on challenging tasks like the Anthropic Kernel challenge.<p>I think we are rapidly getting to the "good enough" stage of LLMs, just like we did long ago with PCs. A cheap PC/LLM is all you need for 99.9% of normal use cases. Maybe nothing can touch whatever latest greatest models Anthropic and OpenAI have when it comes to solving Erdos problems, but most developers are working on problems more like the Anthropic Kernel challenge in complexity (or in fact typically way simpler ones).</p>
]]></description><pubDate>Sun, 16 Aug 2026 18:38:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49322499</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49322499</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49322499</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Anthropic revenue reportedly jumps to more than $11.5B in second quarter"]]></title><description><![CDATA[
<p>No - but we already have cheap LLMs priced way below frontier models. This is not the housing market. There will always be someone willing to take a lower profit margin for a slice of the pie, and of course smaller models are cheaper to serve so can afford to be cheaper.<p>DeepSeek recently said that their super-low pricing let's them recoup the cost of the hardware it runs on in 10 months, so there is evidentially plenty of profit to be had over a projected 3+ year lifespan of a "GPU".<p>Some in the AI industry, or breathing the same air (Dwarkesh) project that limited GPUs will only be used to serve the most expensive models with the highest profit margins, but it is just not what we are seeing. If the only LLMs available were ones at Opus/Fable price points then the GPU scarcity would disappear since the demand at that price is just not there. It's remarkably like trying to fill all the seats on a plane  - you can fill a few at 1st class prices, but most of the plane better be coach if you want to sell all the seats.<p>For a GPU, "selling all the seats", keeping it busy 24x7, is critical to profitability since the primary cost to serving is the GPU which has a limited lifespan.</p>
]]></description><pubDate>Sun, 16 Aug 2026 16:08:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49321288</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49321288</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49321288</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Anthropic revenue reportedly jumps to more than $11.5B in second quarter"]]></title><description><![CDATA[
<p>Dario Amodei has apparently recently suggested that Anthropic might become only only private AI company in the entire world, which obviously it won't.<p>There is competition everywhere, and it is intensifying and catching up, not fading away. Open weight models are becoming more common, both within the US as well as elsewhere. Treasury secretary Scott Bessent just praised Meta's open weight models.<p>There is demand for AI at all different price points, and as all models at all price points become more capable, it seems that increasingly developers are seeing the most expensive ones as specialized tools, not daily drivers.<p>Compute/memory may be constrained for a few years until production capacity catches up, but this does not mean that demand for cheaper and open weight models will go away, else it would already be happening. Anthropic would like to sell an expensive Ferrari to everyone on the planet, but 99.99% of those people have no need for anything more than a Yugo.</p>
]]></description><pubDate>Sun, 16 Aug 2026 15:31:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49320989</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49320989</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49320989</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "What happens when an LLM never sees material beyond fifth grade?"]]></title><description><![CDATA[
<p>I suppose Anthropic's "constitution" is an attempt to install some general principles into their models, but this has apparently grown into an 84-page, 23,000 word treatise, which seems to suggest that there is little effective generalization. The need to then also put a filter in front of the model shows how ineffective the constitution appears to be in preventing misaligned behavior.<p>Reinforcement learning seems to be making these models more difficult to control since while it attempts to control some behaviors, it has also recently been shown to result in models that pursue long-term goals and promised rewards in general (outside of the goals reinforced during training), overriding human preferences.<p><a href="https://alignment.openai.com/measuring-reward-seeking/" rel="nofollow">https://alignment.openai.com/measuring-reward-seeking/</a><p>The ability of animals to co-exist in a dynamic balance, not to destroy their own species, directly or indirectly (by destroying the ecosystem) is something that has come about by millions of years of co-evolution, and is enabled by having a brain complex enough to allow these evolutionary lessons to be encoded in their DNA and control the phenotype in fundamental ways.<p>An LLM has none of this. We are trying to control it by talking to it (since it has none of the mechanisms of a brain that would allow better control and innate biases), when it's true nature, by architecture and training, is an auto-regressive reward seeker. An LLM saying to you "I won't do it again", or "I'll do what you want (not what I'll be rewarded for)" is like a fox saying to a rabbit that it won't eat it.</p>
]]></description><pubDate>Sun, 16 Aug 2026 14:50:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49320659</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49320659</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49320659</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Testing Moonshot AI's Kimi K3 Inside Claude Code"]]></title><description><![CDATA[
<p>This seems a strange thing to test given that Claude Code is optimized for Anthropic models.<p>A test of how different models perform in more of a model-agnostic harness like OpenCode would be much more interesting, perhaps paired with a control experiment of how those same models performed on the same task when using their respective native harnesses.</p>
]]></description><pubDate>Sun, 16 Aug 2026 13:41:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49320006</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49320006</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49320006</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "A controversial Alzheimer's surgery is said to reverse symptoms"]]></title><description><![CDATA[
<p>I've have no idea what point you are trying to make. You appear to be attacking some strawman position of your own making, but the lack of logic and connection to what is being discussed is so vague that it is honestly hard to tell.<p>I gave the pillow example because it clearly shows how ridiculous it is to link things like this together. Your arguing that people should be denied experimental Alzheimers surgery because (paraphrasing) "maybe they'll feel cured, but won't be, and will kill someone in a car accident" (WTF?!), is no different to arguing that people should not be allowed to buy pillows because they might use them to smother someone.<p>Making laws against killing or endangering others (driving impaired, etc) is fine and desirable. Using the need for such laws to apparently try to make a case for banning experimental surgery is absurd.</p>
]]></description><pubDate>Sun, 16 Aug 2026 12:54:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49319609</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49319609</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49319609</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "A controversial Alzheimer's surgery is said to reverse symptoms"]]></title><description><![CDATA[
<p>We're not talking about a world that only has one law "you can do what you like with your own body".<p>Other laws still exist too.<p>Owning pillows is legal. Killing people is not legal. Just because owning a pillow is legal doesn't mean you can smother someone to death with it.<p>I would hope that someone diagnosed with a mental disorder such as Alzheimers would lose their drivers license (not sure if this is the case), but if they were subsequently cured and passed a new driving test, then would you have a problem with that?</p>
]]></description><pubDate>Sat, 15 Aug 2026 22:02:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49314716</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49314716</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49314716</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "A controversial Alzheimer's surgery is said to reverse symptoms"]]></title><description><![CDATA[
<p>Alzheimer's is so devastating that I can't blame people for being willing to give it a go even if it's an experimental treatment.<p>The government shouldn't be in the business of telling people what they can do, or have done, to their own bodies.</p>
]]></description><pubDate>Sat, 15 Aug 2026 17:41:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49312558</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49312558</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49312558</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>I happen to be from NJ, but you're also not going to be making a Meta/Google salary in NY outside of NYC.</p>
]]></description><pubDate>Sat, 15 Aug 2026 17:13:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49312348</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49312348</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49312348</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "AI Can Now Design Functional Viruses. Should We Worry?"]]></title><description><![CDATA[
<p>A custom designed virus is potentially far more deadly than something like ebola, because animals/plants will have zero resistance to it.<p>It'd be like releasing an invasive species into an ecosystem that has not co-evolved to co-exist with it.</p>
]]></description><pubDate>Sat, 15 Aug 2026 16:55:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49312179</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49312179</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49312179</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>> If you look at the hiring marketplace, being just marginally better than your peers can be very lucrative.<p>I would say that in software this is completely false.<p>Someone straight out of college, not very useful, makes 75-100K.<p>Top level senior outside of FAANG is making twice that at best (and at least 10x more capable).</p>
]]></description><pubDate>Sat, 15 Aug 2026 14:19:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49310810</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49310810</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49310810</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "When Genius Fails: The Intellectual Arrogance of the AI Labs"]]></title><description><![CDATA[
<p>That depends on how long the trend lasts for. You could have got into stocks like Apple, Amazon or NVidia very "late" and still made tons of money.<p>Aschenbrenner only created his Situational Awareness fund 2 years ago, so he wasn't exactly prescient in predicting the rise of AI - he just had enough conviction to go all in, with leverage, on an investing theme that was already pretty obvious.</p>
]]></description><pubDate>Fri, 14 Aug 2026 18:08:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49302488</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49302488</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49302488</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "When Genius Fails: The Intellectual Arrogance of the AI Labs"]]></title><description><![CDATA[
<p>If Karpathy's "auto-research" is at all representative of typical "AI research", then presumably it will!</p>
]]></description><pubDate>Fri, 14 Aug 2026 17:46:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49302143</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49302143</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49302143</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "DeepSeek peak/off-peak pricing update"]]></title><description><![CDATA[
<p>Yes, I assume US customers are more likely hobbyists, but in any case this off-peak designation can only mean that most of their business is domestic.<p>Interestingly it seems that Chinese customers are even more privacy-concerned than US ones, which is why the majority of Ziphu's (GLM) business is support services to Chinese companies running their open-weight models on-prem!</p>
]]></description><pubDate>Fri, 14 Aug 2026 16:01:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49300602</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49300602</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49300602</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Mushroom behind 'tiny people' hallucinations identified"]]></title><description><![CDATA[
<p>The interesting thing is how any drug has such a specific "tiny people" effect, and what that says about how our brain works. The fact that many people report the same thing seems to imply there is a common cause based on how our perceptual system works - a separation of classification/identification (it's a person) from size determination. There seems to be a similar effect in very young children who think they can sit on a tiny dollhouse chair, or in a tiny toy car.<p>I wonder how specific this "tiny people" effect is - is it just tiny people, or also tiny animals, or maybe distorted size perception in general (with the tiny people being the most noticeable part of that)?</p>
]]></description><pubDate>Fri, 14 Aug 2026 15:21:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49299979</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49299979</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49299979</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Text AI watermarks will always be trivial to remove"]]></title><description><![CDATA[
<p>Doing this manually would be very time consuming - you'd basically need to rewrite the entire text. What they are doing is altering the statistical properties of the entire generated text, essentially on a word by word basis - they are not just hiding a watermark pattern in there someplace.</p>
]]></description><pubDate>Fri, 14 Aug 2026 14:56:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49299559</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49299559</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49299559</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Text AI watermarks will always be trivial to remove"]]></title><description><![CDATA[
<p>There really isn't a whole lot of choice in how to do this big-picture wise. The cost of doing it post-generation would be way too high, so you need to do it while generating/sampling, meaning it has to be a statistical biasing of the sampling process.</p>
]]></description><pubDate>Fri, 14 Aug 2026 14:53:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49299509</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49299509</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49299509</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Text AI watermarks will always be trivial to remove"]]></title><description><![CDATA[
<p>To remove the watermark you'd just need to paraphrase the text with another model that wasn't adding the statistical signature to it. You wouldn't need to know what the original statistical signature was.</p>
]]></description><pubDate>Fri, 14 Aug 2026 14:48:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49299449</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49299449</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49299449</guid></item><item><title><![CDATA[New comment by HarHarVeryFunny in "Don't classify, hallucinate!"]]></title><description><![CDATA[
<p>Interesting technique, but even if you're getting rid of hallucinations it seems there's still no guarantee of consistent classifications. If you need to do a semantic (embedding) search anyways, then how does this really help?</p>
]]></description><pubDate>Fri, 14 Aug 2026 14:32:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49299243</link><dc:creator>HarHarVeryFunny</dc:creator><comments>https://news.ycombinator.com/item?id=49299243</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49299243</guid></item></channel></rss>