<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: nearbuy</title><link>https://news.ycombinator.com/user?id=nearbuy</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 28 Aug 2026 21:29:59 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=nearbuy" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by nearbuy in "Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users"]]></title><description><![CDATA[
<p>It also looks like they're saturating the test, with one LLM hitting the maximum possible score. (<a href="https://www.trackingai.org/home" rel="nofollow">https://www.trackingai.org/home</a>)<p>The test wasn't made to accurately measure IQs that high.</p>
]]></description><pubDate>Fri, 07 Aug 2026 07:29:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49207035</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49207035</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49207035</guid></item><item><title><![CDATA[New comment by nearbuy in "Position: LLMs Can't Jump"]]></title><description><![CDATA[
<p>The paper also fails to show that their central example, Einstein, relied on sensory experience for his intuition leaps rather than general reasoning. They just kind of claim that thought experiments require sensory experience. But you can easily ask an LLM to perform a thought experiment and simulate an outcome, and the SOTA LLMs generally seem to do about as well as a human. An LLM would certainly know that freefall feels the same as zero gravity, even if they haven't felt the sensation, which was the key intuition the paper talks about for Einstein's General Relativity. The paper's author would probably say any examples of this don't count, but without clear criteria for what would count, their claim is unfalsifiable.</p>
]]></description><pubDate>Wed, 05 Aug 2026 14:58:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49183923</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49183923</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49183923</guid></item><item><title><![CDATA[New comment by nearbuy in "Qwen3.8-Max: A New Bar for Coding and Cowork"]]></title><description><![CDATA[
<p>Only a tiny, tiny fraction of the parameters are encoding information that's specific to a particular programming language. Even if you could remove those without degrading performance, it would have a negligible effect on the model size.</p>
]]></description><pubDate>Mon, 03 Aug 2026 20:12:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49160820</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49160820</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49160820</guid></item><item><title><![CDATA[New comment by nearbuy in "Qwen3.8-Max: A New Bar for Coding and Cowork"]]></title><description><![CDATA[
<p>The problem is Claude Fable is now better than most programmers I know at software architecture and performance optimization as well.</p>
]]></description><pubDate>Mon, 03 Aug 2026 19:00:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49160005</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49160005</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49160005</guid></item><item><title><![CDATA[New comment by nearbuy in "Gemini Robotics 2 brings whole body intelligence to robots"]]></title><description><![CDATA[
<p>That's a good reason to use a dishwasher, and you should keep doing it. But the overall waste is small and people aren't going to care.<p>I'm not saying people who already have a dishwasher will throw it away. But a robot makes installing a new dishwasher much less tempting.<p>At my utility prices, it costs about 5¢ to hand-wash one serving (plate, pan, glass, utensils, and a bowl) with running water. That's not nothing, but it's dwarfed by the energy use of using the oven and the supply-chain energy and water use of the food if it includes meat.<p>People will casually waste 10x that energy without a second thought when they don't need to just because it's convenient, or because they're used to doing things that way.<p>The dishwasher looks more unfavorable if you're considering installing one in a small apartment or condo. Apart from the cost of the dishwasher, you need space to install it, and more cupboard space to keep more dishes, glasses, etc. For an extra 4¢ per person-meal, having your robot do the dishes buys you the convenience of skipping all that and having all your dishes clean and ready all the time.</p>
]]></description><pubDate>Sun, 02 Aug 2026 06:19:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=49141648</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49141648</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49141648</guid></item><item><title><![CDATA[New comment by nearbuy in "Is AI reasoning right for the wrong reasons?"]]></title><description><![CDATA[
<p>Sure, but then there's no such thing as a network that isn't a classifier. Every physically computable function that terminates in finite time will map an input to a fixed set of outputs. And it goes against the common usage, where in machine learning we talk about classifiers, regressors, generative models, etc. as different things. They all become classifiers.<p>The parent commenter was trying to draw some insight from LLMs being classifiers that wouldn't apply equally to everything else.</p>
]]></description><pubDate>Sat, 01 Aug 2026 19:38:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49137664</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49137664</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49137664</guid></item><item><title><![CDATA[New comment by nearbuy in "Is AI reasoning right for the wrong reasons?"]]></title><description><![CDATA[
<p>The LLM has processed two data modalities derived from the apple (text and vision). Your brain processed a third (taste). But it is still just a data stream, sensing compounds and chemical properties of the apple and turning it into a stream of electric signals that reach your brain. You don't have any kind of ground-truth data stream that's inherently more powerful than what could potentially be fed to an AI.</p>
]]></description><pubDate>Sat, 01 Aug 2026 04:47:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49131128</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49131128</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49131128</guid></item><item><title><![CDATA[New comment by nearbuy in "Is AI reasoning right for the wrong reasons?"]]></title><description><![CDATA[
<p>LLMs are not classifiers. A classifier is an algorithm or neural net that assigns a label from a fixed set of labels to an input.<p>You can broaden the definition of classifier to anything that internally divides its input space into regions, but that definition would include every neural network, whether biological or artificial. So it's not very meaningful, and certainly doesn't give any insight into how they differ from humans.</p>
]]></description><pubDate>Sat, 01 Aug 2026 04:35:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49131074</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49131074</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49131074</guid></item><item><title><![CDATA[New comment by nearbuy in "Gemini Robotics 2 brings whole body intelligence to robots"]]></title><description><![CDATA[
<p>I have a dishwasher and I still usually just hand wash. It takes about 10 seconds to wash a dish.<p>The side benefit is all your dishes are always available. With the dishwasher, up to one full dishwasher load are dirty at any time, which means you need more dishes and more cupboard space than someone without a dishwasher.<p>If the robot is going to clear your dishes and wash them for you right away, what's the dishwasher for? To save the robot 10 minutes?</p>
]]></description><pubDate>Fri, 31 Jul 2026 01:07:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117901</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49117901</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117901</guid></item><item><title><![CDATA[New comment by nearbuy in "Gemini Robotics 2 brings whole body intelligence to robots"]]></title><description><![CDATA[
<p>Unitree's R1 humanoid robot is only about $6000, and it's still a nascent, smallish scale technology. They will come down.<p>If the future home robots are any good, it saves you from buying a dishwasher and robot vacuum. It can replace a maid/cleaning service and gardener/lawn-mowing service. For anyone paying for those services, it pays for itself.<p>$1000 isn't that high. Probably about a third of American households have an appliance over $2000, which is still a massive market.</p>
]]></description><pubDate>Thu, 30 Jul 2026 16:29:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49112273</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49112273</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49112273</guid></item><item><title><![CDATA[New comment by nearbuy in "US citizen charged after GrapheneOS phone wipes during airport search"]]></title><description><![CDATA[
<p>Yes, much like that. If he hadn't destroyed evidence, he could have argued it was malicious prosecution.</p>
]]></description><pubDate>Wed, 29 Jul 2026 07:30:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49094421</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49094421</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49094421</guid></item><item><title><![CDATA[New comment by nearbuy in "US citizen charged after GrapheneOS phone wipes during airport search"]]></title><description><![CDATA[
<p>What you're describing is malicious prosecution or abuse of process. It's illegal and it would destroy the prosecution's case. Not only that, but the victim could sue for damages.</p>
]]></description><pubDate>Mon, 27 Jul 2026 04:22:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49065170</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49065170</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49065170</guid></item><item><title><![CDATA[New comment by nearbuy in "OpenAI’s accidental attack against Hugging Face is science fiction that happened"]]></title><description><![CDATA[
<p>The irony is in this case the in-context and classifier "guardrails" would have almost certainly stopped the attack while their attempts at your definition of guardrails (the sandboxing) failed. In general, people keep trying to make secure systems and they fail with surprising regularity. Saying "they should have had better security" every time someone gets hacked is perhaps true, but it's not going to stop hacks from happening. And it's not a sufficient strategy on its own against future LLMs. "The Bitter Lesson" probably applies here.</p>
]]></description><pubDate>Fri, 24 Jul 2026 05:53:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49031652</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=49031652</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49031652</guid></item><item><title><![CDATA[New comment by nearbuy in "$100 AI Music Video: Claude Fable 5 vs. GPT-5.6 Sol"]]></title><description><![CDATA[
<p>It's a strange experiment. Claude and GPT aren't generating the video. They're directing and editing it, and they request video from a generative video model using mainly text-to-video. Neither Claude nor GPT can actually watch video content, and the video model generates shoddy quality clips with artifacts, lack of consistency, and where actions aren't synced to the music.<p>It would be very hard for anyone to make a good music video with the tools Fable and Sol had available to them. They don't have precise control over the clips that get generated, and they can't see or hear the result, other than screenshots. So I'm not sure what the experimenters expected.</p>
]]></description><pubDate>Fri, 17 Jul 2026 03:56:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=48943183</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48943183</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48943183</guid></item><item><title><![CDATA[New comment by nearbuy in "It seems that the age of reading might be a short anomaly in human history"]]></title><description><![CDATA[
<p>The UNESCO/World Bank literacy rate is basically defined how you thought. But high income countries don't usually report this because literacy by this measure is nearly universal. So they often report at higher thresholds (e.g. how many people can read at a grade 9 level), and news headlines often don't make it clear that this is not the same as the UNESCO definition.</p>
]]></description><pubDate>Wed, 08 Jul 2026 14:37:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=48832583</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48832583</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48832583</guid></item><item><title><![CDATA[New comment by nearbuy in "GPT-5.6 Sol Ultra will be in Codex"]]></title><description><![CDATA[
<p>Questions like that cost a tiny fraction of a cent. "What's the capital of Sri Lanka?" cost a fifth of a cent at GPT 5.5 API price, and would cost a fraction of that if the question were routed to a more suitable, cheaper model. The output was 78 tokens.<p>By contrast, when coding, devs typically have hundreds of thousands of tokens in the context window, and may use many millions of input tokens per day.<p>Caching requires the full prefix to match exactly. If a single word differs near the beginning of the prompt, nothing after that can share the cache. So this type of caching would save a few queries that cost virtually nothing, but wouldn't help with the stuff where cost matters.</p>
]]></description><pubDate>Mon, 06 Jul 2026 03:47:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=48800430</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48800430</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48800430</guid></item><item><title><![CDATA[New comment by nearbuy in "GLM-5.2 – How to Run Locally"]]></title><description><![CDATA[
<p>The usage is irrelevant if we're interested in cost per token. If you use it half as much, you get half as many tokens at half the cost. It's still $5.56 in electricity per million output tokens either way (using $0.20/kWh, adjust accordingly if you have cheaper electricity). If you use the API, you also pay half as much if you use half as much.</p>
]]></description><pubDate>Tue, 23 Jun 2026 17:56:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=48648711</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48648711</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48648711</guid></item><item><title><![CDATA[New comment by nearbuy in "Peopleless economy? Not technically impossible"]]></title><description><![CDATA[
<p>Not much point in serfdom when they don't have use for human labor.</p>
]]></description><pubDate>Tue, 16 Jun 2026 04:28:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=48550588</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48550588</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48550588</guid></item><item><title><![CDATA[New comment by nearbuy in "How to earn a billion dollars"]]></title><description><![CDATA[
<p>> I feel like you're also doing something weird—sorta strawmanning and sorta being conveniently inconsistent. You're reframing his argument as something much softer.<p>I'm going to firmly push back on this. A reasonable, straightforward interpretation of AOC's quote is that "you can't earn a billion dollars" without cheating or abusing others. I'm can believe she may have meant that as hyperbole, but even if she did, PG believes "What she meant was that it's impossible to get that rich without doing something bad — without cheating in some way." He's not doing mental gymnastics to get there; it's a straightforward reading of "you can break rules, you can abuse labor laws, you can pay people less than what they’re worth, but you can’t earn that."<p>That makes PG very much not strawmanning. He's arguing directly against her stated position. He doesn't steelman her argument (create a stronger or more generous version of her argument and argue against that), which could have made his argument stronger.<p>That also means I'm not strawmanning. PG's position is exactly what I said: that it's not impossible to become a billionaire without doing something bad. He is quite clear on that. Nowhere does PG use language implying all billionaires don't cheat. Only AOC does that, if you take her words literally.<p>I think you're basically doing what you accused me of doing here: greatly softening AOC's position. Did Dropbox spam users' contact lists to grow? (They didn't, AFAIK, though they did have a referral program to invite friends.) This is a huge downgrade in accusation of wrongdoing compared to breaking rules and abusing labor laws.<p>If you soften your definition of cheating to this level, it's no longer a billionaire thing. It's an everyone thing. Does my gardener report all his cash income? Has a friend ever streamed a TV show less than legitimately? AOC isn't trying to nitpick the smallest infractions. She's trying to paint becoming a billionaire as substantially corrupt.</p>
]]></description><pubDate>Mon, 15 Jun 2026 20:07:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=48546370</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48546370</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48546370</guid></item><item><title><![CDATA[New comment by nearbuy in "How to earn a billion dollars"]]></title><description><![CDATA[
<p>You're not really engaging with PG's argument so much as nitpicking around the edges. Are people becoming billionaires primarily by cheating and stealing, or by making something people want? That's the core argument. Were GitLab, Dropbox and Stripe primarily cheating or offering a service people liked? Rather than cherry-picking one or two examples of what you deem the worst offenses, show that "you can't earn a billion dollars" without cheating or abusing others. PG isn't arguing that startups never do anything bad. He's just arguing that they can sometimes earn a billion dollars without cheating.</p>
]]></description><pubDate>Mon, 15 Jun 2026 04:27:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=48536592</link><dc:creator>nearbuy</dc:creator><comments>https://news.ycombinator.com/item?id=48536592</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48536592</guid></item></channel></rss>