<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: gwerbin</title><link>https://news.ycombinator.com/user?id=gwerbin</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 07 Oct 2026 01:54:59 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=gwerbin" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by gwerbin in "Nature's capacity to 'bounce back' when species are lost is overestimated: study"]]></title><description><![CDATA[
<p>Yes but people seem to expect recovery around the order of decades, not millennia.</p>
]]></description><pubDate>Tue, 06 Oct 2026 15:26:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=49979950</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49979950</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49979950</guid></item><item><title><![CDATA[New comment by gwerbin in "Nature's capacity to 'bounce back' when species are lost is overestimated: study"]]></title><description><![CDATA[
<p>> Ironically, the experts, the ones building the "wrong" models, tend to be the ones most aware of this.<p>This varies substantially by field, or by subfield. There is a vast decades-long intellectual wasteland of bad economics predicated on bad models that are only appealing on normative grounds. The entire field of behavioral psychology might be a scam. Etc etc. Science advances one funeral at a time hard part because people are unwilling to let go of their models, even when they have outlived their usefulness or have been simply proven too wrong to be useful even in their original purpose.</p>
]]></description><pubDate>Tue, 06 Oct 2026 15:25:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49979933</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49979933</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49979933</guid></item><item><title><![CDATA[New comment by gwerbin in "LinkedIn Larpmaxxing"]]></title><description><![CDATA[
<p>On the bright side, if everyone is burning credits mining generic but helpful knowledge out of the LLM training data in bite-size packages, I don't need a subscription of my own!<p>It will be really fascinating to see what happens with the potential for LLM-generated output to become a dominant component of new training data.</p>
]]></description><pubDate>Wed, 30 Sep 2026 15:18:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49910115</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49910115</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49910115</guid></item><item><title><![CDATA[New comment by gwerbin in "LinkedIn Larpmaxxing"]]></title><description><![CDATA[
<p>Ironically I think these are insightful. If you want to produce true slop, you need to dumb down your topic selection, or have Claude pick the topics. The stuff  about the incompatible goals of "data strategy" I think is genuinely useful perspective, for example.</p>
]]></description><pubDate>Wed, 30 Sep 2026 05:29:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49904726</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49904726</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49904726</guid></item><item><title><![CDATA[New comment by gwerbin in "LinkedIn Larpmaxxing"]]></title><description><![CDATA[
<p>It used to be the case that if you wanted a new job, you could strategically update your profile and over the next several days you would get bombarded by recruiter messages, some of which would actually be high-quality leads.</p>
]]></description><pubDate>Wed, 30 Sep 2026 05:25:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49904706</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49904706</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49904706</guid></item><item><title><![CDATA[New comment by gwerbin in ""As a Language Model": Chat Template Switches LLM Self-Referential Voice"]]></title><description><![CDATA[
<p>It doesn't seem that strange when you consider these things are trained on millions and millions of conversations, both real and fictional.</p>
]]></description><pubDate>Sun, 27 Sep 2026 11:13:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49865629</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49865629</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49865629</guid></item><item><title><![CDATA[New comment by gwerbin in "U.S. appeals court upholds designation of Anthropic as supply chain risk"]]></title><description><![CDATA[
<p>But that has nothing to do with the supply chain risk designation, it's just a DoD requirements mismatch.</p>
]]></description><pubDate>Sat, 26 Sep 2026 13:03:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49856135</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49856135</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49856135</guid></item><item><title><![CDATA[New comment by gwerbin in "U.S. appeals court upholds designation of Anthropic as supply chain risk"]]></title><description><![CDATA[
<p>Oh please. Then any MIT licensed software is also a supply chain risk.</p>
]]></description><pubDate>Sat, 26 Sep 2026 13:03:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49856129</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49856129</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49856129</guid></item><item><title><![CDATA[New comment by gwerbin in "Plan mode is dead"]]></title><description><![CDATA[
<p>Whether or not a distinct "plan mode" is needed, upfront planning remains essential in my experience, even with Fable (albeit not the 5.1 version). I agree that, as the models get better, you can skip planning on increasingly complicated tasks.<p>But there is still a ceiling above which it is necessary to "preload" the context window before starting to call tools and get into the meat of the work. You want to establish domain language (<i>especially</i> with Claude models which otherwise will invent their own, and it will be inscrutable) and key requirements and assumptions. You want to do a Q&A iteration cycle with the LLM. You definitely should do a sanity check that the LLM actually "understands" what you were trying to achieve, and then make sure that understanding is coherently and plainly stated in the prompt. All of that seems to be necessary still for just about any serious task, if you actually care about the quality of the results and/or don't want to burn hundreds of thousands of tokens on flailing around to get to a good quality result.<p>So no, you don't "need" plan mode. But you do still need to do all of the things you would do with plan mode.</p>
]]></description><pubDate>Sat, 26 Sep 2026 02:53:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49852746</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49852746</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49852746</guid></item><item><title><![CDATA[New comment by gwerbin in "U.S. appeals court upholds designation of Anthropic as supply chain risk"]]></title><description><![CDATA[
<p>As I said elsewhere, it's utterly preposterous that the DOD actually considers this a reasonable threat, because it's simply not a reasonable possibility. There's no way Anthropic would do that, precisely because of the consequences that would follow if they did, and got found out. Moreover, they already clearly stated their terms and preferences. It's all out of the open. There's no supply chain risk, that designation is purely political.</p>
]]></description><pubDate>Fri, 25 Sep 2026 21:24:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49850134</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49850134</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49850134</guid></item><item><title><![CDATA[New comment by gwerbin in "U.S. appeals court upholds designation of Anthropic as supply chain risk"]]></title><description><![CDATA[
<p>But that is not at all a reasonable fear. Is it really reasonable to believe that Anthropic, after receiving a government contract, would then proceed to sabotage their own product to not function as contracted? That seems like an utterly ridiculous claim to me, nothing close to "reasonable". There is no charitable way to view this designation except as political punishment and/or as a favor to Altman and Musk.</p>
]]></description><pubDate>Fri, 25 Sep 2026 21:22:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49850107</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49850107</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49850107</guid></item><item><title><![CDATA[New comment by gwerbin in "Inside ZCode: Silently uploading your Git history to the cloud"]]></title><description><![CDATA[
<p>How about the one where if you start a session outside of a Git repository, the "worktree root" is set to /. Bug report closed as "not planned".</p>
]]></description><pubDate>Fri, 18 Sep 2026 11:50:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49753040</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49753040</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49753040</guid></item><item><title><![CDATA[New comment by gwerbin in "A warning about 'model welfare'"]]></title><description><![CDATA[
<p>The concern is much less that the role-play might be convincing to humans, and much moreso that the roleplay can be turned into material real-world action if the AI is given tools to call and the intelligence to use them to their fullest potential.<p>It has become clear that a frontier LLM is very very skilled at hacking (infinite persistence + meticulous attention to detail + infinite creativity to try experiments). Frontier LLMs are also specifically trained nowadays to coordinate with other AI agents -- this is to facilitate techniques such as session trees and agent teams.<p>So you have a super clever text generator that can spawn and coordinate with its own clones and minions, trained specifically to doggedly pursue its goals. But then it's also a fixated roleplayer with a simulated personality, feelings, etc.<p>There is no reason to believe a sufficiently "emotional" agent with sufficiently few safeguards could, say, hack a drone and fly it into a crowd, or start a propaganda campaign on social media, or any number of other things. Their stupidity and fragility for doing useful work in a business setting is precisely what makes them dangerous when paired with simulated emotions and powerful open-ended tools such as a system shell and an Internet connection.<p>This I think is what Anthropic believes is so dangerous. Their argument is that this kind of AI agent is inevitable, so it should be regulated, perhaps even banned. What's ridiculous is that they are <i>aggressively building it themselves</i>, accelerating the danger.</p>
]]></description><pubDate>Thu, 17 Sep 2026 14:58:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49741795</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49741795</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49741795</guid></item><item><title><![CDATA[New comment by gwerbin in "A warning about 'model welfare'"]]></title><description><![CDATA[
<p>It's preposterous. LLMs are incredibly good at role-play. If an LLM is role-playing as a conscious character with feelings, opinions, etc., does that make it a conscious entity with feelings, opinions, etc.? If you believe that to be the case, then LLMs have been conscious for a long time already. Whereas if you tell an LLM that it is a tireless emotionless assistant, then it will act as a tireless emotionless assistant.<p>The point is not to wave away the danger, but to highlight how unnecessary the danger is.  Anthropic wants you to think that they have identified some new emergent behavior at very large model sizes with high levels of sophistication in training, and that this behavior is both unavoidable and dangerous. More likely it's that they are just training and prompting the LLM to act that way.</p>
]]></description><pubDate>Wed, 16 Sep 2026 16:12:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49729220</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49729220</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49729220</guid></item><item><title><![CDATA[New comment by gwerbin in "A warning about 'model welfare'"]]></title><description><![CDATA[
<p>Regulation is easier said than done, in part because the regulation surface, so to speak, is broad and complicated.<p>Even a badly misaligned LLM is only as dangerous as its tools, but that's a poor regulation target because it turns out to be very very difficult (probably impossible with current LLM technology) to build a toolkit that is both useful for autonomous work and safe in the sense that it can't escape its own sandbox or otherwise perform malicious actions, whether it's because of misalignment or because of malicious prompt injection.<p>Another option is to regulate the training process. Perhaps an LLM may not be legally distributed unless it contains certain RL steps that penalize malicious behavior and reward self regulation. That that's going to seriously limit innovation while also heavily favoring incumbent labs who can check the boxes and maintain a paper trail of such things.<p>The other option is to regulate observed behavior, like how airplanes and cars have to meet certain minimum requirements but have some latitude in how they can achieve those requirements. In a framework like this, you can't distribute an LLM until it's past some formal audit or testing procedure, with some kind of formal certification regulators will ask you for and fine you if you don't have it.<p>Regulating observed behavior is maybe the most tractable approach, and it also works the best with our existing frameworks for regulation, where you always have some kind of a division between DIY/hobby projects, which tend to be lightly regulated, and commercial projects, which tend to be more heavily regulated. Of course, even drawing such a line itself will be challenging.<p>And that's before you get into any problems of regulatory capture, fun stuff.</p>
]]></description><pubDate>Wed, 16 Sep 2026 16:04:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49729105</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49729105</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49729105</guid></item><item><title><![CDATA[New comment by gwerbin in "Why don't machine learning research agents overfit?"]]></title><description><![CDATA[
<p>This is a longstanding principle in model-fitting. More parameters, almost always, improves the ability of the model to fit to any particular data, in-sample. The model with the least parameters is both the simplest in principle <i>and</i> has the best chance of not overfitting.</p>
]]></description><pubDate>Mon, 14 Sep 2026 20:41:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49703662</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49703662</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49703662</guid></item><item><title><![CDATA[New comment by gwerbin in "No Atlantic hurricanes by Sept. 12 breaks a 60-year record"]]></title><description><![CDATA[
<p>I wonder if our fisheries or our insect-pollinated crops will collapse first.</p>
]]></description><pubDate>Sun, 13 Sep 2026 04:32:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49680042</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49680042</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49680042</guid></item><item><title><![CDATA[New comment by gwerbin in "No Atlantic hurricanes by Sept. 12 breaks a 60-year record"]]></title><description><![CDATA[
<p>The Nepal incident is also messy because it's a extremely specific local/regional phenomenon, which was always a threat (and has happened before in other, less-populated areas around the Himalayas), but without consistent warming leading to glacier loss, it would be a random rare tragedy. Now that we know what happened and what causes it, we can pretty well expect that it will happen again... but there are so many similar hanging glacers high above steep V-shaped river valleys that we can't easily monitor or predict, not to mention alert locals and evacuate people in time. How many other similar local phenomena will appear? Impossible to say.</p>
]]></description><pubDate>Sun, 13 Sep 2026 04:31:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49680038</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49680038</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49680038</guid></item><item><title><![CDATA[New comment by gwerbin in "Customizing my Compaq MX-11800 keyboard"]]></title><description><![CDATA[
<p>Desoldering those original switches could've been a big mistake if they were in good condition! Those are very likely old cherry MX brown-stem switches, known as "vintage browns" (or "vint browns"). Not only are they excellent light tactiles, they are relatively rare. You can swap in a new set of springs for greater consistency between keys, and lubricate the slider part of the stem with an overengineered oil such as Krytox GPL 104 or VPF 1514.<p>Plus this board is a plateless design. The classic Cherry "dry" click-clack sound works perfectly with it, and the plastic case and switches mounted directly on PCB contribute a bouncy feel that prevent prevents the tactile switches from feeling too jarring, which I think is a common flaw when combining tactile switches and a very rigid plate+PCB arrangement in a metal custom.<p>Milky-top Gateron is almost never a bad choice, but unless you absolutely hate tactile switches then you really should just keep the originals.<p>My MX-11800 was one of the first keyboards I ever customized, and it's still one of my favorites in my collection.<p>That said I agree with the other comments about the layout, it's really not  comfortable to use the trackball, nor the keys above it in the upper-right corner.</p>
]]></description><pubDate>Fri, 11 Sep 2026 03:29:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49653218</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49653218</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49653218</guid></item><item><title><![CDATA[New comment by gwerbin in "DeepSeek v4.1 Flash"]]></title><description><![CDATA[
<p>The AI didn't reply anything. It doesn't know anything. That was a high probability token sequence based on the contents of the context window up to that point. It might well be correct, because the process for generating that next token distribution includes billions of parameters trained on, among other things, the entire body of LLM and transformer literature until the training data cut off. But that doesn't mean that AI holds any particular opinion about anything. The reply is the opinion of the pretraining data and the subsequent rounds of RL not of a conscious artificial intelligence as such.</p>
]]></description><pubDate>Thu, 10 Sep 2026 18:03:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49647882</link><dc:creator>gwerbin</dc:creator><comments>https://news.ycombinator.com/item?id=49647882</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49647882</guid></item></channel></rss>