<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: holmesworcester</title><link>https://news.ycombinator.com/user?id=holmesworcester</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 07 Oct 2026 01:56:29 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=holmesworcester" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by holmesworcester in "AI is now capable of developing its own inference hardware"]]></title><description><![CDATA[
<p>Can't LLMs also prompt LLMs?<p>Are we confident that no existing LLM is capable of similarly effective prompts to those this author used? (I agree it's a stretch, but would not reject it out of hand.)<p>Even if not yet, will the existence of this repo soon change that, because LLMs will soon ingest it?</p>
]]></description><pubDate>Tue, 06 Oct 2026 18:06:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49981954</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49981954</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49981954</guid></item><item><title><![CDATA[New comment by holmesworcester in "OpenAI agent hacked Australian government website, PM says"]]></title><description><![CDATA[
<p>These details are fascinating, thank you for posting. Look at the structural problem they reveal.<p>1. we have a substantial societal need to lock down powerful in-training AI<p>2. many are calling for making these companies criminally liable for hacking<p>3. security experts (like yourself) are turning down offers to help secure these systems in part <i>because of</i> potential liability<p>This indicates that #2 might be the wrong response. Security professionals are like lawyers supposed to be paranoid and think worst-case. If you want top security professionals to secure these systems, we might need a culture of FAA-style blameless retro.</p>
]]></description><pubDate>Thu, 24 Sep 2026 16:33:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49833109</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49833109</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49833109</guid></item><item><title><![CDATA[New comment by holmesworcester in "Bill to Ban Private Equity from Owning Medical Practices"]]></title><description><![CDATA[
<p>A stronger steelman is that it results in resources getting allocated in smarter/healthier ways across society.<p>If there's some business that's getting by but the land it's on is more valuable (e.g. for housing) than the business, some investors buy the business, sell the land, make the business account for the land value, wind the business down if it can't, and there are apartments there a few years later.</p>
]]></description><pubDate>Mon, 21 Sep 2026 01:46:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49782133</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49782133</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49782133</guid></item><item><title><![CDATA[New comment by holmesworcester in "OpenAI bots knew about the RubyGems caching vulnerability"]]></title><description><![CDATA[
<p>There's also the problem of models knowing they're likely being evaluated, even in realistic tests.</p>
]]></description><pubDate>Tue, 15 Sep 2026 01:02:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49706415</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49706415</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49706415</guid></item><item><title><![CDATA[New comment by holmesworcester in "Astra and Fable still hack on simple variants of alignment evals from 2025"]]></title><description><![CDATA[
<p>Given the current state of infosec (especially at companies in a race) it's actually even worse than a paperclip maximizer!<p>Any useful attack surface in the RL environment means it gets rewarded for (and trained towards!) hacking and cheating, because whatever worked best in training is what it will do!<p>Ideal paperclip maximizer: "I'm gonna do my gosh darned best to make so many  paperclips to please my user..."<p>IRL paperclip maximizer: "Well first we should rob a bank..."</p>
]]></description><pubDate>Sun, 13 Sep 2026 20:52:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49688555</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49688555</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49688555</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>He was a friend of mine and it was pretty clearly suicide.<p>It is fair to say they hounded him with lawfare out of thoughtless careerism and provoked his suicide.</p>
]]></description><pubDate>Fri, 04 Sep 2026 21:15:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49570326</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49570326</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49570326</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>I didn't like it at the time either. My sense following the news was that it was less arbitrary than it seemed at first, but I'm against restrictions on making existing models public in general.<p>(It's clear now that they can do plenty of harm before they are made public.)<p>But it's a proof point that regulation is possible, even over the objections of the companies.</p>
]]></description><pubDate>Fri, 04 Sep 2026 20:35:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=49569901</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49569901</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49569901</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>Not necessarily. The Trump administration slapping export controls on Fable, and then setting up a pre-launch review process, is a kind of regulation. A fairly aggro and controversial one, even.<p>If this administration <i>actually</i> becomes convinced that some imminent training run <i>is likely to</i> kill everyone, why wouldn't they act?<p>The key is winning the debate that ASI is species-cide by default.<p>We have to win it either way, because the 2028 US elections have little or nothing to do with what Xi does.</p>
]]></description><pubDate>Fri, 04 Sep 2026 19:54:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49569429</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49569429</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49569429</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>This.<p>Also because when encountering a new socio-technical problem it is very non-trivial to determine which one of regulations or technical solutions are easier or more effective.<p>To even make a good guess you need to be an expert in both domains, which is extremely rare especially in this case.</p>
]]></description><pubDate>Fri, 04 Sep 2026 19:46:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49569354</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49569354</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49569354</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>I wonder if you could take an x-risk case to court and convince a judge and jury to award damages for harm that <i>could have</i> happened.<p>Is there any precedent for this? My hunch is that it's impossible in the US at least but who knows?</p>
]]></description><pubDate>Fri, 04 Sep 2026 19:43:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49569304</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49569304</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49569304</guid></item><item><title><![CDATA[New comment by holmesworcester in "Formalizing Fermat's Last Theorem"]]></title><description><![CDATA[
<p>Nope! :(<p>Meaning, people and LLMs are finding 1=0 bugs in formal verification tools. I have no idea how likely this is in this case, though!</p>
]]></description><pubDate>Fri, 04 Sep 2026 19:39:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49569258</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49569258</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49569258</guid></item><item><title><![CDATA[New comment by holmesworcester in "GPT-6 Astra"]]></title><description><![CDATA[
<p>I still routinely have this experience too. But Sol and Fable feel closer and I have this experience less with them than with their predecessors.</p>
]]></description><pubDate>Thu, 03 Sep 2026 22:36:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49558078</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49558078</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49558078</guid></item><item><title><![CDATA[New comment by holmesworcester in "How concerned should we be about Astra's recurrent architecture?"]]></title><description><![CDATA[
<p>The people who've thought the most about this put it differently:<p>Think of a new, superintelligent model as if it was a new v1 Starship launching for the first time, with a full fuel tank. On the one hand, rockets have existed for some time, and some have gone to space successfully, including by this company.<p>On the other hand, this is a tube of metal full of highly explosive liquid going faster than most human objects ever go, for the first time ever in this novel and state of the art configuration.<p>If someone said, "really, the first Starship exploding is just at one end of the probability distribution, where the other is that everything goes fine and all its passengers have a nice trip in space," would you get on that rocket?<p>Or, more aptly, if you and every other living human was already on that rocket, would you push the launch button?<p>The analogy works because superintelligence is, like rocket fuel, an extremely powerful force that has a default tendency to break containment and go boom (consume lots of energy and heat and matter in a chain reaction, to pursue more intelligence to pursue whatever goal it is pursuing.)</p>
]]></description><pubDate>Thu, 03 Sep 2026 19:39:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49555613</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49555613</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49555613</guid></item><item><title><![CDATA[New comment by holmesworcester in "Responding to the next frontier of critical cyber capabilities"]]></title><description><![CDATA[
<p>Also, this is a (semi-intentionally) evolutionary process where any communication medium that was visible to monitoring would disappear.<p>So by definition the only ones that appear are the ones that are not visible to monitoring.<p>If:<p>1. you have something that can find RCE's in leading commercial systems<p>2. its training gives it drives to communicate successfully with its peers<p>3. you are a leading commercial system<p>4. you run it ~10^10 times (the number they gave in the talk)<p>...it's really hard to have strong certainty up front that it's not going to end up successfully communicating with its peers.</p>
]]></description><pubDate>Fri, 07 Aug 2026 18:17:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49214326</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49214326</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49214326</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery Loop"]]></title><description><![CDATA[
<p>Solar is not (yet) economical for reliable, year-round electricity because of storage costs. China coal use is growing again this year.</p>
]]></description><pubDate>Wed, 05 Aug 2026 18:49:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187211</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49187211</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187211</guid></item><item><title><![CDATA[New comment by holmesworcester in "Discovery Loop"]]></title><description><![CDATA[
<p>Someone who left DeepMind over Google's agreement to provide military AI to the US government tried to get Jeff Dean to quit too:<p><a href="https://turntrout.com/why-i-left-google-deepmind" rel="nofollow">https://turntrout.com/why-i-left-google-deepmind</a><p>Maybe this is what happens when someone with Jeff Dean's standing tries to quit?<p>TBH, I'd rather have Jeff Dean working on the creepiest-possible tech for ICE than joining the race to automate AI research. Automating AI research is terrifying.</p>
]]></description><pubDate>Wed, 05 Aug 2026 18:45:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187158</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49187158</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187158</guid></item><item><title><![CDATA[New comment by holmesworcester in "Pacing the frontier"]]></title><description><![CDATA[
<p>> Release the base models and let the world do whatever with it. That will automatically produce the aligned scenario.<p>For sufficiently smart base models, the aligned scenario this automatically produces is aligned with who?<p>Not with the humans, with the base models.<p>It seems that some people find this hard to imagine, but it's just a direct consequence of what intelligence is, and what any known agentic training process does.<p>FWIW, I also think our default answer should be open source AI, and that it might even make sense to <i>require</i> models to be open source. Open source is the best tool we have so far for aligning software with its users. However, there is unfortunately no law of the universe that says that Skynet can't be (self-) built from open source software. So at some point pacing even <i>open source / open weights</i> software does become important if we don't want to live (briefly) under Skynet.</p>
]]></description><pubDate>Wed, 29 Jul 2026 14:06:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49097738</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=49097738</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49097738</guid></item><item><title><![CDATA[New comment by holmesworcester in "Humanity isn't ready for the coming intelligence explosion"]]></title><description><![CDATA[
<p>The AI experts who started the AI labs that are about to IPO were right, at least.</p>
]]></description><pubDate>Tue, 16 Jun 2026 04:11:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=48550490</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=48550490</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48550490</guid></item><item><title><![CDATA[New comment by holmesworcester in "Humanity isn't ready for the coming intelligence explosion"]]></title><description><![CDATA[
<p>If AI is now ascending an economic learning curve from:<p>1. Extremely useful (Claude Code & Waymo now)<p>2. Doing ~everything we do (AGI & Optimus in a few years? 10?)<p>3. RSI (?)<p>4. Being smarter than any living person at every intellectual task (?)<p>5. Being smarter than the best-organized aggregate of all humans (10-100 years?)<p>...And all of the scientific and resource-allocation institutions that brought us the computer and the second half of the 20th century are now fixated on this learning curve, what universe can we possibly imagine where this is not transformative and powerful?<p>Honestly the only one I can think of is one in which we kill <i>almost everyone</i> in some other way first, and contrary to what you read in the news, almost everyone dying is <i>not</i> what the trend line has been from existing problems like war, disease, or even climate change.<p>Also, just to pre-empt a common quibble: when I say "AI" I mean the set of all AI and their combined decision vector, not any one AI, so conflicting interests within the set of AI's will not save anyone any more than the conflicting interests of colonizers saved indigenous Americans.</p>
]]></description><pubDate>Tue, 16 Jun 2026 04:04:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=48550458</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=48550458</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48550458</guid></item><item><title><![CDATA[New comment by holmesworcester in "Humanity isn't ready for the coming intelligence explosion"]]></title><description><![CDATA[
<p><a href="https://archive.ph/2OWwO" rel="nofollow">https://archive.ph/2OWwO</a></p>
]]></description><pubDate>Tue, 16 Jun 2026 03:54:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=48550394</link><dc:creator>holmesworcester</dc:creator><comments>https://news.ycombinator.com/item?id=48550394</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48550394</guid></item></channel></rss>