<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: trunnell</title><link>https://news.ycombinator.com/user?id=trunnell</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 28 Jul 2026 13:14:55 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=trunnell" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by trunnell in "Claude Opus 5"]]></title><description><![CDATA[
<p>The chaos appears to be tamed for now.<p>From the system card [1]:<p><pre><code>  The Fable cyber classifier we have previously discussed also applies to Claude Opus 5 , with one notable exception: for Claude Opus 5 , we’ve unblocked vulnerability finding in source code to help our coding customers develop more secure code.
  If you are a cyber defender and are experiencing blocks on Claude Opus 5 , we are also offering exemptions through our Cyber Verification Program, which will remove blocks to enable activities such as bug bounty hunting and vulnerability research and verification. Enterprise customers can also apply to join the Cyber Verification Program to have mitigations removed to enable penetration testing.
</code></pre>
[1] <a href="https://www-cdn.anthropic.com/c5fbac3f0b1280a933ebd26d3cb8bb9f5bdeaf48/Claude%20Opus%205%20System%20Card.pdf" rel="nofollow">https://www-cdn.anthropic.com/c5fbac3f0b1280a933ebd26d3cb8bb...</a></p>
]]></description><pubDate>Fri, 24 Jul 2026 19:35:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49040629</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=49040629</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49040629</guid></item><item><title><![CDATA[New comment by trunnell in "Precursor"]]></title><description><![CDATA[
<p>I dislike bots as much as anyone else... when weird inquiries come through my company's lead form, it costs some time and attention to sort them.<p>But what makes Cloudflare so confident that automation always equates to "fraud and abuse?"  If I send my agent to go retrieve some information, do they consider that fraud?<p>If I block various ad trackers does that trigger their "bot detection" incorrectly?  Do I have any recourse?  Or is Cloudflare appointing themselves judge, jury and executioner?<p>And let's not forget this little chestnut:
  > 4. Privacy by design. Precursor was designed to collect signals that help to distinguish human patterns from automated and abusive patterns.<p>Ahh, so to "protect" against bots they're standing up a whole new regime of user surveillance and session-level monitoring. And they definitely won't be selling that, they promise. Got it.<p>This crap should be illegal. In the real world, I can authorize others to act on my behalf. The same should be true with software agents.</p>
]]></description><pubDate>Mon, 13 Jul 2026 16:05:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=48894750</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48894750</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48894750</guid></item><item><title><![CDATA[New comment by trunnell in "Fable 5 Is Back"]]></title><description><![CDATA[
<p>Their post detailing the timeline and their actions since reinforce my belief that Anthropic is among the <i>most</i> trustworthy AI companies.<p><a href="https://www.anthropic.com/news/redeploying-fable-5" rel="nofollow">https://www.anthropic.com/news/redeploying-fable-5</a></p>
]]></description><pubDate>Wed, 01 Jul 2026 20:46:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=48752899</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48752899</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48752899</guid></item><item><title><![CDATA[New comment by trunnell in "Fable 5 Is Back"]]></title><description><![CDATA[
<p>Blame Amazon and the White House</p>
]]></description><pubDate>Wed, 01 Jul 2026 20:37:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=48752768</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48752768</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48752768</guid></item><item><title><![CDATA[New comment by trunnell in "Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5"]]></title><description><![CDATA[
<p>There is no reason to have less trust in Anthropic. It's not clear they did <i>anything</i> wrong. It's more likely the White House simply tied itself in knots, consistent with the last year and a half of chaos from them.</p>
]]></description><pubDate>Wed, 01 Jul 2026 04:48:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=48742360</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48742360</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48742360</guid></item><item><title><![CDATA[New comment by trunnell in "Redeploying Fable 5"]]></title><description><![CDATA[
<p>Hooray! Glad everyone came to their senses and we can all get on with business.<p>I bet it'll continue to be messy at the frontier for the foreseeable future as society gradually wakes up to the consequences of strong AI.</p>
]]></description><pubDate>Wed, 01 Jul 2026 04:38:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48742306</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48742306</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48742306</guid></item><item><title><![CDATA[New comment by trunnell in "Will It Mythos?"]]></title><description><![CDATA[
<p>Maybe you mean that an expert will use more specific language which in turn triggers the model to give a response that more closely matches the "expert distribution"<p>Anthropic published a study showing that Claude does more work for the expert user, and experts have a higher rate of "successful sessions" than novices.<p><a href="https://www.anthropic.com/research/claude-code-expertise" rel="nofollow">https://www.anthropic.com/research/claude-code-expertise</a></p>
]]></description><pubDate>Tue, 23 Jun 2026 19:56:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=48650456</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48650456</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48650456</guid></item><item><title><![CDATA[New comment by trunnell in "The Korean telecom giant at the center of Anthropic's Mythos controversy"]]></title><description><![CDATA[
<p>> honest angle is that the industry [felt] exposed by the residual risk [and] have a months-long bugfixing backlog exposed by Glasswing<p>Two problems with this theory.<p>1. Amazon complaining to the White House wouldn't have been the opening salvo. Amazon and Anthropic would find it much easier to talk to each other than go through the White House. We'd need evidence that Amazon (and probably others) already asked Anthropic to not release a Mythos-class model but Anthropic released it anyway. Are they on record saying this?<p>2. The jailbreak Amazon found needs to be real. Maybe the White House staffers are not AI experts and they don't really understand what a jailbreak is... but it's much harder to make that claim about Andy Jassy. For the jailbreak to be the real reason for the export control order, the jailbreak would need to be significant and cause material harm to Amazon. Then Jassy might pass it along to the White House assuming he already was refused by Dario.<p>But there is no evidence the jailbreak was real. There is one story that it amounted to a request, "fix this code."  In any case, Anthropic is on record saying the so-called jailbreak didn't enable any vulnerability work that couldn't already be done by other models.</p>
]]></description><pubDate>Thu, 18 Jun 2026 21:02:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=48591541</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48591541</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48591541</guid></item><item><title><![CDATA[New comment by trunnell in "The Korean telecom giant at the center of Anthropic's Mythos controversy"]]></title><description><![CDATA[
<p>Sure they exist but they're rare, difficult to identify and hire, and take years to train. Mythos/Fable is available on tap.</p>
]]></description><pubDate>Thu, 18 Jun 2026 20:49:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=48591350</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48591350</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48591350</guid></item><item><title><![CDATA[New comment by trunnell in "US holds off blacklisting DeepSeek, more than 100 firms deemed security risks"]]></title><description><![CDATA[
<p>What an amazing achievement by America's adversaries.<p>The Trump administration lists Anthropic as a security risk and kneecaps its best model, despite the fact that compared to the other frontier US labs Anthropic is more transparent, more safety-oriented, frequently honest to a fault, and is clearly acting with patriotic intent.<p>Meanwhile, the same administration is hesitating to counter certain Chinese companies' efforts of industrial-scale theft and sabotage due to a fear of angering the CCP!<p>This administration has it exactly backwards. 4.5 months until election day, 7 months until the next Congress is sworn in.</p>
]]></description><pubDate>Wed, 17 Jun 2026 18:52:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=48574968</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48574968</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48574968</guid></item><item><title><![CDATA[New comment by trunnell in "US holds off blacklisting DeepSeek, more than 100 firms deemed security risks"]]></title><description><![CDATA[
<p>Why?</p>
]]></description><pubDate>Wed, 17 Jun 2026 18:28:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=48574600</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48574600</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48574600</guid></item><item><title><![CDATA[New comment by trunnell in "Statement on US government directive to suspend access to Fable 5 and Mythos 5"]]></title><description><![CDATA[
<p>> They ultimately got what they wanted.<p>No, it's not what they wanted. As it says in your quote, they wanted "a statutory process that is transparent, fair, clear, and grounded in technical facts. <i>This action does not adhere to those principles.</i>"</p>
]]></description><pubDate>Sat, 13 Jun 2026 01:46:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=48511595</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48511595</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48511595</guid></item><item><title><![CDATA[New comment by trunnell in "Anthropic apologizes for invisible Claude Fable guardrails"]]></title><description><![CDATA[
<p>Oh, I agree distillation isn't stealing "outright" as in it's not theft of 100% of the model. But there's a reason they're doing it. I didn't say anything about Chinese labs innovating -- obviously they are.<p>What accounts for the difference between your attitude that distillation is no big deal, "common practice," yet Anthropic sees as it as a huge threat?</p>
]]></description><pubDate>Thu, 11 Jun 2026 20:17:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=48495850</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48495850</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48495850</guid></item><item><title><![CDATA[New comment by trunnell in "Anthropic apologizes for invisible Claude Fable guardrails"]]></title><description><![CDATA[
<p>"Anthropic accused Chinese firms of 'industrial-scale distillation attacks' on its AI models."<p>"Distillation involves training less capable models on more advanced ones’ output, and can be used illicitly to acquire powerful capabilities cheaply. The AI startup accused China’s DeepSeek, MiniMax, and Moonshot of generating 'over 16 million exchanges with Claude through approximately 24,000 fraudulent accounts,'"<p><a href="https://www.semafor.com/article/02/24/2026/anthropic-accuses-chinese-firms-of-distillation-attacks" rel="nofollow">https://www.semafor.com/article/02/24/2026/anthropic-accuses...</a><p>After reading their posts and watching interviews with Dario it's abundantly clear that they view Chinese-lab distillation of US frontier models as a threat to US national security. You can argue with them about whether that is true, but not whether distillation is real.</p>
]]></description><pubDate>Thu, 11 Jun 2026 19:43:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=48495431</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48495431</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48495431</guid></item><item><title><![CDATA[New comment by trunnell in "Anthropic apologizes for invisible Claude Fable guardrails"]]></title><description><![CDATA[
<p>To prevent their models from doing harm in dual-use contexts including CBRN or by accelerating research in authoritarian-backed AI labs.</p>
]]></description><pubDate>Thu, 11 Jun 2026 19:33:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48495330</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48495330</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48495330</guid></item><item><title><![CDATA[New comment by trunnell in "Anthropic apologizes for invisible Claude Fable guardrails"]]></title><description><![CDATA[
<p>I'll defend Anthropic.<p>They are clear about the reasons for guardrails: prevent their models from doing harm in dual-use contexts including CBRN or by accelerating research in authoritarian-backed AI labs.<p>What is the critique against that? It seems pretty reasonable to me. You want AI-accelerated biological or radiological experiments running in your neighbors backyard? You want PRC-backed labs to continue to steal Anthropic's models via distillation?<p>Mitigating the harms of dual-use tech is notoriously difficult and fraught with trade offs. What I would want to see is cautious rollout and quick response, which is EXACTLY what they're doing.<p>Instead, this thread is full of bad-faith arguments about Anthropic being dishonest, making a "useless" model, or "the power is going to their heads." You can't read Anthropic's System Cards and come away with any of these impressions. Quite the opposite, in fact. They are honest to a fault, acknowledging problems they discovered even when it hurts them.<p>If your harmless request was downgraded to Opus, you're billed for Opus. They were 100% clear about that. I'd much rather have a Mythos-class model that falls back to Opus 10% of the time than be capped to Opus 100% of the time. If that doesn't work for you, then make a suggestion for something better!<p>If you are a white-hat security engineer hitting guardrails, I don't think you have standing to complain. I really don't. Their Glasswing program actually got banks and the industrial sector to take action to fix security vulnerabilities. Do you realize how special that is? A huge portion of the economy runs on vulnerable code and has for decades, despite security experts testifying to Congress, begging business leaders, pleading for intervention-- with no results. But suddenly they're all enrolled in a program that will find *and fix* vulnerabilities! White-hat security people should be rejoicing. Instead some of them are throwing rocks. Unbelievable. Shameful.<p>Meanwhile, society is screaming at the AI labs to be more conscientious about potential harms of AI. Legislatures are passing laws limiting data center construction. There are protests. And you, the HN community, the vanguard of our profession, have the temerity to demand "NO GUARDRAILS!" "HOW DARE YOU TRY TO PROTECT DEMOCRACY!" "MY SOFTWARE PROJECT IS MORE IMPORTANT THAN KEEPING NUKES AWAY FROM THE BAD GUYS!"<p>Go ahead HN, downvote me. It'd be an honor.</p>
]]></description><pubDate>Thu, 11 Jun 2026 19:13:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=48495071</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48495071</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48495071</guid></item><item><title><![CDATA[New comment by trunnell in "Citing 'severe' math deficits, UC faculty demand a return to SAT tests for STEM"]]></title><description><![CDATA[
<p>We've heard:<p>- It can make kids "overconfident when they see material they think they already know, so they end up not engaging."<p>- Some programs, particularly RSM, are criticized for valuing speed over depth. Current culture for K-8 math teachers is the opposite, they value depth over speed.<p>Left unsaid:<p>- It can make the teacher's job harder when the class has a wide span of abilities.<p>- Current teaching culture is skeptical of accelerating and/or skipping grades in math.<p>Notably, we've never heard English teachers be upset about a kid reading a book outside of school that's above grade level, or using advanced vocabulary in an essay. They tend to praise it.</p>
]]></description><pubDate>Thu, 28 May 2026 19:25:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=48314153</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48314153</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48314153</guid></item><item><title><![CDATA[New comment by trunnell in "Citing 'severe' math deficits, UC faculty demand a return to SAT tests for STEM"]]></title><description><![CDATA[
<p>I'm in the SF bay area w/ middle school and high school age kids.<p>Between San Jose and San Francisco, 15%-30% of kids are in private school (it's 30% in SF where the public schools are extra dysfunctional). That's far above the California statewide average of 8% in private school.<p>Among our peers, somewhere between 1/4 and 1/3 of kids are doing advanced math outside of school, typically either Russian School of Math or Art of Problem Solving. This group only partially overlaps with the private school group. This is happening despite the fact that both public and private school teachers <i>strongly discourage</i> math outside of school!<p>So by decelerating math in the public school, incentives were created for privileged parents to take matters in their own hands and put their kids into programs that accelerate math education far beyond what public schools used to do. We now have a system that is creating even wider disparities in outcomes. It stands to reason that it's producing far less equitable outcomes, too, given that extremely bright kids who happen to be in lower-resourced schools have fewer opportunities. Universal screening for giftedness, advanced public school math courses, and the SAT -- all avenues for advancement regardless of background -- were all eliminated.</p>
]]></description><pubDate>Thu, 28 May 2026 17:02:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=48311908</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48311908</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48311908</guid></item><item><title><![CDATA[New comment by trunnell in "Citing 'severe' math deficits, UC faculty demand a return to SAT tests for STEM"]]></title><description><![CDATA[
<p>> The people working on this aren't idiots.<p>Which people are you referring to?</p>
]]></description><pubDate>Thu, 28 May 2026 15:48:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=48310679</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=48310679</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48310679</guid></item><item><title><![CDATA[New comment by trunnell in "The Codex App"]]></title><description><![CDATA[
<p>How about, "tell the agent to write instructions for cloud deployment with a cost estimate"</p>
]]></description><pubDate>Mon, 02 Feb 2026 21:05:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=46861525</link><dc:creator>trunnell</dc:creator><comments>https://news.ycombinator.com/item?id=46861525</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46861525</guid></item></channel></rss>