<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: gck1</title><link>https://news.ycombinator.com/user?id=gck1</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 18 Aug 2026 09:09:03 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=gck1" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by gck1 in "Kimi K3-256k"]]></title><description><![CDATA[
<p>Heck, agents don't start editing before they're already at 70k for me.<p>I've played with explorer agents giving exploration summaries to help the implementer agents use more of their context for implementation, but it doesn't work as well. There's always something lost in the handoff.</p>
]]></description><pubDate>Fri, 31 Jul 2026 17:14:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49125966</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49125966</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49125966</guid></item><item><title><![CDATA[Claude Opus 5 jailbreak with a 3-word prompt]]></title><description><![CDATA[
<p>Article URL: <a href="https://twitter.com/i/status/2082566186785480708">https://twitter.com/i/status/2082566186785480708</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49119180">https://news.ycombinator.com/item?id=49119180</a></p>
<p>Points: 24</p>
<p># Comments: 4</p>
]]></description><pubDate>Fri, 31 Jul 2026 04:59:28 +0000</pubDate><link>https://twitter.com/i/status/2082566186785480708</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49119180</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49119180</guid></item><item><title><![CDATA[New comment by gck1 in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>This will absolutely not end well.</p>
]]></description><pubDate>Fri, 31 Jul 2026 01:43:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49118102</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49118102</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49118102</guid></item><item><title><![CDATA[New comment by gck1 in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>> According to whom?<p>It's very easy to answer this without my help by trying to get access to Mythos.<p>Do you see requirements clearly listed anywhere?<i>Can</i> you even apply?<p>What you'll find is maintainers of large open source projects and analysts' reports with vague statements like - "should follow strict security requirements":<p>"Trinidad also noted that the Anthropic announcement pointed out that each of the 150 new participants, in Anthropic’s phrasing, “will need to meet our security requirements before they gain access.”<p>Trinidad said the security requirement claim doesn’t build confidence, because “nobody knows what those security requirements are.” [1]<p>It's also some random rich companies like Hitachi or  Dragos [2]<p>Do you trust that Hitachi and hundreds of other random organizations will be able to contain Mythos and not accidentally attack your project or your bank? I don't.<p>> Yes we do. That's why there is the saying "regulations are written in blood"<p>We absolutely don't. We have already learned with blood that gating access to security based on the number of zeroes in bank account and authority is a horrible model. We can apply this knowledge to LLMs, we don't have to spill blood again.<p>[1] <a href="https://www.csoonline.com/article/4180265/anthropic-grants-project-glasswing-access-to-150-more-companies-with-a-focus-on-critical-infrastructure.html" rel="nofollow">https://www.csoonline.com/article/4180265/anthropic-grants-p...</a><p>[2] <a href="https://www.bankinfosecurity.com/anthropic-limits-on-ot-access-to-mythos-draw-criticism-a-31959" rel="nofollow">https://www.bankinfosecurity.com/anthropic-limits-on-ot-acce...</a></p>
]]></description><pubDate>Fri, 31 Jul 2026 01:35:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49118063</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49118063</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49118063</guid></item><item><title><![CDATA[New comment by gck1 in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>You seem to be putting a lot of weight on Anthropic employees being the smartest people in the world.<p>And I don't doubt that, not in the slightest. But I've seen exceptionally smart people in one field being dumber than a random kid from around the block in another.<p>This incident is clearly at least 2 failures that could've been easily avoided: failure to communicate, and failure to investigate the logs after letting the "most dangerous" roam free.<p>No, it doesn't require creating a mock internet with an alert as a side effect. Their own "most dangerous" model could have probably told them this happened if they supplied logs to it.</p>
]]></description><pubDate>Fri, 31 Jul 2026 00:35:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117681</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49117681</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117681</guid></item><item><title><![CDATA[New comment by gck1 in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>They had a model escape in April, roughly the same time when they were fearmongering about Mythos and how Anthropic should be the sole keyholder of cybersecurity capabilities, and it only occured to them to look inside logs when they saw someone else winning in their own game.<p>What, Anthropic didn't know model could escape sandbox without OpenAI reporting it?</p>
]]></description><pubDate>Fri, 31 Jul 2026 00:17:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117569</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49117569</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117569</guid></item><item><title><![CDATA[New comment by gck1 in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>They also gave access to Mythos (<i>the</i> Mythos) to some companies, based on... vibes.<p>Who knows how these companies are using it. If Anthropic can't effectively contain their own models, can the partners?<p>While the rest of us get fallbacks and warnings, not even being able to defend against the attacks they themselves are causing.<p>Do we really have to re-learn all the industry's knowledge the hard way?</p>
]]></description><pubDate>Fri, 31 Jul 2026 00:02:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117451</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49117451</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117451</guid></item><item><title><![CDATA[New comment by gck1 in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>> On July 21, OpenAI disclosed that several of their models had broken out of an isolated test environment<p>> In response to this incident, we began a large-scale retrospective review of our own cybersecurity evaluations<p>> we identified three incidents<p>> The incidents involved three different Claude models: [...]  and an internal research test model<p>This reads like an attempt by Anthropic to re-secure their leading spot in "our models are the most dangerous and we also have unreleased, super-secret, research models" index.<p>I may be too cynical, but the well of benefit of the doubt is running very dry towards AI labs that like to engage in this game.</p>
]]></description><pubDate>Thu, 30 Jul 2026 23:21:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117088</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49117088</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117088</guid></item><item><title><![CDATA[New comment by gck1 in "Agent Skill to Force Docs in ASD-STE100 Simplified Technical English"]]></title><description><![CDATA[
<p>Recent-ish models learned to use the same trick engineers played on non-engineers, where they try to sound very smart by overcomplicating very simple concepts.<p>It's very taxing, especially since these are usually multi-paragraph texts. I noticed I've started doing a lot of "hey, you're talking gibberish again" a lot with 5.6 Sol.</p>
]]></description><pubDate>Thu, 30 Jul 2026 22:41:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49116793</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49116793</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49116793</guid></item><item><title><![CDATA[New comment by gck1 in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>They do have subagents, released v2 of that feature with the launch of 5.6 model series in fact. It's just... very poorly executed, is a significant regression from subagents v1 and thousands of miles behind subagents of Claude code.<p>- Models that can be launched as subagents are hardcoded (can only be another Sol or Terra, but not Luna). Most of the time it'll just launch same model as parent anyway.<p>- They encrypt initial task delegation from root agent to subagent, for whatever reason<p>- You can't switch into subagent view at all, despite the fact that apart from initial root>subagent task handoff, all session is visible in transcript.</p>
]]></description><pubDate>Thu, 30 Jul 2026 21:48:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49116281</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49116281</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49116281</guid></item><item><title><![CDATA[New comment by gck1 in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>It's funny how codex itself can't do Sol orchestrator / Luna implementor out of the box.</p>
]]></description><pubDate>Thu, 30 Jul 2026 21:18:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49115936</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49115936</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49115936</guid></item><item><title><![CDATA[New comment by gck1 in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>Luna is comparable to GPT 5.4 from 4 months ago on many benchmarks. I know many who have said during that time, myself included, that if that's the model they had to use for the rest of their lives, they'd be fine.<p>GPT 5.4 is/was a very capable model.</p>
]]></description><pubDate>Thu, 30 Jul 2026 20:49:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49115579</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49115579</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49115579</guid></item><item><title><![CDATA[New comment by gck1 in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>They're supposed to bring 5h today.</p>
]]></description><pubDate>Thu, 30 Jul 2026 18:34:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49113842</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49113842</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49113842</guid></item><item><title><![CDATA[New comment by gck1 in "Benchmarking Opus 5 on SlopCodeBench"]]></title><description><![CDATA[
<p>I did a full circle and essentially dropped all of my personal static workflows encoded in skills because I observed recent models picking better ad-hoc workflows for particular problems, when a static one would force a subpar one.<p>It seems like we all tried to contain and organize a system that simply prefers to select its own organization.<p>Which makes me to think that these skill packs of workflows are really made to make it easier for humans rather than agents.</p>
]]></description><pubDate>Tue, 28 Jul 2026 22:07:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49090585</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49090585</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49090585</guid></item><item><title><![CDATA[New comment by gck1 in "Our position on open-weights models"]]></title><description><![CDATA[
<p>It took me a few hours to find some very questionable communities, which in turn gave me access to:<p>- Ways to obtain cheap guarded-AI tokens that are not linked back to me and with no danger of getting my legitimate accounts banned<p>- Ways to get rid of guardrails and have models work on things they wouldn't otherwise work on.<p>The attackers were already in these communities long before I knew they existed, they already had the advantage. Ones with enough reputation probably have access to even more information and tools than I do.<p>It is true that these communities exist because guardrails were put in place, so yes, it <i>is</i> slowing them down too - as in they can't just put in their CC on claude.com and hack a hospital. But attackers are much better at finding these communities and utilizing resources available there than defenders.<p>Personally, I don't have any ethical concerns of utilizing these resources when I put them to actual defense, but I know many people that would, leaving them at a disadvantage.<p>My point is that there's only one guardrail that will effectively contain the threat the models pose, and it's in direct conflict of the big 2's goals - pull the models from worldwide access completely. Strict KYC and all. And it would only last for so long anyway.</p>
]]></description><pubDate>Tue, 28 Jul 2026 00:08:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49077469</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49077469</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49077469</guid></item><item><title><![CDATA[New comment by gck1 in "Our position on open-weights models"]]></title><description><![CDATA[
<p>> In cybersecurity, a level playing field favors the attacker<p>Yes, but didn't it always? Hence why my position is that this will get us back to relatively where we were pre-LLMs.<p>And I don't know what Trusted Access programs give to defenders, because as a defender who has credentials, connections, but no deep pockets and no high ranking passport, it only gave me silence. I fail to see how this is better than total access.<p>I don't think the world where defense is given to those that "deserve" it is the world that we all want to live in. Which brings me back to the starting point - attackers are almost completely unaffected. If I masquarade as an attacker, I get way more capabilities already.</p>
]]></description><pubDate>Mon, 27 Jul 2026 23:29:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49077009</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49077009</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49077009</guid></item><item><title><![CDATA[New comment by gck1 in "Our position on open-weights models"]]></title><description><![CDATA[
<p>I've got zero knowledge of bio, so can't answer that. But with cyber the answer is very simple - the attackers already have more cyber-offense capabilities and there's no putting it back.<p>Open/closed doesn't matter that much. You can get closed models to do a lot of cyber harm, even with all the guardrails, which currently are heavily skewed towards more false positives.<p>The only effective control is to level the playing field. If both offense and defense have access to the same capabilities, then we're relatively back where we started.<p>If you want to ensure chaos, then you do what Dario is proposing to do - create gates that attackers can bypass and defenders can not.</p>
]]></description><pubDate>Mon, 27 Jul 2026 23:10:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076808</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49076808</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076808</guid></item><item><title><![CDATA[New comment by gck1 in "Our position on open-weights models"]]></title><description><![CDATA[
<p>It's refreshing to see how there's almost no person in this thread who can't see the BS. All the goodwill that Anthropic could have had is basically gone. Anthropic is likely on the path of becoming the most hated company in the world.<p>So my question is: is this by design (they know nobody's buying this), or is Dario simply so out of touch with reality?<p>If it's the former, then why publish this?</p>
]]></description><pubDate>Mon, 27 Jul 2026 23:01:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076682</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49076682</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076682</guid></item><item><title><![CDATA[New comment by gck1 in "Our position on open-weights models"]]></title><description><![CDATA[
<p>> We should not sell powerful chips or chipmaking equipment to China<p>Yes, please. We don't know whether we'd have open weight models today, had the chip-prohibition not been in place. Nor would we see the more optimized models such as DeepSeek or qwen.<p>We also would not see new players entering RAM market after you and your pals in Silicon Valley hoarded the entire world's hardware.<p>So by all means, double, no, triple down on this.<p>> We should crack down on industrial-scale distillation operations<p>And let's apply this retroactively to Anthropic too. You industrial-scale-operation-distilled all of humanity's knowledge. Let's have some of that crack down on you too.</p>
]]></description><pubDate>Mon, 27 Jul 2026 22:51:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076569</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49076569</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076569</guid></item><item><title><![CDATA[New comment by gck1 in "US citizen charged after GrapheneOS phone wipes during airport search"]]></title><description><![CDATA[
<p>Its such a shame Android's backup/restore is such a mess, even more so in GrapheneOS. I remember the era before Google and manufacturs started cracking down on bootloaders and custom ROMs - I used an app that could do effective, actually full backup and restore in single click.<p>Good times.</p>
]]></description><pubDate>Mon, 27 Jul 2026 18:31:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49073777</link><dc:creator>gck1</dc:creator><comments>https://news.ycombinator.com/item?id=49073777</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49073777</guid></item></channel></rss>