<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: gr_norm</title><link>https://news.ycombinator.com/user?id=gr_norm</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 29 Jul 2026 15:24:55 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=gr_norm" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by gr_norm in "Pacing the frontier"]]></title><description><![CDATA[
<p>Sorry, I should've used clearer language. I just mean intelligence in the sense humans possess it. Computers have always been able to exceed limited elements of human intelligence (say, at arithmetic), but the problem is to simultaneously match or exceed all of them. I specifically think that the important parts they currently lack make them a no-go for supposed existential risk.</p>
]]></description><pubDate>Wed, 29 Jul 2026 04:06:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49093303</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49093303</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49093303</guid></item><item><title><![CDATA[New comment by gr_norm in "Pacing the frontier"]]></title><description><![CDATA[
<p>Moving the goalposts, so to speak, is how science works! We must of course update the things we think based on an improved understanding of how the world works. Only the deeply incurious could consistently demand that one adhere dogmatically to a prior way of understanding the world in the face of changing knowledge. It's important not to treat science as a game to be won (of which 'goalposts' are evocative), but as the collaborative effort that it is.<p>Note well that I have not said LLMs are not useful, only that no amount of interaction with them has convinced me they remotely approach general intelligence. Their actual function of heuristically regurgitating all information ever known to mankind remains extremely useful in lots of domains.<p>And by the way, it's not obvious an LLM could fool me (or the average person) over a sufficiently long period of time, exactly because they are heuristic machines. That they lack thought guided by underlying cognitive models seems to always leak out, in the end. I suspect pass rates by LLMs on long-range Turing tests (if there have been any conducted) might drop as people become acclimatized to them.</p>
]]></description><pubDate>Wed, 29 Jul 2026 02:52:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49092902</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49092902</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49092902</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>Not only is the DNA sequence for smallpox available, but since 2018 there's been a well-documented end-to-end synthesis procedure for the very closely related horsepox virus [1]! No LLMs needed. Caused quite a stir in the synthetic biology community back then.<p>The world has not come to an end, of course, because even with peer-reviewed and experience-driven (rather than hallucinated and therefore dangerous) instructions detailing obstacles encountered during synthesis and how to overcome them, actually going out and acquiring the materials and ability to use them sufficiently skillfully is another matter entirely. Biosecurity is an important topic, to be sure, but what the AI labs have to say about it (or anything) at this point does not necessarily survive contact with reality.<p>[1] <a href="https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0188453" rel="nofollow">https://journals.plos.org/plosone/article?id=10.1371/journal...</a></p>
]]></description><pubDate>Tue, 28 Jul 2026 08:21:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49080939</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49080939</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49080939</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>The effective altruism/rationalism/AI xrisk people have always had a shockingly poor grasp on subjects outside computer science, despite their attempts to speak on them. I don't blame the actual biologists and chemists working at the frontier labs for wanting to skim a few bucks off all the money flying around, though! I know a couple who've had not-so-kind words to say about their employers' intelligence.<p>I suspect there's at least some "telling the bosses what they want to hear" going on. A massive financial incentive exists to exaggerate and fearmonger even internally to the company, because it makes you and your job seem more important.</p>
]]></description><pubDate>Tue, 28 Jul 2026 02:17:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49078449</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49078449</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49078449</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>The emptiness of AI companies' waxing poetic about the future of humankind is laid bare by simply looking at what they actually do, and who they do business with. Actions speak louder than words, and they've driven the worth of their words into the dirt many times over.</p>
]]></description><pubDate>Tue, 28 Jul 2026 00:28:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49077657</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49077657</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49077657</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>It's baffling that they thought the mental gymnastics in this blog post would make them look better. I'd rather they simply fall silent on the issue; I would respect them more (or at all) for it. Open models obviously threaten fierce competition, if not outright destruction of their bottom line. But no, they needed to try and argue that they have the moral high ground for attempting to singularly consolidate power over all human labor.</p>
]]></description><pubDate>Mon, 27 Jul 2026 23:14:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076843</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49076843</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076843</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>> Distillation does not allow the CCP to obtain equivalent or superior AI capabilities to the US, but it can bring the Chinese frontier to within a few months of the US frontier<p>A message to their investors, it would seem. "They caught up just because they distilled! Obviously they couldn't actually be as good as us!" Really funny thing to say right after an OpenAI higher-up stated point-blank that the performance of K3 can't be chalked up to mere distillation of American models.</p>
]]></description><pubDate>Mon, 27 Jul 2026 23:11:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076810</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49076810</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076810</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>I don't really understand how they can argue the security angle with a straight face. It's not like GLM 5.2 is a slouch. I've seen it do things like exploit an IDOR issue when I was experimenting with a quick-and-dirty web automation task. I simply fixed it, as one does. Open models make the world better to a far greater degree than they set it aflame.<p>Their position is analogous to trying to, say, ensure digital privacy for everyone not by making encryption freely available (because that would let the bad guys use it!), but by making it so you can't use general purpose communications devices that can listen to transmissions not intended for you. Do they hear how moronic that sounds?<p>Each passing frontier-level open model release makes Anthropic's patronizing rhetoric a little more insufferable, because it becomes clearer how unmoored from reality they've become in pursuit of profit.</p>
]]></description><pubDate>Mon, 27 Jul 2026 22:59:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076659</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49076659</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076659</guid></item><item><title><![CDATA[New comment by gr_norm in "Our position on open-weights models"]]></title><description><![CDATA[
<p>> Open-weights models that don’t have dangerous capabilities are a public good<p>Note the hedging against 'dangerous capabilities'. Undoubtedly, all the useful ones trigger this condition in Anthropic's eyes. The rest of the post is filled with similar weasel-wording. Make no mistake, this absolutely confirms that Anthropic is against open models in the sense that any reasonable person understands them.<p>The way the rest of the post unabashedly appeals to the current US administration's China hysteria is hilarious, and not at all subtle.<p>I guess we'll see about all the doomsaying here, won't we? Kimi K3 is frontier-level, and there's no stopping it now. As far as the world is concerned, anyway. If the US wants to kneecap itself that's another matter.</p>
]]></description><pubDate>Mon, 27 Jul 2026 22:38:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49076402</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49076402</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49076402</guid></item><item><title><![CDATA[New comment by gr_norm in "We have proof automation now"]]></title><description><![CDATA[
<p>> I’m not even going to get into how you could provably transform brute force propositional logic into efficient algorithms.<p>> The entire problem is showing that an efficient/reliable program actually implements those rules.<p>Reliability is a standard matter of correctness and captured (partly) by specifications. Efficiency tends to be easy to empirically test, but it is also possible to capture at the specification level [1]. Mind that specifications need not be all-consuming.<p>> But I did spend a grad class with rocq (coq at the time) and a decade working with “systems engineers” and am not convinced that this is a realistic expectation.<p>Agree! But this stuff just got massively more accessible, and the tooling around it is growing quickly. I think we'll end up growing specification systems specific to various domains which will be palatable to those "systems engineers", but I err on the side of optimism here. There's definitely a lot left to do for practicality.<p>[1] See the work of <a href="https://cs.nyu.edu/~shw8119" rel="nofollow">https://cs.nyu.edu/~shw8119</a> for the case of provably-efficient parallelism and garbage collectors</p>
]]></description><pubDate>Mon, 27 Jul 2026 05:43:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49065562</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49065562</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49065562</guid></item><item><title><![CDATA[New comment by gr_norm in "We have proof automation now"]]></title><description><![CDATA[
<p>It is true that finding the correct specification is a formidable task; knowing what correctness even means is arguably most of the difficulty of programming. However! "Moving bugs up from programs to types" isn't how this shakes out in practice, at all. Another commenter already noted that it's often much easier to communicate your intent through specifications, because you can essentially always say what a computation should do much more simply than you can say exactly how to do it.<p>I think it's also important not to miss the forest for the trees: even relatively simple specifications like "the compress and decompress functions must be inverses for all inputs" rules out vast classes of bugs in a compression library. This is not a complete specification; for instance, it does not speak about how the decompressor behaves on malicious input. But in my experience, even partial specifications carry the promise of hitting warp speed with LLMs in a way that I haven't seen anywhere else. After a certain level of specification, you have decent guarantees of being able to whole-heartedly forget about the implementation details of the synthesized program. And you get a better-built, more robust program out of it at the end!<p>The comment at the end of the article about having LLMs directly generate assembly against specifications and letting them rip with finding custom optimizations is the sort of crazy stuff this enables. I really think we're only seeing the tip of the iceberg here. People keep asking what we can do with LLMs that we couldn't before; this is the answer.</p>
]]></description><pubDate>Mon, 27 Jul 2026 05:08:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49065390</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49065390</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49065390</guid></item><item><title><![CDATA[New comment by gr_norm in "Nvidia, Microsoft, Meta warn against overregulating open-weight models"]]></title><description><![CDATA[
<p>To maximize personal influence and wealth, of course. These days, many with power don't seem to care much about uplifting society so long as they get theirs.<p>Hopefully Anthropic eats crow here, lest their wish is fulfilled that we all become slaves to the anointed few who work there. Never getting another dollar from me after pulling these stunts.</p>
]]></description><pubDate>Sat, 25 Jul 2026 04:53:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49044636</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49044636</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49044636</guid></item><item><title><![CDATA[New comment by gr_norm in "Be skeptical of OpenAI's rogue hacker agent story"]]></title><description><![CDATA[
<p>Many who rush to the defense of AI companies' marketing departments seem to take criticism personally, as if not buying into it all hook, line, and sinker is an affront to them. A foreseeable consequence of becoming cognitively dependent on LLMs.<p>And for what it's worth, I'm not an AI skeptic. I fully believe that the frontier models are capable of exploiting (chains of) vulnerabilities, having seen GLM and now Kimi do it myself. What I find no reason to accept is the sci-fi existential risk subtext peddled by the salesmanship around it. We will reach a new equilibrium with more secure software, and LLMs (by finding vulnerabilities, generating proofs, etc) will help us along.</p>
]]></description><pubDate>Sat, 25 Jul 2026 03:43:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49044303</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49044303</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49044303</guid></item><item><title><![CDATA[New comment by gr_norm in "Startup founders urge U.S. government not to shut off Chinese open weight AI"]]></title><description><![CDATA[
<p>Login-walled for me. Is there a Reddit proxy like Nitter?</p>
]]></description><pubDate>Thu, 23 Jul 2026 16:09:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49023942</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49023942</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49023942</guid></item><item><title><![CDATA[New comment by gr_norm in "OpenAI’s accidental attack against Hugging Face is science fiction that happened"]]></title><description><![CDATA[
<p>I could fully see them thinking the incident disclosed yesterday would have made them look good ("wow, OpenAI's models are so capable!"). That it didn't occur to them to discuss specific preventative measures to be taken in the future (airgapping as a foolproof one already familiar to the CTF world, anyone?) indicates to me they're not taking their job seriously; they are the ones treating this as a marketing charade.<p>It's very difficult for me to reconcile belief in the existential risk business with what they actually did. So I agree with you that this makes OpenAI look badly incompetent; but their communication on this makes me think they don't realize it.<p>For what it's worth I don't agree with the xrisk-ness of these models; they're dangerous, but almost certainly only temporarily while a new equilibrium is reached via more secure software. Open models are probably an essential part of the recipe (as you noted) for doing so. I also have a personal suspicion that LM-accelerated formal verification will have no small role to play here, sidestepping the cat-and-mouse game of bug finding-and-fixing.</p>
]]></description><pubDate>Thu, 23 Jul 2026 03:47:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49016695</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49016695</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49016695</guid></item><item><title><![CDATA[New comment by gr_norm in "Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA"]]></title><description><![CDATA[
<p>Yeah, open weights hasn't fully caught up yet, but it's getting very close. And it's certainly passed the point of being reasonably interchangeable with the frontier for (programming) work. Add in the benefits of not being rug-pulled by the frontier labs silently messing with, the knobs on their models or outright denying you the ability to do certain kinds of work (c.f. the HuggingFace fiasco), and they probably come out ahead in several respects.</p>
]]></description><pubDate>Wed, 22 Jul 2026 05:01:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49002078</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49002078</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49002078</guid></item><item><title><![CDATA[New comment by gr_norm in "Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA"]]></title><description><![CDATA[
<p>The idea that the Chinese labs cannot make progress except by copying superior American products is just prejudice against the former and exceptionalism of the latter at play. Even the OpenAI top brass have admitted otherwise [1]. China is an equal match in every respect, and we'd better admit this to ourselves sooner rather than later so as to see the game clearly.<p>[1] <a href="https://xcancel.com/deanwball/status/2078133895766114412" rel="nofollow">https://xcancel.com/deanwball/status/2078133895766114412</a></p>
]]></description><pubDate>Wed, 22 Jul 2026 02:15:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49001022</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=49001022</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49001022</guid></item><item><title><![CDATA[New comment by gr_norm in "OpenAI and Hugging Face address security incident during model evaluation"]]></title><description><![CDATA[
<p>HF need not be party to it at all, beyond being the victim. I suspect the hack is real; I have observed GLM 5.2 being able to discover similar vulnerabilities in web applications I'm hosting (which I've then fixed!). At the same time, it seems very neatly timed at an inflection point in the conversation around open models, and there's questions around the incompetent isolation under which the hacking benchmark appears to have been run.<p>Remember that there is generational wealth on the line for most OpenAI employees, and consider what  people might do to obtain it.</p>
]]></description><pubDate>Tue, 21 Jul 2026 22:49:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=48999404</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=48999404</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48999404</guid></item><item><title><![CDATA[New comment by gr_norm in "OpenAI and Hugging Face address security incident during model evaluation"]]></title><description><![CDATA[
<p>The timing after the release of GLM 5.2 and Kimi K3 is quite convenient, too, as an angle for regulatory quashing of open-weights models just as they're entering the mainstream conversation around usurping the American frontier labs. I accept my thinking here is conspiratorial, but there's also a hell of a lot of money on the line to encourage the unscrupulous.</p>
]]></description><pubDate>Tue, 21 Jul 2026 21:55:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48998895</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=48998895</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48998895</guid></item><item><title><![CDATA[New comment by gr_norm in "OpenAI and Hugging Face address security incident during model evaluation"]]></title><description><![CDATA[
<p>They've been saying so from the beginning, and yet did not take the basic precaution of airgapping their off-the-leash model while it's been instructed to succeed at a hacking benchmark by any means necessary. So which is it? I _want_ to believe them, I do, but there's always these gaps between what they say and their actions on display that give me reason to think otherwise.</p>
]]></description><pubDate>Tue, 21 Jul 2026 21:31:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=48998656</link><dc:creator>gr_norm</dc:creator><comments>https://news.ycombinator.com/item?id=48998656</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48998656</guid></item></channel></rss>