<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: wgd</title><link>https://news.ycombinator.com/user?id=wgd</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 04 Aug 2026 03:36:19 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=wgd" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by wgd in "Don't be a meat proxy"]]></title><description><![CDATA[
<p>There is <a href="https://noslopgrenade.com/" rel="nofollow">https://noslopgrenade.com/</a></p>
]]></description><pubDate>Mon, 03 Aug 2026 07:06:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49152219</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=49152219</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49152219</guid></item><item><title><![CDATA[New comment by wgd in "Claude Code: Anatomy of a Misfeature"]]></title><description><![CDATA[
<p>It's actually pretty straightforward to recover file-states from conversation history. I accidentally deleted the wrong repo on my machine once and recreated all the lost work from agent chat history. It is, ironically, the sort of task which AI agents excel at.</p>
]]></description><pubDate>Fri, 17 Jul 2026 15:33:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=48948621</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48948621</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48948621</guid></item><item><title><![CDATA[New comment by wgd in "Inkling: Our Open-Weights Model"]]></title><description><![CDATA[
<p>You don't hear about them much because their models aren't really competitive. I really wanted to try Trinity Large as a daily-driver in the MiniMax M2 sort of niche but I couldn't make it through a single day. The models need another couple point releases worth of post-training to make useful agents and if memory serves they weren't any less slopped in writing style and those are really the only two things people look for in models.</p>
]]></description><pubDate>Wed, 15 Jul 2026 23:38:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48928644</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48928644</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48928644</guid></item><item><title><![CDATA[New comment by wgd in "The real prices of frontier models"]]></title><description><![CDATA[
<p>> DeepSeek and GLM are left out of the tables entirely: we only have rough characters-divided-by-four estimates for them, not real tokenizer counts, and this post is about measured numbers.<p>lolwut. The open-weight models are inscrutable black boxes for which we can't possibly get real token counts? Typical lazy clanker, BSing their way out of doing the whole job.</p>
]]></description><pubDate>Mon, 13 Jul 2026 23:36:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=48900401</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48900401</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48900401</guid></item><item><title><![CDATA[New comment by wgd in "Backtrack-Free Cursive"]]></title><description><![CDATA[
<p>Yeah, I originally expected this to be about a cursive variant which could be plotted as a single-valued function or something.</p>
]]></description><pubDate>Mon, 13 Jul 2026 21:08:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=48898914</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48898914</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48898914</guid></item><item><title><![CDATA[New comment by wgd in "People Can't Identify AI Poetry, but Enjoy It Less When Told It's by an AI"]]></title><description><![CDATA[
<p>A more accurate title might be "Average University Students Can't Identify Czech AI Poetry". The random-chance performance seen here is reminiscent of the 50-50 nonexpert performance measured in "People who frequently use ChatGPT for writing tasks are accurate and robust detectors of AI-generated text".<p>I would have been really interested to see someone explore whether greater exposure to non-poetry AI text generalizes to greater ability to sniff out AI poetry as well, but sadly this was not that sort of study.</p>
]]></description><pubDate>Mon, 13 Jul 2026 02:56:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=48887312</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48887312</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48887312</guid></item><item><title><![CDATA[New comment by wgd in "Odyssey Linux"]]></title><description><![CDATA[
<p>At first I thought this excerpt was meant to warn people off without directly alleging AI authorship, but I guess that's less likely since I see you're also the submitter</p>
]]></description><pubDate>Sun, 12 Jul 2026 00:28:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=48877154</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48877154</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48877154</guid></item><item><title><![CDATA[New comment by wgd in "Hy3"]]></title><description><![CDATA[
<p>The current Deepseek V4 Pro is still just their initial preview AFAIK, with the "real" model release rumored to come later this month. GLM-5.2 might be outperforming simply because it's had more post-training on top of the GLM-5 base.</p>
]]></description><pubDate>Thu, 09 Jul 2026 16:57:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=48848970</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48848970</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48848970</guid></item><item><title><![CDATA[New comment by wgd in "AI content is everywhere on social media, especially LinkedIn"]]></title><description><![CDATA[
<p>Pangram does work, in the specific sense that when it says something was AI authored it is vanishingly unlikely that it was written by a human (who was not deliberately trying to write like an AI), and IMO getting people to recognize that we actually do have a decent solution in this space now is pretty important if we want the Internet to remain a place for humans and not just bot swarms.<p>> rule of three, em dashes, etc<p>You appear to be misinformed about how Pangram specifically works, it is not based on pattern detection of that sort. I recommend reading their whitepaper, it's a pretty understandable explanation of exactly how they trained their classifier.</p>
]]></description><pubDate>Thu, 09 Jul 2026 16:48:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=48848837</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48848837</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48848837</guid></item><item><title><![CDATA[New comment by wgd in "Anthropic's Method to Losing Goodwill in a Few Easy Steps"]]></title><description><![CDATA[
<p>> The reason that people don't understand why Anthropic wont let the subscription be used with other harnesses<p>Even more specifically, the very fact that people would prefer, if they had the option, to use other harnesses with roughly equivalent feature sets strongly implies that the harness is not bringing them any value they couldn't get from a bunch of other places, including open-source equivalents.<p>Anthropic might want you to use their harness for their own reasons (control over caching, logging your interactions for training data, et cetera), but the idea that the Claude Code harness itself is bringing significant value which would help to lock users into the Anthropic ecosystem more than the Claude models alone do is kind of laughable. So _of course_ it seems like a baffling and arbitrary restriction to many users.</p>
]]></description><pubDate>Mon, 06 Jul 2026 14:53:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48805515</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48805515</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48805515</guid></item><item><title><![CDATA[New comment by wgd in "Aluminum foil (2021)"]]></title><description><![CDATA[
<p>I've read a lot of his other writings so that context might be informing my reading here but it sounds like he's pretty straightforwardly discussing the potential of aluminum foil as a uniform-feedstock-slash-construction-material for a hypothetical self-reproducing microfabricator.</p>
]]></description><pubDate>Mon, 06 Jul 2026 14:41:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=48805337</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48805337</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48805337</guid></item><item><title><![CDATA[New comment by wgd in "Does code cleanliness affect coding agents? A controlled minimal-pair study"]]></title><description><![CDATA[
<p>Yes, those ones would be at least a somewhat-plausible simulation of a real scenario people care about: a once-clean codebase that was allowed to become messy by a succession of insufficiently-careful vibeslop PRs.<p>I'm not a huge fan of their methodology for the AI-degraded cases either (ideally one would set up the mirror pairs by taking some real repositories and rewinding history a month or so and then having a succession of independent agents reimplement each bit of feature work and bugfixes over that period of time), but it's at least a coarse approximation whereas I just don't trust the cleanup methodology to resemble anything real in the first place.</p>
]]></description><pubDate>Mon, 06 Jul 2026 03:25:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=48800325</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48800325</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48800325</guid></item><item><title><![CDATA[New comment by wgd in "Does code cleanliness affect coding agents? A controlled minimal-pair study"]]></title><description><![CDATA[
<p>"agent pipelines that [...] clean a messy [repository]"<p>This feels like a terrible approach, sufficient to condemn the entire study.<p>Apparently half of the "minimal pairs" in this work were constructed in this way. I simply am not going to trust any conclusion that requires assuming these AI "cleaned" repos are in any way representative of actually-good codebases.</p>
]]></description><pubDate>Mon, 06 Jul 2026 01:55:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=48799868</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48799868</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48799868</guid></item><item><title><![CDATA[New comment by wgd in "Half-Baked Product"]]></title><description><![CDATA[
<p>Some dishwashers add a simple timer-based heuristic so if you open it for just a few seconds while you lazily grab something the "clean" indicator stays lit.</p>
]]></description><pubDate>Fri, 03 Jul 2026 13:49:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=48775031</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48775031</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48775031</guid></item><item><title><![CDATA[New comment by wgd in "A Bitter Lesson for Memory"]]></title><description><![CDATA[
<p>I've always been amazed at how terrible most frontier LLMs are at compaction given how embarrassingly easy it is to come up with half a dozen different RL training evals which would teach models to generate useful context summaries. Heck, you could bolt it onto any existing RL eval by just forcing a compaction every three turns.</p>
]]></description><pubDate>Mon, 22 Jun 2026 12:17:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=48629144</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48629144</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48629144</guid></item><item><title><![CDATA[New comment by wgd in "Running local models is good now"]]></title><description><![CDATA[
<p>The problem is that the moment you introduce shared remote hardware there's a slippery slope leading right back down to "just pay an inference host for model tokens". If you're transmitting your prompts over the internet to a trusted host you might as well just let that host be DeepInfra or together.ai or one of the many other providers already in that business.</p>
]]></description><pubDate>Tue, 16 Jun 2026 22:40:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=48563246</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48563246</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48563246</guid></item><item><title><![CDATA[New comment by wgd in "GLM 5.2 Is Out"]]></title><description><![CDATA[
<p>I've got a GLM subscription (mostly because I like supporting open model makers, pretty sure my monthly usage is so low that pay-per-token would be more cost effective), so I generally use GLM-5.1 for any personal projects and I use Opus at work.<p>To be entirely honest I haven't noticed much of a capability gap between the two for the sorts of things I ask of an AI agent. Maybe Opus is _slightly_ smarter or slightly better at long-running tasks but the difference is slim enough it could just be a placebo from the Claude branding / hype.<p>I'm looking forward to giving GLM-5.2 a spin sometime soon and seeing how it stacks up. If nothing else 1M context is a great improvement, feels like between DeepSeek v4, then MiniMax M3, and now GLM-5.2 adding it 1M is rapidly becoming "table stakes" for agentic models.</p>
]]></description><pubDate>Sat, 13 Jun 2026 23:52:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=48522680</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48522680</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48522680</guid></item><item><title><![CDATA[New comment by wgd in "GLM 5.2 Is Out"]]></title><description><![CDATA[
<p>The GLM-5 series is 744B-A40B. This is not a local model for any reasonable definition of local, but it's an open model which means (once they upload the weights in a week or so) there will be a dozen third-party inference providers competing on price per token.</p>
]]></description><pubDate>Sat, 13 Jun 2026 22:39:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=48522209</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48522209</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48522209</guid></item><item><title><![CDATA[New comment by wgd in "Kimi K2.7-Code: open-source coding model with better token efficiency"]]></title><description><![CDATA[
<p>Often in MoE models the experts are quantized while the shared portions, being a much smaller part of the network with greater impact, are kept at higher or full precision. Not familiar with the Kimi QAT approach specifically but it's likely they do this.</p>
]]></description><pubDate>Fri, 12 Jun 2026 21:45:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=48509794</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48509794</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48509794</guid></item><item><title><![CDATA[New comment by wgd in "A Server Called Mercury"]]></title><description><![CDATA[
<p>Yeah, the evidence feature is so terrible that it actively harms the overall reputation of Pangram. The main "is this AI or human?" classification is done with a machine learning model that works very well but has nothing (directly) to do with any of those stylistic cues it surfaces.<p>In any case, the Pangram link was just meant as objective corroboration of what's pretty blatantly obvious if you just read the text.<p>"Not cloud credits, not a managed platform, not a serverless function bobbing in someone else's abstraction."<p>"I used to work at Heroku. That sentence still does a lot of load-bearing work in how I think about computing."<p>"Here's the part that would have sounded like science fiction during my Heroku years: I didn't do most of the migration."<p>If you read these chunks of text and don't immediately feel the AI slop alarms blaring in the back of your head, you are perhaps underprepared for the modern internet.</p>
]]></description><pubDate>Wed, 10 Jun 2026 17:09:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=48479413</link><dc:creator>wgd</dc:creator><comments>https://news.ycombinator.com/item?id=48479413</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48479413</guid></item></channel></rss>