<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: xyzzy123</title><link>https://news.ycombinator.com/user?id=xyzzy123</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sat, 22 Aug 2026 13:35:48 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=xyzzy123" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by xyzzy123 in "Children's stunted lungs show recovery in ultra low emission zone"]]></title><description><![CDATA[
<p>I'm not really sure the thing they were trying to measure was measurable with the study design and level of funding they had. To be fair, the paper is fairly clear on the limitations and much better than the press release.<p>The study is supposed to measure how clearing up pollution in London improved children's lung function. The decrease in London was meaningful — NO2 fell about 22%. But particulates fell faster in Luton and NO2 in London is still roughly double Luton's. The gap is larger than the decrease.<p>By 2022 both cohorts get the same results on blow tests. But how does this happen if we believe that the study's dose-response model is true?<p>Put another way, if the change in London NO2 is so crucial, how come it doesn't matter that the absolute value is still double Luton's?</p>
]]></description><pubDate>Wed, 19 Aug 2026 13:15:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49361223</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49361223</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49361223</guid></item><item><title><![CDATA[New comment by xyzzy123 in "How Compaction Works in Pi"]]></title><description><![CDATA[
<p>In my opinion this is one of the areas where GPUs provide a qualitatively different experience than unified memory boxes.<p>For an EPYC with a 5090 (no layers on CPU) vs an M3 max 128GB, qwen 3.6 27B at 128k context / 7k generation:<p><pre><code>                Cold: prefill + decode    Hot (KV cached)
  5090          40s  + 2-3m  = 3-4 min    2-3 min
  M3 Max 128GB  14m  + 8-10m = 22-25 min  8-10 min
</code></pre>
This is for dense qwen (which I wouldn't run day to day on the mac) - in reality the mac is quite usable with MoEs but you definitely notice a difference.</p>
]]></description><pubDate>Thu, 13 Aug 2026 22:10:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49292477</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49292477</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49292477</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Nvidia doubles RTX PRO 6000 Blackwell's MSRP to a staggering $16,000"]]></title><description><![CDATA[
<p>You can get 5-10x the perf out of the blackwell under the right workloads. It has faster VRAM (> 2x) and can do a lot more matmuls (>> 10x).<p>They're both good value (or crazy expensive) depending on how you look at it.<p>It depends how you price the ability to <i>run a particular model at all</i>, vs run the model quickly and serve several parallel streams.</p>
]]></description><pubDate>Thu, 13 Aug 2026 08:48:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49283258</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49283258</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49283258</guid></item><item><title><![CDATA[New comment by xyzzy123 in "GPT 5.6 Cyber"]]></title><description><![CDATA[
<p>I can't tell if your comment is satire or not, so, bravo :)<p>From my perspective what I always loved about "the profession" was a relative LACK of gatekeeping. I loved offensive security for the same reason, there was a long run where you really just needed to be able to hack, and if you could demonstrate that there was a job for you somewhere (for better or worse).<p>Keeping the industry in its current form frozen in amber would be as weird as, I don't know, keeping horses & carriages in business by regulating scarcity of motor vehicle licenses. Not a great analogy but hopefully you see what I mean.</p>
]]></description><pubDate>Mon, 10 Aug 2026 23:35:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49251334</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49251334</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49251334</guid></item><item><title><![CDATA[New comment by xyzzy123 in "GPT 5.6 Cyber"]]></title><description><![CDATA[
<p>Great, the start of model segmentation where I'm gonna need a legal license to ask about legal problems, a nutritionist license to create a meal plan, a medical license to ask about an x-ray, a pilots license to ask about a flight plan, be a registered electrician to ask how to wire something, etc etc. The licensing of allowed thoughts.<p>Apparently I can pay for partial solutions to the Riemann hypothesis but if my question involves a crackme or something that is an existential risk somehow.</p>
]]></description><pubDate>Mon, 10 Aug 2026 22:40:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49250875</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49250875</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49250875</guid></item><item><title><![CDATA[New comment by xyzzy123 in "DeepMind's WeatherNext model achieves breakthrough forecasting cyclones"]]></title><description><![CDATA[
<p>I feel like the details of this are highly dependent on the confidence of the warning; moving large numbers of people (particularly elderly) will result in some deaths regardless. I guess more time to do it should help though.</p>
]]></description><pubDate>Sat, 08 Aug 2026 14:00:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49221900</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49221900</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49221900</guid></item><item><title><![CDATA[New comment by xyzzy123 in "DeepMind's WeatherNext model achieves breakthrough forecasting cyclones"]]></title><description><![CDATA[
<p>I know this is uncharitable and I am wrong but I am having trouble coming up with concrete scenarios where you die with 2 days notice but survive with 3. I am nonethless a believer that more accurate forecasting has value.</p>
]]></description><pubDate>Sat, 08 Aug 2026 13:48:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49221817</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49221817</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49221817</guid></item><item><title><![CDATA[New comment by xyzzy123 in "DeepSeek V4 Flash 0731"]]></title><description><![CDATA[
<p>There are a lot of tasks that are hard for organisations to run consistently but require some intelligence - monitoring logs and metrics for anomalies and security events, backup audits, audit processes in general, ensuring document quality and consistency, database advice and tuning, customer experience management, process optimisation - that are not "long horizon" in the classical sense of each step depending on the last, but are the result of consistency and attention over a long period of time and a large amount of data.<p>For this genre of task execution can run with limited horizon and is independent but would be too expensive to do with "us frontier tokens", I think for these, there is value in availability of cheaper tokens.</p>
]]></description><pubDate>Fri, 07 Aug 2026 23:07:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49217292</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49217292</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49217292</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Nashville uses eminent domain to block data center near zoo"]]></title><description><![CDATA[
<p>The value prop for the crazy sounding idea of putting GPUs into orbit.</p>
]]></description><pubDate>Thu, 06 Aug 2026 04:14:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49192351</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49192351</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49192351</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Ask HN: How do you correct spatial reasoning of LLMs?"]]></title><description><![CDATA[
<p>I guess my v0 would be notches around the mating end to allow for compression, and a hose clamp. Maybe split shaft collars off aliexpress for a fancier look. The design space of barbell collars is a big set of worked solutions to a very similar problem.</p>
]]></description><pubDate>Thu, 06 Aug 2026 03:23:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49192049</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49192049</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49192049</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Ask HN: How do you correct spatial reasoning of LLMs?"]]></title><description><![CDATA[
<p>Since it's the morning I thought I would ask Claude to create an example. This is just a first pass; it will be totally wrong but demonstrates the general idea. I don't know anything about barbells or your sensor: <a href="https://claude.ai/code/artifact/1f1a8a12-e5b2-4ee8-b20c-7da06720e58c" rel="nofollow">https://claude.ai/code/artifact/1f1a8a12-e5b2-4ee8-b20c-7da0...</a><p>I'm not sure from your description if you control the geometry of the sensor part (i.e, can the clamp be integral to the sensor housing or does it need to be a separate part) also ignores the internals, etc etc.<p>One thing I noticed is that off the bat, it did think about the assembly in general terms but would need guidance to think harder about FDM limitations and layer orientation etc. These are not good designs for printing. A good dev loop and git history help with these kinds of revisions.<p>The general principle is that if your domain is verifiable at all, give the model tools and a workflow that constrain and check its output. You want the LLM arguing with the geometry kernel instead of you.<p>To fully close the loop you print parts as fast and possible and concretely see where they suck.</p>
]]></description><pubDate>Wed, 05 Aug 2026 23:06:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49190278</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49190278</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49190278</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Ask HN: How do you correct spatial reasoning of LLMs?"]]></title><description><![CDATA[
<p>I would consider asking it to model its proposed designs in openscad or build123d (ideally something query-able). Then have it render and examine plausibility / suitability from different angles. Get it to render the part in use also and give instructions to think about forces and motion.<p>Recommend doing this in a coding harness not a chat box.<p>The reason I think you might have more success with this is that the model is mostly thinking about the part in words, which it can convert to a part design in CAD <i>in code</i>. LLMs are really good at coding. Also means it can use relative positioning and relationships.<p>You will be able to iterate more easily, compare things, compute properties, commit to git etc. The process is more reproducible and steerable than generative production of images.<p>When the LLM can look at renders of the geometry it generated, it’s easier for it to discriminate when it’s producing nonsense like misaligned parts, things that don’t fit, etc. It’s still going to kind of suck, but it will be better. The whole process of code -> render -> inspect forces the model to put up or shut up and provides grounding. Meshes > bloviating.<p>As far as I know today's LLMs don't have a "visual imagination" but a process like this could be a slow approximation of one. They clearly do have SOME spatial understanding (pelican tests show us that!) but it feels really non-human.<p>One thing missing from this is kinesthetics. Personally I am mostly not thinking in accurate visuals in mechanical design. I am imagining how the parts <i>feel</i> and kind of how they move and what slips first and what bends and what feels heavy. Imagining what my hands would feel. But I don't think I trust LLMs to evaluate that stuff by writing simulation code yet.</p>
]]></description><pubDate>Wed, 05 Aug 2026 12:46:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49182080</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49182080</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49182080</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Oxide Computer raises $445M (SEC Form D)"]]></title><description><![CDATA[
<p>I do feel like Oxide are qualitatively different from the others pjlmp mentioned because they're staying commodity in key interfaces like CPU ISA / OS / ecosystem (thing with big network effects) and mainly focused on fixing the control plane / management / architecture mess.<p>The others suffered "ecosystem collapse". With Oxide you won't be stuck on a "burning platform", your main risk is that the value prop for the hardware & management experience doesn't play out.</p>
]]></description><pubDate>Wed, 05 Aug 2026 07:25:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49179637</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49179637</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49179637</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Why Airplanes Use 400 Hz Power"]]></title><description><![CDATA[
<p>The other aspect is that the electrical parts of the power system are lighter at the higher frequency too. A 400hz transformer is meaningfully smaller and lighter than a 50hz one.<p>This is the same idea as switching power supplies but planes had to solve that before power semiconductors were cheap enough.<p>The power frequency also has an impact on the size of all the induction motors.<p>I guess if you were designing it from scratch today you would let the generator produce power at whatever frequency the mechanical engineers tell you they want to and the power electronics could convert it without problems. You would do the same thing the other way around on the motors.</p>
]]></description><pubDate>Tue, 04 Aug 2026 01:26:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49163400</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49163400</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49163400</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>You're right of course, LLMs provide a partial, unsound oracle.<p>The "halting problem is unsolvable" argument relies on the oracle not being able to output "not sure". But adding that option admits trivial oracles, like ones which output "not sure" for everything, so some are better than others.<p>The "real world" use most people have for halting oracles is as part of software safety, where if the checker outputs "not sure" you modify the software until the checker can decide if it halts.</p>
]]></description><pubDate>Fri, 31 Jul 2026 01:07:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117893</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49117893</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117893</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>Its funny because you can write a halting problem oracle by calling out to an LLM and have it return yes / no / not sure and get it to work reliably for almost all real code, like that is an entirely practical thing to do in 2026.<p>All we need now is some sort of program to evaluate halting problem oracles...</p>
]]></description><pubDate>Thu, 30 Jul 2026 23:13:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117027</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49117027</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117027</guid></item><item><title><![CDATA[New comment by xyzzy123 in "AI's top startups are barely publishing their research"]]></title><description><![CDATA[
<p>Hi, shredding the book is to reduce potential legal liability of format shifting - if you shred the book after scanning, the theory is you are not increasing the number of copies. This has come up as a factor in legal rulings. The legality of format shifting is still murky though, and the legality of <i>training</i> on the work are a separate question.</p>
]]></description><pubDate>Wed, 29 Jul 2026 23:13:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49104300</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49104300</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49104300</guid></item><item><title><![CDATA[New comment by xyzzy123 in "What if useful AI is a fantasy?"]]></title><description><![CDATA[
<p>> Practically speaking, is there really a difference between prompting an LLM to create a CMS for a website and installing something like WordPress or Ghost?<p>There is some difference; when you install WordPress or Ghost, your problems are shared. There's a community that goes along with the software and a "shared understanding in the world" of how it works and where the rough edges and limitations are. A lot of the time, LLMs will actually be better at modifying WordPress or Ghost to do what you want than they will be at fixing the CMS you built last Tuesday.<p>When you custom build, you can get an exact fit for your needs but your misery is yours alone.</p>
]]></description><pubDate>Wed, 29 Jul 2026 00:51:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49092050</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49092050</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49092050</guid></item><item><title><![CDATA[New comment by xyzzy123 in "What if useful AI is a fantasy?"]]></title><description><![CDATA[
<p>The problem he's describing comes up any time you have a team of developers. That doesn't mean the usefulness of teams is a fantasy because you didn't write all the code yourself.<p>If you don't have a useful mental model of the architecture that's a communication / review / documentation problem, not a problem with the entire idea of codegen.</p>
]]></description><pubDate>Wed, 29 Jul 2026 00:43:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49091989</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49091989</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49091989</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Be skeptical of OpenAI's rogue hacker agent story"]]></title><description><![CDATA[
<p>The incident is funny on 2 levels; a) OpenAI thought so little of the model's _actual_ security capabilities that they gave it a "wet paper bag" sandbox and b) OpenAI failed to use AI to accelerate their security processes.</p>
]]></description><pubDate>Fri, 24 Jul 2026 22:03:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49042164</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49042164</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49042164</guid></item></channel></rss>