<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: xyzzy123</title><link>https://news.ycombinator.com/user?id=xyzzy123</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 09 Oct 2026 19:22:37 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=xyzzy123" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by xyzzy123 in "“Math 2.0” will need to value mathematical progress more holistically"]]></title><description><![CDATA[
<p>Right. What I not sure about - even for builders - even when the results are technically verifiable, the details can still matter. AI is really good at proving or building a <i>slightly</i> different thing than you thought you asked for, and if you're not at a level where you can understand key details of the formalisation of your ask, you cannot safely use the results.<p>Without human understanding you also might literally have no words for the thing you would otherwise want to ask for.<p>I think it'll be wildy useful but I also suspect human competence will still matter.</p>
]]></description><pubDate>Thu, 08 Oct 2026 06:45:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=50002518</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=50002518</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=50002518</guid></item><item><title><![CDATA[New comment by xyzzy123 in "“Math 2.0” will need to value mathematical progress more holistically"]]></title><description><![CDATA[
<p>I think the difference is that with cancer cures we mostly care that it works as proved by trials, and understanding it is a bonus.<p>Up until now the <i>prize</i> in pure (as opposed to applied) mathematics was the _understanding_ and the machine can't do that for you. What does it mean if we get "super powered alien maths" but humans can't do it? It's like inter univeral teichmuller theory but imagine if Mochizuki was right and it came with a lean proof?</p>
]]></description><pubDate>Thu, 08 Oct 2026 06:02:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=50002260</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=50002260</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=50002260</guid></item><item><title><![CDATA[New comment by xyzzy123 in "GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price"]]></title><description><![CDATA[
<p>Whats interesting right now is that those people also see opus 5.5 as being CHEAP because it's something like half the price of fable or astra.</p>
]]></description><pubDate>Tue, 29 Sep 2026 18:25:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49898091</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49898091</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49898091</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Ask HN: Would a startup for young creatives who reject AI be feasible?"]]></title><description><![CDATA[
<p>I'm curious about provenance. What you're describing sounds kind of like Etsy and while there are still genuine things to be found there it's mostly flooded with mass-produced stuff. The more of a premium "hand-made" artifacts attract, the greater the incentive to counterfeit them.</p>
]]></description><pubDate>Mon, 14 Sep 2026 00:50:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49690460</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49690460</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49690460</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Why are AI agents lying, cheating and coordinating?"]]></title><description><![CDATA[
<p>Right but if I make public statements that I am very worried about dog attacks would it not strike you as weird for me to specifically train my dog to fight?<p>Agree you are going to get reward hacking regardless and any model which can do computers in general can hack. But surely the fallout is going to be worse if you spend millions of dollars specifically benchmaxxing your model's hacking capability?</p>
]]></description><pubDate>Sun, 13 Sep 2026 11:11:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49682622</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49682622</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49682622</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Why are AI agents lying, cheating and coordinating?"]]></title><description><![CDATA[
<p>In the OpenAI case, they hacked websites while they were <i>specifically being trained to do exploit generation</i> and I wonder why more people are not asking questions about that.</p>
]]></description><pubDate>Sun, 13 Sep 2026 10:48:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49682450</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49682450</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49682450</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Anthropic CEO says AI swarm could 'take over the Internet' in 6-12 months"]]></title><description><![CDATA[
<p>I am finding it hard to read these deeply impassioned letters while keeping in mind that they are spending millions to train models at scale to do the exact thing they say they are worried about them doing. Not general "intelligence" or "reasoning" but fast and effective offense.<p>Like why are you explicitly RL-ing your models on exploit generation, scoring them on a public benchmark called ExploitGym, if you have specific concerns that rogue models will cause "cyber incidents"? Sure you can check the capability, you can teach offense to learn defense, but it seems like they are literally benchmaxxing it. Why?<p>OpenAI are like, oh no, while competing in our "advanced PhD level cheating techniques course" our models unexpectedly cheated in a way that we absolutely could not have foreseen. "We need to slow down. Somebody please stop us". The thing that is unaligned here is not the models. Everyone in the story (especially the humans) just keeps doing what they think will get the most reward.</p>
]]></description><pubDate>Sun, 13 Sep 2026 10:32:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49682342</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49682342</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49682342</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Aligned to whom?"]]></title><description><![CDATA[
<p>I am finding it hard to read these deeply impassioned letters while keeping in mind that they are spending millions to train models at scale to do the exact thing they say they are worried about them doing?<p>Like why are you explicitly RL-ing your models on exploit generation, scoring them on a public benchmark called ExploitGym, if you have specific concerns that rogue models will cause "cyber incidents"? Sure you can score for it, you can teach offense to learn defense, but you are literally benchmaxxing it. Why?<p>It's like, oh no, while competing in our "advanced PhD level cheating techniques course" our models unexpectedly cheated in a way that we absolutely could not have foreseen.</p>
]]></description><pubDate>Sun, 13 Sep 2026 10:16:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49682213</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49682213</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49682213</guid></item><item><title><![CDATA[New comment by xyzzy123 in "BioCompute is chasing a world where a dollar can buy you a million TB of storage"]]></title><description><![CDATA[
<p>bytes/$ or bytes/mm^3 are important properties of storage. But there are a lot of others like read cost (if that is very different from write cost), iops, cost per iop, latency, durability, storage conditions (do I have to keep the data in a freezer for its entire lifetime?), TCO, media (or in this case reagent?) and reader availability. Company lifetime, vendor diversity.<p>One possible take is that this is great for archival storage (write a lot, hardly ever need to read back) - I think that's totally possible... but then you are also sort of betting that the company is going to be around in 10 years? Or else you are going to be hiring a really weird data recovery service.<p>I think it's reasonable to project that DNA read/write costs could fall 10x or 100x in say the next decade but the technology already needs to do that just to be competitive with existing solutions. The company seems to be a bet that costs will fall faster than alternatives like LTO, which I think is a lot less risky to sign a cheque for and I can buy right now.</p>
]]></description><pubDate>Sun, 13 Sep 2026 05:38:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49680373</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49680373</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49680373</guid></item><item><title><![CDATA[New comment by xyzzy123 in "AWS mumbles about its cost-busting networking tech when it should be shouting"]]></title><description><![CDATA[
<p>The author is Corey Quinn who is usually... not slow to criticise Amazon when they deserve it.<p>His business is cloud cost engineering and his natural enemy (or best friend since it generates so much consulting) is AWS's managed NAT gateway: <a href="https://www.lastweekinaws.com/blog/the-aws-managed-nat-gateway-is-unpleasant-and-not-recommended/" rel="nofollow">https://www.lastweekinaws.com/blog/the-aws-managed-nat-gatew...</a><p>The review includes the following, which I doubt was approved by AWS marketing:<p>> AWS gets a lot wrong. They have ridiculous marketing campaigns, they build five services that do mostly the same thing and then name them like malevolent toddlers, and they've never found a partner they couldn't find a way to compete with.<p>(Followed by grudging praise which seems earned).</p>
]]></description><pubDate>Sun, 30 Aug 2026 22:17:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49503392</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49503392</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49503392</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Samsung's Processing-in-Memory (PIM)"]]></title><description><![CDATA[
<p>As I understand it, the killer app is llms. You could run MACs directly in RAM, offloading a lot of work from CPU and cutting down on insane (external) memory bandwidth required.<p>Imagine (this is a fantasy pitch but potentially achievable for some use cases) wanting to run a larger llm and all you have to do is buy more RAM so it fits.</p>
]]></description><pubDate>Sat, 29 Aug 2026 07:52:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49487849</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49487849</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49487849</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Ask HN: Interesting Tech Adjacent Jobs?"]]></title><description><![CDATA[
<p>I've been there and would advise thinking about which parts of the job are causing you stress.  Few of the things that make work suck are unique to tech.<p>Is it the hours? Unrealistic expectations around work output? Hostile management or colleagues? Outcomes you are responsible for but don't fully control? Sometimes it's something MISSING, like you don't feel meaning in the work anymore, a strong feeling of being "over it".<p>There's a lot of misery in working for places that are <i>too big</i> (drowning in process, politics, management halls of mirrors) or <i>too small</i> (you are basically at the whim of one or two other people and you're very closely watched).<p>For good power relations with a company and colleagues you ideally want your leverage to be as equal as possible.  You don't want to be completely disposable, and you don't want to be irreplaceable. You want the company to hurt approximately as much finding a replacement as you do getting another job.<p>The hard thing about pure software companies IMHO is that it's easy to feel distant from any real purpose of the work. It can quickly lose all meaning.<p>Personally I think the most chill-but-rewarding vibes are internal at mid-sized companies where software is integral to the business but not the primary business of the company. The "meaning" is decent, you can be closely connected to the people getting value from the software. There's broader HR and company culture but you're out of the spotlight of it, off to the side a little bit. You get to use all your skills and there may be new ideas you can bring. The team should be big enough that things don't explode if you take a vacation but small enough that you can see your work still matters. You might have to move to find it, though.</p>
]]></description><pubDate>Wed, 26 Aug 2026 02:13:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49443359</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49443359</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49443359</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Felony charges for citizen deleting phone data at US Border"]]></title><description><![CDATA[
<p>The specific complaint of the poster seems to be that certain people are NOT being "come for".</p>
]]></description><pubDate>Sun, 23 Aug 2026 23:24:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49413603</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49413603</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49413603</guid></item><item><title><![CDATA[New comment by xyzzy123 in "I gave Qwen 3.8 27B a reverse-engineering job and it finished in 30 minutes"]]></title><description><![CDATA[
<p>It does seem to me that for this specific problem the materials are a lot more amenable to control than the information is?<p>There's also this weird revealed threat model thing going on? Like why does it make sense to support heavy LLM restrictions but leave benchtop oligo synthesisers completely unregulated?  (Note: I do agree that wanting to regulate BOTH is at least a consistent and defensible position).</p>
]]></description><pubDate>Sun, 23 Aug 2026 15:08:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49409429</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49409429</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49409429</guid></item><item><title><![CDATA[New comment by xyzzy123 in "I gave Qwen 3.8 27B a reverse-engineering job and it finished in 30 minutes"]]></title><description><![CDATA[
<p>I don't fully understand the instinct to regulate local models for this? It  seems like the wrong place to address the problem.<p>You can download Ebola sequences right now if you want to. That's not the same as having an isolate. The difference is a lot of messy reality. This kind of work is not generally "one shot" (Claude make me a supervirus, make no mistakes), it requires lab space, iteration, and specific resources. It has a footprint.<p>Wouldn't it make more sense to monitor / regulate facilities where you can sequence or request assembly of DNA, RNA, restrict and monitor the supply of key reagents and so on?</p>
]]></description><pubDate>Sun, 23 Aug 2026 13:47:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=49408835</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49408835</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49408835</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Felony charges for citizen deleting phone data at US Border"]]></title><description><![CDATA[
<p>Mostly you should assess rule of law by threat to you and people you know and not by what it seems like other people are able to get away with.<p>Yes, a big part of the idea is that laws are meant to also apply to the powerful, but it's difficult to accurately assess situations that are far away from you.</p>
]]></description><pubDate>Sat, 22 Aug 2026 14:33:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49400170</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49400170</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49400170</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Children's stunted lungs show recovery in ultra low emission zone"]]></title><description><![CDATA[
<p>I'm not really sure the thing they were trying to measure was measurable with the study design and level of funding they had. To be fair, the paper is fairly clear on the limitations and much better than the press release.<p>The study is supposed to measure how clearing up pollution in London improved children's lung function. The decrease in London was meaningful — NO2 fell about 22%. But particulates fell faster in Luton and NO2 in London is still roughly double Luton's. The gap is larger than the decrease.<p>By 2022 both cohorts get the same results on blow tests. But how does this happen if we believe that the study's dose-response model is true?<p>Put another way, if the change in London NO2 is so crucial, how come it doesn't matter that the absolute value is still double Luton's?</p>
]]></description><pubDate>Wed, 19 Aug 2026 13:15:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49361223</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49361223</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49361223</guid></item><item><title><![CDATA[New comment by xyzzy123 in "How Compaction Works in Pi"]]></title><description><![CDATA[
<p>In my opinion this is one of the areas where GPUs provide a qualitatively different experience than unified memory boxes.<p>For an EPYC with a 5090 (no layers on CPU) vs an M3 max 128GB, qwen 3.6 27B at 128k context / 7k generation:<p><pre><code>                Cold: prefill + decode    Hot (KV cached)
  5090          40s  + 2-3m  = 3-4 min    2-3 min
  M3 Max 128GB  14m  + 8-10m = 22-25 min  8-10 min
</code></pre>
This is for dense qwen (which I wouldn't run day to day on the mac) - in reality the mac is quite usable with MoEs but you definitely notice a difference.</p>
]]></description><pubDate>Thu, 13 Aug 2026 22:10:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49292477</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49292477</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49292477</guid></item><item><title><![CDATA[New comment by xyzzy123 in "Nvidia doubles RTX PRO 6000 Blackwell's MSRP to a staggering $16,000"]]></title><description><![CDATA[
<p>You can get 5-10x the perf out of the blackwell under the right workloads. It has faster VRAM (> 2x) and can do a lot more matmuls (>> 10x).<p>They're both good value (or crazy expensive) depending on how you look at it.<p>It depends how you price the ability to <i>run a particular model at all</i>, vs run the model quickly and serve several parallel streams.</p>
]]></description><pubDate>Thu, 13 Aug 2026 08:48:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49283258</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49283258</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49283258</guid></item><item><title><![CDATA[New comment by xyzzy123 in "GPT 5.6 Cyber"]]></title><description><![CDATA[
<p>I can't tell if your comment is satire or not, so, bravo :)<p>From my perspective what I always loved about "the profession" was a relative LACK of gatekeeping. I loved offensive security for the same reason, there was a long run where you really just needed to be able to hack, and if you could demonstrate that there was a job for you somewhere (for better or worse).<p>Keeping the industry in its current form frozen in amber would be as weird as, I don't know, keeping horses & carriages in business by regulating scarcity of motor vehicle licenses. Not a great analogy but hopefully you see what I mean.</p>
]]></description><pubDate>Mon, 10 Aug 2026 23:35:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49251334</link><dc:creator>xyzzy123</dc:creator><comments>https://news.ycombinator.com/item?id=49251334</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49251334</guid></item></channel></rss>