<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: foltik</title><link>https://news.ycombinator.com/user?id=foltik</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 08 Oct 2026 19:08:51 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=foltik" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by foltik in "“Math 2.0” will need to value mathematical progress more holistically"]]></title><description><![CDATA[
<p>It does during reasoning. You could even put it in a loop and force it to reflect at every step. Or spawn a bunch of review agents.<p>And yet you still just get slop.</p>
]]></description><pubDate>Thu, 08 Oct 2026 15:20:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=50006901</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=50006901</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=50006901</guid></item><item><title><![CDATA[New comment by foltik in "“Math 2.0” will need to value mathematical progress more holistically"]]></title><description><![CDATA[
<p>Exactly. In my experience, even giving VERY specific design guidelines and aesthetic criteria, these models always produce subpar overcomplicated code (and writing). Unless excruciatingly spoonfed at every step.</p>
]]></description><pubDate>Thu, 08 Oct 2026 15:13:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=50006813</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=50006813</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=50006813</guid></item><item><title><![CDATA[New comment by foltik in "GPT‑6 and Intelligent UI for everyone"]]></title><description><![CDATA[
<p>Ciechanowski’s works of art are in a completely different league than this AI slop.<p>It’s all just surface level complexity with no intention behind it. A clumsy approximation at best.</p>
]]></description><pubDate>Thu, 08 Oct 2026 05:26:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=50002056</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=50002056</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=50002056</guid></item><item><title><![CDATA[New comment by foltik in "What Meta got right with Muse"]]></title><description><![CDATA[
<p>So we should judge each piece of evidence in a vacuum?<p>Facebook has broken its promises about user data many times before. I don't buy anything they say about Muse, and nobody in their right mind should. Unless they’re getting paid to, of course.<p>Aside, if “a skill file organizing purchase history in markdown format” is the only thing you took away from that thread, you might want to take a closer look.</p>
]]></description><pubDate>Mon, 05 Oct 2026 01:04:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49959612</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49959612</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49959612</guid></item><item><title><![CDATA[New comment by foltik in "What Meta got right with Muse"]]></title><description><![CDATA[
<p>We all know exactly what it’s for, don’t pretend you don’t.</p>
]]></description><pubDate>Sun, 04 Oct 2026 18:42:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49956631</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49956631</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49956631</guid></item><item><title><![CDATA[New comment by foltik in "Micron CEO Says Memory Supply Will Be Much Tighter in 2027 and 2028 Than in 2026"]]></title><description><![CDATA[
<p><a href="https://en.wikipedia.org/wiki/DRAM_industry_price_fixing" rel="nofollow">https://en.wikipedia.org/wiki/DRAM_industry_price_fixing</a></p>
]]></description><pubDate>Thu, 01 Oct 2026 14:46:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49922378</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49922378</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49922378</guid></item><item><title><![CDATA[New comment by foltik in "U.S. appeals court upholds designation of Anthropic as supply chain risk"]]></title><description><![CDATA[
<p>Again, the statute says nothing about "dictating terms." It covers an adversary who might sabotage or subvert a system. And no, Anthropic did not cancel anything, did not pull any capabilities, and has no backdoor into deployed models. There was no "um actually." It wrote two prominent usage restrictions into a contract, the military read and signed it, and then later demanded the terms be removed.<p>As a sanity check, I’d highly recommend trying out your own line of reasoning in cases which you think you might feel differently about. For example, imagine it's instead a certain authoritarian state saying "obey the Party or get blacklisted as a national security threat."<p>It might feel good in the moment, like "yeah, Anthropic, don't tell the military what to do," because you agree with this particular call. But the government doesn't give that power back, and you might not be so happy about the next way it gets used.</p>
]]></description><pubDate>Sat, 26 Sep 2026 14:33:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49856972</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49856972</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49856972</guid></item><item><title><![CDATA[New comment by foltik in "U.S. appeals court upholds designation of Anthropic as supply chain risk"]]></title><description><![CDATA[
<p>Where did you get "trying to dictate" from? The text is "the risk that an adversary may sabotage, maliciously introduce unwanted function, or otherwise subvert [...]"<p>A mutually agreed upon "you may not use our product for X or Y purpose" in a contract clearly doesn't pose any such risk.</p>
]]></description><pubDate>Sat, 26 Sep 2026 00:13:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49851793</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49851793</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49851793</guid></item><item><title><![CDATA[New comment by foltik in "SAML: A fractal of bad design"]]></title><description><![CDATA[
<p>> I'm an IT consultant working for a private company<p>Makes sense for GP locally, I guess, but globally we’re all worse off when nobody is motivated to break free from the path of least resistance.</p>
]]></description><pubDate>Tue, 22 Sep 2026 22:00:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808775</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49808775</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808775</guid></item><item><title><![CDATA[New comment by foltik in "SAML: A Fractal of Bad Design"]]></title><description><![CDATA[
<p>That seems like a bit of an arbitrary line to draw, no? Would you say the same about webhooks?</p>
]]></description><pubDate>Tue, 22 Sep 2026 21:50:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808659</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49808659</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808659</guid></item><item><title><![CDATA[New comment by foltik in "Claude Opus 5.5"]]></title><description><![CDATA[
<p>Let’s not kid ourselves, Anthropic would be running their own distillation “attacks” too if _they_ were the ones playing catch-up. They’ve already shown as much with their illegal scraping of pirated books ($1.5B settlement).<p>I say just let them duke it out. After a decade of regulatory capture and enshittification, it’s nice to see some actual competition again.</p>
]]></description><pubDate>Tue, 22 Sep 2026 21:08:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808130</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49808130</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808130</guid></item><item><title><![CDATA[New comment by foltik in "NASA’s Mars Sample Return mission is dead"]]></title><description><![CDATA[
<p>Not that I have a solution, but isn’t that kind of disappointing? Why don’t we just do it directly, rather than only as a side effect of needless conflicts?</p>
]]></description><pubDate>Tue, 22 Sep 2026 03:46:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49796589</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49796589</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49796589</guid></item><item><title><![CDATA[New comment by foltik in "NASA’s Mars Sample Return mission is dead"]]></title><description><![CDATA[
<p>Whatever gets the humans to work harder I guess.</p>
]]></description><pubDate>Tue, 22 Sep 2026 03:28:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49796487</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49796487</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49796487</guid></item><item><title><![CDATA[New comment by foltik in "I had Gemini train its own replacement for $9"]]></title><description><![CDATA[
<p>I found it much more useful to go to a knife shop and handle a whole bunch of knives for myself. They’re all pretty similar besides material, so not much signal you’re going to be able to glean from people arguing on reddit.</p>
]]></description><pubDate>Thu, 17 Sep 2026 13:59:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49740873</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49740873</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49740873</guid></item><item><title><![CDATA[New comment by foltik in "Compressing a flag to 11 bits"]]></title><description><![CDATA[
<p>I’d give it a bit more credit than that. They encoded 128 flags with their scheme, only 5 of which have the union jack. The rest are truly built from scratch.<p>While in fairness the remaining flags seem like they would need more hard coded symbols, it’s still interesting as a procedural plausible flag generator.</p>
]]></description><pubDate>Tue, 15 Sep 2026 02:15:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49706896</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49706896</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49706896</guid></item><item><title><![CDATA[New comment by foltik in "What do Visa and Mastercard do? An intro to card networks"]]></title><description><![CDATA[
<p>Which is obviously not the case in this context?</p>
]]></description><pubDate>Thu, 10 Sep 2026 04:34:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49638440</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49638440</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49638440</guid></item><item><title><![CDATA[New comment by foltik in "Large language models develop novel social biases through adaptive exploration"]]></title><description><![CDATA[
<p>I'm also making a statistical observation. Saying a model "picks up on" a concept is standard shorthand, same as saying it has "learned.” What I meant is that the model has trained on plenty of neutral proper nouns that have negligible influence on the distribution of the following tokens, so the model is already <i>conditioned</i> towards treating them neutrally.<p>Not perfectly neutrally, as you said. But by your definition the only "unbiased" model is one whose output distribution perfectly matches the training distribution, i.e. one that memorized it. All LLMs have some amount of "bias” on literally every possible input.<p>The tribe names are no different. In the paper they run the same game again, and the bias is different every time. There's no innate preference between them trained into the model, just noise that's revealed due to a lack of any other signal. In a real situation with actually relevant information about the candidate in context, that noise is drowned out.<p>The more interesting thing to look for would be a bias that's strong enough to persist across different contexts. For example, is "banananow" consistently followed by positive tokens more than "pearian" across a diverse set of realistic prompts, by enough that someone could actually exploit it? The paper shows that’s explicitly not the case for made up tribe names.<p>What it does show, from what I can gather, is that bias can form inside a feedback loop. The model gets a success or failure result after each hire, and if a hire from one tribe happens to fail early on, the model steers that tribe away from that job for the rest of the game, even though every candidate had the same odds.</p>
]]></description><pubDate>Thu, 10 Sep 2026 03:34:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49638081</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49638081</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49638081</guid></item><item><title><![CDATA[New comment by foltik in "Large language models develop novel social biases through adaptive exploration"]]></title><description><![CDATA[
<p>But these scenarios are obviously ambiguous nonsense, which an LLM will pick up on.<p>And given to the lack of training data on such scenarios, surely the activations are mostly random noise?<p>It seems much more interesting to look for biases that appear robustly across different realistic scenarios that would actually be influenced by the training data</p>
]]></description><pubDate>Tue, 08 Sep 2026 23:40:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49618727</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49618727</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49618727</guid></item><item><title><![CDATA[New comment by foltik in "Saving 100 terabytes of memory by optimizing 1.1.1.1's DNS cache"]]></title><description><![CDATA[
<p>If you squint, even their TLV encoding in a Box<[u8]> is kind of a specialized arena allocator.</p>
]]></description><pubDate>Sat, 29 Aug 2026 04:57:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49487000</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49487000</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49487000</guid></item><item><title><![CDATA[New comment by foltik in "Show HN: Talos – An AI agent with a permission kernel between model and shell"]]></title><description><![CDATA[
<p>Bahahaha, come on man. The Claudish is so obvious it hurts.<p>Doubt you even read your own “645 line policy kernel.” Can you explain what that is and how it works in YOUR OWN words? If you paste Claude at us again we’re gonna know.</p>
]]></description><pubDate>Fri, 28 Aug 2026 23:00:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49485258</link><dc:creator>foltik</dc:creator><comments>https://news.ycombinator.com/item?id=49485258</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49485258</guid></item></channel></rss>