<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: 405error</title><link>https://news.ycombinator.com/user?id=405error</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 09 Sep 2026 12:30:48 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=405error" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by 405error in "How An AI math breakthrough ignited a controversy"]]></title><description><![CDATA[
<p>But consider they <i>could</i> decide that they want to scoop more regular research work too. They could automate it with just a few LoC. Even if you opted out in the ToS, you'd have to file a massive lawsuit just to enforce it. And the actual fine would be inconsequential to OpenAI.<p>I think going forward, any researcher should consider anything submitted to an LLM to be copied/stolen.</p>
]]></description><pubDate>Wed, 09 Sep 2026 11:39:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49624808</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49624808</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49624808</guid></item><item><title><![CDATA[New comment by 405error in "How An AI math breakthrough ignited a controversy"]]></title><description><![CDATA[
<p>Personally this is a watershed moment for researchers and grad students I know. All of them are close sourcing WIP repos, not putting their progress in LLMs, or have lab level initiatives to self host models.</p>
]]></description><pubDate>Wed, 09 Sep 2026 11:36:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49624778</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49624778</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49624778</guid></item><item><title><![CDATA[New comment by 405error in "The Navier–Stokes Millennium Prize Problem"]]></title><description><![CDATA[
<p>I can give you some context.
1. Terence Tao's mastodon explains the way this problem was solved does not in itself contribute much. LLMs (and in this case) produce massive, often unintelligible proofs that do not further understanding. It is often that in pursuit of solving these problems, many other discoveries are made. 
2. There is a more serious question about scooping. If OAI is using chat data from researchers to make discoveries, essentially <i>every</i> researcher who chats with an LLM can get scooped. You could be 80% of your way to solving a problem, and LLM could solve the remaining 20%, and get all the credit. Years of your work could be scooped in an instant. If you're a PhD student, this is even worse. Here it's a world famous problem. But imagine you're a PhD student, working on your small but extremely career/progression critical problem, and you get scooped by an AI you talk to. No one is even going to care.</p>
]]></description><pubDate>Wed, 09 Sep 2026 07:05:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49622382</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49622382</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49622382</guid></item><item><title><![CDATA[New comment by 405error in "Navier-Stokes – Tristan Buckmaster [pdf]"]]></title><description><![CDATA[
<p>We <i>know</i> they are training on private chats. It's listed in the ToS.</p>
]]></description><pubDate>Wed, 09 Sep 2026 06:21:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49621931</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49621931</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49621931</guid></item><item><title><![CDATA[New comment by 405error in "Navier-Stokes – Tristan Buckmaster [pdf]"]]></title><description><![CDATA[
<p>Tristian's allegations are <i>much</i> more serious than academic slap-fighting. If what he suggests is true, every academic using AI is going to get scooped. Yes AI can do non-trivial work, but the situation is that you could be a PhD student 90% of a way to make a major breakthrough. Then OAI scoops up your chats, dumps ten million tokens, and claims it for itself.</p>
]]></description><pubDate>Wed, 09 Sep 2026 06:19:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49621918</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49621918</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49621918</guid></item><item><title><![CDATA[New comment by 405error in "Navier-Stokes – Tristan Buckmaster [pdf]"]]></title><description><![CDATA[
<p>Many academics and grad students I know have closed source their in progress work, and started being really careful about what they chat with LLMs (or using local ones) because of the drama around this. No one wants four years of their life getting sniped by ten million dollars worth of tokens.</p>
]]></description><pubDate>Wed, 09 Sep 2026 06:11:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49621852</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49621852</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49621852</guid></item><item><title><![CDATA[New comment by 405error in "Navier-Stokes – Tristan Buckmaster [pdf]"]]></title><description><![CDATA[
<p>It would not be difficult to write a pipeline to remove 99% of low quality posts, especially about specific subjects. It would be very easy to identify accounts as researchers based on their chat logs.</p>
]]></description><pubDate>Tue, 08 Sep 2026 23:40:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49618728</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49618728</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49618728</guid></item><item><title><![CDATA[New comment by 405error in "4.5B Posts Scraped from TikTok"]]></title><description><![CDATA[
<p>It's probably AI coded and hallucinated many things. That drumming up the importance of a minor thing is a real tell. Another hallucination - it hasn't found any video APIs (despite statements that it has and uploaded it). It has video <i>metadata</i>.</p>
]]></description><pubDate>Thu, 03 Sep 2026 12:46:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49549268</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49549268</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49549268</guid></item><item><title><![CDATA[New comment by 405error in "4.5B Posts Scraped from TikTok"]]></title><description><![CDATA[
<p>More directly, there simply aren't any video files uploaded. It's just parquet files, which contain no video columns (I'm not even sure if it supports it).</p>
]]></description><pubDate>Thu, 03 Sep 2026 12:40:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49549214</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49549214</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49549214</guid></item><item><title><![CDATA[New comment by 405error in "4.5B Posts Scraped from TikTok"]]></title><description><![CDATA[
<p>It's that mix of dense, impressive sounding jargon, but even scanning across it raises glaring problems. Like, if you have 4.5 billion videos on HF, and it's 289GB, it's about 60 bytes per video. Checking the column fields as well, there doesn't seem to be any video files*.</p>
]]></description><pubDate>Thu, 03 Sep 2026 12:28:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49549106</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49549106</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49549106</guid></item><item><title><![CDATA[New comment by 405error in "4.5B Posts Scraped from TikTok"]]></title><description><![CDATA[
<p>The first immediate smell is that if you have 4.5B rows and 289GB in data, you have ~60 bytes per row.</p>
]]></description><pubDate>Thu, 03 Sep 2026 12:15:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49548992</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49548992</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49548992</guid></item><item><title><![CDATA[New comment by 405error in "4.5B Posts Scraped from TikTok"]]></title><description><![CDATA[
<p>I cannot verify whether it is technically correct, but it's about how to defeat Tiktok's bot filters to scrape it.</p>
]]></description><pubDate>Thu, 03 Sep 2026 12:09:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49548945</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49548945</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49548945</guid></item><item><title><![CDATA[New comment by 405error in "Cultivating a state of mind where new ideas are born (2023)"]]></title><description><![CDATA[
<p>I have a vague recollection of a point someone made before that most ideas are usually wrong or incorrect at the start. For example, Nvidia was built on the thesis that non-triangular polygons would dominate, which was 100% wrong (today all polygons are triangular). But they often develop into something true.</p>
]]></description><pubDate>Tue, 18 Aug 2026 09:30:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49343325</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49343325</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49343325</guid></item><item><title><![CDATA[New comment by 405error in "A quick look at zero-knowledge proofs"]]></title><description><![CDATA[
<p>Aren't several cryptocurrencies built on ZKPs? Their business model aside, ZKPs do look like one of the rare examples of theoretical elegance and real world use (even if not widespread).</p>
]]></description><pubDate>Tue, 18 Aug 2026 09:28:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49343312</link><dc:creator>405error</dc:creator><comments>https://news.ycombinator.com/item?id=49343312</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49343312</guid></item></channel></rss>