<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: wrsh07</title><link>https://news.ycombinator.com/user?id=wrsh07</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 25 Aug 2026 05:52:51 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=wrsh07" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by wrsh07 in "Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)"]]></title><description><![CDATA[
<p>There are other methods, including asking them to eg speak using the Google style guide rules, or to speak at 2/10 verbosity, or to speak as if they're talking to a very intelligent middle schooler, or to use simple sentences in subject verb object form</p>
]]></description><pubDate>Fri, 21 Aug 2026 14:21:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49388556</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49388556</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49388556</guid></item><item><title><![CDATA[New comment by wrsh07 in "Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (2025)"]]></title><description><![CDATA[
<p>To quote Sarah Constantin:<p>> Humans Who Are Not Concentrating Are Not General Intelligences^<p>But yeah, you should still treat them with humanity. (Related: you should treat LLMs well, not because they're human but because you are^^)<p>^
<a href="https://srconstantin.github.io/2019/02/25/humans-who-are-not-concentrating.html" rel="nofollow">https://srconstantin.github.io/2019/02/25/humans-who-are-not...</a><p>^^ hmm couldn't find this tweet but didn't look too hard</p>
]]></description><pubDate>Fri, 21 Aug 2026 13:33:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49387828</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49387828</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49387828</guid></item><item><title><![CDATA[New comment by wrsh07 in "Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing"]]></title><description><![CDATA[
<p>The scheme that Scott Aaronson describes essentially uses a specific prng, and you can then check a certain function with relatively few tokens to get a sense of whether or not a model using that scheme generated the text<p>A few caveats: you need to know the key to the function (used when generating the text) and you need to know the bias it would introduce<p>The point is that you do not need to know the full prefix, just a modest sample set of contiguous tokens</p>
]]></description><pubDate>Mon, 17 Aug 2026 17:55:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49335010</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49335010</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49335010</guid></item><item><title><![CDATA[New comment by wrsh07 in "Stripe will reportedly acquire OpenRouter for $7B+"]]></title><description><![CDATA[
<p>It's pretty valuable to own the customer touch point<p>Also when people ask questions like this it usually means they're asking questions about a point in time, eg given today's numbers<p>But valuation should account for trajectory (where will they be in five years?)<p>In the hypothesized bull case for ai, they benefit dramatically from the secular tailwinds (ie unrelated to their own business strategy) of AI<p>So one way to consider this is it's a hedge by stripe in case AI becomes as big as some people think it is. And 7b to get in on the ai wave is a good deal in that case. And in all the other cases, it goes to zero and stripe is fine</p>
]]></description><pubDate>Mon, 17 Aug 2026 03:10:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49326169</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49326169</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49326169</guid></item><item><title><![CDATA[New comment by wrsh07 in "Accelerating GPT-5.6 Sol Ultrafast"]]></title><description><![CDATA[
<p>Seems like they will do Sol first while capacity constrained? I can't imagine the margins they'll be charging</p>
]]></description><pubDate>Thu, 13 Aug 2026 19:52:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49291042</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49291042</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49291042</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>It depends on how it does watermarking!!<p>Note, there are many ways to represent words visually on computers that look identical</p>
]]></description><pubDate>Tue, 11 Aug 2026 19:15:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49263077</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49263077</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49263077</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>That's actually much worse because it fundamentally changes the output, whereas this doesn't change the output, it just changed the prng</p>
]]></description><pubDate>Tue, 11 Aug 2026 19:12:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49263030</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49263030</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49263030</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>No you don't need to do that, the prng is detectable if you know what bias to look for and have the key</p>
]]></description><pubDate>Tue, 11 Aug 2026 19:11:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49263022</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49263022</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49263022</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>That's what I excerpted, although I had seen it presented from his talk at Stony Brook</p>
]]></description><pubDate>Tue, 11 Aug 2026 19:10:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49263014</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49263014</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49263014</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>Scott Aaronson talks about his project at OpenAI to do this^<p>You can carefully select which pseudorandom number generator (prng) you use to be able to id text of a certain length. I expect there is some performance characteristic you have to manage since you're doing this on every inference, but once you do that it doesn't change the output in any meaningful way (the prng is still a statistically valid prng, it just happens to let you check if the output used that prng)<p>The point is that you can do this simply by swapping to a different RNG, which isn't noticeable to the end user, and while it changes the output, it's not any different from how using a different seed or being lumped in a different batch will change the output.<p>^ excerpt:<p>> So then to watermark, instead of selecting the next token randomly, the idea will be to select it pseudorandomly, using a cryptographic pseudorandom function, whose key is known only to OpenAI. That won’t make any detectable difference to the end user, assuming the end user can’t distinguish the pseudorandom numbers from truly random ones. But now you can choose a pseudorandom function that secretly biases a certain score—a sum over a certain function g evaluated at each n-gram (sequence of n consecutive tokens), for some small n—which score you can also compute if you know the key for this pseudorandom function.</p>
]]></description><pubDate>Tue, 11 Aug 2026 13:11:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49257743</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49257743</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49257743</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>I'm curious about your thoughts on pangram. I only really see posts on Reddit claiming it falsely labels their content as ai generated but nobody will actually post examples of "textbook from twenty years ago" or upload screenshots of a journal (also those posts usually feel deeply ai generated without an ai detector)<p>Do you think this is an impossible task and we shouldn't try to solve it? Or do you think it's doable and that some ai detectors might be better than others?</p>
]]></description><pubDate>Tue, 11 Aug 2026 13:01:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49257618</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49257618</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49257618</guid></item><item><title><![CDATA[New comment by wrsh07 in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>Somewhat trivially, if I ask Claude to transcribe an image and then check if that transcription is ai generated it will likely say yes.<p>Many users are not smart enough to realize that the transcription step is where the ai (watermarks) were necessarily injected.</p>
]]></description><pubDate>Tue, 11 Aug 2026 12:59:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49257596</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49257596</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49257596</guid></item><item><title><![CDATA[New comment by wrsh07 in "Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs"]]></title><description><![CDATA[
<p>This is probably true in a smaller way for OpenAI<p>When you have a lot of free users the business demands that you serve them with the best cheap model you can build<p>And time spent building that may provide dividends (eg OpenAI has very good RL and reasoning) but it might take resources away from the larger model training<p>(I have no inside knowledge, so please consider this to all be speculation)</p>
]]></description><pubDate>Wed, 05 Aug 2026 19:24:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187679</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49187679</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187679</guid></item><item><title><![CDATA[New comment by wrsh07 in "Ten advances in mathematics and theoretical computer science"]]></title><description><![CDATA[
<p>It seems like they threw it a decently large battery of open math problems and probably limited it to something like $200-500 per problem:<p><a href="https://x.com/polynoamial/status/2083478171975082334" rel="nofollow">https://x.com/polynoamial/status/2083478171975082334</a><p>As a complete guess, it seems like they tested hundreds to thousands of problems with a relatively low per-problem budget<p>--<p>The linked tweet from Noam Brown at OpenAI reads:<p>> And yes we did try other major problems without success. Sadly no Millennium Prize problems (yet).<p>> But also, we didn’t spend a lot on each problem. It’s possible to push test-time compute much further.</p>
]]></description><pubDate>Sat, 01 Aug 2026 18:18:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49136915</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49136915</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49136915</guid></item><item><title><![CDATA[New comment by wrsh07 in "Show HN: tale.fyi, we deserve a home for fiction"]]></title><description><![CDATA[
<p>This is great! I would love an explicit "set bookmark to here" button if I scroll up from the high watermark<p>I discovered the reset, but it took a minute and I expect given the way I read being able to explicitly walk the high watermark back a bit would be nice (maybe that's doable if you select text? I didn't try!)<p>The format is nice, I love standard ebooks<p>I opened Moby Dick but didn't see how to listen to it on my phone, not sure what the status of the "listen" effort is.<p>I listened to some books from standard ebooks using eleven labs, but I'm really excited about a better implementation</p>
]]></description><pubDate>Tue, 28 Jul 2026 13:49:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49083868</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49083868</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49083868</guid></item><item><title><![CDATA[New comment by wrsh07 in "The Dark Night of Mathematics"]]></title><description><![CDATA[
<p>I have two strong reactions to pieces like this:<p>First, this is about not just a job or career, but someone's identity. And we must have compassion that they are losing something that is fundamental to who they are. To them, it doesn't just feel like they're losing it, it is being taken by these companies that have so often acted in ways we despise<p>And that is a tragedy! And there are so many tragedies like this that will happen regularly as the technology advances<p>But my second reaction is one of this great shared experience. I studied math. Many of my friends are mathematicians. And it is now possible for anybody to access mathematical insights or think deeply about strange conjectures and theorems that were previously incomprehensible to anyone who hadn't at least studied mathematics in college<p>I don't know if 3b1b's Grant Sanderson considers himself a mathematician or science communicator, but I think of him as both. And while I expect he is someone with the mental capacity to find and prove new things in the world, I am grateful for the time he spends instead understanding things at a fundamental level and explaining and celebrating those concepts with his audience.<p>Not every mathematician can be or wants to be _that_ kind of mathematician. But for the moment, it is enough for me that higher mathematics is more accessible than ever.<p>And with some trepidation, I predict that the "traditional" job of a mathematician will change (of course it will). And it will change in big obvious ways and also subtle little ones. How will it change, though?<p>Importantly, math is actually going to be one of the fields where the fundamental things we thought we knew are shaken, because in the next five years we are going to start to see connections between things that were previously considered completely separate.<p>And so one way the job will change is that anyone who discusses math regularly will need to learn new things.<p>To me, this is exciting. It's almost like finding a bunch of new dinosaur fossils that fundamentally reshape our understanding. There's going to be a lot of work to do!</p>
]]></description><pubDate>Sat, 25 Jul 2026 17:57:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49049927</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49049927</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49049927</guid></item><item><title><![CDATA[New comment by wrsh07 in "Terence Tao's ChatGPT conversation about the Jacobian Conjecture counterexample"]]></title><description><![CDATA[
<p>It's also much better to distribute the challenge of identifying problems amenable to which prompts<p>Could they do it? Sure, but to what end? It would make more people hate them and feel even more "take our interesting work." Pitching it as a useful tool just makes more sense on all levels</p>
]]></description><pubDate>Thu, 23 Jul 2026 13:02:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49020915</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49020915</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49020915</guid></item><item><title><![CDATA[New comment by wrsh07 in "Terence Tao's ChatGPT conversation about the Jacobian Conjecture counterexample"]]></title><description><![CDATA[
<p>Note, in posts like the following, the author indicates they are able to get free subscriptions from OpenAI<p>As an aside, there's this idea in math that when you create a new field you shouldn't solve all the easy problems - you need to entice other people to learn about the field!<p>I expect there's some element of that here. It's much better for OpenAI and Anthropic if their users are the ones discovering and writing up the results of the AI solving hard math. Look at the high school and college age students who have become ai power users and potentially learned how to use git to contribute ai generated solutions<p>(Related: I believe Terry has also gotten all of the subscriptions gifted to him)<p><a href="https://xenaproject.wordpress.com/2026/07/20/human-mathematicians-are-being-outcounterexampled/" rel="nofollow">https://xenaproject.wordpress.com/2026/07/20/human-mathemati...</a></p>
]]></description><pubDate>Thu, 23 Jul 2026 13:00:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49020877</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=49020877</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49020877</guid></item><item><title><![CDATA[New comment by wrsh07 in "Faster binary search: from compiled code to mechanical sympathy"]]></title><description><![CDATA[
<p>Can you say more about what happens for bucket assignment? Eg is it appending to a file or writing to a 2d array?<p>Agreed on linear vs n log n, ofc, and streaming considerations might also be relevant</p>
]]></description><pubDate>Sat, 18 Jul 2026 01:33:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=48954244</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=48954244</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48954244</guid></item><item><title><![CDATA[New comment by wrsh07 in "Faster binary search: from compiled code to mechanical sympathy"]]></title><description><![CDATA[
<p>I'm somewhat curious about the initial problem:<p>> Consider the following real problem, one of the steps in scikit-learn’s gradient histogram boosting algorithm:<p>> You have a large array of floating point numbers.<p>> You want to assign them to the integer range 0-254, spread out evenly.<p>Naively I would consider sorting the initial array and then using something like `batched` from itertools to chunk them into the 255 buckets - binary search will give you a bunch of random accesses, and sorting can be cache-oblivious (eg efficient for arbitrary data sizes)<p>But I'm somewhat concerned I don't fully understand the underlying problem being solved with this step, so I might be misunderstanding the intended result</p>
]]></description><pubDate>Fri, 17 Jul 2026 15:31:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=48948596</link><dc:creator>wrsh07</dc:creator><comments>https://news.ycombinator.com/item?id=48948596</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48948596</guid></item></channel></rss>