<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: kingstnap</title><link>https://news.ycombinator.com/user?id=kingstnap</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 17 Aug 2026 07:19:12 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=kingstnap" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by kingstnap in "Abdominal fat predicts heart disease risk better than BMI"]]></title><description><![CDATA[
<p>On the individual level it is a discipline problem. But the fact it has become an individual discipline problem is exactly the core issue.<p>If you turn things into individual discipline problems then surprise you get population level issues. As you pointed out people have other things to focus willpower on thats not this.<p>The fault, as it often does, lies in marketing. Turns out heavily marketed, hyper palatable food, designed to be minimally satiating so you maximally over eat is great for profits and terrible for obesity.</p>
]]></description><pubDate>Sun, 16 Aug 2026 13:59:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49320156</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49320156</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49320156</guid></item><item><title><![CDATA[New comment by kingstnap in "Abdominal fat predicts heart disease risk better than BMI"]]></title><description><![CDATA[
<p>Counting macros is extremely effective. The general population doesn't have the discipline for it, but amoung sport coaches, actors, models, bodybuilders, this is *the* technique for achieving goals.<p>Anyway what's cute about this is that this isn't novel phenomenon. There are lots of parallels to this in other fields where often effective solutions do exist, but at a system level don't seem to work.<p>Telling people to diet doesn't fix population level obesity. Telling people about personal financial management doesn't stop people from accruing too much high interest debt. Telling teenagers to stop idolizing instagram influences doesn't fix body anxiety issues. Telling people to stop smoking/drinking doesn't fix addictions. 3-2-1 data backups absolutely work but people lose files all the time.</p>
]]></description><pubDate>Sun, 16 Aug 2026 05:48:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49317245</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49317245</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49317245</guid></item><item><title><![CDATA[New comment by kingstnap in "DeepSeek Harness"]]></title><description><![CDATA[
<p>You would think raw performance wouldn't be a problem given most of whats happening is waiting for network calls and streaming tokens.<p>But modern bloat manages perfectly well to make apps that wait for network calls run poorly enough to give you a bad experience.</p>
]]></description><pubDate>Thu, 13 Aug 2026 14:22:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49286422</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49286422</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49286422</guid></item><item><title><![CDATA[New comment by kingstnap in "Grok 4.6"]]></title><description><![CDATA[
<p>The $60 billion cursor option that SpaceX bought was exercised on June 16th. The deal is closed.</p>
]]></description><pubDate>Wed, 12 Aug 2026 22:35:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49279511</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49279511</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49279511</guid></item><item><title><![CDATA[New comment by kingstnap in "How Claude marks AI-generated content"]]></title><description><![CDATA[
<p>It could also partly be a byproduct of examples of claude writing being in the dataset, which of course anthropic has lots and lots of and they do train on.</p>
]]></description><pubDate>Mon, 10 Aug 2026 22:19:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49250627</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49250627</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49250627</guid></item><item><title><![CDATA[New comment by kingstnap in "Learning more about Claude's mathematical capabilities"]]></title><description><![CDATA[
<p>Lets play over/under on an AI model proving (or counter exampling) the Riemann hypothesis?<p>I'm not sure what a good mark would be, but considering this result lets put it at 2027-08-10 (One year from today).</p>
]]></description><pubDate>Mon, 10 Aug 2026 18:27:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49247688</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49247688</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49247688</guid></item><item><title><![CDATA[New comment by kingstnap in "Mistral Patent for “Code implemented tool calls”"]]></title><description><![CDATA[
<p>> The USPTO has a strange insistence on granting them even though they aren't legally valid<p>I recently learned [0] that the USPTO makes it money from patents, its not government funded. Not only that but checking patents loses them net money while maintenance fees are the real cash cow.<p>The whole system is similar to the revenue model of a shitty journal that just publishes whatever research as long as the author pays. Except the office doesn't even need to care about their reputation in granting dubious patents because they have legal backing.<p>[0] It was a comment on hacker news, that I checked.</p>
]]></description><pubDate>Mon, 10 Aug 2026 15:54:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49245340</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49245340</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49245340</guid></item><item><title><![CDATA[New comment by kingstnap in "Auto Mode will be the default in Claude Code – because humans can't be trusted"]]></title><description><![CDATA[
<p>The appeal of auto mode is obvious.<p>Non yolo mode where you manually approve each command is literally pure security theatre. No human on earth has the patience to discerne the huge stream of commands agents run.<p>So even if false negatives are technically zero, having 100% false positives is not acceptable.<p>So a fundamental premise is you have to have filtration. Where only very few things are surfaced for humans to look<p>The first option which a decent number of harnesses do is being able to setup an auto approve list. Like `ls` is fine, `cat` is fine etc.<p>Now maintaining this list is in and of itself a giantic pita. But the real issue is that it still has way too many false positives. Fundamentally it comes down to the halting problem where you can't really include important things like `bash python <<PY` and what not which agents like to use. But regex can't solve the halting problem to figure out if the Python is safe.<p>So naturally the next best option is to use an LLM. Which isn't that stupid because even if its non deterministic at least it can dramatically reduce the false positives from the regex auto approve list.</p>
]]></description><pubDate>Sat, 08 Aug 2026 17:44:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49224037</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49224037</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49224037</guid></item><item><title><![CDATA[New comment by kingstnap in "Improving GPT-5.6 Sol in ChatGPT—and expanding access for free users"]]></title><description><![CDATA[
<p>You can't even pick 5.6 luna on the chat app with a paid subscription. It just gives you sol (or use older models) with what seems like basically as much usage as you want. And sol is considerably smarter than luna.<p>All of this is of pretty minor importance though. You can't read as many tokens as a subcription can produce so more chat is not the value add nor super important.<p>I mean there are literally so many providers for free chat if you are willing to use several seperate apps.<p>The real value in these subs is using codex cli, much like the real point of anthropic subs is using claude code. Because agentic work actually does require a lot of tokens.</p>
]]></description><pubDate>Thu, 06 Aug 2026 19:58:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49201596</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49201596</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49201596</guid></item><item><title><![CDATA[New comment by kingstnap in "Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users"]]></title><description><![CDATA[
<p>Its always fun to try to read between the lines here to speculate why they are doing this.<p>Maybe Luna efficiency gain was actually significant enough that putting all the free users and giving them super generous limits makes sense.<p>They might be doing this to improve the messaging of AI among causal users since right now there is a huge amount of datacenter backlash in the US due to AI grievances.<p>Maybe they have too much excess capacity or they really want to juice token numbers and market share on their dashboards for marketing.<p>I also wonder if being given access to an actually a decent model like luna with actual thinking budget instead of brainless "instant" modes will start to make causal users understand the real capabilities of these models.</p>
]]></description><pubDate>Thu, 06 Aug 2026 19:31:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49201242</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49201242</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49201242</guid></item><item><title><![CDATA[New comment by kingstnap in "OpenAI says my prepaid credits were consumed, refuses to show any record"]]></title><description><![CDATA[
<p>I mean it could be problematic if taxes and the credit card fees are involved right.<p>It would be interesting to see something like:<p>$100 credits and $13 in taxes refunded as<p>$97 in refund after merchent fee + $12.61 as reversed taxes.<p>It's objectively better than nothing for sure. But the thing about this is such complicated nuanced solutions tend to get not done.</p>
]]></description><pubDate>Wed, 05 Aug 2026 22:41:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49190044</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49190044</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49190044</guid></item><item><title><![CDATA[New comment by kingstnap in "OpenAI says my prepaid credits were consumed, refuses to show any record"]]></title><description><![CDATA[
<p>They need better clairty on it for sure, ideally with a clear warning on deposit. But apparently a non-expiring credit is a considered an accounting problem since you have income (deposit) associated with a contract liability (credits) that could end up making a headache if it splays across a bunch of fiscal years.<p>You can setup limits though to minimize loss. For example you can say auto reload $5 when your account hits $5 with an absolute monthly limit of $40. Then you don't have risk of losing more than the roughly ~$5 you have in an account.</p>
]]></description><pubDate>Wed, 05 Aug 2026 22:04:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49189672</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49189672</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49189672</guid></item><item><title><![CDATA[New comment by kingstnap in "Microsoft Is Quietly Deleting Mentions of 32 GB RAM Recommendation for Win 11"]]></title><description><![CDATA[
<p>If you are planning to use vibe coded programs then you better plan for a lot of RAM. Well I say planning when its really being thrust upon you in a lot of cases.<p>But be extremely mindful when looking at anything touched by Anthropic or OpenAI.<p>These coding agents have terrible feedback cycles for lag and memory leaks. A particularly egregious one is GPU usage which is absurdly high in most of these apps due to terrible design and react render loop issues. Someone needs to make RL enviroments than punish this.<p>I will caveat this with saying being wasteful with resources is not new at all. But at this point it seems like no one is doing any amount of basic due diligence.</p>
]]></description><pubDate>Wed, 05 Aug 2026 18:55:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187300</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49187300</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187300</guid></item><item><title><![CDATA[New comment by kingstnap in "We accidentally built an LLVM compiler for Jax"]]></title><description><![CDATA[
<p>Just for receipts I span up a basic test for XLA:<p>Can you figure out that 16 matrix vector multiplications into a concatenate is the same as concatenating first into a matrix matrix operation which can go on the GEMM. This is like the most basic thing you can imagine doing.<p>Turns out no it doesn't and there is a 23% performance difference by moving the concatenate up in the python code in my specific test.<p>Not to say that it didn't recognize it. I had an agent look at the XLA and the graph actually does a partial fusion into a sum and stack. But does not realize the whole thing is just a GEMM.<p>In more complicated examples the differences you can get can be much larger.</p>
]]></description><pubDate>Sun, 02 Aug 2026 05:43:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49141473</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49141473</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49141473</guid></item><item><title><![CDATA[New comment by kingstnap in "We accidentally built an LLVM compiler for Jax"]]></title><description><![CDATA[
<p>This is actually really neat, I had an agent create some experimental probes and check out whats going on here and it seems like a pretty cool project. I'm gonna need to digest how this might be useful but I think this could be cool for some sort of FPGA uses perhaps.<p>> this is not going to beat XLA for standard deep learning workloads. XLA has years of hyper-specific optimizations for linear algebra on GPUs and TPUs.<p>I'm not so convinced about this.<p>It's actually really easy to write two mathematically equivalent formulations in Python of something basic that have over a 2x performance difference in them after jax jit.<p>XLA is not that smart. And I'm not talking some niche nonsense I mean simple matrix multiplication graphs and residual connections on the CUDA backend.</p>
]]></description><pubDate>Sun, 02 Aug 2026 05:17:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49141342</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49141342</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49141342</guid></item><item><title><![CDATA[New comment by kingstnap in "Ten advances in mathematics and theoretical computer science"]]></title><description><![CDATA[
<p>I didn't argue that knowing the total cost is uninteresting. What I was saying is that realistically the total cost is:<p>Hours needed for prompt + Hours needed to check result + API costs.<p>You don't say "well let's add together the total yearly compensation of all the engineers and mathematicians at OpenAI that were involved" and throw that into the total cost. That's simply nonsense accounting.<p>The actual comparison you are making is some university researcher weighing between getting a grad student (several tens of thousands of dollars) vs typing up a prompt and sending a request to OpenAI for inference (as mentioned in the article, around $2000 in API and maybe a few hours for the prompt and harness).</p>
]]></description><pubDate>Sat, 01 Aug 2026 16:35:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49135883</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49135883</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49135883</guid></item><item><title><![CDATA[New comment by kingstnap in "Ten advances in mathematics and theoretical computer science"]]></title><description><![CDATA[
<p>Why would you factor in salary unless they had to baby it through. You would only count the hours for setting up the harness and prompt and checking the result.<p>Training the model is going to be amortized over other uses.</p>
]]></description><pubDate>Sat, 01 Aug 2026 11:27:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49133458</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49133458</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49133458</guid></item><item><title><![CDATA[New comment by kingstnap in "Ten advances in mathematics and theoretical computer science"]]></title><description><![CDATA[
<p>It's remarkable how you can manage to get these models to produce remarkable breakthroughs like an explicit construction of a non-sofic group.<p>And yet this is the exact same company that has screwed up their android app so bad that the latex N^3 rendering problem makes it so having it explain it to me crashes the app.<p>Truly jagged beyond belief.</p>
]]></description><pubDate>Sat, 01 Aug 2026 11:00:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49133251</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49133251</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49133251</guid></item><item><title><![CDATA[New comment by kingstnap in "I flagged two research papers for fake authors and both were accepted as orals"]]></title><description><![CDATA[
<p><a href="https://arxiv.org/stats/monthly_submissions" rel="nofollow">https://arxiv.org/stats/monthly_submissions</a><p>They should consider swapping this for a log plot.<p>I can imagine in 2027 academia looking like Moltbook.</p>
]]></description><pubDate>Thu, 30 Jul 2026 23:54:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117381</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49117381</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117381</guid></item><item><title><![CDATA[New comment by kingstnap in "Advancing the price-performance frontier with GPT‑5.6"]]></title><description><![CDATA[
<p>Those prices on luna are killer.<p>Haiku was already in a ditch.<p>But this is coming straight for the jugular of a ton of models on openrouter.</p>
]]></description><pubDate>Thu, 30 Jul 2026 17:52:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49113333</link><dc:creator>kingstnap</dc:creator><comments>https://news.ycombinator.com/item?id=49113333</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49113333</guid></item></channel></rss>