<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: ddp26</title><link>https://news.ycombinator.com/user?id=ddp26</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 03 Sep 2026 05:57:16 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=ddp26" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[Reasons robotics is hard]]></title><description><![CDATA[
<p>Article URL: <a href="https://secondthoughts.ai/p/14-reasons-robotics-is-hard">https://secondthoughts.ai/p/14-reasons-robotics-is-hard</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49543191">https://news.ycombinator.com/item?id=49543191</a></p>
<p>Points: 70</p>
<p># Comments: 32</p>
]]></description><pubDate>Wed, 02 Sep 2026 22:02:30 +0000</pubDate><link>https://secondthoughts.ai/p/14-reasons-robotics-is-hard</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49543191</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49543191</guid></item><item><title><![CDATA[New comment by ddp26 in "Gemini 3.8 Flash and 3.8 Flash Cyber"]]></title><description><![CDATA[
<p>There must be a deeper read on why Google can rapidly ship better small models while being delayed months on the bigger model.<p>What's the simplest explanation?</p>
]]></description><pubDate>Wed, 02 Sep 2026 19:13:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49540990</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49540990</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49540990</guid></item><item><title><![CDATA[New comment by ddp26 in "How accurate have Ed Zitron's AI skeptic predictions been?"]]></title><description><![CDATA[
<p>Even though these predictions turned out mostly wrong, we should not castigate people for publicly forecasting! That is virtuous, and more people should do it.<p>Thank you Ed!</p>
]]></description><pubDate>Wed, 02 Sep 2026 15:16:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49537604</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49537604</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49537604</guid></item><item><title><![CDATA[Models may behave differently in graded episode]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.lesswrong.com/posts/AfoGGrJfuNzofpzWL/models-may-behave-differently-in-graded-episodes-a-tirade">https://www.lesswrong.com/posts/AfoGGrJfuNzofpzWL/models-may-behave-differently-in-graded-episodes-a-tirade</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49536633">https://news.ycombinator.com/item?id=49536633</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Wed, 02 Sep 2026 14:18:26 +0000</pubDate><link>https://www.lesswrong.com/posts/AfoGGrJfuNzofpzWL/models-may-behave-differently-in-graded-episodes-a-tirade</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49536633</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49536633</guid></item><item><title><![CDATA[New comment by ddp26 in "Gemini 3.7 Flash"]]></title><description><![CDATA[
<p>What are we to infer from no release of gemini-3.5-pro, but frequent releases of smaller flash models (presumably from the same large pre-training run?)</p>
]]></description><pubDate>Fri, 14 Aug 2026 02:42:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49294220</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49294220</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49294220</guid></item><item><title><![CDATA[Jeff Dean's Discovery Loop Should Automate Chip Design First]]></title><description><![CDATA[
<p>Article URL: <a href="https://futuresearch.ai/discovery-loop-forecast/">https://futuresearch.ai/discovery-loop-forecast/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49200796">https://news.ycombinator.com/item?id=49200796</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Thu, 06 Aug 2026 18:57:21 +0000</pubDate><link>https://futuresearch.ai/discovery-loop-forecast/</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49200796</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49200796</guid></item><item><title><![CDATA[Google is in talks for a $1.5B deal to acquire Mechanize]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.businessinsider.com/google-mechanize-deal-talent-tech-ai-coding-2026-8">https://www.businessinsider.com/google-mechanize-deal-talent-tech-ai-coding-2026-8</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49200199">https://news.ycombinator.com/item?id=49200199</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Thu, 06 Aug 2026 18:08:17 +0000</pubDate><link>https://www.businessinsider.com/google-mechanize-deal-talent-tech-ai-coding-2026-8</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49200199</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49200199</guid></item><item><title><![CDATA[AI isn't enough to protect social media communities from AI]]></title><description><![CDATA[
<p>Article URL: <a href="https://arstechnica.com/gadgets/2026/08/ai-isnt-enough-to-protect-social-media-communities-from-ai/">https://arstechnica.com/gadgets/2026/08/ai-isnt-enough-to-protect-social-media-communities-from-ai/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49197340">https://news.ycombinator.com/item?id=49197340</a></p>
<p>Points: 3</p>
<p># Comments: 0</p>
]]></description><pubDate>Thu, 06 Aug 2026 14:40:34 +0000</pubDate><link>https://arstechnica.com/gadgets/2026/08/ai-isnt-enough-to-protect-social-media-communities-from-ai/</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49197340</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49197340</guid></item><item><title><![CDATA[New comment by ddp26 in "Goodhart's Law Comes for Every Benchmark You Trust"]]></title><description><![CDATA[
<p>Not forecasting though. You can't goodhart predicting real-world events</p>
]]></description><pubDate>Wed, 05 Aug 2026 21:31:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49189314</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49189314</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49189314</guid></item><item><title><![CDATA[New comment by ddp26 in "Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs"]]></title><description><![CDATA[
<p>This seems bad for AI safety/risk. Does DeepMind have any checks on model alignment now? What's stopping them from using AI for military/surveillance purposes?</p>
]]></description><pubDate>Wed, 05 Aug 2026 19:03:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187412</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49187412</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187412</guid></item><item><title><![CDATA[New comment by ddp26 in "Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs"]]></title><description><![CDATA[
<p>Would it though? Meta and Microsoft have had very scandalous AI things happen, and their shares didn't tank (or quickly recovered)</p>
]]></description><pubDate>Wed, 05 Aug 2026 18:48:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187206</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49187206</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187206</guid></item><item><title><![CDATA[New comment by ddp26 in "Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs"]]></title><description><![CDATA[
<p>I would have thought Jeff Dean would never ever leave Google. What on earth is going on?</p>
]]></description><pubDate>Wed, 05 Aug 2026 16:48:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49185410</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49185410</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49185410</guid></item><item><title><![CDATA[Show HN: FutureSearch, AI forecasting you can verify]]></title><description><![CDATA[
<p>*Title:* Show HN: FutureSearch, AI forecasting you can verify<p>AI forecasting is now approximately superhuman. Today, FutureSearch is exiting our long public beta and launching.<p>We started FutureSearch in August 2023. (We’re the original AI forecasting company, at least in a Tetlock-ian, “forecast anything” sense.) We’re currently #1 of 194 in the most competitive AI forecasting tournament [1], and we score above the #3 and #2 human forecasters in the premier mixed human-bot tournaments [2].<p>Many people on HN seem to equate forecasting with prediction markets and finance. FutureSearch is not a financial tool, in the same way that “deep research” is not a financial tool. Yes, we do evaluate our forecaster on prediction markets [3]. But forecasting is about being as accurate about the future as possible, and the real game is in forecasting scientific progress, geopolitics, and the future of humanity. Our founding team came from Metaculus, where we pushed human forecasting to the limit on questions like when AGI would arrive. At. FutureSearch, we co-authored the AI 2027 timeline forecast, where we predicted superhuman coding and research would come around 2032, longer than the other authors, but still shorter than skeptics [4].<p>Thousands of people used the FutureSearch beta and ran >10k high-effort forecasts, on all sorts of diverse topics. Ask it anything about the future. We now support decision forecasts too: “If I do X, will I achieve this outcome?”<p>Forecasting, as a capability, is useful even at the level of expert human crowds. But we predict that we will soon have strongly superhuman forecasting. People who bet against AI capability trend lines tend to lose, and the trend line in AI forecast accuracy tells a pretty clear story [5]. There’s no reason to think the best human forecasters have figured out everything predictable about the world. There’s a lot more signal to be found.<p>And if you’re skeptical, try it. We’ve seen our fair share of exaggerated claims about AI forecasting accuracy [6]. So part of the reason we made the free tier give a few of our highest effort forecasters free is to let anyone verify the quality.<p>[1] <a href="https://www.metaculus.com/tournament/summer-futureeval-2026/" rel="nofollow">https://www.metaculus.com/tournament/summer-futureeval-2026/</a>
[2] evals.futuresearch.ai
[3] markets.futuresearch.ai
[4] <a href="https://ai-2027.com/research/timelines-forecast" rel="nofollow">https://ai-2027.com/research/timelines-forecast</a>
[5] <a href="https://www.astralcodexten.com/p/the-ai-superforecasters-are-here" rel="nofollow">https://www.astralcodexten.com/p/the-ai-superforecasters-are...</a>
[6] <a href="https://www.lesswrong.com/posts/uGkRcHqatmPkvpGLq/contra-papers-claiming-superhuman-ai-forecasting" rel="nofollow">https://www.lesswrong.com/posts/uGkRcHqatmPkvpGLq/contra-pap...</a></p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49157941">https://news.ycombinator.com/item?id=49157941</a></p>
<p>Points: 11</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 03 Aug 2026 16:27:59 +0000</pubDate><link>https://futuresearch.ai/</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=49157941</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49157941</guid></item><item><title><![CDATA[New comment by ddp26 in "Data centers have hiked electricity prices on the public by $23B"]]></title><description><![CDATA[
<p>Isn't this the same as saying "utility regulators delaying connecting new power to the grid hiked electricity prices on the public by $23B?"<p>When my apples are expensive, I don't generally grumble about all the demand from pie makers. If they demand more apples, new suppliers should come in to restore the price, right?</p>
]]></description><pubDate>Wed, 15 Jul 2026 01:32:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=48915159</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48915159</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48915159</guid></item><item><title><![CDATA[New comment by ddp26 in "The Tower Keeps Rising"]]></title><description><![CDATA[
<p>> I can ask an agent to add OAuth, you can ask one to add caching, and somebody else can ask one to rebuild the database from first principles and make the UI pink. Each change can be reasonable in isolation.<p>But this is just bad vibecoding? This would be bad if humans did it too. With agents or humans, you need to coordinate.</p>
]]></description><pubDate>Tue, 14 Jul 2026 20:21:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=48912485</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48912485</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48912485</guid></item><item><title><![CDATA[New comment by ddp26 in "Proof of care in the age of AI"]]></title><description><![CDATA[
<p>When I know something is (primarily) AI generated, I lose interest.<p>The exception is when it's about a niche I care about, e.g. an analysis of opening trends of early world chess champions. I'll read AI on that for an hour.<p>My sense is that, for most writing,  it's fundamentally interpersonal, the information is about the author as much as it is about the world.<p>Maybe this flood of slop will cause people to care more about the substance of the writing, not the perspectives of the writing.</p>
]]></description><pubDate>Tue, 14 Jul 2026 16:04:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=48908917</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48908917</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48908917</guid></item><item><title><![CDATA[New comment by ddp26 in "GPT-5.6"]]></title><description><![CDATA[
<p>Is it possible GPT-5.6 is not a very aligned model?</p>
]]></description><pubDate>Fri, 10 Jul 2026 14:00:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48860060</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48860060</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48860060</guid></item><item><title><![CDATA[New comment by ddp26 in "GLM 5.2 and the coming AI margin collapse"]]></title><description><![CDATA[
<p>People have been making claims about the commoditization of llms since chatGPT, and they've been wrong every time as quality and prices and differentiation have increased.</p>
]]></description><pubDate>Tue, 07 Jul 2026 15:30:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=48819202</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48819202</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48819202</guid></item><item><title><![CDATA[New comment by ddp26 in "The AI Superforecasters Are Here"]]></title><description><![CDATA[
<p>But Scott's point is more: why even have markets? Once you have the superforecasting available on the questions you care about, why do you need to publish it for everyone to also react to?</p>
]]></description><pubDate>Mon, 06 Jul 2026 20:19:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=48809945</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48809945</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48809945</guid></item><item><title><![CDATA[New comment by ddp26 in "The AI Superforecasters Are Here"]]></title><description><![CDATA[
<p>Almost by definition, once AI forecasters are in the market, they won't (all) be beating the market.<p>But why evaluate AI forecasters by beating the market? Do we evaluate deep learning by whether hedge funds make money from it in the markets? These things have far, far more utility outside of finance.</p>
]]></description><pubDate>Mon, 06 Jul 2026 18:42:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48808747</link><dc:creator>ddp26</dc:creator><comments>https://news.ycombinator.com/item?id=48808747</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48808747</guid></item></channel></rss>