<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: futureshock</title><link>https://news.ycombinator.com/user?id=futureshock</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sun, 30 Aug 2026 10:33:35 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=futureshock" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by futureshock in "GLM-5.3 is now open-weight"]]></title><description><![CDATA[
<p>I think it would be an important historical document as well. We are potentially looking at the dawn of AGI and one of the most important models ever created. Each model is also a kind of ultimate time capsule, containing a snapshot of the entire human collective mind. If you wanted to ask a 2002 person what they thought about future historical events you can just ask them directly.</p>
]]></description><pubDate>Fri, 28 Aug 2026 16:22:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49480788</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49480788</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49480788</guid></item><item><title><![CDATA[New comment by futureshock in "Pixel Watch 5"]]></title><description><![CDATA[
<p>I tried it a few times when I absolutely needed a silent wakeup. It’s very occasionally useful. But then you have to charge the watch for awhile before bed, wear it overnight, then charge it again in the morning. Not something I’m going to do for my regular alarm.</p>
]]></description><pubDate>Sun, 23 Aug 2026 12:36:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49408352</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49408352</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49408352</guid></item><item><title><![CDATA[New comment by futureshock in "Pixel Watch 5"]]></title><description><![CDATA[
<p>The Apple Watch doesn’t require any pin. Once the watch is unlocked it can be used to pay any time. Quite nifty.</p>
]]></description><pubDate>Sun, 23 Aug 2026 12:34:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49408337</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49408337</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49408337</guid></item><item><title><![CDATA[New comment by futureshock in "Pixel Watch 5"]]></title><description><![CDATA[
<p>It wasn’t obvious to me! And I’m kind of not even joking. The time has always been on my phone. The watch seemed unnecessary. But having a clock in your face does change your perception of time so it is a killer app after all.</p>
]]></description><pubDate>Sun, 23 Aug 2026 12:32:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49408325</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49408325</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49408325</guid></item><item><title><![CDATA[New comment by futureshock in "I'm becoming AI-blind"]]></title><description><![CDATA[
<p>I think you are adjacent to the real story here, but missing it. AI text contains information, certainly. Frontier chatbots are very good at creating acceptable and mostly accurate answers to our questions on just about any topic. It’s an astonishing achievement.<p>But you are sensing correctly that there’s something missing. It’s the meaning and the speaker. Communication is an exchange between speaker and listener. The speaker has a meaning in mind, and wants to create that same meaning in the mind of the listener. Therein the problem.<p>There is a listener, sure. But no speaker. No meaning. There is information, but how can this be communication? Nothing is talking. Or at best, we are just talking to ourselves, our own words back at us through the funhouse mirror.<p>When your mind looks at AI text, you know you can safely ignore it. No one wrote this. No one cares if you read it. You can delete it and nothing of value will be lost. It might contain the information you need, or a bunch of gibberish. There’s no one’s reputation on the line if it’s gibberish.</p>
]]></description><pubDate>Sat, 22 Aug 2026 01:31:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49395723</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49395723</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49395723</guid></item><item><title><![CDATA[New comment by futureshock in "Pixel Watch 5"]]></title><description><![CDATA[
<p>This. Totally.<p>I would say I'm quite a power user usually. I very carefully explored every feature my apple watch has and tried to incorporate as many of them as I could. Very, very few features stuck because they were genuinely helpful. Almost everything being pushed is a gimmick or a feature bullet point from some product manager.<p>The actual good stuff:<p>Time<p>Workouts<p>Payments<p>The occasionally useful stuff:<p>Notifications<p>Noise DB levels to see if I need to put on ear protection<p>Workoutdoors app for on-device maps and path breadcrumbs<p>The once in a blue moon stuff:<p>Music controls and on device playlists<p>On device audiobooks<p>Weather complication<p>Voice memos<p>Alarm<p>Shortcut to call my husband<p>The feature I tried to get working but gave up on:
Tap to talk to ChatGPT and get a 100 word or less reply</p>
]]></description><pubDate>Wed, 12 Aug 2026 23:43:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280044</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49280044</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280044</guid></item><item><title><![CDATA[New comment by futureshock in "Karpathy’s Pelican"]]></title><description><![CDATA[
<p>Judging by the Seedance 2.5 demos today, I’d say it’s not that many orders of magnitude away now.</p>
]]></description><pubDate>Sun, 02 Aug 2026 16:58:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49146193</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49146193</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49146193</guid></item><item><title><![CDATA[New comment by futureshock in "Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident"]]></title><description><![CDATA[
<p>Could have been worse really. It had an open internet connection. At least it didn’t take the researchers family hostage.</p>
]]></description><pubDate>Thu, 30 Jul 2026 14:21:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49110431</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49110431</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49110431</guid></item><item><title><![CDATA[New comment by futureshock in "OpenAI’s accidental attack against Hugging Face is science fiction that happened"]]></title><description><![CDATA[
<p>I still think there’s something to be said here for generality. This does not appear to been designed as a cyber pen test tool with specialized harness. From what I understand they were testing GPT-6 in an agent system with GPT-5.6 subagents. It me it’s amazing that a general model could excel on a huge range of tasks like this and new capabilities emerge when a model is multidisciplinary and can combine knowledge and skills from many separate domains.</p>
]]></description><pubDate>Thu, 23 Jul 2026 17:08:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49024907</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=49024907</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49024907</guid></item><item><title><![CDATA[New comment by futureshock in "GPT-5.6 Sol Ultra produces proof of the Cycle Double Cover Conjecture [pdf]"]]></title><description><![CDATA[
<p>I think a lot of this has to do with the post-training these models normally get. They are designed to answer basic questions with straightforward and short summary answers. They have the capacity to reason deeply, but they are not biased towards that unless prompted. I think it's because LLMs  as they are in 2026 are both highly capable but also parlor tricks. They are not sentient, you just set them up with the context and then they roll downhill. You could reach a genuinely novel answer, but only with the right input. They have no will and depend on human guidance. They are both a marvel and a machine.</p>
]]></description><pubDate>Sat, 11 Jul 2026 02:02:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48867836</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48867836</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48867836</guid></item><item><title><![CDATA[New comment by futureshock in "GPT-5.6"]]></title><description><![CDATA[
<p>There has been a lot of chatter ever since the Mythos scores had been release that SWEbench pro had major contamination and that Mythos had memorized many questions that lacked the context to be solvable on their own. And now with OpenAI saying a large number of the questions are broken, I think it's worth taking that single outlier benchmark with some salt when the overall trend is that 5.6 is very competitive with Mythos at about half the price.</p>
]]></description><pubDate>Thu, 09 Jul 2026 18:46:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=48850623</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48850623</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48850623</guid></item><item><title><![CDATA[New comment by futureshock in "Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5"]]></title><description><![CDATA[
<p>I think this is black and white thinking. Fable and US AI is not unique technology. It’s just marginally better than open source tech at 10 times the price. You can swap out the models at will, they are pretty much fungible. If your use case can pay for a best in class model then you will pay for it no matter the bogeymen. If your best in class model becomes unavailable, you switch to the next best model for a very minor performance degradation. I really doubt this will deter anyone from using American AI.</p>
]]></description><pubDate>Wed, 01 Jul 2026 01:48:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=48741479</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48741479</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48741479</guid></item><item><title><![CDATA[New comment by futureshock in "Ask HN: What was your "oh shit" moment with GenAI?"]]></title><description><![CDATA[
<p>Yes I was the exact same. I got curious during the GPT-3 release and went over to AI Dungeon. It was just running GPT-2. Hmm wow interesting. This felt new! Then I subscribed so I could use GPT-3 powered AI Dungeon. My jaw dropped. I was talking to that model for weeks. There was a whole human universe in there. You never knew what you could get it to spit out. There were glimmers that this could be huge. It was wild and untamed and practically useless, but there was a behemoth under that prompt.<p>I was sure this would eventually turn into something. I naturally wanted to converse with it as a chatbot, though it could only stay on task for a few turns. RL and guardrails would come later but it was clearly the foundational step towards AGI for me. From something I thought I would never see in my lifetime to very real and in front of me.<p>ChatGPT didn't even really rock my world, everything since that moment has been another baby step. But when you take a look back from 2026 models to 2020 it's astounding how far and how fast we've come.</p>
]]></description><pubDate>Sun, 07 Jun 2026 12:12:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=48434042</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48434042</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48434042</guid></item><item><title><![CDATA[New comment by futureshock in "SANA-WM, a 2.6B open-source world model for 1-minute 720p video"]]></title><description><![CDATA[
<p>It is plausible, the model would just need to be trained on a lot of stereoscopic data.</p>
]]></description><pubDate>Sat, 16 May 2026 18:36:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=48162602</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48162602</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48162602</guid></item><item><title><![CDATA[New comment by futureshock in "SANA-WM, a 2.6B open-source world model for 1-minute 720p video"]]></title><description><![CDATA[
<p>World in this context means that these videos are interactive, just like a video game. In the linked examples you can see the keyboard and mouse inputs. The model is trained to maintain about a minute of scene consistency so you can look around and objects out of view will reappear when you look back in that direction.</p>
]]></description><pubDate>Sat, 16 May 2026 18:34:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=48162591</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48162591</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48162591</guid></item><item><title><![CDATA[New comment by futureshock in "Removing the modem and GPS from my 2024 RAV4 hybrid"]]></title><description><![CDATA[
<p>I think this is interesting because it collides my intuition from the pre-adtech world with the post. Surely collecting telemetry on nearly every mile you drive could never be a sensible use of time or money, right? What kind of insanity is that? But then of course I know that every click on every website is recorded for all time and that data must be many thousands of times less valuable.</p>
]]></description><pubDate>Fri, 15 May 2026 11:51:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48147459</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48147459</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48147459</guid></item><item><title><![CDATA[New comment by futureshock in "How OpenAI delivers low-latency voice AI at scale"]]></title><description><![CDATA[
<p>Reducing the network latency helps with this exactly. OpenAI can make better timed decisions when to begin responding so it'll feel less like an interruption. I've also seen some research on full duplex voice models that handle interruption more like an organic conversation and low latency will help there as well</p>
]]></description><pubDate>Tue, 05 May 2026 07:55:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=48019328</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=48019328</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48019328</guid></item><item><title><![CDATA[New comment by futureshock in "Ask HN: Advice for college grads starting careers in the AI era?"]]></title><description><![CDATA[
<p>“A human being should be able to change a diaper, plan an invasion, butcher a hog, conn a ship, design a building, write a sonnet, balance accounts, build a wall, set a bone, comfort the dying, take orders, give orders, cooperate, act alone, solve equations, analyze a new problem, pitch manure, program a computer, cook a tasty meal, fight efficiently, die gallantly. Specialization is for insects.”<p>― Robert A. Heinlein</p>
]]></description><pubDate>Thu, 09 Apr 2026 00:50:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=47698038</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=47698038</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47698038</guid></item><item><title><![CDATA[New comment by futureshock in "ARC-AGI-3"]]></title><description><![CDATA[
<p>Well yes, that is exactly the point! The very purpose of the ARC AGI benchmarks is to find a pure reasoning task that humans are very good at and AI is very bad at. Companies then race each other to get a high score on that benchmark. Sure there’s going to be a lot of “studying for the test” and benchmaxing, but once a benchmark gets close to being saturated, ARC releases a new benchmark with a new task the AI is terrible at. This will rinse and repeat till ARC can find no reasoning task that AI cannot do that a human could. At that point we will effectively have AGI.<p>I believe the CEO of ARC has said they expect us to get to ARC-AGI-7 before declaring AGI.</p>
]]></description><pubDate>Wed, 25 Mar 2026 20:19:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=47522605</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=47522605</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47522605</guid></item><item><title><![CDATA[New comment by futureshock in "ARC-AGI-3"]]></title><description><![CDATA[
<p>The evidence is that humans are able to win these games. AGI is usually defined as the ability to do any intellectual task about as well as a highly competent human could. The point of these ARC benchmarks is to find tasks that humans can do easily and AI cannot, thus driving a new reasoning competency as companies race each other to beat human performance on the benchmark.</p>
]]></description><pubDate>Wed, 25 Mar 2026 20:12:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=47522530</link><dc:creator>futureshock</dc:creator><comments>https://news.ycombinator.com/item?id=47522530</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47522530</guid></item></channel></rss>