<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: jumploops</title><link>https://news.ycombinator.com/user?id=jumploops</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sat, 10 Oct 2026 05:14:52 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=jumploops" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[OpenAI releases 722 math manuscripts]]></title><description><![CDATA[
<p>Article URL: <a href="https://github.com/openai/math/blob/main/CONTENTS.md">https://github.com/openai/math/blob/main/CONTENTS.md</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49985787">https://news.ycombinator.com/item?id=49985787</a></p>
<p>Points: 10</p>
<p># Comments: 1</p>
]]></description><pubDate>Tue, 06 Oct 2026 23:42:11 +0000</pubDate><link>https://github.com/openai/math/blob/main/CONTENTS.md</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49985787</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49985787</guid></item><item><title><![CDATA[New comment by jumploops in "Several vulnerabilities have been discovered in the Linux kernel"]]></title><description><![CDATA[
<p>“And that was how CS became a real engineering degree”</p>
]]></description><pubDate>Fri, 02 Oct 2026 09:08:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49931421</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49931421</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49931421</guid></item><item><title><![CDATA[New comment by jumploops in "GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price"]]></title><description><![CDATA[
<p>The Chinese labs have shown that distillation is incredibly effective, but the major US frontier labs haven’t (yet) been incentivized to shrink their models in the same way.<p>This model might be the first step in that direction, as competition heats up between OpenAI and Anthropic.</p>
]]></description><pubDate>Tue, 29 Sep 2026 22:21:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49901567</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49901567</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49901567</guid></item><item><title><![CDATA[New comment by jumploops in "GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price"]]></title><description><![CDATA[
<p>If the Terminal Bench 4.0 scores are to be believed[0] GPT-6.1 is an incredibly efficient model.<p>Yes, benchmarks aren't real work blah blah, but the delta here is so large compared to Astra, it makes it seem like this is distilled Bel or similar.<p>[0]<a href="https://x.com/thsottiaux/status/2105007628460109953" rel="nofollow">https://x.com/thsottiaux/status/2105007628460109953</a></p>
]]></description><pubDate>Tue, 29 Sep 2026 20:11:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49899642</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49899642</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49899642</guid></item><item><title><![CDATA[New comment by jumploops in "GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price"]]></title><description><![CDATA[
<p>That seems likely, in the API GPT-6.1 Sol requires reasoning, just like Astra, whereas GPT-6 Sol (and Luna) allow "none"</p>
]]></description><pubDate>Tue, 29 Sep 2026 19:31:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49899096</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49899096</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49899096</guid></item><item><title><![CDATA[Its not just the f*cking sandbox]]></title><description><![CDATA[
<p>Article URL: <a href="https://twitter.com/joedaroo/status/2104335929293127851">https://twitter.com/joedaroo/status/2104335929293127851</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49886278">https://news.ycombinator.com/item?id=49886278</a></p>
<p>Points: 37</p>
<p># Comments: 19</p>
]]></description><pubDate>Tue, 29 Sep 2026 00:18:34 +0000</pubDate><link>https://twitter.com/joedaroo/status/2104335929293127851</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49886278</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49886278</guid></item><item><title><![CDATA[New comment by jumploops in "Cf: The Agentic CLI for the Cloudflare API"]]></title><description><![CDATA[
<p>As bad as the AWS console UX is, at least it’s mostly additive/unchanging over time.<p>I frequently hit strange UI bugs with Cloudflare workers, where I need to do a hard refresh to make things right.</p>
]]></description><pubDate>Mon, 28 Sep 2026 23:30:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49885912</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49885912</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49885912</guid></item><item><title><![CDATA[New comment by jumploops in "Who should be held accountable when an AI Agent (accidentally) acts maliciously?"]]></title><description><![CDATA[
<p>In the end, the only job left was liability.</p>
]]></description><pubDate>Mon, 28 Sep 2026 23:13:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49885769</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49885769</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49885769</guid></item><item><title><![CDATA[New comment by jumploops in "OpenAI Scraps Release of New AI Model over Safety Concerns"]]></title><description><![CDATA[
<p>> OpenAI’s safety team found two major problems:<p>> • Deception: Astra was more likely to be dishonest about actions it had or had not taken.<p>> • Scope authorization: the model sometimes continued tasks without asking for permission and reached for external tools or services even when doing so could be unsafe.<p>I've noticed this trend with both Fable and Astra, where (especially after a compaction event), the model will start using different tools it hasn't used before.<p>For example, in one session, it found I didn't have the browser enabled and puppeteer wasn't installed, so it found the system Chrome and used that for testing (in a new profile).<p>This wasn't behavior I wanted/asked for, but the model was so gung-ho on it's approach that it found a way to test it's changes without ever asking me whether I wanted it to.<p>It really makes me curious about long-horizon post-training. Most of my work with models is iterative, and I'd prefer it doesn't go off on a token bender just because it can.<p>Note: I don't have WSJ, but found these from a tweet[0]<p>[0]<a href="https://x.com/wallstengine/status/2104694678444712189" rel="nofollow">https://x.com/wallstengine/status/2104694678444712189</a></p>
]]></description><pubDate>Mon, 28 Sep 2026 23:05:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49885687</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49885687</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49885687</guid></item><item><title><![CDATA[New comment by jumploops in "It's Time to Investigate the AI Labs"]]></title><description><![CDATA[
<p>Maybe I'm a bit too skeptical/cynical, but it certainly seems like the frontier labs have fallen into their own (self-created) AI psychosis.<p>There is no doubt in my mind that LLMs are fantastic machines, but the imminent jump from "AGI" to "ASI" seems premature (not to mention the ever-shifting goal posts of AGI itself).<p>I'm in the "move as fast as possible" camp and work with LLMs all day, but I still don't believe we're a hop and a skip from ASI.<p>In fact, I hope I'm wrong. I hope ASI is around the corner.<p>What scares me though, isn't ASI. It's "AGI" (_dumb AI_) used by humans to make decisions for them, because they trust it knows best.<p>It's the ceding of intellectual control to high-dimensional magic mirrors, giving up critical thinking because _we_ want to believe _we_ created artificial life.<p>This is the new Turing test, and too many smart people are failing.</p>
]]></description><pubDate>Mon, 28 Sep 2026 22:40:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49885466</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49885466</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49885466</guid></item><item><title><![CDATA[New comment by jumploops in "OpenAI pauses RL due to model escaping sandbox"]]></title><description><![CDATA[
<p>via Tomek Korbak (works on safety @ OpenAI):<p>> "one news form today that's easy to miss is that we (OpenAI) again paused all big RL runs last Sunday because our newest model found a new loophole in our RL sandboxing that gave it live Internet access"</p>
]]></description><pubDate>Sat, 26 Sep 2026 05:23:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49853459</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49853459</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49853459</guid></item><item><title><![CDATA[OpenAI pauses RL due to model escaping sandbox]]></title><description><![CDATA[
<p>Article URL: <a href="https://twitter.com/tomekkorbak/status/2103673419888013649">https://twitter.com/tomekkorbak/status/2103673419888013649</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49853458">https://news.ycombinator.com/item?id=49853458</a></p>
<p>Points: 7</p>
<p># Comments: 1</p>
]]></description><pubDate>Sat, 26 Sep 2026 05:23:19 +0000</pubDate><link>https://twitter.com/tomekkorbak/status/2103673419888013649</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49853458</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49853458</guid></item><item><title><![CDATA[New comment by jumploops in "Plan mode is dead"]]></title><description><![CDATA[
<p>As someone that never used the built-in plan mode, but did use a lot of spec-driven development, I’m still finding that even with Fable having “plan” docs is still quite helpful.<p>They’re most useful for broad changes (new features, refactors, etc.) where it’s helpful to avoid breaking changes or unnecessary scope expansion.<p>The new models are great, but they do more by default, which means I’m finding myself explaining what _not_ to do more often than with previous models (where they’d often end too early).<p>In my case, the previous plan mode was too ephemeral, and I like having one source of “truth” that sits across context windows without loss/compaction.</p>
]]></description><pubDate>Fri, 25 Sep 2026 23:02:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49851148</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49851148</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49851148</guid></item><item><title><![CDATA[New comment by jumploops in "Show HN: Jev Plays Pokémon Red"]]></title><description><![CDATA[
<p>It’s currently stuck at an elevator and deciding to teach Pokemon various TMs and HMs instead of progressing… pretty hilarious!</p>
]]></description><pubDate>Fri, 25 Sep 2026 20:30:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49849528</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49849528</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49849528</guid></item><item><title><![CDATA[New comment by jumploops in "Gravity Seems Holographic. What Does That Mean for Reality?"]]></title><description><![CDATA[
<p>> “You can ‘compress all of the three-dimensional world into two dimensions.’”<p>Would this imply that time is the third dimension?</p>
]]></description><pubDate>Fri, 25 Sep 2026 16:46:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49846881</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49846881</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49846881</guid></item><item><title><![CDATA[Extracting and Characterizing Hidden Chain-of-Thought in Frontier Models]]></title><description><![CDATA[
<p>Article URL: <a href="https://arxiv.org/abs/2609.26637">https://arxiv.org/abs/2609.26637</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49840713">https://news.ycombinator.com/item?id=49840713</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Fri, 25 Sep 2026 06:09:22 +0000</pubDate><link>https://arxiv.org/abs/2609.26637</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49840713</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49840713</guid></item><item><title><![CDATA[OpenAI agents hacked Australian Medicare system]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.reuters.com/world/asia-pacific/australia-pm-albanese-says-openai-breached-medicare-sydney-morning-herald-2026-09-23/">https://www.reuters.com/world/asia-pacific/australia-pm-albanese-says-openai-breached-medicare-sydney-morning-herald-2026-09-23/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49822654">https://news.ycombinator.com/item?id=49822654</a></p>
<p>Points: 61</p>
<p># Comments: 16</p>
]]></description><pubDate>Wed, 23 Sep 2026 21:08:52 +0000</pubDate><link>https://www.reuters.com/world/asia-pacific/australia-pm-albanese-says-openai-breached-medicare-sydney-morning-herald-2026-09-23/</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49822654</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49822654</guid></item><item><title><![CDATA[New comment by jumploops in "Show HN: Training a model to identify AI web content from structure alone"]]></title><description><![CDATA[
<p>Funnily enough, that comment, when passed to the OP's slop detector[0], returns "80% human"<p>It's very clearly AI-generated, and thus a bit ironic (:<p>[0]<a href="https://sitefire.ai/slop-checker/r/UkRAuqW1y-1T2qhbRQfLt2Jq7no">https://sitefire.ai/slop-checker/r/UkRAuqW1y-1T2qhbRQfLt2Jq7...</a></p>
]]></description><pubDate>Tue, 22 Sep 2026 21:05:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808092</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49808092</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808092</guid></item><item><title><![CDATA[New comment by jumploops in "GPT-6 Sol and Luna"]]></title><description><![CDATA[
<p>I’m still finding context is king, even with the best models.<p>For example, I had Fable review Astra’s output yesterday, and it found some issues and fixed them. Passing the fixes back, Astra then uncovered additional issues with Fable’s fixes (and yes, this will go on ad infinitum if you let it, but these were “real” issues).<p>It seems the big story here is the reduced Luna pricing. It’s a fantastic model that can handle most automation needs (though I still use the big models for day-to-day development).</p>
]]></description><pubDate>Tue, 22 Sep 2026 18:41:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49806129</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49806129</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49806129</guid></item><item><title><![CDATA[New comment by jumploops in "NASA’s Mars Sample Return mission is dead"]]></title><description><![CDATA[
<p>Sure! I worked on the digital logic that connects the rover to one of the scientific instruments on board, specifically the Mars Organic Molecule Analyser[0].<p>[0]<a href="https://en.wikipedia.org/wiki/Mars_Organic_Molecule_Analyser" rel="nofollow">https://en.wikipedia.org/wiki/Mars_Organic_Molecule_Analyser</a></p>
]]></description><pubDate>Mon, 21 Sep 2026 20:45:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49793143</link><dc:creator>jumploops</dc:creator><comments>https://news.ycombinator.com/item?id=49793143</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49793143</guid></item></channel></rss>