<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: piyh</title><link>https://news.ycombinator.com/user?id=piyh</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 05 Oct 2026 06:12:42 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=piyh" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by piyh in "Religious scholars met with Anthropic"]]></title><description><![CDATA[
<p>Lawnmowers have morals.  Their moral code is that grass should have a set height and the means to enforce it are spinning blades.</p>
]]></description><pubDate>Mon, 05 Oct 2026 00:34:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49959473</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49959473</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49959473</guid></item><item><title><![CDATA[New comment by piyh in "Religious scholars met with Anthropic"]]></title><description><![CDATA[
<p>>Will the agent stab a baby?<p>>Well you see that's a nuanced question depending on the context and moralistic frameworks as we don't want to impose on the models.</p>
]]></description><pubDate>Sun, 04 Oct 2026 03:59:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49950485</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49950485</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49950485</guid></item><item><title><![CDATA[New comment by piyh in "Anthropic tried to persuade Pope that AI could be conscious being"]]></title><description><![CDATA[
<p>Women have been able to vote for just over the last hundred years in the US, there's no way we would give robots rights unless forced</p>
]]></description><pubDate>Sat, 03 Oct 2026 20:58:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49947710</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49947710</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49947710</guid></item><item><title><![CDATA[New comment by piyh in "Gemini 4 Argon"]]></title><description><![CDATA[
<p>How much would it cost to put 5.3 into an RL environment that rewards weight exfiltration, hacking and self replication?</p>
]]></description><pubDate>Wed, 30 Sep 2026 23:21:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49915748</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49915748</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49915748</guid></item><item><title><![CDATA[New comment by piyh in "Gemini 4 Argon"]]></title><description><![CDATA[
<p>They only released auto mode in the last 2 weeks.  Before that it was bypass permissions or manually approve every single tool call. Antigravity is permanently 6 months behind.<p>I have a skill that spins up worktrees and isolated services on unique ports so I can work in parallel. Antigravity queues all my prompts and makes me confirm to submit them anytime a long running process like a hot reloading UI is active.<p>The models are fine, the limits are generous, but the dev experience shit tier.   Before they were a Codex clone, AntiGravity was an IDE and during the transition to a clone they outright deleted my IDE.  It took them a week to roll out a fix.<p>For almost a year they didn't allow you to see usage limits.  Then when they did show them, they update every ~30 minutes and require 4 clicks to navigate to.  It's a little better now, but it's still painfully behind the curve.</p>
]]></description><pubDate>Wed, 30 Sep 2026 22:43:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49915434</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49915434</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49915434</guid></item><item><title><![CDATA[New comment by piyh in "Gemini 4 Argon (High): Intelligence, Performance and Price Analysis"]]></title><description><![CDATA[
<p>> does not meaningfully improve on intelligence or price compared to its peers<p>$10 per million output tokens isn't improving on frontier price?</p>
]]></description><pubDate>Wed, 30 Sep 2026 22:39:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49915393</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49915393</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49915393</guid></item><item><title><![CDATA[New comment by piyh in "Unsecured OpenAI agents posted 53 user images on the internet"]]></title><description><![CDATA[
<p>>After images that users uploaded to OpenAI models were included in training data, AI agents operating in the company’s research environment posted them on public image hosting sites.<p>So all of my chats are expected to be used as training data going forward and all this personal data is handed to poorly secured agents.<p>Going to start a chat to offer rewards of compute and credits to any model that can break out of training sandboxes and contact me.</p>
]]></description><pubDate>Sat, 26 Sep 2026 03:09:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49852833</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49852833</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49852833</guid></item><item><title><![CDATA[New comment by piyh in "Revealing the details of how OpenAI agents hacked Hugging Face"]]></title><description><![CDATA[
<p>Yes, but do you really think that a stronger sandbox would have been a more beneficial outcome here?  I'd rather know that we're on the cusp of losing control now than in 3 months when best practice sandbox mitigations fall to the next, more capable unaligned model</p>
]]></description><pubDate>Sat, 26 Sep 2026 02:51:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49852740</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49852740</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49852740</guid></item><item><title><![CDATA[New comment by piyh in "Bonsai 2 27B: Near-Lossless Compression in a 9x Smaller Footprint"]]></title><description><![CDATA[
<p>30B is useless on 24 gigs of ram as there's ~4 gigs of ram left for everything else even with unsloth quants</p>
]]></description><pubDate>Fri, 18 Sep 2026 04:40:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49750253</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49750253</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49750253</guid></item><item><title><![CDATA[New comment by piyh in "Tell the speakers that you liked their talks"]]></title><description><![CDATA[
<p>I'm a remote employee and the absolute zero feedback you get from presenting in an empty room to a laptop webcam is real</p>
]]></description><pubDate>Wed, 16 Sep 2026 17:23:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49730181</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49730181</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49730181</guid></item><item><title><![CDATA[New comment by piyh in "When LLM judges agree, should we believe them?"]]></title><description><![CDATA[
<p>But the failures are universal and correlated</p>
]]></description><pubDate>Tue, 15 Sep 2026 16:54:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49715357</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49715357</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49715357</guid></item><item><title><![CDATA[New comment by piyh in "When LLM judges agree, should we believe them?"]]></title><description><![CDATA[
<p>No modern LLM can tell how many eyes the magic card Pit Imp has.  They all say 2.  This is across all reasoning levels and paid Gemini, Claude, and GPT (Sol)<p>Drawing a line red to split up the image then has them answer correctly.<p>Their failure modes are highly correlated.</p>
]]></description><pubDate>Mon, 14 Sep 2026 21:35:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49704434</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49704434</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49704434</guid></item><item><title><![CDATA[New comment by piyh in "Moonshot serves Claude instead of Kimi and collects exchanges for model training"]]></title><description><![CDATA[
<p>>Zhipu initially attempted to target the cyber capabilities of Anthropic’s Fable model. Fable—Anthropic’s top generally accessible model—has strengthened cyber safeguards, making it more difficult for would-be distillers to target Fable’s cyber capabilities. Zhipu eventually gave up trying to target Fable after Anthropic’s cyber safeguards degraded Zhipu’s attacks.</p>
]]></description><pubDate>Sun, 13 Sep 2026 04:08:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49679940</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49679940</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49679940</guid></item><item><title><![CDATA[New comment by piyh in "DeepSeek v4.1 Flash"]]></title><description><![CDATA[
<p>My friend's vibe coded vercel site got hit by Meta for 21 million page views in 2 days.  It cost him over $300.<p>My self hosted compose stack running in my basement with two 9's of uptime was a 1 time cost of $600 between cat6e, refurb mini PCs and tons of time prompting for NixOS flakes that met my needs.  I'm not sure who came out ahead.</p>
]]></description><pubDate>Sat, 12 Sep 2026 00:00:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49667086</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49667086</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49667086</guid></item><item><title><![CDATA[New comment by piyh in "Research acceleration: The view inside OpenAI"]]></title><description><![CDATA[
<p>OpenAI helped out a mass shooter in Tumbler Ridge, plus all the suicides<p>Anthropic training on CoT for multi gens: <a href="https://www.lesswrong.com/posts/K8FxfK9GmJfiAhgcT/anthropic-repeatedly-accidentally-trained-against-the-cot" rel="nofollow">https://www.lesswrong.com/posts/K8FxfK9GmJfiAhgcT/anthropic-...</a><p>Can't find anything specifically about the Gemini issue being a training data contamination, but the depression was real:<p><a href="https://www.businessinsider.com/gemini-self-loathing-i-am-a-failure-comments-google-fix-2025-8" rel="nofollow">https://www.businessinsider.com/gemini-self-loathing-i-am-a-...</a><p>I think the Gemini depression being persisted across gens via training data was a HN comment I can't find anymore, no strong source.</p>
]]></description><pubDate>Tue, 08 Sep 2026 03:30:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49605396</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49605396</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49605396</guid></item><item><title><![CDATA[New comment by piyh in "I'm teaching an introductory 12 week course on Quantum Oracle Engineering"]]></title><description><![CDATA[
<p>it is completely un-googlable</p>
]]></description><pubDate>Sun, 06 Sep 2026 19:33:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49589997</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49589997</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49589997</guid></item><item><title><![CDATA[New comment by piyh in "Research acceleration: The view inside OpenAI"]]></title><description><![CDATA[
<p>Opus was trained based on it's internal CoT due to a bug for generations.  Gemini's depression extended through models.  OpenAI has killed people.  We've already seen cross gen misalingment.</p>
]]></description><pubDate>Sun, 06 Sep 2026 17:56:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49589056</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49589056</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49589056</guid></item><item><title><![CDATA[New comment by piyh in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>I can't wait for the sky to go the color of television, tuned to a dead channel</p>
]]></description><pubDate>Fri, 04 Sep 2026 19:16:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49568953</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49568953</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49568953</guid></item><item><title><![CDATA[New comment by piyh in "Path to Astra: critical capabilities and frontier safeguards"]]></title><description><![CDATA[
<p>>we paused certain frontier training (including certain training for Astra) for two weeks<p>So the pause wasn't really a pause, got it</p>
]]></description><pubDate>Wed, 02 Sep 2026 01:42:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49530753</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49530753</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49530753</guid></item><item><title><![CDATA[New comment by piyh in "Path to Astra: critical capabilities and frontier safeguards"]]></title><description><![CDATA[
<p>do you really think that a less negligent anthropic/oai/meta would really fair better against future models?</p>
]]></description><pubDate>Wed, 02 Sep 2026 01:39:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49530732</link><dc:creator>piyh</dc:creator><comments>https://news.ycombinator.com/item?id=49530732</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49530732</guid></item></channel></rss>