<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: 6thbit</title><link>https://news.ycombinator.com/user?id=6thbit</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 25 Aug 2026 03:04:49 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=6thbit" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by 6thbit in "OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)"]]></title><description><![CDATA[
<p>Same here. Thought Luna was for “moonshots” and sol for.. sunshots? While keeping Terra earthly.<p>But alas</p>
]]></description><pubDate>Mon, 24 Aug 2026 23:11:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49427084</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49427084</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49427084</guid></item><item><title><![CDATA[New comment by 6thbit in "OpenRouter is joining Stripe"]]></title><description><![CDATA[
<p>This has to be all about cashflow, right?<p>Surely stripe if anyone have learned to harness the cash flowing through their system. Hell, they could be emitting bonds on expected token consumption bills!</p>
]]></description><pubDate>Wed, 19 Aug 2026 18:17:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49365141</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49365141</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49365141</guid></item><item><title><![CDATA[New comment by 6thbit in "The price of a Costco hot dog has gone up"]]></title><description><![CDATA[
<p>Clickbaity title. The price tag didn’t change and still sits at 1.50.<p>Shrinkflation is a thing and the bun has been its victim. They could slowly shrink the sausage now until it fits the bun again.</p>
]]></description><pubDate>Sat, 15 Aug 2026 17:06:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49312278</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49312278</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49312278</guid></item><item><title><![CDATA[New comment by 6thbit in "Muse Code and Muse Spark 1.2"]]></title><description><![CDATA[
<p>So what Meta believes fair for "paying" for your data is $0.1/Mtok plus the opportunity cost of $3/Mtok in output?</p>
]]></description><pubDate>Thu, 06 Aug 2026 17:57:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49200073</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49200073</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49200073</guid></item><item><title><![CDATA[New comment by 6thbit in "Muse Code and Muse Spark 1.2"]]></title><description><![CDATA[
<p>is this their way to push their harness?<p>people interested in the discounted -contributor model would increase their harness beta testers. But it'd be a terrible strategy to capture the better paying customer base through openrouter and similar...</p>
]]></description><pubDate>Thu, 06 Aug 2026 17:52:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49200008</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49200008</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49200008</guid></item><item><title><![CDATA[New comment by 6thbit in "Muse Code and Muse Spark 1.2"]]></title><description><![CDATA[
<p>This is likely an legitimate risk, not from you exactly, but if there's unethical competitors with unlimited budgets, their training could indeed be poisoned.<p>Not sure if anyone would bother.</p>
]]></description><pubDate>Thu, 06 Aug 2026 17:44:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49199916</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49199916</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49199916</guid></item><item><title><![CDATA[New comment by 6thbit in "LLMs reward expertise"]]></title><description><![CDATA[
<p>So we could run a lighter LLM in front of humans, which translates from 'no domain knowledge' to 'domain expert' and in turn prompts over to the larger LLM.<p>Then the larger LLM gets all the right lights on, yields better outputs and we translate back into user domain.<p>I kinda thought the chain-of-thought reasoning already did this, no?</p>
]]></description><pubDate>Mon, 03 Aug 2026 23:41:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49162770</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49162770</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49162770</guid></item><item><title><![CDATA[New comment by 6thbit in "Prevent cognitive debt by manually retyping LLM-generated code"]]></title><description><![CDATA[
<p>Use it not just as an output tool but as an input as well.<p>You could try doing the high level design yourself at least. Ask for its review and iterate without asking it to do it all.<p>Once it has generated some implementation, critique it and ensure you understand its approach And you agree with it, steer it otherwise.<p>If there’s anything unclear to you say so and have it rewrite it in an easier way to understand.</p>
]]></description><pubDate>Mon, 03 Aug 2026 17:40:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49159027</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49159027</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49159027</guid></item><item><title><![CDATA[New comment by 6thbit in "Tailscale didn't stop the Hugging Face intrusion"]]></title><description><![CDATA[
<p>I've always found tailscale's json config a bit intimidating. I greatly appreciate the new UI that makes it easier to define rules, alas, both going to relevant docs straight from it and determining 'is this rule just lazy/bad/unsafe' is hard and frustrating most of the time.<p>That's where I'd like to see this sort of checkup. 
Yell at me please if i just said anyone can ssh as root from any node!</p>
]]></description><pubDate>Fri, 31 Jul 2026 20:22:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49128123</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49128123</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49128123</guid></item><item><title><![CDATA[New comment by 6thbit in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p>There was no rush for this disclosure on their side. And they publish at a point where they have not yet taken corrective actions:<p><pre><code>    > Some of the solutions here may even be simple fixes;
</code></pre>
They are still throwing ideas. Why have they not made those simple fixes yet before disclosing?</p>
]]></description><pubDate>Fri, 31 Jul 2026 00:07:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117490</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49117490</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117490</guid></item><item><title><![CDATA[New comment by 6thbit in "Investigating three real-world incidents in our cybersecurity evaluations"]]></title><description><![CDATA[
<p><p><pre><code>    > closer to a harness and operational failure than a model alignment failure. Our models were told they had no internet access and to capture the flag, while in fact being misconfigured to have internet access. 
    > This led them to believe—arguably reasonably—that the real environments they encountered were simulations.

</code></pre>
That the AI lab most typically preaching for alignment does not consider this an obvious misalignment is a clear red flag.</p>
]]></description><pubDate>Thu, 30 Jul 2026 23:59:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49117427</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49117427</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49117427</guid></item><item><title><![CDATA[New comment by 6thbit in "Codex Security"]]></title><description><![CDATA[
<p>How does it fare against its own codebase?</p>
]]></description><pubDate>Tue, 28 Jul 2026 22:41:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49090965</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49090965</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49090965</guid></item><item><title><![CDATA[New comment by 6thbit in "Claude Opus 5"]]></title><description><![CDATA[
<p>Honestly that's the simplest explanation and thus likely the correct one.</p>
]]></description><pubDate>Fri, 24 Jul 2026 19:44:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49040724</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49040724</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49040724</guid></item><item><title><![CDATA[New comment by 6thbit in "Claude Opus 5"]]></title><description><![CDATA[
<p>Anyone has an insight into how much money labs are putting into benchmarks?<p>Just Arg-AGI-3 is quoted above 20K USD and footnote says average of 5 runs (!!).
Likely just a drop in the bucket to the training budget but still..</p>
]]></description><pubDate>Fri, 24 Jul 2026 17:39:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49039135</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49039135</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49039135</guid></item><item><title><![CDATA[New comment by 6thbit in "Claude Opus 5"]]></title><description><![CDATA[
<p>"although Opus 5 shows improvements in its ability to identify software vulnerabilities, it is substantially behind Mythos 5 in its ability to exploit them."<p>"Opus 5’s safeguards match those of Claude Fable 5’s, with one change: it now permits source-code vulnerability discovery at all access levels".<p>This is probably great news, but then again, where does this leave Fable as a choice?</p>
]]></description><pubDate>Fri, 24 Jul 2026 17:35:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49039084</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49039084</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49039084</guid></item><item><title><![CDATA[New comment by 6thbit in "Claude Opus 5"]]></title><description><![CDATA[
<p>Their communication is confusing. They say "Opus 5 is not more capable overall than Fable 5", but their blog post proceeds to list how much better Opus 5 is than Fable 5 on __most__ benchmarks listed.<p>Then system card goes on to "Its AI R&D capabilities are comparable to those of Claude Mythos 5", which is supposed to be fable minus restrictions.</p>
]]></description><pubDate>Fri, 24 Jul 2026 17:33:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49039052</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49039052</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49039052</guid></item><item><title><![CDATA[New comment by 6thbit in "Are AI Labs Pelicanmaxxing?"]]></title><description><![CDATA[
<p>I think svg is a balanced test because of the level of indirection and the required 'conceptualization' of physical elements then expressed through code.</p>
]]></description><pubDate>Wed, 22 Jul 2026 22:51:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=49014543</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49014543</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49014543</guid></item><item><title><![CDATA[New comment by 6thbit in "Are AI Labs Pelicanmaxxing?"]]></title><description><![CDATA[
<p>What would be an alternative format or process with a similar effort to drawing SVGs?</p>
]]></description><pubDate>Wed, 22 Jul 2026 22:50:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49014541</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49014541</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49014541</guid></item><item><title><![CDATA[New comment by 6thbit in "Terence Tao's ChatGPT conversation about the Jacobian Conjecture counterexample"]]></title><description><![CDATA[
<p>Presumably this was Sol on xhigh, then over to Pro (as per his indication on chat)?<p>Is there any way to tell a conversation's model and thinking level?</p>
]]></description><pubDate>Wed, 22 Jul 2026 18:30:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49011286</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49011286</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49011286</guid></item><item><title><![CDATA[New comment by 6thbit in "Does creatine make you smarter?"]]></title><description><![CDATA[
<p>and that scientific evidence is new? like a new study boosted this or just resurfacing?</p>
]]></description><pubDate>Wed, 22 Jul 2026 17:31:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49010358</link><dc:creator>6thbit</dc:creator><comments>https://news.ycombinator.com/item?id=49010358</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49010358</guid></item></channel></rss>