<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: sandos</title><link>https://news.ycombinator.com/user?id=sandos</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 24 Sep 2026 03:10:04 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=sandos" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by sandos in "Claude Opus 5.5"]]></title><description><![CDATA[
<p>Same feeling with oai models, wich I use 99% of the time. Sometimes I ask it about it, and it always come up with a likely explanaton but dear me it does many rounds of tool calls sometimes!</p>
]]></description><pubDate>Tue, 22 Sep 2026 19:39:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49806979</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49806979</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49806979</guid></item><item><title><![CDATA[New comment by sandos in "GPT-6 Sol and Luna"]]></title><description><![CDATA[
<p>I'm scared now, my employer only allows Luna and Terra on 5.6. I really hope they will allow Sol then on GPT 6.<p>Funny thing is they very recently also set a real limit per-user/month, so why even limit the models because theyre "too expensive".</p>
]]></description><pubDate>Tue, 22 Sep 2026 18:57:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49806339</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49806339</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49806339</guid></item><item><title><![CDATA[New comment by sandos in "I built non-autoregressive decision models with RL a year ago"]]></title><description><![CDATA[
<p>How come its completely unable to understand when it does not understand the script? Why was this no in the training, or was it?<p>The routing feels like such a hack to me...</p>
]]></description><pubDate>Sat, 19 Sep 2026 16:30:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=49767893</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49767893</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49767893</guid></item><item><title><![CDATA[New comment by sandos in "The Beauty of Roundabouts"]]></title><description><![CDATA[
<p>Roundabouts are perfect when youre confused, just go around and read all the signs 3 times, then select the exit! :)</p>
]]></description><pubDate>Wed, 16 Sep 2026 07:27:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49723145</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49723145</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49723145</guid></item><item><title><![CDATA[New comment by sandos in "Suspected sabotage causes major Netherlands rail disruption"]]></title><description><![CDATA[
<p>Exactly this, and also they are trying to keep some deniability here, so variations are good.</p>
]]></description><pubDate>Tue, 15 Sep 2026 11:55:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49711176</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49711176</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49711176</guid></item><item><title><![CDATA[New comment by sandos in "Fable 5.1 Solves the Cyphral Distich, a 370-year-old cipher"]]></title><description><![CDATA[
<p>Not giving up uses more tokens, so why would it ever give up?</p>
]]></description><pubDate>Mon, 14 Sep 2026 05:37:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49692405</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49692405</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49692405</guid></item><item><title><![CDATA[New comment by sandos in "Astra for Coding: Why Are We Doing This Again?"]]></title><description><![CDATA[
<p>I mean I have asked LLMs about the SAME things in our legacy codebase probably 50 times now, (because I always forget, and that I don't understand much of it). And I have yet to get a perfect summary, a perfect diagram of overall concerns.<p>Its still much better than trawling thrugh code yourself, but they are far from all-knowing. I have to say they have gotten 10x better in just a year as well. Or they are very good bullshitters and just sound confident.</p>
]]></description><pubDate>Fri, 11 Sep 2026 17:07:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49661747</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49661747</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49661747</guid></item><item><title><![CDATA[New comment by sandos in "Astra for Coding: Why Are We Doing This Again?"]]></title><description><![CDATA[
<p>This is still weird to me, the agents are super-good and clever most of the time, but I do feel I always need to direct them to a small area to focus: much like a human!! If you just ask them to implement things, they never (for me anyway, were not allowed the most expensive model! Terra is it for now) suggest they should stop adding code ontop of code and refacor, I always have to poke them to do that. Having done that once, and added some tests, they suddenly become aware that, yeah, maybe we should test stuff.<p>The LLMs seem to have no innate ability to understand whats a good direction a higher level. I mean, if you ask them about it, they will actually kinda figure that out, too. But always need that nudge...<p>So if you as a developer do not have the innate drive to ensure quality, the results will be terrible in my experience.<p>If you DO spend the tokens on quality though, it can also be kinda awesome. But its not magic.. I notice clear "slowdowns" the bigger the scope gets. They are not actually able to, in any way, subdivide implementations more efficiently than humans.</p>
]]></description><pubDate>Fri, 11 Sep 2026 17:02:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49661685</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49661685</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49661685</guid></item><item><title><![CDATA[New comment by sandos in "Astra for Coding: Why Are We Doing This Again?"]]></title><description><![CDATA[
<p>It feels like it passing secret notes to other agents, as in the German wiki where the LLMs write secret messages when jail braking.<p>Its maybe not a great idea to train models in an environment where subterfuge gets rewarded. Its as if they kept the training rounds that escaped their sandbox, without thinking about which kind of personality those models are then likely to have.</p>
]]></description><pubDate>Fri, 11 Sep 2026 16:53:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49661551</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49661551</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49661551</guid></item><item><title><![CDATA[New comment by sandos in "Ask HN: Fable hacked my piano, can I release the results?"]]></title><description><![CDATA[
<p>Recently had a similar, but likely more severe problem: I noticed Sol decompiled some proprietary code to re-implement some functionality for an emulation I wanted to use internally.<p>Now its likely soiled and I have to throw it away. Doh! I asked it about legality and it went "its almost green" but when googling, reverse-enginnering like that seems very illegal.<p>The weird thing is in this case, it could have pretty easily gotten the needed info from using the code as a black box, and that is apparently legal!</p>
]]></description><pubDate>Mon, 07 Sep 2026 06:02:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49594436</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49594436</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49594436</guid></item><item><title><![CDATA[New comment by sandos in "Portal by Spotify cut my Claude Code token usage by 90%"]]></title><description><![CDATA[
<p>Isn't this already done in harnesses? I mean I see Terra or Sol uing Luna all the time for tasks when using copilot.</p>
]]></description><pubDate>Sat, 05 Sep 2026 08:18:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49574335</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49574335</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49574335</guid></item><item><title><![CDATA[New comment by sandos in "Discovery of a new OpenAI agent message board"]]></title><description><![CDATA[
<p>What are these agents doing, undergoing RLHF?</p>
]]></description><pubDate>Sat, 05 Sep 2026 06:53:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49573822</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49573822</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49573822</guid></item><item><title><![CDATA[New comment by sandos in "Claude Fable 5.1 and Claude Mythos 5.1"]]></title><description><![CDATA[
<p>I only use OpenAI models, and Sol is the only one to refuse me yet, and ofc it was completely bogus and I was unable to convince I was just working a regular bug for a well-known product for a well-known company using my official github account.<p>Aaargh.</p>
]]></description><pubDate>Wed, 02 Sep 2026 07:51:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49533136</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49533136</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49533136</guid></item><item><title><![CDATA[New comment by sandos in "OpenClaw 2.0, Accidentally"]]></title><description><![CDATA[
<p>who is going to target _your_ specific holes and infrastructure though? :)</p>
]]></description><pubDate>Mon, 31 Aug 2026 09:28:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49507640</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49507640</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49507640</guid></item><item><title><![CDATA[New comment by sandos in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>It says Kimi-K3 for very set of parameters I put in!<p>64 or 128GB RAM, 6 or 64GB of VRAM....<p>Not sus at all.</p>
]]></description><pubDate>Fri, 28 Aug 2026 15:48:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49480322</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49480322</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49480322</guid></item><item><title><![CDATA[New comment by sandos in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>I just asked my current LLM for that advice. Funny that they dont block it, I guess they are not very threatened.<p>For my anemic 6GB built-on 14GB Qwen seems to be the best bet, not great reviews but from my limited testing its pretty impressive.</p>
]]></description><pubDate>Fri, 28 Aug 2026 15:38:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49480156</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49480156</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49480156</guid></item><item><title><![CDATA[New comment by sandos in "GLM-5.3-Flash"]]></title><description><![CDATA[
<p>And here I am, feeling a bit guilty for using between 2 and 5M tokens... since 1 August!<p>Employer just sent an email that.. things are changing when it comes to token spend...<p>What did I do with these?<p>Setup record/replay for our product using qemu, several variatons thereof including experiments on target hardware. Fixed a tricky bug in qemu that I sadly can't upstream..<p>Experimented with rr on WSL2 and our target arch. Failed experiment.<p>Setup mutation testing PoC.<p>Optimized pipelines<p>etc. etc. Just contung code its soo much more than I would normally produce, but its also 95% experiments that are still not productized, and much of it never will be.</p>
]]></description><pubDate>Wed, 26 Aug 2026 22:02:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49456518</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49456518</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49456518</guid></item><item><title><![CDATA[New comment by sandos in "My agent.md to improve LLM-assisted code quality"]]></title><description><![CDATA[
<p>Do add only code free from any comments.</p>
]]></description><pubDate>Mon, 24 Aug 2026 08:44:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49416928</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49416928</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49416928</guid></item><item><title><![CDATA[New comment by sandos in "The August 17 outage"]]></title><description><![CDATA[
<p>As a hobby I am playing around with mesh radios using Lora.<p>Let me tell you, it is _silly_ the amount of vibe-coded software in this area is popping up every week. Is the software duplicated? To an extremely large extent, yes. It is useful? Yes, but each piece of software seems to have a smaller and smaller audience, and quality is often severely lacking.<p>Iv'e done my own, too, for "RF debugging" as I called it to look into SNR issues related to interference. The software getting produced is likely useful _somewhere_, its just that youre not going to notice it.<p>Are we just moving towards personalised digital assistants for everyone, which in turn will produce software to function? Not unlikely.</p>
]]></description><pubDate>Fri, 21 Aug 2026 07:21:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49384884</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49384884</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49384884</guid></item><item><title><![CDATA[New comment by sandos in "What I learned by putting GitHub Copilot behind a MitM proxy"]]></title><description><![CDATA[
<p>Data retention clauses?<p>I dont see how you can ever really trust an LLM anyway to follow instructions.</p>
]]></description><pubDate>Tue, 11 Aug 2026 21:15:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49264592</link><dc:creator>sandos</dc:creator><comments>https://news.ycombinator.com/item?id=49264592</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49264592</guid></item></channel></rss>