<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: saberience</title><link>https://news.ycombinator.com/user?id=saberience</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 04 Aug 2026 12:23:12 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=saberience" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by saberience in "Try sending "see the below –" to Opus 5"]]></title><description><![CDATA[
<p>One of my responses:<p>=0.0000
t=0.0001 [A] sensor_read ok
t=0.0002 [B] sensor_read ok
t=0.0005 [A] compute begin
t=0.0009 [B] compute begin
t=0.0011 [A] compute end
t=0.0012 [B] compute end
t=0.0013 [A] write buffer[0]=1
t=0.0014 [B] write buffer[0]=2
t=0.0015 [A] read buffer[0] -> 2
t=0.0016 [B] read buffer[0] -> 2
t=0.0017 [A] ERROR checksum mismatch expected 1 got 2<p>—<p>what do you think</p>
]]></description><pubDate>Fri, 31 Jul 2026 08:30:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49120501</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49120501</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49120501</guid></item><item><title><![CDATA[New comment by saberience in "Try sending "see the below –" to Opus 5"]]></title><description><![CDATA[
<p>Has to be Opus 5 and has to be the right form of hyphen symbol, I had to play around for a while and it works 100% of the time now.</p>
]]></description><pubDate>Fri, 31 Jul 2026 08:29:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49120491</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49120491</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49120491</guid></item><item><title><![CDATA[New comment by saberience in "Show HN: Brolly, a plain-text weather forecast site"]]></title><description><![CDATA[
<p>It gets my location totally wrong</p>
]]></description><pubDate>Sun, 26 Jul 2026 09:58:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49056423</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49056423</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49056423</guid></item><item><title><![CDATA[New comment by saberience in "ARC-AGI Leaderboard"]]></title><description><![CDATA[
<p>ARC-AGI is a terrible benchmark for testing LLMs because LLMs are not made, trained, or tuned for playing games.<p>They are trained on text to respond well to text based questions and do tasks involving modifying text files.<p>They are not designed for playing games, looking at games, or visual puzzles. Also translating games into text input for the LLM skews the test completely.<p>Imagine trying to get a human to solve visual puzzle but they can’t look at the puzzle but it has to be explained to them in textual format, we would be terrible at it.<p>But yet we persist in wasting time on this benchmark. It doesn’t mean anything.</p>
]]></description><pubDate>Sat, 25 Jul 2026 16:06:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49048788</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49048788</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49048788</guid></item><item><title><![CDATA[New comment by saberience in "Show HN: Echo – Fable-level results at 1/3 the cost using open-weight models"]]></title><description><![CDATA[
<p>Sounds very scam-like.<p>Literally promising frontier-equal results but at 1/3 price, i.e. cheaper than Kimi K3?<p>Doesn't offer any real benchmarks or explanation of how this magic trick is accomplished. No credible team or notable scientists behind it...</p>
]]></description><pubDate>Fri, 24 Jul 2026 14:37:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49036362</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49036362</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49036362</guid></item><item><title><![CDATA[New comment by saberience in "LLMs Are Still Toxic, Stuck in the Past, and Bad at Math"]]></title><description><![CDATA[
<p>It's a standard trope that anti-ai folk love to draw upon.<p>Hey look! I managed to get the AI to fail at some basic thing (after trying 1000s of times), it means AI is failing!!!!!<p>It's a total red herring and mispresents the current state of the models completely. It's like meeting a child prodigy in maths and then saying he's actually dumb because can't drive a car yet.</p>
]]></description><pubDate>Fri, 24 Jul 2026 13:21:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49035160</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49035160</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49035160</guid></item><item><title><![CDATA[New comment by saberience in "LLMs Are Still Toxic, Stuck in the Past, and Bad at Math"]]></title><description><![CDATA[
<p>Is there a term for bloggers who try to get attention by participating in this sort of dated, inaccurate "AI doomerism"?<p>The idea that AIs are "bad at math" because if you send it 100s of basic arithmetic questions, it can get one wrong, is frankly laughable.<p>It's like me saying Terence Tao is bad at math because I sent him 100 long division problems and he made a mistake in one. Yes, we know models think in tokens, and if you give it a bunch of math problems (and its not using tools like Python to deterministically work on the problems) then of course you cannot guarantee accuracy.<p>But no one thought tool-less LLM calls were a solution for arithmetic in the first place. So the author is constructing a great big straw-man and attacking it vigorously.<p>The reality is this, the frontier models are as good as (OR BETTER THAN) the leading mathematicians in the world right now. Leading mathemeticians (like Terence Tao) are using Fable and GPT5.6 as partners in doing research.<p>As for being stuck in past? Again, weird Anti AI/AI Doomerism because models have fixed weights. So what? They can use tools (and do so very well) if they need up to date data and information.<p>Again it's like the author wants to paint a picture, based on their own biased beliefs and chooses to represent the current state of AI in an entirely inaccurate fashion.</p>
]]></description><pubDate>Fri, 24 Jul 2026 13:18:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49035124</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49035124</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49035124</guid></item><item><title><![CDATA[New comment by saberience in "Claude Cookbook"]]></title><description><![CDATA[
<p>I mean, the point stands.<p>The models are beyond expert level in many areas at this point.<p>Do you really believe that adding extra junk to your prompt is going to make the model write code better than it does already?<p>Again, imagine going to Terrence Tao and "prompting" him to get better at Maths, do you think you can do it? What prompt would you give to him to make him produce better maths. Unless you're already a world-leading Mathematician I think you would find it hard.</p>
]]></description><pubDate>Fri, 24 Jul 2026 12:56:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49034903</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49034903</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49034903</guid></item><item><title><![CDATA[New comment by saberience in "Claude Cookbook"]]></title><description><![CDATA[
<p>It's not really a skill. The models are at this point smarter than you are, so the idea that you can prompt them "better" is laughable really when discussing frontier models.<p>It's like imagining you could "prompt" Richard Feynman to be smarter at Physics.<p>That is, for 99.9% of engineers, if you want the model to do a code review of your project, the best solution is to just ask Fable, "Hey Fable, do a code review of this project." Throwing in extra text like "think like a senior engineer", "ensure you focus on DRY principles, KISS, self documenting code, etc", doesn't make a difference.<p>These sorts of tricks used to work with dumber models, but now, like I said before, it's like thinking you can prompt Linus Torvalds into writing better C than he already can do.</p>
]]></description><pubDate>Fri, 24 Jul 2026 11:23:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49033992</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49033992</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49033992</guid></item><item><title><![CDATA[New comment by saberience in "Alphabet's cash burn raises alarm for Big Tech as AI spending climbs"]]></title><description><![CDATA[
<p>Since when is investing in infrastructure burning money?<p>If there is a huge demand for shipping goods internationally, investing in ships and planes isn't burning money.<p>There is massive demand for compute in the world right now, Google is investing in that area. That's a good thing.</p>
]]></description><pubDate>Thu, 23 Jul 2026 13:45:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49021511</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49021511</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49021511</guid></item><item><title><![CDATA[New comment by saberience in "Alphabet's cash burn raises alarm for Big Tech as AI spending climbs"]]></title><description><![CDATA[
<p>What problem? What alarms?<p>I see everyone around me doing way more work, of way more depth, than they ever did before using AI models. I see my company and friends of mine all paying large sums of money to Anthropic, Google, OpenAI to use AI models, and do more work than we did before.<p>So Google is investing in infrastructure which is HIGHLY in demand, there is much more demand than supply, and then they are making money from this infrastructure...<p>That's a good thing for Google, and as an investor in Google, I am glad they are making these investments.</p>
]]></description><pubDate>Thu, 23 Jul 2026 13:44:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49021487</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49021487</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49021487</guid></item><item><title><![CDATA[New comment by saberience in "Intel Starts Shipping High-NA EUV Silicon"]]></title><description><![CDATA[
<p>Bingo!<p>Yes the OP is completely biased.</p>
]]></description><pubDate>Wed, 22 Jul 2026 13:30:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49006508</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49006508</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49006508</guid></item><item><title><![CDATA[New comment by saberience in "Intel Starts Shipping High-NA EUV Silicon"]]></title><description><![CDATA[
<p>Most of those technologies being European, including the most difficult to make components in the system.<p>Why isn't there an American company doing this? Why isn't an American company making the lasers, mirrors, metrology, vacuum handling tech, etc?</p>
]]></description><pubDate>Wed, 22 Jul 2026 13:30:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49006492</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49006492</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49006492</guid></item><item><title><![CDATA[New comment by saberience in "Intel Starts Shipping High-NA EUV Silicon"]]></title><description><![CDATA[
<p>The most absurdly biased and inaccurate take on ASML I've ever seen, congrats.<p>If the machines are so easy to build and integrate, why isn't it happening in America?</p>
]]></description><pubDate>Wed, 22 Jul 2026 13:28:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49006470</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=49006470</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49006470</guid></item><item><title><![CDATA[New comment by saberience in "Mythologizing AI makes it more likely that we’ll fail to operate it well (2023)"]]></title><description><![CDATA[
<p>Humans make all these mistakes too, in fact, humans make more mistakes than AI does in coding at the moment.</p>
]]></description><pubDate>Tue, 21 Jul 2026 14:04:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=48992532</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=48992532</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48992532</guid></item><item><title><![CDATA[New comment by saberience in "Show HN: I've built a words game based on binary search"]]></title><description><![CDATA[
<p>Had a similar issue, several of my words weren't accepted.</p>
]]></description><pubDate>Thu, 16 Jul 2026 14:05:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=48934795</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=48934795</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48934795</guid></item><item><title><![CDATA[New comment by saberience in "Inkling: Our Open-Weights Model"]]></title><description><![CDATA[
<p>I prefer 5.6 Sol to Fable personally and I've used both extensively.</p>
]]></description><pubDate>Thu, 16 Jul 2026 11:10:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=48932921</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=48932921</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48932921</guid></item><item><title><![CDATA[New comment by saberience in "Show HN: I RL-trained an agent that trains models with RL (for ~$1.3k)"]]></title><description><![CDATA[
<p>Why should anyone put serious time and effort into using/understanding a product when the author hasn't put serious time and effort into making it?<p>I could make this exact same thing over a weekend and post it on Hackernews. But I won't because I would be embarrassed to do so.<p>The bar for posting something to HN should be high, the bar for wanting people to read your code, your writing, should be putting serious effort and thought into it. Not just vibe coding something up with a vibe coded README and 100% vibe coded code and not even a novel idea or implementation.</p>
]]></description><pubDate>Wed, 15 Jul 2026 14:38:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=48921548</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=48921548</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48921548</guid></item><item><title><![CDATA[New comment by saberience in "Show HN: I RL-trained an agent that trains models with RL (for ~$1.3k)"]]></title><description><![CDATA[
<p>The AI generated README?</p>
]]></description><pubDate>Tue, 14 Jul 2026 14:03:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=48907073</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=48907073</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48907073</guid></item><item><title><![CDATA[New comment by saberience in "Show HN: I RL-trained an agent that trains models with RL (for ~$1.3k)"]]></title><description><![CDATA[
<p>Well, you say that, but when "measuring" anything in RL, that measurement itself is not always obvious.<p>That is, creating the scoring system/judge models etc for RL is not easy at all. You can easily create an RL loop which is getting better and improving its scores, but actually the result is totally garbage, because you're measuring the wrong thing.</p>
]]></description><pubDate>Tue, 14 Jul 2026 13:58:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=48906991</link><dc:creator>saberience</dc:creator><comments>https://news.ycombinator.com/item?id=48906991</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48906991</guid></item></channel></rss>