<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: randomblock1</title><link>https://news.ycombinator.com/user?id=randomblock1</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 09 Oct 2026 03:29:31 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=randomblock1" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by randomblock1 in "LLM Ass Bench"]]></title><description><![CDATA[
<p>Now put them on a bike.</p>
]]></description><pubDate>Tue, 22 Sep 2026 22:03:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808815</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49808815</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808815</guid></item><item><title><![CDATA[New comment by randomblock1 in "Claude Opus 5.5"]]></title><description><![CDATA[
<p>>  On our benchmarks, Claude Opus 5.5 leads in agentic coding, computer use, and knowledge work. That said, at these levels of capability we’ve found that benchmark margins have become a less reliable guide to real-world differences. In our own use, the gap between Opus 5.5 and Claude Fable 5.1 is narrower than these scores suggest.</p>
]]></description><pubDate>Tue, 22 Sep 2026 16:45:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49804188</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49804188</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49804188</guid></item><item><title><![CDATA[New comment by randomblock1 in "I don't like passkeys"]]></title><description><![CDATA[
<p>It's amazon, I have 1password and it always asks me to create another security key</p>
]]></description><pubDate>Fri, 18 Sep 2026 18:46:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49758515</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49758515</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49758515</guid></item><item><title><![CDATA[New comment by randomblock1 in "Bend – A language that blocks AI mistakes via proof, on CPU and GPU"]]></title><description><![CDATA[
<p>Multiple times, even. Still no real reason why. <a href="https://github.com/bendlang/bend/activity?ref=main" rel="nofollow">https://github.com/bendlang/bend/activity?ref=main</a><p>One time they force pushed and erased everything except a 2-line README... on purpose.<p>Pre-obliteration version: <a href="https://github.com/bendlang/bend/tree/814453670d0e0d6777c1313c972764dba0491b7f" rel="nofollow">https://github.com/bendlang/bend/tree/814453670d0e0d6777c131...</a></p>
]]></description><pubDate>Thu, 17 Sep 2026 21:24:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49746751</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49746751</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49746751</guid></item><item><title><![CDATA[New comment by randomblock1 in "TSMC revealing details about next gen A14 node"]]></title><description><![CDATA[
<p>It's not quite SRAM, it needs to be refreshed like DRAM. That's the main downside compared to SRAM but it's still very interesting.</p>
]]></description><pubDate>Thu, 17 Sep 2026 19:18:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49745289</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49745289</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49745289</guid></item><item><title><![CDATA[New comment by randomblock1 in "Hister: A private search engine for the pages you visit and the files you keep"]]></title><description><![CDATA[
<p>Nothing about this depends on the provider, you could spin up a local Qwen and give it a search tool like SearXNG or something. At this point local models are more than good enough for simple tasks like that. Using ChatGPT is just (usually) faster and simpler</p>
]]></description><pubDate>Thu, 17 Sep 2026 17:51:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49744228</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49744228</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49744228</guid></item><item><title><![CDATA[New comment by randomblock1 in "Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra"]]></title><description><![CDATA[
<p>No it still requires Pro in the CLI. Only free model is SWE-1.6 slow</p>
]]></description><pubDate>Mon, 14 Sep 2026 02:16:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49691071</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49691071</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49691071</guid></item><item><title><![CDATA[New comment by randomblock1 in "Why is the x86 undefined instruction called ud2? Why 2?"]]></title><description><![CDATA[
<p>Schrodinger's instruction: simultaneously defined and undefined</p>
]]></description><pubDate>Mon, 14 Sep 2026 01:57:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49690940</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49690940</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49690940</guid></item><item><title><![CDATA[New comment by randomblock1 in "Why is the x86 undefined instruction called ud2? Why 2?"]]></title><description><![CDATA[
<p>0 undoubtedly comes first, but we all know 1 == true...</p>
]]></description><pubDate>Mon, 14 Sep 2026 01:55:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49690923</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49690923</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49690923</guid></item><item><title><![CDATA[New comment by randomblock1 in "So you want to use OpenRouter?"]]></title><description><![CDATA[
<p>You know how there's a router mode to use the cheapest provider? That only takes into account uncached rates, last I checked. Make another one that takes into account effective rates (the ones that include cache).</p>
]]></description><pubDate>Fri, 11 Sep 2026 23:41:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49666939</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49666939</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49666939</guid></item><item><title><![CDATA[New comment by randomblock1 in "Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra"]]></title><description><![CDATA[
<p>Oh I tried in cloud. I'll give it another shot</p>
]]></description><pubDate>Fri, 11 Sep 2026 09:35:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49655658</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49655658</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49655658</guid></item><item><title><![CDATA[New comment by randomblock1 in "More questions about whether researchers can trust OpenAI with unpublished math"]]></title><description><![CDATA[
<p>So then they DIDN'T "learn the secret to cracking the problem". They simply knew that part of the problem was solved. Knowing a problem can be solved and knowing the solution are not the same thing.</p>
]]></description><pubDate>Thu, 10 Sep 2026 17:35:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49647477</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49647477</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49647477</guid></item><item><title><![CDATA[New comment by randomblock1 in "Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra"]]></title><description><![CDATA[
<p>I just gave it a try and it doesn't appear to be free, it used up some of my on demand usage. It does say 75% off though. Seems like for Pro subscribers SWE-1.7 is free, maybe SWE-2 is free for them?</p>
]]></description><pubDate>Thu, 10 Sep 2026 17:28:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49647379</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49647379</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49647379</guid></item><item><title><![CDATA[New comment by randomblock1 in "Searching for the best silicone USB cable"]]></title><description><![CDATA[
<p>It's the website's "grid-bg" effect. It's glitching out and also normally barely visible at all.</p>
]]></description><pubDate>Wed, 09 Sep 2026 16:06:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49628779</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49628779</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49628779</guid></item><item><title><![CDATA[New comment by randomblock1 in "Protecting Engineers' Skills in the AI Era"]]></title><description><![CDATA[
<p>The writer is the CEO of... whatever this is: <a href="https://auraspark.com/" rel="nofollow">https://auraspark.com/</a><p>Completely slopped up website, zero human touch. I would even go so far as to call them a hyprocrite, for getting AI to do all that for them instead of, you guessed it, training a junior engineer to do it.</p>
]]></description><pubDate>Thu, 03 Sep 2026 23:54:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49558732</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49558732</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49558732</guid></item><item><title><![CDATA[New comment by randomblock1 in "Show HN: FrontierHarness Eval – 9 harness, same model, cost per pass varies 17x"]]></title><description><![CDATA[
<p>I think it's still a useful data point. For example, omp, which is pi with some default extensions, scores worse. I do agree that adding more configurations of Pi would help though.</p>
]]></description><pubDate>Wed, 02 Sep 2026 19:45:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49541413</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49541413</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49541413</guid></item><item><title><![CDATA[New comment by randomblock1 in "Can I opt out of my input or output data being used for training?"]]></title><description><![CDATA[
<p>The retain it for "the duration necessary to achieve the intended purposes", which could mean forever.</p>
]]></description><pubDate>Wed, 02 Sep 2026 19:40:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49541324</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49541324</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49541324</guid></item><item><title><![CDATA[New comment by randomblock1 in "Apple caught off guard by AI demand for Mac Mini and Mac Studio"]]></title><description><![CDATA[
<p>It's not that far off anymore. On my 7900 XTX 24GB, I can run Qwen3.8 27B with 131K context at Q4_K_M (55 tok/s with MTP). Excluding hardware cost, it's about $0.02 tok/M in and $0.40 tok/M out (cached in $0.0001). On OpenRouter, that would cost more than 10x what it actually costs me.<p>Of course, 131k context at 4-bit quant is a trade off, but even then, it's VERY capable. It doesn't feel that far behind something like GPT 5.6 Luna.</p>
]]></description><pubDate>Mon, 31 Aug 2026 21:22:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49514992</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49514992</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49514992</guid></item><item><title><![CDATA[New comment by randomblock1 in "Previewing the Model Hardware Standard"]]></title><description><![CDATA[
<p>To me, this seems like OPC UA / SiLA but instead of being software <-> machine semantics/control it's AI <-> machine semantics/control. Or in simpler terms, an AI-facing hardware abstraction / device-description standard.<p>This is how I understand it:<p>Agent <-> MCP/CLI/code <-> MHS <-> vendor API/SCPI/OPC UA/ROS/etc <-> CAN/Modbus/USB/etc <-> hardware<p>MHS can describe capabilites, metadata, safety limits, as well as provide read/write control and discovery. Things like "can measure temperature", "arm weighs X kg", "never exceed X RPM", that agents can easily understand. (as opposed to that being buried in a datasheet somewhere, or having to be included in the prompt)<p>Also see: <a href="https://en.wikipedia.org/wiki/OPC_Unified_Architecture" rel="nofollow">https://en.wikipedia.org/wiki/OPC_Unified_Architecture</a>, <a href="https://en.wikipedia.org/wiki/Standardization_in_Lab_Automation" rel="nofollow">https://en.wikipedia.org/wiki/Standardization_in_Lab_Automat...</a>, <a href="https://xkcd.com/927/" rel="nofollow">https://xkcd.com/927/</a></p>
]]></description><pubDate>Thu, 27 Aug 2026 21:15:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49471381</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49471381</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49471381</guid></item><item><title><![CDATA[New comment by randomblock1 in "YouTube Format IDs"]]></title><description><![CDATA[
<p>Anything with motion. Sports, games, vlogs, and so on experience immense improvements. And it's not like it takes 2x the bandwidth, because inter-frame compression can be smarter about it. It's like 30-50% more bandwidth.</p>
]]></description><pubDate>Wed, 26 Aug 2026 19:31:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49454544</link><dc:creator>randomblock1</dc:creator><comments>https://news.ycombinator.com/item?id=49454544</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49454544</guid></item></channel></rss>