<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: thduabmd</title><link>https://news.ycombinator.com/user?id=thduabmd</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 17 Sep 2026 15:09:02 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=thduabmd" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by thduabmd in "Introducing System One Models and Jev"]]></title><description><![CDATA[
<p>No. Your launch post puts “0%” on a hallucination chart, then explains that the number comes from guaranteed schema matching.<p>You’ve already agreed that this doesn’t establish correctness. An approve for an unauthorized action still meets the schema guarantee.<p>That’s why I find the messaging misleading. You’re acknowledging the limitations in these replies while defending the broader reliability pitch.<p>Even granting that each answer is calibrated individually, that doesn’t establish calibration of the decision that combines them.<p>Sure, I can threshold a composite score, but there may be many wrong answers with the same score. An unauthorized action doesn’t become acceptable because it scores highly on the other dimensions.<p>I still have to define the constraints and test which wrong actions get through the complete workflow on my own data. That’s a substantial part of the work being pushed back onto the developer.</p>
]]></description><pubDate>Wed, 16 Sep 2026 00:34:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49720703</link><dc:creator>thduabmd</dc:creator><comments>https://news.ycombinator.com/item?id=49720703</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49720703</guid></item></channel></rss>