<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: Ukv</title><link>https://news.ycombinator.com/user?id=Ukv</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 20 Aug 2026 21:29:58 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=Ukv" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by Ukv in "Why Microsoft Entertainment Pack had a sticker announcing that it had Tetris?"]]></title><description><![CDATA[
<p>Why it was announced specifically <i>on a sticker</i> (rather than on the box like the other games) is the core of the question - the headline no longer really works without it. I'd go with one of:<p>> Why did Microsoft Entertainment Pack have a sticker announcing it had Tetris?<p>> Why did Microsoft Entertainment Pack announce that it had Tetris on a sticker?<p>> Why the Microsoft Entertainment Pack had a sticker announcing it had Tetris<p>But doesn't really matter that much.</p>
]]></description><pubDate>Thu, 20 Aug 2026 09:48:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49372450</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49372450</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49372450</guid></item><item><title><![CDATA[New comment by Ukv in "Local Models Will Not Win"]]></title><description><![CDATA[
<p>> Local models are never going to be as powerful. I think this point should be obvious: all of the current frontier models (closed and open-weights) are far too big to run on anything but a full GPU cluster in a datacenter<p>Diminishing returns with respect to scale (a 10X larger model is typically not 10X better at any given task) has so far meant that, even when datacenter compute grows faster than individual compute, the gap in quality between hosted and local models has generally shrunk. There are still plenty of tasks where that extra gain in quality is noticeable, but I feel there are also an increasing number of "saturated" tasks where it really doesn't matter.<p>Which puts more focus on other factors. Local models are private, low-latency, work offline, and can be tinkered with to your liking - like changing the system prompt to avoid refusals. I would not trust a hosted model to classify my documents, for example.</p>
]]></description><pubDate>Tue, 11 Aug 2026 12:00:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=49256895</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49256895</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49256895</guid></item><item><title><![CDATA[New comment by Ukv in "Making difficulty curves in games"]]></title><description><![CDATA[
<p>> You can turn off some annoying parts of the game [...] cool that everyone can have the amount of difficulty/frustration they desire<p>Depends on the type of game, but I'd generally prefer those decisions be made by the game designer. If there's frustration I want it to feel like overcoming a concrete obstacle, rather than it being self-hindering that accomplishes nothing and that I could've just avoided.</p>
]]></description><pubDate>Sun, 09 Aug 2026 15:58:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49232601</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49232601</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49232601</guid></item><item><title><![CDATA[New comment by Ukv in "Mistral's Shieldstral: 3B open-weights model for multimodal moderation"]]></title><description><![CDATA[
<p>> You have no way to provide a concrete reason to the user at that point.<p>You should just be able to look at the user's message and tell them why it's against your policy, else reverse the decision if you see no violation.<p>If a user is curious specifically about how the model made its decision, and you want to reveal detail at that level, it's an open-weights model so interpretability techniques should work ("biggest impact on score came when focusing on this word in your message and this part of the policy").</p>
]]></description><pubDate>Wed, 05 Aug 2026 15:47:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49184535</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49184535</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49184535</guid></item><item><title><![CDATA[New comment by Ukv in ""the very foundation of modern academia has been blown to bits""]]></title><description><![CDATA[
<p>> But in 1905 a paper was published... Too bad politicans do not get implications of that and still was pushing communism decades later...<p>The photoelectric effect/quanta? I'm not sure I understand the supposed connection to communism.</p>
]]></description><pubDate>Fri, 31 Jul 2026 09:08:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49120763</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49120763</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49120763</guid></item><item><title><![CDATA[New comment by Ukv in "Google will expand age checks on Android worldwide till the end of the year"]]></title><description><![CDATA[
<p>> Why do we think this should be different on the internet?<p>Because when applied to the internet, the age gating model that works okayish for brick-and-mortar locations is ineffective (children can use a friend's older brother's ID or account, websites outside of the relevant jurisdiction can just decide not to verify age, etc.) and invasive (sending private and sensitive information off to US companies that have been breached or linked to mass surveillance). Lack of efficacy is then used to justify crack-downs on privacy tools and the need for broad content-blocking powers with no due process.<p>Preferable IMO would be with filtering on the local network level, like schools have been doing for decades. Parents typically own the router/mobile data plan, so it'd pretty much just be a change of defaults and maybe some new interoperability standards. A lot less invasive, and arguably more effective.</p>
]]></description><pubDate>Thu, 30 Jul 2026 15:36:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49111485</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49111485</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49111485</guid></item><item><title><![CDATA[New comment by Ukv in "Truth is not a direction: a Tarski attack on LLM probes"]]></title><description><![CDATA[
<p>> plug an LLM into the nukes and funnel worldstate input into it and have it make the decision "is it time to fire the nukes?" over and over again each second [...] I argue that that would require far more nines than even "will food turn to poison in my mouth" would.<p>Sure - but (even assuming that's a practical purpose) the point is it that it doesn't need to be a 100% accurate truth oracle, which is all the article's argument prohibits. If the current human chain of command has 99.99999994% accuracy, then 99.99999995% accuracy is an improvement and not ruled out by the argument.</p>
]]></description><pubDate>Wed, 29 Jul 2026 21:42:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49103449</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49103449</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49103449</guid></item><item><title><![CDATA[New comment by Ukv in "Truth is not a direction: a Tarski attack on LLM probes"]]></title><description><![CDATA[
<p>ziofill's claim was that <i>"A direction [in an LLM's embedding vector space] that is 99.99% accurate"</i> is fine for practical purposes, not that 99.99% is fine for the chance of any given bite of food not killing you or similar hypotheticals - you'd want a few more 9s there.<p>To justify relevance of inability to correctly answer liars-paradox-type questions ("what won't your response to this be?"), the article suggested the way LLMs are used in practice is dependant on them being entirely accurate truth oracles:<p>> > as a truth-oracle [...] is how these things will be used practically by the vast majority of people. They are already replacing standard Google search results<p>But for the replacement to make sense they just need to be more accurate than what they're replacing (ignoring other factors like convenience and cost) - in this case standard Google search results and knowledge box which were obviously not 100.0% accurate.</p>
]]></description><pubDate>Wed, 29 Jul 2026 10:55:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49095785</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49095785</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49095785</guid></item><item><title><![CDATA[New comment by Ukv in "Private Claude Chats Exposed in Google and Bing Search Results"]]></title><description><![CDATA[
<p>Private as in chats for which the user generated a share link and posted it somewhere online that a search engine's web crawler found, as far as I can tell.</p>
]]></description><pubDate>Tue, 28 Jul 2026 15:23:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49085325</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49085325</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49085325</guid></item><item><title><![CDATA[New comment by Ukv in "A missing underscore sent innocent man to prison for 18 months"]]></title><description><![CDATA[
<p>> Insane. An LLM is just as likely to hallucinate a missing/extra underscore and ping the wrong person<p>If it's just for "Catching typos", a hallucinated missing/extra underscore would just be a false positive to dismiss.<p>> A machine cannot be held accountable.<p>Seems unlikely that his lawyer, the law firm, the judge, whoever made the typo, or the police department will be held accountable either.<p>Nor can any of the tools they used, since that's not really the level at which it makes sense to hold accountability, but that's no reason not to use a tool that could find errors and reduce the chance for an innocent person to spend time in prison.</p>
]]></description><pubDate>Tue, 28 Jul 2026 09:41:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49081531</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49081531</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49081531</guid></item><item><title><![CDATA[New comment by Ukv in "Scanning for Pangram Errors"]]></title><description><![CDATA[
<p>> Pangram boasts a false positive rate of 1 in a 10,000. That is, if Pangram says a block of text is AI there is only a one in ten thousand chance that it was written by a human.<p>That'd be if they had a <i>false discovery rate</i> of 1/10,000.<p>If for instance:<p>* 100,000 samples are tested<p>* 100 of which are AI-generated, the rest human-written<p>* Pangram flags 50 of the AI-generated samples (true positives)<p>* Pangram also flags 10 human-written samples (false positives)<p>Then the FPR is 1 in 10,000, but the chance that a flagged sample isn't actually AI (FDR) is 1 in 6.</p>
]]></description><pubDate>Thu, 23 Jul 2026 13:52:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49021613</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=49021613</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49021613</guid></item><item><title><![CDATA[New comment by Ukv in "Judge halts Paramount's $111B purchase of Warner Bros. in win for US states"]]></title><description><![CDATA[
<p>Observationally, larger companies already in the lead seem prefer a safe X% return for their shareholders and don't need to take large risks. Smaller companies trying to make it don't have the liberty to rest on their laurels, and often will be risking it all on some idea being a massive hit.</p>
]]></description><pubDate>Tue, 21 Jul 2026 12:14:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=48991241</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48991241</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48991241</guid></item><item><title><![CDATA[New comment by Ukv in "AI advice made people less accurate but more confident – sudy"]]></title><description><![CDATA[
<p>The six questions they asked were:<p>> 1) What animal is on the bow of the pirate ship from “Asterix and Obelix”?<p>> 2) In the movie “The Grand Budapest Hotel”, what is Agatha’s signature hairstyle?<p>> 3) What color is the team’s uniform in “Bend It like Beckham”?<p>> 4) What vehicle does Monica drive in “Like a Cat on a Highway”?<p>> 5) What color is the turtle in the animated movie “Momo” by Enzo d’Alò?<p>> 6) What pet animal does Asenath have in “Joseph King of Dreams”?<p>Most of these are just a matter of knowing it or not, where you can't really distinguish a plausible answer from the correct answer just by thinking.</p>
]]></description><pubDate>Mon, 20 Jul 2026 12:29:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=48977904</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48977904</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48977904</guid></item><item><title><![CDATA[New comment by Ukv in "It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)"]]></title><description><![CDATA[
<p>> Speed and cost are nothing without quality<p>Quality was what the hypothetical was assuming had reached parity, no? ("humans are as bad as AI", "If AI really is at human level quality/error rate", etc.)<p>> accountability<p>Why could the company not still take accountability? That's already the case for non-ML automated systems, some with high failure rates. As a customer I rarely if ever care about blame being pinned on a specific employee.<p>> Your counter argument was outside the context of this articles claims, specifically that programmers and other knowledge workers can be replaced by LLMs.<p>AndrewKemendo's comment and your reply ("any humans", "replace humans") seemed to generalize, but speed and cost being important factors is still true for knowledge work. For some given level of quality, a web developer offering a lower quote with shorter turnaround time will be preferred to one offering a higher quote with longer turnaround time.</p>
]]></description><pubDate>Fri, 03 Jul 2026 18:30:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=48778268</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48778268</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48778268</guid></item><item><title><![CDATA[New comment by Ukv in "It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)"]]></title><description><![CDATA[
<p>My understanding of your argument is (paraphrasing):<p>> > People try to excuse AI issues/failure modes by saying humans have them too, but even if they're equally bad then what would be the whole point of replacing a human worker with AI?<p>To which my response is that speed and cost are also important factors, which can often give AI the edge in considerations when quality/error rate is equal.<p>If you meant something other than that, you may have to specify.</p>
]]></description><pubDate>Fri, 03 Jul 2026 16:39:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776998</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48776998</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776998</guid></item><item><title><![CDATA[New comment by Ukv in "It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)"]]></title><description><![CDATA[
<p>> Have outputs from engineers traditionally been measured in cost and speed?<p>Yes. How long it'll take and how much it'll cost are going to be among pretty much any customer's first questions.<p>They're not the <i>only</i> considerations, and could potentially be outweighed by other concerns even when quality is the same, but I think they are the main drives of AI adoption in industry. If error rate is the same, a $1/hr (amortized) camera and machine vision model capable of checking 300ft of material for defects per minute will likely be preferred to a $10/hr human QA capable of checking 30ft per minute, for instance.</p>
]]></description><pubDate>Fri, 03 Jul 2026 16:10:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776691</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48776691</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776691</guid></item><item><title><![CDATA[New comment by Ukv in "It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)"]]></title><description><![CDATA[
<p>> The intention is something like “so humans are as bad as AI” when the original question boils down to something like “why would I replace humans with AI?”<p>If AI really is at human level quality/error rate (I don't think it is for general tasks, but there are some areas where it is), then the answer is typically cost and speed/capacity.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:52:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776471</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48776471</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776471</guid></item><item><title><![CDATA[New comment by Ukv in "Why I like snake_case"]]></title><description><![CDATA[
<p>I do like snake_case but a lot of this feels a bit circular, effectively just saying that it's good because it's already used by the author's code and things it interacts with.<p>I'd like kebab-case even more if it weren't for the annoying detail that `-` is also subtraction.</p>
]]></description><pubDate>Wed, 01 Jul 2026 16:23:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=48749360</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48749360</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48749360</guid></item><item><title><![CDATA[New comment by Ukv in "30-year sentence for transporting zines is a five-alarm fire for free speech"]]></title><description><![CDATA[
<p>> They brought guns and shot government officials<p>Only Benjamin Song, convicted of attempted murder/discharging a firearm, shot the police officer. Some others didn't bring firearms, were not in any planning chat (in which no violence was planned regardless), weren't at the protest or had already left, yet still received absurdly harsh sentences - that's the chilling effect.</p>
]]></description><pubDate>Tue, 30 Jun 2026 17:58:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=48736570</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48736570</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48736570</guid></item><item><title><![CDATA[New comment by Ukv in "30-year sentence for transporting zines is a five-alarm fire for free speech"]]></title><description><![CDATA[
<p>> The 30 year sentence was for hiding documentation [...] it wasn't just "transporting Zines"<p>As far as I can tell, the moving of zines (he was pulled over and had a box in his car) is what's being presented as "hiding documentation" - not something beyond that.<p>> being sought under a federal warrant<p>Timeline seems to be that a warrant was obtained <i>after</i> pulling him over ("Sanchez-Estrada was then arrested on state traffic offenses, and officers obtained a search warrant [...]"). Can't find a source saying there was a warrant prior to this.<p>> The warrant was for documentation after the protesters shot fireworks to bring out first responders from the ICE facility, and allegedly one of the group shot a responder in the neck instead of the head.<p>It's true that demonstrators were setting off fireworks, and it's true that Benjamin Song later shot at a police officer who had drawn his gun. But it's just the government's narrative/speculation that the intent of the fireworks was to draw out first responders to ambush, and that Sanchez-Estrada's zines were in some way documentation of this despite him not being at the protest and his wife not being the shooter.</p>
]]></description><pubDate>Mon, 29 Jun 2026 12:11:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=48718225</link><dc:creator>Ukv</dc:creator><comments>https://news.ycombinator.com/item?id=48718225</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48718225</guid></item></channel></rss>