<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: yetanotherjosh</title><link>https://news.ycombinator.com/user?id=yetanotherjosh</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 06 Aug 2026 06:58:14 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=yetanotherjosh" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by yetanotherjosh in "Meta Ran Ads That Contained AI-Generated Child Sexual Abuse Imagery"]]></title><description><![CDATA[
<p>10^9 is the standard definition of a billion, it's very wrong to call 10^12 a "standard billion." You can say "long scale billion" if you want to refer to the definitively non-standard 10^12 billion.</p>
]]></description><pubDate>Wed, 05 Aug 2026 22:54:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49190170</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=49190170</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49190170</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "The state of open source AI"]]></title><description><![CDATA[
<p>I'm not sure what kind of point you're trying to make. There are projects to train competent modern LLMs in which the entire pipeline (data, training process, final weights) is all completely transparent, shared, and reproducible by anyone with the compute to try it out.<p>Or is your definition of "open source" mean that a small indie dev should be able to reproduce the entire pipeline? Because that would disqualify more than just LLMs, but also hardware platforms like Arduino where you need to pay for manufacturing to get the underlying stuff built... is Arduino "open source?"</p>
]]></description><pubDate>Fri, 17 Jul 2026 23:02:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=48953150</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48953150</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48953150</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "The state of open source AI"]]></title><description><![CDATA[
<p>Olmo 3? K2 V2? There are definitely LLMs with very compelling capabilities where the dataset, training process, and final weights are all open. There are also initiatives in the EU and various national government levels (e.g. Switzerland) to develop LLMs in an open and transparent way, not just "open weights."</p>
]]></description><pubDate>Fri, 17 Jul 2026 22:56:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48953094</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48953094</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48953094</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Towards a harness that can do anything"]]></title><description><![CDATA[
<p>I experience this too, but the danger is is making assumptions based on "tells" which might just be how the user wrote the post. It's a situation now where if someone authentically writes something, but happens to rub up against some arbitrary "tell," they flip a bit in the reader and get their work rejected. For example, I used em dashes fairly regularly in my writing before AI started convincing everyone that the mere presence of an em dash is the AI tell, so I actually stopped using them, just to avoid this backlash.</p>
]]></description><pubDate>Fri, 17 Jul 2026 05:51:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=48943746</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48943746</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48943746</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "How to stop Claude from saying load-bearing"]]></title><description><![CDATA[
<p>The real problem is not terms like "load-bearing," which communicate clearly enough. It's the constant invention of cryptic shorthand terms and phrases that have no referent, and end up acting like a puzzle to be decoded. This is often paired with hyphenation, but not always:<p>"The current behavior paper" -> The behavior in the running system that was previously described as papered over.<p>"Marker transport over-claim" -> The inaccurate review finding on the object's sentinel flag in the API response.<p>I suppose the cryptic/invented language problem is about token efficiency? But this sort of token efficiency is extremely difficult to deal with when it comes to conversation with a human about complex system. It might be efficient inside reasoning blocks, but when the model generates the final turn text, it should avoid this, as it's brutally inefficient due to the time spent wondering what each uniquely coined phrase means and having to ask for constant clarifications, which then you have to wait for another turn, eating up time and context while it burns more xhigh reasoning just thinking about how to explain its own awful language.</p>
]]></description><pubDate>Tue, 14 Jul 2026 15:30:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=48908379</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48908379</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48908379</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "GPT-5.5 Codex reasoning-token clustering may be leading to degraded performance"]]></title><description><![CDATA[
<p>There is nothing called "GPT5.5 Codex" unless I've completely misunderstood OpenAI's product line?<p>Codex is a harness, while GPT-5.5 is a model. The last codex-branded model was 5.3. Codex as a harness ships as a CLI, a desktop app, and a web product (and I'm not at all sure how similar the underlying harness is between them.)<p>Is the bug here supposed to be with the CLI harness, or the model? Does it also happen in pi, opencode, etc while running GPT-5.5?</p>
]]></description><pubDate>Sun, 05 Jul 2026 15:18:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=48794923</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48794923</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48794923</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "What happened after 2k people tried to hack my AI assistant"]]></title><description><![CDATA[
<p>Kinda reads to me like: "I'm not worried about prompt injection anymore because I setup a test where my agent could just ignore the input channel as noise, and a bunch of comically simple  attacks thrown at it didn't succeed."<p>To be fair I appreciate the effort of running and sharing the test. It will hopefully lead to better ones. But this is not a great test. Super interesting to think about what would constitute a better test.<p>For one, I think the agent would have to be expected to have productive interaction through the email channel, in a way the user depends on it generally working for some real world use case / value prop. In other words, needing emails to actually have the agent really do work, respond with results, etc. Also, most requests should be legit and the real attacks should be intelligently disguised, not pitiful/joke-level spam (although those would be arguably realistic to have in the stream, but, perhaps only as deflection so that the real attack is mischaracterized.)</p>
]]></description><pubDate>Fri, 26 Jun 2026 19:45:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=48691099</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48691099</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48691099</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "What happened after 2k people tried to hack my AI assistant"]]></title><description><![CDATA[
<p>Well said. This experiment is extremely unrealistic and gave the model the opportunity to simply refuse to deal with the channel outright. If he had built it to be a functional agent that depends on real interaction via email and occasional mixed attacks (and attacks that were better designed than the pitiful examples given), this would have gone differently.</p>
]]></description><pubDate>Fri, 26 Jun 2026 19:39:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=48691032</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48691032</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48691032</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Is Grep All You Need? How Agent Harnesses Reshape Agentic Search"]]></title><description><![CDATA[
<p>From the article:<p>> LongMemEval rewards recovering literal witnesses: exact dates, counts, preferences, and spans that often remain stable under tokenization.<p>Is this saying they chose a benchmark that is biased towards doing well against literal string matching, thus works well with grep, and then (gasp) showed that grep did well, finally declaring "grep is all you need"?<p>The examples in the benchmark's demo image(1) are all examples you could see grep doing well on. A conversation about bikes, then a query about bike(s) where "bike" is a common token hit. But not stuff like a conversation about a Beethoven sonata, then a question about classical music, where embedding based approach would shine.<p>(1) <a href="https://github.com/xiaowu0162/LongMemEval/blob/main/assets/longmemeval_examples.png" rel="nofollow">https://github.com/xiaowu0162/LongMemEval/blob/main/assets/l...</a></p>
]]></description><pubDate>Tue, 09 Jun 2026 23:21:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=48469179</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48469179</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48469179</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Postmortem: TanStack NPM supply-chain compromise"]]></title><description><![CDATA[
<p>How is this not a Github P0? Can anyone explain?<p>When I read that, I thought they must be using 'fork' wrong, and actually mean branch on the official repo, as that can't be right!?" Good lord.</p>
]]></description><pubDate>Mon, 11 May 2026 23:33:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48102206</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=48102206</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48102206</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Buteyko Method"]]></title><description><![CDATA[
<p>I practiced the Buteyko method for many years when I was in states of high anxiety and frequent panic attacks, and it was incredibly helpful. I had a syndrome called new daily persistent headache, which means a sudden-onset headache that becomes constant from that instant forward, as in 24/7, for years. It's hard for people who haven't experienced that to understand I mean that literally. Buteyko breathing was the only thing that ever cut down on that headache, and a couple times I was able to suspend it for an hour or two, which when you have a literally constant headache for years, is a big deal.<p>My big takeaway from it was that breathing and neurological state are deeply connected and <i>actually</i> relaxed, natural, healthy breathing (and the corresponding state of the brain and nervous system) is something that most people have probably never even experienced unfortunately. We all think our state is normal, but I assure you, it is very far from the state where your control pause is 40s-60s or more, it's a radically different experience.<p>Also the nuance of what the control pause and how to measure it correctly is lost on I would say, even most people who attempt to learn Buteyko. The control pause is how long you can, under <i>normal</i> breathing, suspend breath with <i>zero</i> discomfort, and then <i>prefectly resume normal</i> breathing without any change from before. If you took a bigger breath to start, or when you start breathing again it's even slightly heavier, it's not a control pause measurement, it becomes an ego metric juiced to make you feel better about a number while avoiding the disappointing facts.<p>Buteyko claimed that healthy breathing had a 40+ second control pause. Which if you think about the real meaning and how to measure it, is a super long time. And I got there sometimes, it's a major learning experience about what deep alignment and relaxation of the brain/nerves can really feel like.</p>
]]></description><pubDate>Sat, 20 Dec 2025 19:51:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=46338984</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=46338984</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46338984</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Buteyko Method"]]></title><description><![CDATA[
<p>Buteyko practioners build up the ability to work very hard while only nasal breathing over the long term. The point is to learn to modulate breathing in a way that keeps a certain kind of blood chemistry (CO2 levels) and cellular oxygenation.<p>If you commit to nasal breathing for exercise as a constraint, it does force you to modulate your exertion while also increasing your CO2 and developing, according to the Buteyko folks, a new baseline for respiratory health and capability.<p>If you're running for your life from a tsunami, by all means mouth breathe.  If your purpose is maximum exertion, of course mouth breathe. But that's not the only possible purpose of exercise. It can also be about respiratory training. Nasal breathing becomes a natural guideline/modulator for long term improvement in that regard.</p>
]]></description><pubDate>Sat, 20 Dec 2025 19:33:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=46338855</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=46338855</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46338855</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Be Careful with GIDs in Rails"]]></title><description><![CDATA[
<p>I struggle to understand what this specifically has to do with rails or global IDs. In ANY framework or query system, if you are asking an LLM to produce IDs which you are then passing to a database for lookup, you need to understand those identifiers could be hallucinated or incorrect in surprising or malicious ways, and can lead to data leaks or exfiltration.<p>It's like writing an article about "the dangers of PostgreSQL" ... when generating SQL from an LLM. It has nothing to do with Postgres specifically, it's that you're generating queries to run in a trusted context from an untrustable origin.</p>
]]></description><pubDate>Tue, 16 Dec 2025 17:03:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=46291036</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=46291036</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46291036</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "8M users' AI conversations sold for profit by "privacy" extensions"]]></title><description><![CDATA[
<p>I don't understand how code review would catch this. The extension advertises itself as an AI protection tool, that monitors your AI interactions. The code is basically consistent with the stated purpose. That it doesn't stop collecting data when you turn of the UI alerting is perhaps an inconsistency, but I think that's debatable (is there a rule in google's terms that says data collection is contingent on UI alerts being enabled?). I'm curious what workflow or decision tree you'd expect a code review process to follow here that results in this being rejected? The problem here doesn't seem like code related, it's policy related, as in, what are they doing with the information, not that the extension has code to collect it.</p>
]]></description><pubDate>Tue, 16 Dec 2025 16:39:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=46290715</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=46290715</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46290715</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Valdi – A cross-platform UI framework that delivers native performance"]]></title><description><![CDATA[
<p>So you can get native behaviors when it’s critical. Like share sheets, push  and many other critical features that only apps get even if the bulk of the experience can be done in a webview. This is because mobile OS platforms choose not to make these available to web apps, because app store profits are better for them than an open ecosystem where sites can do the same things as apps.</p>
]]></description><pubDate>Sat, 08 Nov 2025 10:13:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=45855624</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=45855624</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45855624</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "The security paradox of local LLMs"]]></title><description><![CDATA[
<p>.95 is quite generous here</p>
]]></description><pubDate>Wed, 22 Oct 2025 14:57:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=45670121</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=45670121</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45670121</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "Caching is an abstraction, not an optimization"]]></title><description><![CDATA[
<p>Yes "good" caching - a consistent storage interface - is an abstraction over "bad" caching - multiple different storage interfaces with different speeds. But caching overall is not an abstraction over not having caching.</p>
]]></description><pubDate>Fri, 04 Jul 2025 15:47:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=44465541</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=44465541</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44465541</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "OpenAI Codex CLI: Lightweight coding agent that runs in your terminal"]]></title><description><![CDATA[
<p>Why?</p>
]]></description><pubDate>Tue, 22 Apr 2025 05:54:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=43759414</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=43759414</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43759414</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "OpenAI Codex CLI: Lightweight coding agent that runs in your terminal"]]></title><description><![CDATA[
<p>Astroturfing alert. This comment author is also the author of cursor-agent-tools.</p>
]]></description><pubDate>Tue, 22 Apr 2025 05:51:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=43759393</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=43759393</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43759393</guid></item><item><title><![CDATA[New comment by yetanotherjosh in "DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL"]]></title><description><![CDATA[
<p>ollama is stating there's a difference: <a href="https://ollama.com/library/deepseek-r1">https://ollama.com/library/deepseek-r1</a><p>"including six dense models distilled from DeepSeek-R1 based on Llama and Qwen. "<p>people just don't read? not sure there's reason to criticize ollama here.</p>
]]></description><pubDate>Sun, 26 Jan 2025 18:07:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=42832230</link><dc:creator>yetanotherjosh</dc:creator><comments>https://news.ycombinator.com/item?id=42832230</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42832230</guid></item></channel></rss>