<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: chaboud</title><link>https://news.ycombinator.com/user?id=chaboud</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 07 Oct 2026 02:22:56 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=chaboud" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by chaboud in "Salesforce Global Outage"]]></title><description><![CDATA[
<p>From what I can tell, the business model is "it doesn't work, but you can pay folks exorbitant fees to 'customize' it for you"...<p>Also see: Oracle</p>
]]></description><pubDate>Wed, 16 Sep 2026 14:31:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49727654</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49727654</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49727654</guid></item><item><title><![CDATA[New comment by chaboud in "LRU is harder to beat than the KV-cache papers suggest"]]></title><description><![CDATA[
<p>I've been building latency-sensitive LLM systems for a while, and I've come to rely heavily on pre-fill-considerate mechanics like ping-pong overlapped async context construction.  For interactive mechanics, the worst case, even if rare, is problematic.<p>A toy/simplified version lives here:
<a href="https://github.com/chaboud/goulash" rel="nofollow">https://github.com/chaboud/goulash</a><p>Consideration of mutation rate (a sort of temporal Shannon-ish coding/ordering) lives in there (with some RoPE-friendly structuring).  Note: That was a vacation project, not the day job, but similar principles apply even with larger models.</p>
]]></description><pubDate>Sat, 12 Sep 2026 18:31:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49675541</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49675541</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49675541</guid></item><item><title><![CDATA[New comment by chaboud in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Wait, wait, wait... you're telling me that polars, something made in this decade, is better at modern problems than something made two decades ago?!<p>I'm shocked!  Shocked, I say!<p>Sometimes you just want something that works with all the things. That's pandas. But, like the bamboo eaters, it wont be long...</p>
]]></description><pubDate>Sat, 12 Sep 2026 06:30:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669500</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49669500</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669500</guid></item><item><title><![CDATA[New comment by chaboud in "Recreating Minecraft Is Not a Benchmark"]]></title><description><![CDATA[
<p>"When a measure becomes a target, it ceases to be a good measure."<p>Goodhart's law strikes again.<p><a href="https://en.wikipedia.org/wiki/Goodhart%27s_law" rel="nofollow">https://en.wikipedia.org/wiki/Goodhart%27s_law</a><p>However, what <i>is</i> meaningful is whether something is able to create usefully adjacent output, like "let's make Minecraft, but with marching cubes, subdivision surfaces, and global illumination... and behaviorally accurate pandas..." (or something like that).<p>I have an 11 year old, and most of his game ideas are adjacent to other games he's played.  He can make those now, or, at least, enough that he can see where it works and where it doesn't.<p>Compared to a few years ago, that's pretty cool.</p>
]]></description><pubDate>Sun, 06 Sep 2026 19:36:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49590022</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49590022</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49590022</guid></item><item><title><![CDATA[New comment by chaboud in "Claude Fable 5.1 and Claude Mythos 5.1"]]></title><description><![CDATA[
<p>Your observation is game-changing, and it reverses my suggested priority completely.</p>
]]></description><pubDate>Wed, 02 Sep 2026 09:04:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49533704</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49533704</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49533704</guid></item><item><title><![CDATA[New comment by chaboud in "Claude Code May–August 2026 weekly limits promotion"]]></title><description><![CDATA[
<p>Codex is fabulous at work, where token use is near limitless and ultra-thorough tool use is welcome.  Go ahead and fire off searches for look-alike terms on my 96-core cloud instance.<p>By contrast, Claude Code's bias to make assumptions of reasonableness about underlying systems has proven to be immensely frustrating over the last month or two, both personally and at work. I've wasted days on "that was my mistake.  I've been reporting numbers on the old architecture because I hadn't enabled the new one in the config" both at work and home.  It's immensely frustrating.<p>But here we are.  Wrestling with energetic idiots in model form, wrangled by over-specific harnesses that struggle to stay off of deranged side-quests.<p>What a time to be alive!</p>
]]></description><pubDate>Wed, 19 Aug 2026 04:12:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49356709</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49356709</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49356709</guid></item><item><title><![CDATA[New comment by chaboud in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>Keep in mind that the model is thinking in a token space, itself a compressive representation of language.<p>(Note: there's still a huge grammar penalty, so, ugh do think small.)</p>
]]></description><pubDate>Mon, 17 Aug 2026 04:35:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49326586</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49326586</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49326586</guid></item><item><title><![CDATA[New comment by chaboud in "Advice for First Ten Customers"]]></title><description><![CDATA[
<p>Listen to them, but don't <i>listen</i> to them.<p>You are building to solve their problems or open up their capabilities, and if they knew how to do it, they would have already.  So listen to your customers for the <i>"why"</i>, but use your own judgment for the <i>"what"</i> and the <i>"how"</i> of it all.<p>And be honest about being differentially valuable.  I once had a prospective consultancy customer ask me for a very specific and elaborate piece of software to be built.  He'd been thinking about it for years. I dug in on what he was actually trying to achieve, and I realized that small modifications to an existing in-market product would be able to satisfy his actual needs. I connected him with that company, he ended up with exactly what he needed, and they ended up with a high-end expert user as a resource.<p>Of course, I ended up putting myself out of a lucrative consultancy job, but I don't regret the decision at all.  I suspect that most folks on HN feel the same way.  Find a way to make a difference, and be honest with your customers (and leads) about what you're bringing to the table.  That kind of integrity pays dividends in the long run if you have something truly valuable.</p>
]]></description><pubDate>Sat, 01 Aug 2026 17:07:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49136229</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49136229</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49136229</guid></item><item><title><![CDATA[Show HN: Goulash – heckler/helper to local-LLM your shell as little as possible]]></title><description><![CDATA[
<p>I use the terminal a lot, but I don't remember every command, flag, and syntax available.  ffmpeg, piping to awk, find, etc...  I know I'm not cool enough, but it's crazy-making.<p>I get into a regular pattern of asking Claude, ChatGPT, or a local model how I'd write a command or do something with some files or a prompt, then going back to the terminal with a mad-lib command and reasoning through it.<p>So I built goulash.  It takes up the bottom four lines of your terminal as a place to comment on your session and suggest commands that you can select by hitting <i>down arrow</i> to select your "<i>future history</i>", staying in flow with <i>up arrow</i> for history in most shells.  Want to ask it a direct command question? Just type "# your question" and hit enter, and it will take that inline comment and make a suggestion.<p>Running a suggested command is always up to the user.  Hit down arrow to scroll through suggestions, modify if you want, hit enter to execute.<p>This is made for running with local engines hosting local-friendly models, like Ollama or LM Studio with gemma4:e4b on a Macbook Air.  It relies heavily on context caching to stay snappy, but goulash doesn't block as it thinks or runs.  It's a ride-along in the passenger seat.<p>Have a tweaky command set that your tiny model doesn't know?  Use "#@ YourCommandReference.md" to give that gemma4:e4b (or whatever) a guide as a pin for the session, or use it to temporarily style your life (e.g., I have a FarmAnimals.md file that adds a barnyard joke to any comment. Hey... the tokens are local...).<p>Written in rust, largely with Fable and Opus 5.  Tested with zsh and bash on Mac.  Check out the repo, or install with "cargo install goulash".</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49135861">https://news.ycombinator.com/item?id=49135861</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Sat, 01 Aug 2026 16:33:52 +0000</pubDate><link>https://github.com/chaboud/goulash</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49135861</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49135861</guid></item><item><title><![CDATA[New comment by chaboud in "The new rules of context engineering for Claude 5 generation models"]]></title><description><![CDATA[
<p>I've renovated houses (and I have a bad habit of buying 100 year old hacked up chaos-boxes - my current house was <i>moved</i> from one hill in San Francisco to another 80 years ago, so it's a puzzle box), so yes, I've been using "load-bearing" as a shorthand for decades, reduced to be less jargon-y by adopting "structurally essential" for increasingly international team composition, where English is a second or third language.<p>However, I've been <i>hearing</i> "load-bearing" at least two orders of magnitude more often over the last few months, particularly after uncorking Claude Code for the team.<p>I don't think it's a dead give-away of AI usage, and I don't think AI usage is a problem.  I just think we can introduce phrases into common use by having them be used by common tools.  So let's train the models on obscure/archaic terms and see what happens. Heck, we can just prompt it...</p>
]]></description><pubDate>Sun, 26 Jul 2026 17:34:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49060328</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49060328</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49060328</guid></item><item><title><![CDATA[New comment by chaboud in "The new rules of context engineering for Claude 5 generation models"]]></title><description><![CDATA[
<p>I have heard so many of my co-workers use "load-bearing" over the last couple of months.  It's truly comical.  Maybe this is a way that we can make "fetch" happen.</p>
]]></description><pubDate>Sun, 26 Jul 2026 05:25:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49054993</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49054993</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49054993</guid></item><item><title><![CDATA[New comment by chaboud in "Claude Opus 5"]]></title><description><![CDATA[
<p>I don't think it is. It's really easy to get a persistent, clever, hacky model to dial that down a bit and just come back to chat before stomping off into the woods.<p>If a model couldn't ever do that in the first place, it'll just get stuck.<p>I work in the "ZeroOne" space, working on concepts and prototypes for things that don't exist in market yet.  Sometimes these models crank hard and immolate tokens while grounding themselves on expensive-to-ingest self-developed frameworks.  If the results are well judged and the crank-turn latency is low, I'm okay with the cost as long as the model isn't wasting <i>my</i> time.<p>But when I want to do more boilerplate work, I turn down the model and thinking level and get more traditional about restraining action.  For the really hard stuff, I reach for the models that will start a token bonfire in the back yard.</p>
]]></description><pubDate>Sat, 25 Jul 2026 21:13:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49051638</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49051638</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49051638</guid></item><item><title><![CDATA[New comment by chaboud in "Freeze-Casting"]]></title><description><![CDATA[
<p>It's a carrier medium... and it's pretty thick. I would have just committed to "slurry".<p>Sandy water at the beach that looks a bit brown?  Suspension.<p>Mining run off that looks like soup and coats everything?  Slurry.<p>What's the Half & Half between milk and cream but for tweaky techniques for fabrication?  Slurspension?</p>
]]></description><pubDate>Fri, 24 Jul 2026 05:12:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49031431</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=49031431</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49031431</guid></item><item><title><![CDATA[New comment by chaboud in "Claude Code: Anatomy of a Misfeature"]]></title><description><![CDATA[
<p>(Edit: BTW, if you haven't, read Nudge by Thaler and Sunstein.  It's a somewhat long-winded exploration on choice architectures, but you're neck deep in that space now.  I'll give you my copy if you're in SF.)<p>Thanks for the openness.  I got bit by this one and was, frankly, pretty surprised.<p>The funny thing about user-facing interaction mechanics is that everyone is part of some minority, and everyone comes with their own sense of what "natural" or "obvious" is.  With something this impactful, communicated clarity of behavior will important.  Your feature is also doing double-duty, serving as a last net against prompt-injection attacks by giving the user the final say.<p>(Also, BTW, folks outside of Anthropic are unlikely to be as tooled-up for long-running unsupervised Claude jaunts as you guys.  The cost of wild success is wide adoption.)<p>One thing I'll suggest is that the mechanics of permissions and asking are presently pretty hacker/nerd friendly but simultaneously too-scary and not-scary-enough for non-coders.<p>Examples:<p>- Wild-cards on always approve is awesome, but, with prefixes like timeout and nohup, the "thing" that is getting done is buried and largely unexplained to the user.<p>- Auto is actually kind of a sweet spot (sometimes goes off into the weeds), but the designers and PM's I've been working with might as well YOLO.  They have no idea if they're breaking things, but they gravitate between plan and auto mode.<p>- Fewer permission prompts is great, but it comes <i>after</i> a user has slogged through generation of a data-set to work against, like battle scars for paper cuts.  It's the thermostat problem.  The signal comes when the user is uncomfortable.  And it's a way to learn me, but not me <i>now</i>.<p>I've had good fortune with Opus 4.8 and Fable just telling the system what phase of my life it's in.  Things like "I'm going to go make dinner... Go profile the matrix or configurations and build the dataset for the next two hours while I'm away" have a pretty good hit rate.  On the flip side "keep me in the loop and bring me your results before making structural changes" also articulates well with Fable.  It will tread more carefully.<p>And these approaches are the ones we'd use with someone transitioning from SDE1 to SDE2.  A little more autonomy, and the grounding in the bigger picture.  Can we eventually translate to perfectly judging what the user wants in the moment based on incomplete signal?<p>No, but I'm glad you're trying.  Keep the interaction model clear to your broad set of users, and we'll come along for the ride.</p>
]]></description><pubDate>Fri, 17 Jul 2026 16:56:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=48949542</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48949542</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48949542</guid></item><item><title><![CDATA[New comment by chaboud in "LARP – Revenue infrastructure for serious founders"]]></title><description><![CDATA[
<p>Clicks on link.  Reads this thinking it's gotta be a joke... continues reading not sure if it's a joke.  Relieved to find out it's a joke.<p>I guess that's the joke.<p>Heavies have been playing this "pretend internet money" game off and on for decades.  In investment analysis, details matter.</p>
]]></description><pubDate>Mon, 13 Jul 2026 00:24:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=48886356</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48886356</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48886356</guid></item><item><title><![CDATA[New comment by chaboud in "Since Chromium 148, Math.tanh is now fingerprintable to link underlying OS"]]></title><description><![CDATA[
<p>I'd rather penalize the application than the technique.  Windows was rumored to long have "quirks" that would do better things for apps that had bugs that the OS ended up fixing instead of the app.<p>Javascript systems have long had polyfills for varied browser feature comparability gaps.<p>Whether you agree with these, making probing detection via fingerprinting illegal would take away this lever. Making surreptitious tracking via fingerprinting illegal?  Even for state actors?<p>Yeah, that's probably reasonable.  If someone is going to wear a tracking collar in exchange for "free" services, a little disclosure makes sense.</p>
]]></description><pubDate>Sun, 12 Jul 2026 21:45:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=48885139</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48885139</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48885139</guid></item><item><title><![CDATA[New comment by chaboud in "The short leash AI coding method for beating Fable"]]></title><description><![CDATA[
<p>This could also read as "how to be a horrible people manager for junior engineers".<p>Techniques that work for inexperienced engineers with high ability but limited judgment often work well with agentic coding systems.<p>- Give them clarity of purpose. Why are they doing what they're doing?<p>- Make explaining it back to you part of the job.<p>- Give them two-way doors.  Make mistakes reversible.<p>- Put effort into thoughtful refactoring as an actual sub-task instead of just accepting piled on hacks.<p>- Make your operating rules crisp and make sure they store them in their memories.<p>- Be accountable for their work.  It's not okay to crank out AI Slop and then say "Claude's fault".<p>We're all Software Development Managers now.<p>So, micromanage the LLMs if you want to, but you'll be missing out on chances to improve them for your purposes and, more importantly, to improve <i>yourself</i> as a manager.</p>
]]></description><pubDate>Sat, 04 Jul 2026 06:44:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=48783182</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48783182</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48783182</guid></item><item><title><![CDATA[New comment by chaboud in "Claude-real-video － any LLM can watch a video"]]></title><description><![CDATA[
<p>It's going to make itself unavailable again.  Actually... that's probably a litmus test for sentience.</p>
]]></description><pubDate>Fri, 03 Jul 2026 06:37:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=48771632</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48771632</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48771632</guid></item><item><title><![CDATA[New comment by chaboud in "Tokenmaxxing is dead, long live tokenmaxxing"]]></title><description><![CDATA[
<p>This is more likely the junior camper version of "not everything that counts can be counted, and not everything that can be counted counts."<p>In the early days of LLMs, we saw the classic hype-driven bi-modality of opinions.  Folks were in the "fake news, fad" camp, or they were in the "omg, take over the world" camp.<p>Those of us closer to the space, with the awareness to know that there was some truth (and a lot of misjudgment) to go around, were in the middle of nowhere. When I co-wrote some driver code with Chat GPT, other engineers (and even one of our directors) told me to keep it quiet.  At the same time I had directors and VPs asking me how we could accelerate adoption.  For a while, I had access to a cheat code just because I had the audacity to not ask for permission.  Folks were sure I would get in trouble for spending thousands per month in LLM operation, but a handful came along for the ride, burning tokens like firewood and learning along the way.<p>Tokenmaxxing is probably coming from at least a few things:<p>1. A course-correction for the practiced frugality that kept folks from jumping in and just learning at the ragged edge.<p>2. A willful and deliberate recognition that the best innovations in the later phases of a disruptive introduction often come from sparks of ideation in concentrations of activity. In other words, we don't know where good is, and we need to find it.  (Charitable interpretation from the article)<p>3. Recognition that, even if they don't know why, leaders and product owners will get punished for not jumping in and, because of bullets 1 and 2, won't get punished for trying and missing. Even if they have no idea what they're doing, they're going to fake it until they make it (or slide into another job).<p>This last set is where the pain lives. An organization with healthy and increasing AI tool
usage will see elevated token counts, but so too will one using LLMs to rewrite wikipedia articles without the letter "m" to keep token counts high.  These are pathological behaviors brought on by conflated metrics.<p>We had discussions about this in the early LLM days, where my old team was looking to ship new capabilities for older products. There was a lengthy VP-level discussion about getting to "80% usage" of the new system vs the old.  Because the new system was a superset of the old, I eventually said "we can do that immediately, but it's a <i>cost</i> goal, where we're just aiming to make our business more expensive to operate, rather than a <i>value</i> goal for our users".  We didn't adopt the target, but folks were understandably frustrated that they didn't have a straightforward way to measure and report progress.<p>Tokenmaxxing is, inevitably, a conflated goal, but it's what we have right now. Take advantage of the moment, learn, build, and keep an eye on levers for efficiency.</p>
]]></description><pubDate>Mon, 29 Jun 2026 00:40:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=48713395</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48713395</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48713395</guid></item><item><title><![CDATA[New comment by chaboud in "Anonymous GitHub account mass-dropping undisclosed 0-days"]]></title><description><![CDATA[
<p>I don't want to "see" any of it...</p>
]]></description><pubDate>Sat, 27 Jun 2026 17:44:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=48700177</link><dc:creator>chaboud</dc:creator><comments>https://news.ycombinator.com/item?id=48700177</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48700177</guid></item></channel></rss>