<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: sdrinf</title><link>https://news.ycombinator.com/user?id=sdrinf</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 21 Jul 2026 23:51:20 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=sdrinf" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by sdrinf in "Show HN: Smol machines – subsecond coldstart, portable virtual machines"]]></title><description><![CDATA[
<p>hi, great project! Windows support is sorely lacking, though. As someone working a lot with sandboxed LLMs right now, the options-space on windows for sandboxing is _extremely lacking_. Any plans to support it?</p>
]]></description><pubDate>Fri, 17 Apr 2026 18:35:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=47809103</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=47809103</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47809103</guid></item><item><title><![CDATA[New comment by sdrinf in "Can I run AI locally?"]]></title><description><![CDATA[
<p>Just want to echo the recommendation for qwen3.5:9b. This is a smol, thinking, agentic tool-using, text-image multimodal creature, with very good internal chains of thought. CoT can be sometimes excessive, but it leads to very stable decision-making process, even across very large contexts -something we haven't seen models of this size before.<p>What's also new here, is VRAM-context size trade-off: for 25% of it's attention network, they use the regular KV cache for global coherency, but for 75% they use a new KV cache with linear(!!!!) memory-token-context size expansion! which means, eg ~100K token -> 1.5gb VRAM use -meaning for the first time you can do extremely long conversations / document processing with eg a 3060.<p>Strong, strong recommend.</p>
]]></description><pubDate>Fri, 13 Mar 2026 20:35:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=47369502</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=47369502</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47369502</guid></item><item><title><![CDATA[New comment by sdrinf in "California's Digital Age Assurance Act, and FOSS"]]></title><description><![CDATA[
<p>Counterpoint to peeps on this thread:<p>* This approach is the _most consistent_ with retaining anonymity on the internet, while actually helping parents with their issues. If any age-relevant gatekeeping needs to be made on the internet at all, this is the one I find acceptable.<p>* this is because the act very specifically does NOT require age _verification_ ie using third-parties to verify whether the claimed age is correct. Rather, it is piggybacking on the baked-in assumption, that parents will set up the device for their kids, indicating on first install what the age/DoB is, then handing over the device -a setting which can, presumably, only be modified with parental consent<p>* yes, there are edge cases, esp in OSS, and yes, it would be nice to iron those out -but the risk = probability x impact calculus on this is very very low.<p>* If retaining anonymity on the internet is of value to you, don't let the perfect be the enemy of good enough.</p>
]]></description><pubDate>Wed, 04 Mar 2026 04:47:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=47243189</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=47243189</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47243189</guid></item><item><title><![CDATA[New comment by sdrinf in "How will OpenAI compete?"]]></title><description><![CDATA[
<p>Taking the opposite side of that bet, here is why:<p>* even if an openweight model appears on huggingface today, exceeding SOTA, given my extensive experience with a wide variety of model sizes, I would find it highly surprising the "99% of use cases" could be expressed in <100B model.<p>* Meanwhile: I pulled claude to look into consumer GPU VRAM growth rates, median consumer VRAM went 1-2GB @ 2015 to ~8GB @ 2026, rougly doubles every 5 years; top-end isn't much better, just ahead 2 cycles.<p>* Putting aside current ram sourcing issues, it seems very unlikely even high-end prosumers will routinely have >100GB VRAM (=ability to run quantized SOTA 100b model) before ~2035-2040.</p>
]]></description><pubDate>Thu, 26 Feb 2026 08:33:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=47163466</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=47163466</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47163466</guid></item><item><title><![CDATA[New comment by sdrinf in "Ask HN: Has anyone achieved recursive self-improvement with agentic tools?"]]></title><description><![CDATA[
<p>I'm working on something like this. Specifically, I'm doing recursive self-improvement via autocatalysis -but predominantly in writing/research / search tasks. It's very early, but shows some very interesting signs.<p>The purely code part you described is a bit of an "extra steps" -you can just... vscode open target repo, "claude what does this do, how does it do it, spec it out for me"  then paste into claude code for your repo "okay claude implement this".    This sidesteps the security issue, the deadly trifecta, and the accumulation of unused cruft.</p>
]]></description><pubDate>Thu, 12 Feb 2026 07:14:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=46985724</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=46985724</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46985724</guid></item><item><title><![CDATA[New comment by sdrinf in "Discord Alternatives, Ranked"]]></title><description><![CDATA[
<p>can someone please try running the experiment of "but what if just forking&spinning up an OSS clone, scaling up to take in the migrants, acquire network effects, collect roughly same subscription revenue, but run on just, like, 10 people?"<p>Discord has a financially and politically vulnerable posture that is downstream of having to operate a very large team, raise funding, be exposed to investor market pressure. However, it is also one of the rare instances of successful consumer freemium subscription monetization. A clone does not have to pay the tuition of "what makes this specific space compelling, and want-to-pay-for"; it just have to _exists_, passively soaking up migrants from each platform shift.<p>ITT WTB 3rd place for my frens.</p>
]]></description><pubDate>Tue, 10 Feb 2026 06:41:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=46956140</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=46956140</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46956140</guid></item><item><title><![CDATA[New comment by sdrinf in "Claude is a space to think"]]></title><description><![CDATA[
<p>Besides the editorial control -which openai openly flagged to want to remain unbiased- there is a deeper issue with ads-based revenue models in AI: that of margins. If you want ads to cover compute & make margins -looking at roughly $50 ARPU at mature FB/GOOG level- you have two levers: sell more advertisement, or offer dumber models.<p>This is exactly what chatgpt 5 was about. By tweaking both the model selector (thinking/non-thinking), and using a significantly sparser thinking model (capping max spend per conversation turn), they massively controlled costs, but did so at the expense of intelligence, responsiveness, curiosity, skills, and all the things I've valued in O3. This was the point I dumped openai, and went with claude.<p>This business model issue is a subtle one, but a key reason why advertisement revenue model is not compatible (or competitive!) with "getting the best mental tools" -margin-maximization selects against businesses optimizing for intelligence.</p>
]]></description><pubDate>Wed, 04 Feb 2026 18:02:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=46889218</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=46889218</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46889218</guid></item><item><title><![CDATA[New comment by sdrinf in "GLM-4.7-Flash"]]></title><description><![CDATA[
<p>Note: I strongly recommend against using Novita -their main gig is serving quantized versions of the model to offer it for cheaper / at better latency; but if you ran an eval against other providers vs novita, you can spot the quality degradation. This is nowhere marked, or displayed in their offering.<p>Tolerating this is very bad form from openrouter, as they default-select lowest price -meaning people who just jump into using openrouter and do not know about this fuckery get facepalm'd by perceived model quality.</p>
]]></description><pubDate>Mon, 19 Jan 2026 20:39:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=46684187</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=46684187</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46684187</guid></item><item><title><![CDATA[New comment by sdrinf in "Ask HN: Has anyone deployed your own MCP server connector to ChatGPT?"]]></title><description><![CDATA[
<p>You have two options:<p>* Use it as a "source": chatgpt -> settings -> apps & connectors -> add it as your connector. This supports only 2 functions: search, and fetch; details: <a href="https://help.openai.com/en/articles/11487775-connectors-in-chatgpt" rel="nofollow">https://help.openai.com/en/articles/11487775-connectors-in-c...</a>    ; in business / edu version there is support for "full MCP mode": <a href="https://help.openai.com/en/articles/12584461-developer-mode-and-full-mcp-connectors-in-chatgpt-beta" rel="nofollow">https://help.openai.com/en/articles/12584461-developer-mode-...</a><p>* Enable "developer mode" chatgpt -> settings -> apps & connectors -> advanced settings -> developer mode. Available on paid&pro levels only. This can do full MCP access, but can't (currently) use your memory settings.<p>The option that works under all conditions is to use the API, and add it as a function directly (no MCP) -this works regardless what plan you have on openai.</p>
]]></description><pubDate>Mon, 27 Oct 2025 14:52:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=45721670</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=45721670</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45721670</guid></item><item><title><![CDATA[New comment by sdrinf in "Claude Fraud? Or Just an Anomaly?"]]></title><description><![CDATA[
<p>The specific "anomaly" is that claude 4 / opus model _does not know_ because it is _not in its' training data_ what its own model version is; AND because it's training data amalgamates "claude" of previous versions, the non-system-prompted model _thinks_ that it's knowledge cut-off date is April 2024.
However, this is NOT a smoking gun in different model serving. The web version DOES know because it's in its prompt (see full system prompts here: <a href="https://docs.claude.com/en/release-notes/system-prompts" rel="nofollow">https://docs.claude.com/en/release-notes/system-prompts</a> )<p>Specific repro steps: set system prompt to:
"Current date: 2025-09-28
Knowledge cut-off date: end of January 2025"<p>Then re-run all your tests through the API, eg "What happened at the 2024 Paris Olympics opening ceremony that caused controversy? Also, who won the 2024 US presidential election?"  -> correct answers on opus / 4.0, incorrect answers on 3.7. This fingerprints consistently correctly, at least for me.</p>
]]></description><pubDate>Sun, 28 Sep 2025 19:35:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=45407214</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=45407214</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45407214</guid></item><item><title><![CDATA[New comment by sdrinf in "FFmpeg moves to Forgejo"]]></title><description><![CDATA[
<p>I actually _like_ this, and so does the comfyweb & weebs who are a very significant portion of the driving force behind calm, decade-long projects.</p>
]]></description><pubDate>Sun, 17 Aug 2025 04:07:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=44928764</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=44928764</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44928764</guid></item><item><title><![CDATA[New comment by sdrinf in "Wikipedia loses challenge against Online Safety Act"]]></title><description><![CDATA[
<p>Follow-up question is big lulz:
<a href="https://yougov.co.uk/topics/society/survey-results/daily/2025/07/31/91334/3" rel="nofollow">https://yougov.co.uk/topics/society/survey-results/daily/202...</a><p>"And how effective do you think the new rules will be at preventing those younger than 18 from gaining access to pornography?"<p>-> 64% "not very effective / not at all effective"</p>
]]></description><pubDate>Mon, 11 Aug 2025 23:36:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=44870696</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=44870696</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44870696</guid></item><item><title><![CDATA[New comment by sdrinf in "Visa and Mastercard are getting overwhelmed by gamer fury over censorship"]]></title><description><![CDATA[
<p>This absolutely works... until, and when network effects kick in.<p>Payment processors have major network effects in that infra setup is expensive, banks need to be onboarded one-by-one, and whichever network has the most consumers, businesses will gravitate towards it. Iterate this over 20 years, and this always results in natural monopolies / duopolies. This creates a natural chokepoint/linchpin over which millions of people's mutually exclusive needs are getting banged at; including consumers at large, govs at large, and special-interest groups at large.<p>Absent crystal clear legislation -and porn is anything, but- this will always be arbitrary, and leave one side in the dust.</p>
]]></description><pubDate>Mon, 28 Jul 2025 19:03:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=44714191</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=44714191</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44714191</guid></item><item><title><![CDATA[New comment by sdrinf in "Ask HN: Hackathons feel fake now"]]></title><description><![CDATA[
<p>Early-40s here who still does all-nighters. How long is recovery time for you? What does it entails -ie what doesn't work as much as it should / takes longer while you recover?</p>
]]></description><pubDate>Mon, 05 May 2025 03:37:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=43891757</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=43891757</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43891757</guid></item><item><title><![CDATA[New comment by sdrinf in "Google does not want rights to things you do using Chrome (2008)"]]></title><description><![CDATA[
<p>Mozilla is sooo fucked here. On one hand, it would take them approx ~1 sentence of blog to say "We won't sell your input info to anyone" and this drama goes away.<p>OTOH: if the currently pending court case on anti-monopoly bars google from making payments to mozilla (which is about ~90%++ of their revenue), mozilla truly, and well is fucked. Meaning -they need to diversify, and they know it; they can't sell browsers, related services are heavily competed for, so ads & selling user data is broadly the only viable strat that can underwrite their existence.<p>Of course, the community won't have it. And therein lies the rub: by going with google's bribe, on this long term, they wrote themselves into a corner they can't exit.</p>
]]></description><pubDate>Sun, 02 Mar 2025 21:16:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=43235214</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=43235214</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43235214</guid></item><item><title><![CDATA[New comment by sdrinf in "Ask HN: Claude 3.5 Sonnet vs. o1 vs. <other> for coding. Let's talk!"]]></title><description><![CDATA[
<p>O1 for collabing on design docs, o1 for overall structure, break it into tasks per preference / sort; sonnet/o1 for executing each small tasks.<p>O1 is higher quality, more nuanced, and has deeper understanding; the biggest downside rn is the significantly higher latency (both due to thinking, and also, continue.dev doesn't support o1 streaming currently, so you're waiting until it's all done), and higher cost.<p>In terms of tools: either vscode with continue.dev / cline, or cursor<p>Languages: node.js / javascript, and lately c# / .net / unity</p>
]]></description><pubDate>Mon, 25 Nov 2024 02:32:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=42232685</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=42232685</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42232685</guid></item><item><title><![CDATA[New comment by sdrinf in "Ask HN: What are your favorite tools for taming your email?"]]></title><description><![CDATA[
<p>Cease and desist letters.<p>There are many, many people, and companies who operate under the false belief that the CAN-SPAM act does not apply to them; and eg create new mailing lists to blast many people with their spam. Some of these unfortunately includes corps I have business relationship with (looking at you, Google), so "mark as spam" doesn't work well. Cease and desisting their legal department does. I have changed marketing strat of multiple largecorps by being a dangerous professional.</p>
]]></description><pubDate>Tue, 19 Nov 2024 02:06:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=42179454</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=42179454</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42179454</guid></item><item><title><![CDATA[New comment by sdrinf in "Human drivers are to blame for most serious Waymo collisions"]]></title><description><![CDATA[
<p>Because hypotheticals have a process of moving into "stuff that's happening", on the timescale of years, decades.<p>Once it's starts happening, speak; if speaking doesn't work, fight; if fighting doesn't work, move. This works.</p>
]]></description><pubDate>Thu, 12 Sep 2024 03:24:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=41517285</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=41517285</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41517285</guid></item><item><title><![CDATA[New comment by sdrinf in "With more legal action on the horizon, how long before Archive.org closes?"]]></title><description><![CDATA[
<p>* IA's most important function (at least for me) is holding copies of the World Wide Web as it was.<p>* Given an annually compounding 30% linkrot, 99.92% of all the content ever published on the Internet is no longer available.<p>* This has been litigated, see Field v. Google Inc., 412 F (2006), and held to be "fair use" due to safe harbor of Section 512(b) of the DMCA<p>* This exemption does not apply to books, music, videos, or any of the other pirated material.</p>
]]></description><pubDate>Sun, 08 Sep 2024 22:45:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=41483886</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=41483886</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41483886</guid></item><item><title><![CDATA[New comment by sdrinf in "Google Chrome warns uBlock Origin may soon be disabled"]]></title><description><![CDATA[
<p>For the moment, the best possible solution seems to me simply disabling auto-updates. On long-term, if supermium can port over the critical fixes from chromium, ubo v2 may still survive with chrome-ish packaging.<p>For larger context, the ecosystem is fragmenting, and I have ~10 browser extensions that are critical to me. I don't think I will prioritize chrome's software cadence over my own preferences, thank you.</p>
]]></description><pubDate>Sat, 03 Aug 2024 17:40:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=41148007</link><dc:creator>sdrinf</dc:creator><comments>https://news.ycombinator.com/item?id=41148007</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41148007</guid></item></channel></rss>