<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: rnxrx</title><link>https://news.ycombinator.com/user?id=rnxrx</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 09 Oct 2026 04:32:56 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=rnxrx" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by rnxrx in "Why isn't the industry freaking out about DeepSeek 4.1 Flash?"]]></title><description><![CDATA[
<p>It depends hugely on what "rent usage of this model through one of the many LLM hosting providers" means.  If you're asking them to host the model privately then yes, all of that 1.6T of RAM is likely in use holding weights, activations and KV cache by an inference engine that's only getting/answering requests from you alone.  When you aren't actively using the model the hosting process is still active and waiting with all of that memory still wired to it.<p>As background: For the most part VRAM oversubscription/paging/swapping isn't a thing in the same way that RAM for a VM often is.  There <i>are</i> some approaches to it, but (to my knowledge) not at that sort of scale.<p>There are some systemic reasons for this, but very broadly speaking the GPU vendors are building toward the highest bandwidth and lowest latency possible, and the overhead/complexity of something like protected memory modes serves neither of those priorities.</p>
]]></description><pubDate>Thu, 08 Oct 2026 21:49:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=50012857</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=50012857</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=50012857</guid></item><item><title><![CDATA[New comment by rnxrx in "I could've accessed 17T Microsoft records"]]></title><description><![CDATA[
<p>$5K seems like an absolute steal compared to paying contracted security experts to find such bugs.  I'm surprised these programs aren't pushed harder, as the potential ROI seems fantastic.</p>
]]></description><pubDate>Thu, 01 Oct 2026 02:33:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49917012</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=49917012</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49917012</guid></item><item><title><![CDATA[New comment by rnxrx in "Gemini 4 Argon"]]></title><description><![CDATA[
<p>I do something similar - as I approach the context limits I have a pre-compact flush skill that extracts anything useful from the context, updates the MEMORY.md and my Obsidian vaults (set up as a poor man's graph DB) and so forth.  Once everything's been stored I run /compact to keep the general session flow intact.  Recently I added a small embedder/vector search setup to the same skill, which seems promising so far.<p>On another environment I've been doing something roughly similar, but have integrated Hindsight as a kind of all-in-one of the above and am still trying to suss out the best compaction strategy.</p>
]]></description><pubDate>Thu, 01 Oct 2026 02:20:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49916922</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=49916922</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49916922</guid></item><item><title><![CDATA[New comment by rnxrx in "MiMo v2.6"]]></title><description><![CDATA[
<p>The cost issue is obviously of prime importance, but I'd also argue that the transparency of the innovations creates a tremendous cross-pollination, and not only within the Chinese communities but in the US/Europe as well.  How many of us are learning the practical aspects of actually running and building AI based primarily on open models?  As an example - how far would the work of vLLM or SGLang or even NVIDIA itself (all random examples) be without these models and the challenges they pose?</p>
]]></description><pubDate>Tue, 22 Sep 2026 00:50:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49795475</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=49795475</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49795475</guid></item><item><title><![CDATA[New comment by rnxrx in "Show HN: Docx-CLI: agents read/edit Word docs using 1/2 the time and tokens"]]></title><description><![CDATA[
<p>This is great - and another example of how much more efficient CLI tool use ends up being in actual day-to-day use.  Claude Code and Hermes took it in and it runs great in my initial tries at it.  Thanks for making and sharing it!</p>
]]></description><pubDate>Tue, 07 Jul 2026 23:14:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=48825299</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48825299</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48825299</guid></item><item><title><![CDATA[New comment by rnxrx in "Local, CPU-Friendly, High-Quality TTS (Text-to-Speech) with Kokoro"]]></title><description><![CDATA[
<p>Another endorsement - I used Kokoro pretty extensively with an app I was developing over the last year and it's been excellent, both on- and off- GPU.  Even with Elevenlabs (long time subscriber) the comparative quality of Kokoro keeps up really well until you get to their larger models with their professional voices.<p>I do wish there were better support for SSML, as well as deeper documentation of how to influence inflection in-line, but the default does well with standard emphasis (e.g. putting asterisks around text elements).  Both asks are getting outside the zone of reasonable asks for this sort of distribution, though, and I remain incredibly grateful for the quality of what hexgrad and nazdridoy have put out in the world.</p>
]]></description><pubDate>Tue, 07 Jul 2026 23:07:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=48825228</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48825228</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48825228</guid></item><item><title><![CDATA[New comment by rnxrx in "OpenWrt One – Open Hardware Router"]]></title><description><![CDATA[
<p>It's worth being careful here: a lot of the affordable enterprise-class routers from 10+ years ago aren't as fast as cheap consumer hardware - like the OP or just a decent mini PC.  The primary arguments for the enterprise-class gear are around feature availability and certain aspects of reliability (redundant power/fans, better heat tolerance, higher quality components). It's also worth remembering that this kind of gear tends to be built for dedicated environments: loud fans, higher power draw/lack of power saving features, etc.<p>Beside the potential performance and environmental issues the other big downsides tend to include firmware availability - either because download from the site requires a login on the vendor's site or, increasingly commonly, the gear has hit LDOS and images just aren't posted.  Obviously there are other "unofficial" places for such images, but the risk/legality are a whole other (potentially serious) question.<p>There's an additional issue mapping the requirements of home networking to enterprise gear: Ethernet switches are lousy firewalls (little or no NAT, primitive built-in security, DLNA/mDNS and friends aren't really sane options, etc).  Finally, even at "1/5" the price the gear may still be quite a bit more expensive than other options. And if it's <i>not</i> expensive, it's usually because nobody wants it any more because of the issues mentioned above (being near- or beyond- LDOS).<p>FWIW this is from someone who literally built a commercial-class machine room in his house with dedicated AC, subpanels, commercial UPS, etc for data center class Ethernet, Fibre Channel and Infiniband switching as well as carrier-grade routers and still runs "enterprise" grade WLAN and switching and can lay hands on as much as I could reasonably want without too much drama or cost...  Going down this road can absolutely be amazing if you either have a.) the background to properly source and run the hardware/software or b.) have a driving desire to learn how to do so or c.) have some very atypical requirements for home networking.  Otherwise it tends to not be something to be done lightly.</p>
]]></description><pubDate>Mon, 06 Jul 2026 23:52:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=48812011</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48812011</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48812011</guid></item><item><title><![CDATA[New comment by rnxrx in "Qwen 3.6 27B is the sweet spot for local development"]]></title><description><![CDATA[
<p>It depends on what's meant by "fully utilized" but fp8 quants of Nemotron 3 Super, the latest Minimax, Cohere A+ and the Mistral small and (especially) medium variants all sit in that 128-256 category, especially with full context or even moderate concurrency. In fact, in a 192GB environment I work with (Hopper GPUs, fwiw) I was pushed into using 4-bit quants with a couple of those to get the model working with a reasonable context window (..but 256 would have rocked out).</p>
]]></description><pubDate>Mon, 29 Jun 2026 21:10:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=48725280</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48725280</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48725280</guid></item><item><title><![CDATA[New comment by rnxrx in "Qwen 3.6 27B is the sweet spot for local development"]]></title><description><![CDATA[
<p>There are also nvfp4 quants of Qwen 3.6 27/35 floating around.  I've done benchmarks of both and the quality difference vs fp8/bf16 was barely notable.  Honestly the nvfp4 capability is the most interesting feature of the Spark (at least for me).</p>
]]></description><pubDate>Mon, 29 Jun 2026 21:03:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=48725190</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48725190</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48725190</guid></item><item><title><![CDATA[New comment by rnxrx in "Show HN: OpenKnowledge – open source AI-first alternative to Obsidian/Notion"]]></title><description><![CDATA[
<p>There is at least one MCP server in Obsidian's community plugins, plus the REST API access capability which is already addressed in several open source MCP plugins.<p>I use Obsidian as a persistent context store and knowledge graph (..loosely defined, i.e. link/back-link) for both Claude Code and Hermes, while also using it to generate live Wiki pages for working documentation.  The native replication and the Git integrations work well keeping it all synchronized across multiple harnesses, as well.  I use the native MCP server mentioned above, plus just letting the agent work with the markdown files directly.<p>That said, having built out all of this manually I'm excited to try out something that addresses much of this out of the box.  I'd also be curious about the integration with Hermes/OpenClaw/etc.</p>
]]></description><pubDate>Thu, 25 Jun 2026 20:48:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48678992</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48678992</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48678992</guid></item><item><title><![CDATA[New comment by rnxrx in "Apple raises prices of MacBooks, iPads"]]></title><description><![CDATA[
<p>It *is* hard to make memory, especially HBM (...which is what the AI market wants, and is what the manufacturers are focusing on) and bringing on new capacity takes <i>years</i>.  There's the additional wrinkle that the manufacturers we have left are the ones who survived periods where the market was glutted with oversupply in the wake of previous shortages.<p>These decisions play out on the order of trillions of dollars and 3+ year horizons. They're also incredibly sensitive to other geopolitical issues (Taiwan, issues with Chinese tech capability vs export/import controls, etc).<p>There are a lot of valid discussions to be had about how we got to this state of oligopoly: Taiwan's consistent sponsorship of its semiconductor capabilities and the subsequent concentration of technology (expertise, capacity, etc), the lack of investment/support (and ceding of technical leadership) in Western countries, the various rivalries with China and the implications of it becoming a first-class producer of semiconductors at scale, etc.  None of those discussions and none of their potential outcomes can substantively change that we're going to continue in this situation (massive price increases, spotty availability, etc) for at least the next 18-24 months.</p>
]]></description><pubDate>Thu, 25 Jun 2026 16:36:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=48675917</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48675917</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48675917</guid></item><item><title><![CDATA[New comment by rnxrx in "45°C cooling design cuts data center water use to near zero"]]></title><description><![CDATA[
<p>It’s usually open loop - closed loop, so closed loop goes through CRACs or liquid cooled equipment manifolds.  That heated water circulates through an heat exchanger on the roof that uses open loop cooling to shed the heat to the surrounding environment.</p>
]]></description><pubDate>Wed, 24 Jun 2026 22:44:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48666488</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48666488</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48666488</guid></item><item><title><![CDATA[New comment by rnxrx in "The deadly rise of giant trucks and SUVs"]]></title><description><![CDATA[
<p>In terms of platform variants the Taurus outsold the Crown Vic by a fair amount.  The latter was certainly hugely popular as fleet vehicles, but the Taurus was Ford's best selling sedan for quite a while.  The Taurus was never rated for anything more than fairly light-duty towing (<= 1500lbs).<p>More to the point - at least for automatic transmission variants in the early 90s - even the base F150 could certainly out-tow Crown Vics of the time, unless the latter had a tow package, which would push the rating up to 5000 lbs (..same as the lower end F150).  Later on the delta was a lot greater (in the truck's favor) as the tow packages for the cars were phased out.<p>Nowadays the <i>minimum</i> ratings of F150s (even with shorter beds) is 5000lbs, or more.  The smaller CUV/SUV platforms are usually rated ~1500lbs *max*, with some of the larger truck-based models obviously running in the same range as the trucks.<p>TL;DR - Even back in the day the "typical" sedan was still only rated for towing small loads, unless specially ordered with towing packages.  Now..all that said, the frequency with which typical F150 owners ever actually tow <i>anything</i> is a whole other question.</p>
]]></description><pubDate>Tue, 23 Jun 2026 22:23:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=48652320</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48652320</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48652320</guid></item><item><title><![CDATA[New comment by rnxrx in "AI Hiring Tools Yield Racial Bias and Systemic Rejection; 26% Black & 15% Asian"]]></title><description><![CDATA[
<p>Genuine curiosity:  Is there any speculation as to what these tools are keying on to reject those particular applicants?  It seems like it just being the applicant's name is too easy an answer, but I could be overthinking it.</p>
]]></description><pubDate>Tue, 23 Jun 2026 21:48:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=48651939</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48651939</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48651939</guid></item><item><title><![CDATA[New comment by rnxrx in "San Diego photologs from the 1970s"]]></title><description><![CDATA[
<p>The commentary on hand-created signage was especially fascinating.  The observations about the use of computers "enshitifying" design sort of eerily echo a lot of the commentary about AI now, including the (not unfounded) fear of the loss of human inconsistency, and the beauty it can bring.</p>
]]></description><pubDate>Tue, 23 Jun 2026 18:22:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=48649146</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48649146</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48649146</guid></item><item><title><![CDATA[New comment by rnxrx in "Claude Fable 5"]]></title><description><![CDATA[
<p>I have the $100 plan and had almost never run out of credits until I started using the ultracode / workstreams feature w/Opus 4.8..at which point I managed to blow the full 6 hour allocation in like 20 minutes, or so.  In fairness, it did some amazing things with the extracted information, but it also strongly suggested that I'd need the $200 subscription *plus* a budget for extra usage.</p>
]]></description><pubDate>Tue, 09 Jun 2026 18:26:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=48465354</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48465354</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48465354</guid></item><item><title><![CDATA[New comment by rnxrx in "Albania Is Not for Sale: Kushner's $4B Resort Triggers'Flamingo Revolution'"]]></title><description><![CDATA[
<p>Maybe it’s naive, but there’s something incredibly hopeful that there are folks not only protesting this kind of corruption, but also that there’s a government actually responding to the voice of the people. That the EU’s own legal frameworks might positively (if indirectly) affect things is even better.</p>
]]></description><pubDate>Tue, 09 Jun 2026 16:45:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=48463569</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48463569</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48463569</guid></item><item><title><![CDATA[New comment by rnxrx in "Cleaning up after AI rockstar developers"]]></title><description><![CDATA[
<p>I don’t think the analogy between rock star developers and LLMs bears out here.  Like any other tool, AI abides by the basic reality of garbage in/garbage out.  If you don’t make sure the LLM has sufficient context then you get what you get.  Same point with vague prompts.<p>AI tools are producing code on behalf of a developer.  If that developer is fine putting their name on code that they don’t understand, applied to a code base they also don’t fully understand then you have a very human problem.  The technology just magnifies this.</p>
]]></description><pubDate>Tue, 09 Jun 2026 14:38:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=48461755</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48461755</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48461755</guid></item><item><title><![CDATA[New comment by rnxrx in "Fidonet: Technology, Use, Tools, and History (1993)"]]></title><description><![CDATA[
<p>I always knew the "point" was there, but never saw it actually used in my travels (..which finished in '90 or '91).  I seem to recall the node lists at some point didn't have decimals, but perhaps my recollection is inaccurate.</p>
]]></description><pubDate>Wed, 03 Jun 2026 02:06:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=48378981</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48378981</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48378981</guid></item><item><title><![CDATA[New comment by rnxrx in "Fidonet: Technology, Use, Tools, and History (1993)"]]></title><description><![CDATA[
<p>99:9008/206 / 1:137/206</p>
]]></description><pubDate>Wed, 03 Jun 2026 02:01:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48378947</link><dc:creator>rnxrx</dc:creator><comments>https://news.ycombinator.com/item?id=48378947</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48378947</guid></item></channel></rss>