<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: abalashov</title><link>https://news.ycombinator.com/user?id=abalashov</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sun, 27 Sep 2026 06:56:46 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=abalashov" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by abalashov in "One Month Without AI"]]></title><description><![CDATA[
<p>Sadly, a great deal of applied business programming in modern capitalism doesn't give you the option of enjoying the creative process.<p>In effect, everything is classified as: "low value work with dead lines, where the customers don't really care about the result either."</p>
]]></description><pubDate>Sat, 26 Sep 2026 16:03:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49857809</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49857809</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49857809</guid></item><item><title><![CDATA[New comment by abalashov in "One Month Without AI"]]></title><description><![CDATA[
<p>I've found it helpful to think of models as precisely what they are: bundles of probabilistic textual associations, convolutions of their training, more search engine than "intelligence".<p>Just having that mental paradigm will lead to more correct uses, and avoid a lot of the most ones most likely to lead one down the primrose path. This cultural moment of SF AI psychosis will pass, but we'll be stuck with the resulting code for a long time.<p>Also, +1 to the folks saying that code is a liability. More is almost never better. Aim for succinctness and brevity. That cuts directly against the thrust of LLMs' tendency, but will give you plenty to curate.</p>
]]></description><pubDate>Sat, 26 Sep 2026 15:56:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49857727</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49857727</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49857727</guid></item><item><title><![CDATA[New comment by abalashov in "The current balance of power in open models"]]></title><description><![CDATA[
<p>> which AWS supports, and doesn't share any data with Anthropic<p>Ah... oh.<p>Well, it's a nice thought.</p>
]]></description><pubDate>Wed, 23 Sep 2026 13:01:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49815424</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49815424</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49815424</guid></item><item><title><![CDATA[New comment by abalashov in "People hooked on vapes try a new way to quit: cigarettes"]]></title><description><![CDATA[
<p>Yup. This was my main takeaway from vaping in the 2010s, after I 'quit' smoking.<p>However, the vast majority of my nicotine usage after quitting was chewing nicotine gum -- for 11 years! It also has the qualities you mention. You can be chewing all the time, anywhere, everywhere, and without any of the vaping restrictions places have.<p>Nicotine gum isn't for everyone, but if you're one of the people for whom it really takes as an addiction of its own, it's that much more difficult to quit because it's discreet, there's essentially no stigma, and you can do chew absolutely anywhere, doing practically anything.</p>
]]></description><pubDate>Wed, 23 Sep 2026 12:59:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49815394</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49815394</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49815394</guid></item><item><title><![CDATA[New comment by abalashov in "The current balance of power in open models"]]></title><description><![CDATA[
<p>Yeah, that's broadly accurate, although details matter and local models are surprisingly capable. Local models are viable for a lot of small use-cases, but no, there's no deluding oneself that frontier quality doesn't require frontier size. However, everyone who thinks they need frontier "intelligence" for internal CRUD type tasks should look in the mirror and ask themselves if that's really true.<p>Still, even if you need other people's industrial-class hardware, open models offer a lot more freedom and options. They allow you to use GPU capacity from entities who are not themselves building or training models and are ostensibly disinterested.<p>There's a big range of possibilities here in terms of data sovereignty and so forth.<p>1) You can use OpenRouter to route your open model requests to US-based inference providers with ZDR (zero data retention), as far as you can believe anything in this world. If you look at who actually serves open Chinese models on OpenRouter, you'll see a lot of folks like Digital Ocean, etc. I suppose I can't vouch for their purity, no, but I'd much rather send data there than send it to Dario.<p>2) Or, you can rent GPUs from companies like Runpod or Vast.ai and serve some very sizable models to yourself (e.g. using their pre-built vLLM images). You can't serve a model like Kimi K3 to yourself that way, at least not in any economically reasonable way. However, a private H100SXM or B200 can go a long way. You could serve the big Qwens, or DeepSeek-V4-Flash--you could do a lot if you're willing to spend on a rented GPU with sizable VRAM.<p>3) Finally, if you have and want to spend $750K-$1MM+ (I suspect I'm low-balling at this point), and if you can get them in the current demand climate, you can absolutely buy 16 x H200SXMs, with the appropriate boards to take them, pay for 20-25 kW of cooling, etc., and run one of these models yourself, on your kitchen floor if you like. You simply cannot do that with Anthropic or OpenAI.</p>
]]></description><pubDate>Wed, 23 Sep 2026 12:55:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49815355</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49815355</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49815355</guid></item><item><title><![CDATA[New comment by abalashov in "The current balance of power in open models"]]></title><description><![CDATA[
<p>Data points from my own usage, for whatever they're worth[1]. I'm just an individual contributor and hardly represent enterprise or large orgs:<p>I've been on Kimi, with a little DeepSeek-V4-Pro, GLM 5.2/5.3, and MiMo thrown in, for probably about a year now. It's great here!<p>1) For DeepSeek, I recommend their Reasonix harness strongly, due to its alignment to DeepSeek's prefix cache. It means mostly (95%+) cache hit input tokens, so very cheap large-scale code analyses and things that require mega context windows (at the cost of some attentional drift, yes). Reasonix does require that you send data to China. This is fine. I mostly use this for big, expansive ingestion of open-source codebases to figure out how something really works, usually something that documentation doesn't quite reach.<p>The economy of doing it this way versus American frontier model companies' token pricing cannot be overstated. I think I topped up $10 in June (2.5 months ago) and have still not burned through it, despite cycling untold tens of millions of tokens through it.<p>My biggest annoyance is that DeepSeek does seem to be considerably rate-limited of late, at least during working hours in Beijing, which is a range that I gather to be quite expansive there. I'm not blasting it with anything, I'm just noting that the agent takes 10-20 minutes to do stuff that takes much less time if I'm willing to pay the OpenRouter premium.<p>2) For most everyday stuff outside of where Reasonix + DeepSeek just makes overwhelming sense, I use OpenCode/Maki/Pi/whatever harness I feel like using today with Kimi K3, via OpenRouter. This does not require sending data to China.<p>I also use Kimi K3 in Zed via OpenRouter quite a bit, but sometimes like to mix it up with the other models.<p>3) Because I have the most experience with it, I can say with confidence that I would generally consider the SWE capabilities of Kimi to be on par with Claude, at least for the bottom 99% of purposes--and certainly, any routine business programming.<p>I think this has been true for a long time, well before K3. I've been using Kimi since K2.5.<p>4) For local hardware experiments on my MacBook Pro (M4 Max, 128 GB unified memory), Qwen3.6-35B-A3B (speed) and Qwen3.8-27B (intelligence, but slow). As has been widely noted, this amount of unified memory isn't as useful as it seems, due to memory bandwidth and decoding constraints, lack of tensor cores (on the M4 Max, anyway), etc.<p>A giant bag of memory isn't fast, but it'll let you load some impressively big models.<p>The future M5 Studio Macs will continue in this general vein, but will of course be somewhat faster, particularly due to the apparition of tensor cores in the M5 -- excuse me, "Neural Accelerators".<p>Still, if you really want to cook, get a real GPU. Real GPU running quantisations is still a lot better than a big slow bag of unified memory.<p>5) Overall, the Chinese models are simply excellent, and cater to lots of use-cases and tastes. However, I'll still tend to use Claude ($20/mo subscription) for general Q&A, whether of a technical nature or otherwise, particularly where web research and worldly knowledge is required.<p>[1] Disclaimer: this comment is an elaboration of <a href="https://news.ycombinator.com/item?id=49809605">https://news.ycombinator.com/item?id=49809605</a>, which is not something I'd normally do. However, it seems a lot more relevant here than where I had originally posted it, in an article about gauging MiMo Pro v2.6 capabilities.</p>
]]></description><pubDate>Wed, 23 Sep 2026 12:40:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=49815221</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49815221</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49815221</guid></item><item><title><![CDATA[New comment by abalashov in "MiMo-v2.6-Pro: Intelligence, Performance and Price Analysis"]]></title><description><![CDATA[
<p>Welcome to the dark side. I've been on Kimi, with a little DeepSeek-V4-Pro, GLM 5.2/5.3, and MiMo thrown in, for probably about a year now. It's great here!<p>For DeepSeek, I recommend their Reasonix harness strongly, due to its alignment to DeepSeek's prefix cache. It means mostly (95%+) cache hit input tokens, so very cheap large-scale code analyses and things that require mega context windows (at the cost of some attentional drift, yes). Reasonix does require that you send data to China.<p>For most everyday stuff outside of where Reasonix + DeepSeek just makes overwhelming sense, I use OpenCode/Maki/Pi/whatever harness I feel like using today with Kimi K3, via OpenRouter. This does not require sending data to China.<p>I also use Kimi K3 in Zed via OpenRouter quite a bit, but sometimes like to mix it up with the other models.<p>For local hardware experiments on my MacBook (128 GB unified memory), Qwen3.6-35B-A3B (speed) and Qwen3.8-27B (intelligence, but slow). As has been widely noted, this amount of unified memory isn't as useful as it seems, due to memory bandwidth and decoding constraints, lack of tensor cores (on the M4 Max, anyway), etc. A giant bag of memory isn't fast, but it'll let you load some impressively big models. The future M5 Studio Macs will continue in this general vein, but will of course be somewhat faster, particularly due to the apparition of tensor cores in the M5 -- Neural whateverApplecallsthem.<p>The Chinese models are simply _excellent_, and cater to lots of use-cases and tastes.</p>
]]></description><pubDate>Tue, 22 Sep 2026 23:17:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49809605</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49809605</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49809605</guid></item><item><title><![CDATA[New comment by abalashov in "If AI coding is lowering your code quality, you're not managing quality right"]]></title><description><![CDATA[
<p>I'm not sure how literally you mean "most people". This might be true in a purely volumetric sense, but that's not really the bar around these parts...</p>
]]></description><pubDate>Mon, 21 Sep 2026 13:32:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49787082</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49787082</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49787082</guid></item><item><title><![CDATA[New comment by abalashov in "If AI coding is lowering your code quality, you're not managing quality right"]]></title><description><![CDATA[
<p>> I know my codebase well<p>I'll bet you know it because you wrote and/or worked on it manually, likely over a period of years. The odds of you knowing a slop codebase that well, or even particularly at all, are much lower.</p>
]]></description><pubDate>Mon, 21 Sep 2026 13:21:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49786955</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49786955</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49786955</guid></item><item><title><![CDATA[New comment by abalashov in "How to Write with an LLM"]]></title><description><![CDATA[
<p>It seems to me if you're thinking about this for more than 500 milliseconds, maybe you should just write it yourself.</p>
]]></description><pubDate>Sun, 20 Sep 2026 03:40:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49772332</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49772332</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49772332</guid></item><item><title><![CDATA[New comment by abalashov in "Cloudflare Quick Tunnels"]]></title><description><![CDATA[
<p>Ah, an ngrok competitor. It's what we elderly (~40 myself), yelling-at-clouds types call "reverse SSH tunnels", but this is an anachronism from the era of knowing how to do things rather than paying for managed SaaS products to do them. Don't mind me, just yelling at clouds...</p>
]]></description><pubDate>Sat, 19 Sep 2026 14:49:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49767009</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49767009</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49767009</guid></item><item><title><![CDATA[New comment by abalashov in "We must pace the frontier"]]></title><description><![CDATA[
<p>> It’s a beautiful Saturday morning with my family here in the East Bay - I don’t know how many more Saturdays I’ll be able to play outside with my kids in the sunshine so I’m going to make the most of the time we have.<p>I don't know how else to say this. Put... the... peace pipe... down!</p>
]]></description><pubDate>Sat, 12 Sep 2026 17:50:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49675055</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49675055</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49675055</guid></item><item><title><![CDATA[New comment by abalashov in "We must pace the frontier"]]></title><description><![CDATA[
<p>Allow me to recommend switching to Kimi K3 / GLM 5.3 / DeepSeek. I use all three for the better part of a year now (DeepSeek via the Reasonix harness) and haven't touched Claude Code in at least as long.</p>
]]></description><pubDate>Sat, 12 Sep 2026 17:48:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49675021</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49675021</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49675021</guid></item><item><title><![CDATA[New comment by abalashov in "We must pace the frontier"]]></title><description><![CDATA[
<p>The best take is, of course, from Jason Gorman, who, when the fearsome "capabilities" of Mythos originally dropped, said this:<p>"Claude Mythos is that guy down the pub who is so good at karate that if he used it on you, you'd die instantly, and that's why you'll never see him using karate.”<p>In this case, it's more: "I'm having to act with great restraint because my karate is so good. Everyone should do likewise."</p>
]]></description><pubDate>Sat, 12 Sep 2026 17:44:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49674977</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49674977</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49674977</guid></item><item><title><![CDATA[New comment by abalashov in "Has the hallucination problem in AI been solved?"]]></title><description><![CDATA[
<p>The problem hasn't been solved, and, by the very nature of what LLMs are, can't be. However, it has gotten a lot better, as models have got much larger and pretraining more sophisticated.<p>Part of the difficulty--not in solving, but in discussing--is in defining what a hallucination is. On the face of it, it seems straightforward: an obviously counterfactual claim or manifest error of reasoning. However, it's not always that simple. A lot of what people consider to be hallucinations are misattributions, specious diagnoses, strangely lopsided preoccupations, eccentric design choices, needlessly verbose or circuitous output or explanations, a kind of metaphysical conflation of the trivial with the significant, etc.</p>
]]></description><pubDate>Sun, 16 Aug 2026 06:17:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49317358</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49317358</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49317358</guid></item><item><title><![CDATA[New comment by abalashov in "Our position on open-weights models"]]></title><description><![CDATA[
<p>Market failure! Market failure is the term you're looking for.<p>Also see: childcare, healthcare, many forms of public transport and social infrastructure. Some things markets simply cannot deliver well, and that's okay.<p>Daycare is the most easily reachable example because it's quite simple compared to the others. Daycare workers are simultaneously some of the lowest paid workers in the US, yet daycare costs are famously high and prohibitive, yet childcare centres are very far from money-printing machines. Margins in the daycare sector are most commonly < 1%.</p>
]]></description><pubDate>Tue, 28 Jul 2026 16:43:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49086559</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49086559</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49086559</guid></item><item><title><![CDATA[New comment by abalashov in "US Government targets Cop City protester over phone operating system"]]></title><description><![CDATA[
<p>Evidence of what?</p>
]]></description><pubDate>Mon, 27 Jul 2026 18:12:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49073457</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49073457</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49073457</guid></item><item><title><![CDATA[New comment by abalashov in "US Government targets Cop City protester over phone operating system"]]></title><description><![CDATA[
<p>I wish I shared your optimism. For the sake of the accused, I hope you're right, but I could see them really doubling down on this.</p>
]]></description><pubDate>Mon, 27 Jul 2026 18:12:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49073448</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49073448</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49073448</guid></item><item><title><![CDATA[New comment by abalashov in "How is the Bun rewrite in Rust going?"]]></title><description><![CDATA[
<p>Oh, I 100% agree. I'm imagining the case where someone asks an LLM to do trivial things just for the novelty, or because they are fatigued of thinking + typing, but conceptualise the thing they are putting off as primarily a problem of typign and not thinking (I often do that).</p>
]]></description><pubDate>Mon, 27 Jul 2026 13:15:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49069279</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49069279</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49069279</guid></item><item><title><![CDATA[New comment by abalashov in "How is the Bun rewrite in Rust going?"]]></title><description><![CDATA[
<p>Based on the way they live-tweeted (as it were) their rewrite and their other publicity stunts? Yes, actually.</p>
]]></description><pubDate>Mon, 27 Jul 2026 13:11:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49069215</link><dc:creator>abalashov</dc:creator><comments>https://news.ycombinator.com/item?id=49069215</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49069215</guid></item></channel></rss>