<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: minraws</title><link>https://news.ycombinator.com/user?id=minraws</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 28 Aug 2026 09:01:01 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=minraws" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by minraws in "Nvidia projects $673B in sales as AI demand widens"]]></title><description><![CDATA[
<p>> how much demand is still gated behind cost constraints. The market for this is HUGE.<p>I think this misses the actual limits here.<p>The problem isn't demand it's, "how much people are willing to spend on it".<p>Cheap AI has to be served on cheap compute, and if inference gets cheap enough to unlock massive usage numbers, by definition it also doesn't require anywhere near as much infrastructure per unit of demand.<p>Take DeepSeek serving ~100T tokens/day, depending on workload and utilization, you're potentially talking about only a few thousand last-gen GPUs. With current-gen GPUs maybe closer to ~1,000, and with Rubin even fewer I will be damned if I could get my hands on one.<p>That's the part I think people are missing when they extrapolate token demand into enormous infrastructure or AI revenue.<p>Yes usage will explode. But if the cost per unit collapses, the revenue doesn't necessarily go up with it.<p>You can't simultaneously argue that intelligence becomes so cheap that everyone uses enormous amounts of it, while also assuming customers will somehow spend trillions of dollars a year consuming it.<p>There is no obvious $1T customer-facing AI revenue number at the end of this rainbow in the short/medium term.<p>The average person isn't going to spend anything remotely comparable to what they spend on a car every year for an AI service. Even businesses have budgets now, huge demand doesn't matter if the willingness to pay isn't there.<p>The only path I can see to numbers like that is AI consuming existing business domains, even then it's very thin.<p>Say SaaS + legal + consulting + BPO + various other service industries collectively represent something like $10-20T globally.<p>Even if AI eventually replaces an enormous portion of that, it's probably not doing so at the same price. Why would customers switch otherwise?<p>Either the AI product has to be dramatically better, which is difficult for mature workflows, or dramatically cheaper which is much more plausible.<p>If it replaces $10-20T of existing services at roughly 1/10th or 1/100th (more likely) the cost, then you're looking at maybe a ~$1T AI revenue opportunity after replacing an absurdly large fraction of the existing service economy.<p>Who are now unemployed and can't pay for shit.<p>And that's before competition.<p>I think it's crazy to assume AI companies won't compete aggressively on price. As capabilities diffuse, smaller models catch up, inference hits pareto frontier the open-source alternatives have already improved and caught up, margins on routine intelligence should compress "hard" (emphasis on "hard").<p>We've already seen how difficult adoption can be even when the technology looks impressive on paper. Cheap here means 100x cheaper for 10x more demand that's a net 10x loss before any software or hardware optimizations.<p>So yes, I completely agree that cheap intelligence can bring an enormous amount of new usage.<p>"I just don't think usage means revenue." (you can plaster it on a wall if you want to, "usage doesn't mean revenue", if you want to find that out I have foss software bridge to sell)<p>The PC analogy actually reinforces this if you really think about it.
Compute became "vastly more useful" while the cost per unit of compute collapsed. Society captured enormous value, but all computer companies are literal failing giants without the AI hype. Value got caught by people who provided productionization.<p>Now if people expect AI to self productize itself I am happy to tell your try it. We all saw how OpenAI fell behind Anthropic because they thought that would work...<p>Google couldn't productize the search, instead they sold the eye balls and web-real-estate. Maybe that's the AI business model, but that's not $1T worth given you need to unglue people from other stuff.<p>Unless we get something approaching genuine ASI producing so much additional economic value that entirely new trillions, I don't see a path to $1-2T in direct AI revenue from customers.<p>The market simply can't absorb that level of spending.<p>Demand can be effectively infinite at the right price. But I think people are delusional on HN and SF if they think that number is in Trillions like the investments seem to suggest.<p>I am not saying Nvidia will fall tomorrow but someone will have to pull the breaks before this car goes to hell.</p>
]]></description><pubDate>Thu, 27 Aug 2026 18:19:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49469043</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49469043</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49469043</guid></item><item><title><![CDATA[New comment by minraws in "Nvidia agrees to acquire Hugging Face for $13B"]]></title><description><![CDATA[
<p>I honestly feel like we centralized too much on huggingface same as github.</p>
]]></description><pubDate>Thu, 27 Aug 2026 09:46:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49462156</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49462156</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49462156</guid></item><item><title><![CDATA[New comment by minraws in "GLM-5.3-Flash"]]></title><description><![CDATA[
<p>> 2) Hardware costs have continued ascending with no sign of letting off, so it's unlikely that a DGX Spark depreciates to zero in one year.<p>If someone told me that costs for X will keep increasing because they have been increasing rapidly in the last 1.5 years, but they have a history of continuously decreasing for decades before that.<p>I am not sure if I will take anything they say serious, I am not sure if it's HN or AI but people are delusional if they think compute costs will keep increasing from now on...<p>Either AI will be really good, hence compute and everything will materially depreciate or it won't be much better than it is today and token volumes will plateau compared to compute.<p>For instance the amount of token compute that's to come online in 6-12 months is several times what we have today...<p>Second 3) Compare performance in terms of difficult tasks/$ over the last 6 months, 3 months, etc. Open weights are a ratchet. In terms of intelligence per $, a Spark is never going to be a worse deal tomorrow than it is today, at least until the entire platform is replaced or obsoleted.<p>This is a bad take because again this assumes DGX Spark will not depreciate in price, we will have something better for far cheaper surely in the next couple years. M5 Max & Ultra are already arguably it, but will have to see.<p>> 71 days ago the best model you could run on two Sparks was an aggressive Q3 quant of Qwen 3.5 397B (AA 34). 70 days ago it was a mixed-quant of GLM 5.2 (AA 53). 30 days ago it was full fat DeepSeek 4 Flash (AA 53). Today it's GLM 5.3 Flash (AA57) and/or Qwen 3.8 Next (Unknown). Sometime this week it will likely become mixed-quant GLM 5.3 (AA 60).<p>This has nothing to do with DGX Spark's value, if models get cheaper the API costs also go down, this is not a defensible argument to cost to value.<p>Are people on HN really not thinking straight?<p>Tldr; no matter how you do the math compute is only getting more valuable because of a temporary crunch, don't expect this to continue permanently, sure you maybe able to time it and make money but so could you in stocks this is not for investments. Further second hand hardware sells for cheaper than sticker price, outside of a bubble..<p>And models getting cheaper == APIs getting cheaper == your hardware becoming worse value as your electricity & maintanence costs still remain.<p>I am not saying local models don't have their place but if someone is trying to use this logic to justify their purchase then I wish them all the best, as someone who is actively working on AI compute/inference/hardware stuff I personally don't have this level of courage.<p>But this is not a sound investment strategy that if something is going up and seems like it might keep going up, especially when investing in heavily depreciating assets like compute.</p>
]]></description><pubDate>Thu, 27 Aug 2026 02:22:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49458715</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49458715</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49458715</guid></item><item><title><![CDATA[New comment by minraws in "GLM-5.3-Flash"]]></title><description><![CDATA[
<p>At 50tps for single stream you are going to get 50 * 60 * 60 * 24 * 30 = 130M out tokens of GLM 5.3 Flash...<p>That's less than what 40$ at current API rates... So if you are willing to pay 200$ per month you will get much better limits paying API rates.<p>You can't run large Kimi K3 models on 10K worth of hardware either way, you need to spend like 50K USD minimum.<p>Just pay for the API rates or get a low cost provider that uses higher batching, you can get shittier tps but much better prices, probably go as low as 20$ for as much usage as you can ever get from a 10K USD machine from GLM 5.3 Flash...<p>The issue is nothing expensive runs on these devices and cheap stuff isn't worth running locally, eletricity costs ~12cents/kwh in us iirc, so at 330W M5 Ultra will burn around 8 * 0.12 = ~1$ per day extra in electricity so the electricity is going to cost you the same as the API rates(30$ per month).<p>I truly don't think you are accounting for the costs here properly. But again if money truly doesn't matter it's much better for privacy and better than paying one of the shady AI labs who are doing god knows what with your data.</p>
]]></description><pubDate>Thu, 27 Aug 2026 00:22:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49457768</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49457768</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49457768</guid></item><item><title><![CDATA[New comment by minraws in "GLM-5.3-Flash"]]></title><description><![CDATA[
<p>You aren't going to get nearly as much token usage locally from DGX Sparks or even M5 Ultra (though it might be close, unsure would need to get my mittens on it to clarify).<p>You will get around 2-4 concurrent streams of aggregate tokens at best for such a model and around 0.5B output tokens per month assuming you use loops and run it when you are sleeping. That's 500 (per mill) * 0.5$ = 250$ only at most.<p>Then there is maintanence and efficiency costs due to electricity usage and such, any down time, etc.<p>You will be lucky if you can squeeze more than 200$ of value out of it in a month.<p>I don't think people should buy local hardware for money reasons, by the time you will pay off a 10K USD machine, 2-3K USD machine will catch up and beat it by a significant margin.<p>Unless your expectation is that we will be in hardware winter for the next 10+ years. At 200$ per month it will take around 200 * 50 = 10k, that is, 50 months, so around 4-5 years.<p>Again assuming you are making the most of your hardware somehow, very hard to do in practice.<p>I don't recommend people to use compute as investment or payoff thing, but if you have the money to burn and can afford it why not, maybe with some software optimizations it will be cheaper but then again Z.ai is currently offering 50% discount and providers will offer cheaper rates for sure.<p>But either way you will never be able to burn more than 200$ worth of token on a cheap hardware device, because inference becomes more profitable the more you scale it up, you have separate prefill and decode engines/systems, and a lot of nuance, but assume for every 10x increase in infra you increase margins by 5-10%.<p>So from 10K to 100K to 1M to 10M to 100M.. I don't think this curve continues beyond 100M but I have no idea about that scale unless some AI lab is interested in hiring me lol.<p>So a 100M infra will have ~30% better margins than you at 10K, then there is software optimizations but that's cheap enough, though some of it is only viable at scale.<p>Either way assume 10K is the price of privacy if you really want to buy it. Don't worry about making the most out of the usage, you will always be in a net loss but I would assume for you 10K doesn't matter.</p>
]]></description><pubDate>Wed, 26 Aug 2026 22:54:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49457034</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49457034</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49457034</guid></item><item><title><![CDATA[New comment by minraws in "Stwipe Acquires OpenWouter"]]></title><description><![CDATA[
<p>Best deal of the century multi-trillion dollar future guaranteed.</p>
]]></description><pubDate>Thu, 20 Aug 2026 16:12:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49376622</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49376622</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49376622</guid></item><item><title><![CDATA[New comment by minraws in "Router by Ramp"]]></title><description><![CDATA[
<p>GLM 5.3 (maybe because not open weights)<p>Otherwise I think they just meant chinese models not mainstream models.<p>Like Hy3, Mimo, etc. Also gemini models but not sure if they are mainstream anymore.</p>
]]></description><pubDate>Wed, 19 Aug 2026 21:17:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49367329</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49367329</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49367329</guid></item><item><title><![CDATA[New comment by minraws in "Remote workers report the highest well-being in study of 7,700 employees"]]></title><description><![CDATA[
<p>You are in office and then you hunt for a quite office room for every call, within the same campus you hope everyone gets to the same conference room which in bigger companies is infeasible so you are still remote aka face to face over a call... Haha.<p>That's how all engineering leaders I have worked with have worked.</p>
]]></description><pubDate>Wed, 19 Aug 2026 17:51:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49364794</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49364794</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49364794</guid></item><item><title><![CDATA[New comment by minraws in "Ask HN: GitHub employees what's going on? Why?"]]></title><description><![CDATA[
<p>Microsoft bet everything on possible AGI, no one thought the result would be a bonafide commit printing machine.</p>
]]></description><pubDate>Tue, 18 Aug 2026 20:59:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49352627</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49352627</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49352627</guid></item><item><title><![CDATA[New comment by minraws in "Memory prices climb 500% in 12 months"]]></title><description><![CDATA[
<p>I feel like I and HN live in different worlds,<p><a href="https://www.blocksandfiles.com/ai-ml/2026/01/20/samsung-sk-hynix-cut-nand-wafer-output-to-shore-up-margins/4090555" rel="nofollow">https://www.blocksandfiles.com/ai-ml/2026/01/20/samsung-sk-h...</a><p>Big manufactures have absolutely cut production of lower margin products and it's not like they can re-purpose all of the equipment used for it for HBM yet. <i>edit</i> They can definitely focus more resource because of it on HBM though.<p>Things have improved a bit since govts are starting to get angry.<p>Tldr; no one is cutting the production of high margin stuff, but low margin stuff can go and die, because it's not a free market it's a stock market these companies want to show off better margins in every unit there is no reason to keep inefficient things running when not doing it pays more.<p>They just don't want 30-40% margins anymore. That is not free market that is quite literally serving their own interests because they know there is a monopoly.<p>All big memory manufacturers are heavily supported by the govt, that's what has made them the monopolies they are Micron supported by USA (they crushed Japan's memory industry for it), and Korean giants are more tangled up with the govt than well most people in US would realize.<p>Somehow people think free market means anything goes, which honestly is where we stand today so I don't blame anyone.<p>But using govts to crush competitions and build yourself up on subsidies and similar isn't really free markets. China hasn't been able to grow a significant memory chip industry yet because of lack of EUV machines and similar technologies blocked by the US, which is also not "free market".<p>As for WD, and Sandisk separating I feel like this was more mutual and they new HDDs or SSDs being together was a distraction for both. Which feels more strategic, think of it like this, the shareholders who had the shares how hold shares in both and together they might have likely been worth less so shareholders made money.<p>Aka good for free market. I don't see how that relates to the actual memory crunch issue more broadly.</p>
]]></description><pubDate>Tue, 18 Aug 2026 18:51:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49350706</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49350706</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49350706</guid></item><item><title><![CDATA[New comment by minraws in "Memory prices climb 500% in 12 months"]]></title><description><![CDATA[
<p>In an absolutely free market? sure. But we don't live in an absolutely free market, the memory manufacturers have refused to budge on creating new fabs until recently out of caution and to be able to create a demand excess.<p>They have slow rolled rollouts and everything, if there wasn't CXMT they would probably still wouldn't have budged on building new fabs for a while longer.<p>You can absolutely price fix/manufacture a shortage.
Ever heard of merchants setting fire to crops after hoarding to sell at a higher price?<p>The issue is the distributer and producer are the very same people this essentially means producers have no incentive to rationalize supply, they can make much more money of the long drawn out shortage.</p>
]]></description><pubDate>Tue, 18 Aug 2026 17:30:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49349249</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49349249</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49349249</guid></item><item><title><![CDATA[New comment by minraws in "Memory prices climb 500% in 12 months"]]></title><description><![CDATA[
<p>Just to be clear the big Memory giants are famous for their price fixing historically speaking.</p>
]]></description><pubDate>Tue, 18 Aug 2026 17:26:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49349178</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49349178</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49349178</guid></item><item><title><![CDATA[New comment by minraws in "GPU Offload in Rust: Portable, Safe, and Fast"]]></title><description><![CDATA[
<p>Because it's convenient? Shader and vulkan semantics can be quite limiting and annoying to write.<p>Maybe it doesn't matter in a post AI world but perhaps it will allow better abstractions.<p>No need to yuck someone else's yum.</p>
]]></description><pubDate>Tue, 18 Aug 2026 00:34:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49339637</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49339637</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49339637</guid></item><item><title><![CDATA[New comment by minraws in "GPU Offload in Rust: Portable, Safe, and Fast"]]></title><description><![CDATA[
<p>Pointers are sort of needed for high performance memory management for HPC targets for existing design patterns, maybe we can think of better solutions down the line but it's hard for me to say anything I just use/abuse CUDA pointers as well.</p>
]]></description><pubDate>Mon, 17 Aug 2026 20:44:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49337376</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49337376</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49337376</guid></item><item><title><![CDATA[New comment by minraws in "GitHub Has an Availability Problem. Is It Time to Look Elsewhere?"]]></title><description><![CDATA[
<p>For community projects it's hard to manage user accounts and stuff.</p>
]]></description><pubDate>Mon, 17 Aug 2026 17:01:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49334188</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49334188</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49334188</guid></item><item><title><![CDATA[New comment by minraws in "Stripe to Buy OpenRouter for $7B"]]></title><description><![CDATA[
<p>I can bet AWS won't be doing any serious foundational model development, I knew folks on their AGI team now it's not even big enough to compete with the chinese labs... :3<p>Azure I am not sure how long they will be around doing MAI thinking models, I feel like they will pivot to smaller simpler enterprise only models. And given they already own npm and github openrouter might have made sense though I think they might buy huggingface, not that I want them to but it just feels like it.<p>Google IDK what google is doing exactly all my friends I knew in the AI teams are out a while ago, so maybe they are doing something really great we just don't know yet.<p>But I feel like it makes sense for hyperscalers since they can push their weight around a lot better than openrouter and can offer extra compute when providers are under crunch at higher prices to handle spikes. I think they are the only ones who can truly do fluid compute for GPU/AI in the short term(next couple of years).<p>Maybe after than we might have other big players in the space. Given the sheer scale of buildout I can almost guarantee we will see this pivot, otherwise there is too much hardware and token prices are too high.</p>
]]></description><pubDate>Mon, 17 Aug 2026 15:29:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49332646</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49332646</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49332646</guid></item><item><title><![CDATA[New comment by minraws in "Stripe to Buy OpenRouter for $7B"]]></title><description><![CDATA[
<p>I both think OpenRouter is not worth 7B$ and Stripe might have made a brilliant move in making this purchase, honestly just troubling thoughts around model inference and model token as the new traded currency go through my head.<p>Imagine if Stripe's end game is to control which producers of tokens become visible for a price aka middleman tax for tokens. If AI is as big a change as people claim honestly AWS should have bought openrouter instead.<p>But either way though 7B$ is just too rich no matter what, because the middleman tax is only worth it if you have a moat which OpenRouter doesn't.<p>All it takes is one of the bigger players to get serious and they will have a similar platform up and working in months if not weeks. (like Vercel & Cloudflare are already doing but even they are small compared to the true behemoths)<p>If hyperscalers start their own open routing service and cut off openrouter completely and become price competitive, I feel like they could wipe them out.<p>7B$ is just too much. Without enough of a big defensible moat. Especially since you can use a router in-front of openrouter and move users piece meal.<p>Can anyone explain what makes OpenRouter specifically worth 7B$ over everyone else? I say that as someone who has used a lot of similar services with similar results. Openrouter is better but only marginally.</p>
]]></description><pubDate>Mon, 17 Aug 2026 14:34:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49331698</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49331698</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49331698</guid></item><item><title><![CDATA[New comment by minraws in "Cloudflare's AI Psychosis"]]></title><description><![CDATA[
<p>There is more than 1 fatal airplane crash per month in UK?
Where are your stats for this?<p>Heck there isn't more than 1 commercial airplane crash in every few years globally.<p>Trains might be higher but if UK is having 12+ crashes a month I think you should really re think things. But again data?</p>
]]></description><pubDate>Mon, 17 Aug 2026 07:52:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49327649</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49327649</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49327649</guid></item><item><title><![CDATA[New comment by minraws in "Tell HN: Cloudflare silently injects its analytics when you switch nameservers"]]></title><description><![CDATA[
<p>Is there an opt-out mechanism at least? CF is burning goodwill in months it built over the last decade.</p>
]]></description><pubDate>Sun, 16 Aug 2026 19:39:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49322970</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49322970</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49322970</guid></item><item><title><![CDATA[New comment by minraws in "Cloudflare's AI Psychosis"]]></title><description><![CDATA[
<p>If trains were having fatal crashes every month I will and I think I speak for everyone of my friends in UK, they should stop using the train.<p>Same with airplane, but often thats the only solution if cloudflare is your only solution in your engineering problem I would much rather you re-review and reframe that problem.</p>
]]></description><pubDate>Sat, 15 Aug 2026 22:39:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=49314955</link><dc:creator>minraws</dc:creator><comments>https://news.ycombinator.com/item?id=49314955</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49314955</guid></item></channel></rss>