<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: trsohmers</title><link>https://news.ycombinator.com/user?id=trsohmers</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sat, 12 Sep 2026 11:53:07 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=trsohmers" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by trsohmers in "John Ternus to become Apple CEO"]]></title><description><![CDATA[
<p>Software people, in my very direct experience, are terrible at hardware... While in jest, I do think most software engineer's understanding of hardware abstractions is pretty poor and does disservice to the hardware they run on.<p>I know between Moore's Law and Gate's Law which one I would prefer to be the industry standard... <a href="https://en.wikipedia.org/wiki/Andy_and_Bill%27s_law" rel="nofollow">https://en.wikipedia.org/wiki/Andy_and_Bill%27s_law</a></p>
]]></description><pubDate>Mon, 20 Apr 2026 20:53:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=47840475</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=47840475</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47840475</guid></item><item><title><![CDATA[New comment by trsohmers in "Running Tesla Model 3's computer on my desk using parts from crashed cars"]]></title><description><![CDATA[
<p>It actually stands for "lizard brain"... it is (or at least was) an Infineon Aurix control and monitoring microcontroller, they may have changed to a newer one.</p>
]]></description><pubDate>Wed, 25 Mar 2026 22:22:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=47524066</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=47524066</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47524066</guid></item><item><title><![CDATA[New comment by trsohmers in "Positron's $230M Funding Led by Financial Trading Firms"]]></title><description><![CDATA[
<p>Feel free to ask me any questions!<p>Website: <a href="https://positron.ai" rel="nofollow">https://positron.ai</a></p>
]]></description><pubDate>Wed, 04 Feb 2026 14:09:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=46885982</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=46885982</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46885982</guid></item><item><title><![CDATA[Positron's $230M Funding Led by Financial Trading Firms]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.eetimes.com/positron-230-million-funding-led-by-financial-trading-firms/">https://www.eetimes.com/positron-230-million-funding-led-by-financial-trading-firms/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=46885964">https://news.ycombinator.com/item?id=46885964</a></p>
<p>Points: 1</p>
<p># Comments: 1</p>
]]></description><pubDate>Wed, 04 Feb 2026 14:07:55 +0000</pubDate><link>https://www.eetimes.com/positron-230-million-funding-led-by-financial-trading-firms/</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=46885964</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46885964</guid></item><item><title><![CDATA[The New Chips Designed to Solve AI's Energy Problem]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.wsj.com/tech/ai/the-new-chips-designed-to-solve-ais-energy-problem-1ba9cac1">https://www.wsj.com/tech/ai/the-new-chips-designed-to-solve-ais-energy-problem-1ba9cac1</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=44713225">https://news.ycombinator.com/item?id=44713225</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 28 Jul 2025 17:36:24 +0000</pubDate><link>https://www.wsj.com/tech/ai/the-new-chips-designed-to-solve-ais-energy-problem-1ba9cac1</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=44713225</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44713225</guid></item><item><title><![CDATA[New comment by trsohmers in "Show HN: Buckaroo – Data table UI for Notebooks"]]></title><description><![CDATA[
<p>Only with the oscillation overthruster flag enabled.</p>
]]></description><pubDate>Sun, 18 May 2025 21:46:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=44024530</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=44024530</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44024530</guid></item><item><title><![CDATA[New comment by trsohmers in "Charlie Javice convicted of defrauding JPMorgan in $175M startup sale"]]></title><description><![CDATA[
<p>I was put on it in 2015 after an acquaintance of mine that was previously on the list recommended me… I only heard from Forbes a few days before the list came out, they asked me for a photo and asked if I approved the 2 sentence blurb they prepared, and that was it. For years afterwards they would try to get me to come to their events, but I never had any interest, and I assume that was how they made money… but I never paid anything to be on the list or had real interest in being on it, and I don’t think it led to anything other than my technically illiterate parents thinking that it was impressive.</p>
]]></description><pubDate>Tue, 01 Apr 2025 04:32:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=43542871</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=43542871</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43542871</guid></item><item><title><![CDATA[New comment by trsohmers in "Llama 3.1 405B now runs at 969 tokens/s on Cerebras Inference"]]></title><description><![CDATA[
<p>Based on their S1 filing and public statements, the average cost per WSE system for their (~90% of their total revenue) largest customer is ~$1.36M, and I’ve heard “retail” pricing of $2.5M per system. They are also 15U and due to power and additional support equipment take up an entire rack.<p>The other thing people don’t seem to be getting in this thread that just to hold the weights for 405B at FP16 requires 19 of their systems since it is SRAM only… rounding up to 20 to account for program code + KV cache for the user context would mean 20 systems/racks, so well over $20M. The full rack (including support equipment) also consumes 23kW, so we are talking nearly half a megawatt and ~$30M for them to be getting this performance on Llama 405B</p>
]]></description><pubDate>Tue, 19 Nov 2024 06:21:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=42180527</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=42180527</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42180527</guid></item><item><title><![CDATA[New comment by trsohmers in "Meta's open AI hardware vision"]]></title><description><![CDATA[
<p>Do you think that the 16k GPUs get used once and then are thrown away? Llama 405B was trained over 56 days on the 16k GPUs; if I round that up to 60 days and assume the current mainstream hourly rate of $2/H100/hour from the Neoclouds (which are obviously making margin), that comes out to a total cost of ~$47M. Obviously Meta is training a lot of models using their GPU equipment, and would expect it to be in service for at least 3 years, and their cost is obviously less than what the public pricing on clouds is.</p>
]]></description><pubDate>Tue, 15 Oct 2024 19:56:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=41852476</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=41852476</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41852476</guid></item><item><title><![CDATA[New comment by trsohmers in "Visit Bletchley Park"]]></title><description><![CDATA[
<p>+1 this commenter. I just visited the UK for the first time at the beginning of this month and had a fantastic ~3 hours at Bletchley Park, but felt I had to cram TNMOC and the amazing Colossus live demonstration (where I asked a million questions) and everything else in the museum in the 90 minutes I was there. If I assume other HN readers are like me, I would dedicate at least 2.5-3 hours for TNMOC to actually get a chance to actually see and play around with their extensive collection of vintage machines.</p>
]]></description><pubDate>Fri, 30 Aug 2024 01:48:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=41397084</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=41397084</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41397084</guid></item><item><title><![CDATA[An "Observatory" for a Shy Super AI?]]></title><description><![CDATA[
<p>Article URL: <a href="https://substack.com/@robreid/p-147352947">https://substack.com/@robreid/p-147352947</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=41202970">https://news.ycombinator.com/item?id=41202970</a></p>
<p>Points: 3</p>
<p># Comments: 0</p>
]]></description><pubDate>Fri, 09 Aug 2024 15:49:38 +0000</pubDate><link>https://substack.com/@robreid/p-147352947</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=41202970</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41202970</guid></item><item><title><![CDATA[New comment by trsohmers in "Codestral Mamba"]]></title><description><![CDATA[
<p>They meant that there is no support for Codestral Mamba for llama.cpp yet.</p>
]]></description><pubDate>Tue, 16 Jul 2024 17:06:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=40978304</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=40978304</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40978304</guid></item><item><title><![CDATA[New comment by trsohmers in "Rex Computing"]]></title><description><![CDATA[
<p>We had a basic LLVM backend that supported a slightly modified clang frontend and a basic ABI. We tried to make it drastically easier for both the programmer and compiler to handle memory by having all memory (code+data) be part of a global flat address space across the chip, with guarantees being made to the compiler by the NoC on the latency of all memory accesses across one or multiple chips. We tested this with very small programs that could fit in the local memory of up to two chips (128KB of memory), but in theory it could have scaled up to the 64 bit address space limit. Compilation time for programs was long, but fully automated, specifically to improve upon problems faced by Cell and other scratchpad memory architectures… some of our original funding in 2015 from DARPA was actually for automated scratchpad memory management techniques on Texas Instruments DSPs and Cell (our paper: <a href="https://dl.acm.org/doi/pdf/10.1145/2818950.2818966" rel="nofollow">https://dl.acm.org/doi/pdf/10.1145/2818950.2818966</a>)<p>This was all designed a decade ago, and REX has been in effectively hibernation since the end of 2017 after successfully taping out our 16 core test chip back in 2016, but being unable to raise additional funding to continue. I have continued to work on architectures that have leveraged scratchpad memories in different ways, including on cryptocurrency and machine learning ASICs, including at my current startup, Positron AI (<a href="https://positron.ai" rel="nofollow">https://positron.ai</a>)</p>
]]></description><pubDate>Thu, 30 May 2024 02:27:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=40519531</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=40519531</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40519531</guid></item><item><title><![CDATA[New comment by trsohmers in "Rex Computing"]]></title><description><![CDATA[
<p>Founder of REX Computing here; I highly recommend checking out my interview on the Microarch Club podcast linked elsewhere on the thread; will also answer questions on this thread if anyone has them.</p>
]]></description><pubDate>Thu, 30 May 2024 01:06:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=40519031</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=40519031</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40519031</guid></item><item><title><![CDATA[New comment by trsohmers in "Putting a Wet Towel on a Tesla Supercharger Handle Gets Faster Charging Speeds"]]></title><description><![CDATA[
<p>This is a lesson that like all good Hitchhikers, you should always carry a towel.</p>
]]></description><pubDate>Thu, 09 May 2024 22:41:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=40313677</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=40313677</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40313677</guid></item><item><title><![CDATA[New comment by trsohmers in "Building Meta's GenAI infrastructure"]]></title><description><![CDATA[
<p>Significantly more than that; MFN pricing for NVIDIA DGX H100 (which has been getting priority supply allocation, so many have been suckered into buying them in order to get fast delivery) is ~$309k, while a basically equivalent HGX H100 system is ~$250k, coming to a price per GPU at the full server level being ~$31.5k. With Meta’s custom OCP systems integrating the SXM baseboards from NVIDIA, my guess is that their cost per GPU would be in the ~$23-$25k range.</p>
]]></description><pubDate>Tue, 12 Mar 2024 18:22:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=39682984</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=39682984</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39682984</guid></item><item><title><![CDATA[New comment by trsohmers in "Who uses Google TPUs for inference in production?"]]></title><description><![CDATA[
<p>The quote from the linked press release is that they do training on TPUv4, while inference is running on GPUs. I have also heard this separately from people associated with Midjourney recently, and that they solely do training on TPUs.</p>
]]></description><pubDate>Mon, 11 Mar 2024 21:57:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=39673698</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=39673698</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39673698</guid></item><item><title><![CDATA[New comment by trsohmers in "Dune: Part Two Is the Best Sci-Fi Film of the Decade"]]></title><description><![CDATA[
<p>I’m right on the millenial/gen Z divide and an inner selfish purpose for me working on AI/ML is just to enable a creation of Jodorowsky’s 10 hour version of Dune with soundtrack by Pink Floyd.</p>
]]></description><pubDate>Thu, 29 Feb 2024 19:20:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=39553790</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=39553790</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39553790</guid></item><item><title><![CDATA[New comment by trsohmers in "Groq runs Mixtral 8x7B-32k with 500 T/s"]]></title><description><![CDATA[
<p>Long story, but technically REX is still around but has not been able to continue to develop due to lack of funding and my cofounder and I needing to pay bills. We produced initial test silicon, but due to us having very little money after silicon bringup, most of our conversations turned to acquihire discussions.<p>There should be a podcast release (<a href="https://microarch.club/" rel="nofollow">https://microarch.club/</a>) in the near future that covers REX's history and a lot of lessons learned.</p>
]]></description><pubDate>Tue, 20 Feb 2024 06:31:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=39438579</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=39438579</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39438579</guid></item><item><title><![CDATA[New comment by trsohmers in "Groq runs Mixtral 8x7B-32k with 500 T/s"]]></title><description><![CDATA[
<p>I thought that was clear through my profile, but yes, Positron AI is focused on providing the best performance per dollar while providing the best quality of service and capabilities rather than just focusing on a single metric of speed.<p>A guarantee to match the cheapest per token prices is sure a great way to lose a race to the bottom, but I do wish Groq (and everyone else trying to compete against NVIDIA) the greatest luck and success. I really do think that the great single batch/user performance by Groq is a great demo, but is not the best solution for a wide variety of applications, but I hope it can find its niche.</p>
]]></description><pubDate>Mon, 19 Feb 2024 19:02:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=39433379</link><dc:creator>trsohmers</dc:creator><comments>https://news.ycombinator.com/item?id=39433379</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39433379</guid></item></channel></rss>