<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: yarri</title><link>https://news.ycombinator.com/user?id=yarri</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sun, 13 Sep 2026 06:54:38 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=yarri" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by yarri in "We Must Pace the Frontier"]]></title><description><![CDATA[
<p>>> To be clear, pacing does not mean halting model training or technical progress, but ensuring companies take adequate time to align and safeguard their models, and for third party evaluators to confirm this<p>> if slowing this down were possible<p>Why is embedded alignment evaluation not possible?<p>I agree with Dario, this did work globally in banking and did encourage a race to the top. Didn’t prevent the GFC but also we did survive the GFC.</p>
]]></description><pubDate>Sat, 12 Sep 2026 15:13:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49673159</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=49673159</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49673159</guid></item><item><title><![CDATA[New comment by yarri in "Helion: A high-level DSL for performant and portable ML kernels"]]></title><description><![CDATA[
<p>Pallas is the Triton equivalent in JAX land. There are some old auto tuning prototypes if you search for Pallas, like this <a href="https://github.com/jax-ml/jax-triton/pull/108" rel="nofollow">https://github.com/jax-ml/jax-triton/pull/108</a></p>
]]></description><pubDate>Sat, 08 Nov 2025 00:33:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=45852956</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=45852956</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45852956</guid></item><item><title><![CDATA[New comment by yarri in "Steve Wozniak: Life to me was never about accomplishment, but about happiness"]]></title><description><![CDATA[
<p>Many of us grew up in the PLD era, k-maps, etc. Woz pushed early HW to the limit, with SW APIs that delivered real value. Woz made astute design trade-offs based on full stack knowledge that his peers lacked. The world’s moved on to the GPU (low precision, accelerated parallel compute?) era, but the Woz view point still holds. You can see it in the AI kernel optimizations, or rematerialization methods to push GPU HW to the new limits, and trade-offs need to be made. GPU HW for 4-bit QAT or even 2-bit will dramatically affect the SW (AI) of this era. What trade-offs do you make?<p>I saw Woz on Northbound 280 “driving” his cherry red Model S, using FSD. He was looking down at the screen the whole time I watched him. Swear he had ssh’d into it.</p>
]]></description><pubDate>Fri, 15 Aug 2025 15:38:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=44913714</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=44913714</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44913714</guid></item><item><title><![CDATA[New comment by yarri in "Cloud Run GPUs, now GA, makes running AI workloads easier for everyone"]]></title><description><![CDATA[
<p>[edit - Gabe responded]. See this Cloud Run spending cap recommendation [0] to disable billing, which potentially irreversibly deletes resources but does cap spend!<p>[0] <a href="https://cloud.google.com/billing/docs/how-to/disable-billing-with-notifications#functions_cap_billing_dependencies-python" rel="nofollow">https://cloud.google.com/billing/docs/how-to/disable-billing...</a></p>
]]></description><pubDate>Wed, 04 Jun 2025 15:21:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=44181694</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=44181694</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44181694</guid></item><item><title><![CDATA[New comment by yarri in "AlphaEvolve: A Gemini-powered coding agent for designing advanced algorithms"]]></title><description><![CDATA[
<p>Not sure what “official” means but would direct you to the GCP MaxText [0] framework which is <i>not</i> what this GDM paper is referring to but rather this repo contains various attention implementations in MaxText/layers/attentions.py<p>[0] <a href="https://github.com/AI-Hypercomputer/maxtext">https://github.com/AI-Hypercomputer/maxtext</a></p>
]]></description><pubDate>Thu, 15 May 2025 15:58:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=43996374</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=43996374</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43996374</guid></item><item><title><![CDATA[New comment by yarri in "AlphaEvolve: A Gemini-powered coding agent for designing advanced algorithms"]]></title><description><![CDATA[
<p>I assume the Gemini results are JAX/PAX-ML/Pallas improvements for TPUs so would look there for recent PRs</p>
]]></description><pubDate>Wed, 14 May 2025 20:14:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=43988702</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=43988702</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43988702</guid></item><item><title><![CDATA[New comment by yarri in "Tokyo released point cloud data of the entire city for free"]]></title><description><![CDATA[
<p>Background on the Tokyo government’s digital twin program, including sourcing and maintenance efforts<p><a href="https://github.com/tokyo-digitaltwin/roadmap_v1.0/blob/main/Roadmap%20for%20the%20Social%20Implementation%20of%20Digital%20Twin%20First%20Edition.md">https://github.com/tokyo-digitaltwin/roadmap_v1.0/blob/main/...</a></p>
]]></description><pubDate>Tue, 24 Dec 2024 18:18:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=42503727</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=42503727</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42503727</guid></item><item><title><![CDATA[New comment by yarri in "YC is wrong about LLMs for chip design"]]></title><description><![CDATA[
<p>Please don’t do this, Zach. We need to encourage more investment in the overall EDA market not less. Garry’s pitch is meant for the dreamers, we should all be supportive. It’s a big boat.<p>Would appreciate the collective energy being spent instead towards adding to amor refining Garry’s request.</p>
]]></description><pubDate>Sat, 16 Nov 2024 17:04:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=42157445</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=42157445</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42157445</guid></item><item><title><![CDATA[New comment by yarri in "Sohu – first specialized chip (ASIC) for transformer models"]]></title><description><![CDATA[
<p>This is a datacenter chip. HVAC requirements are more interesting IMO, they seem to be targeting air cooled air edge deployments with that card. They’ll probably wind up with a baseboard design similar to the early v4i TPUs.<p><a href="https://ieeexplore.ieee.org/document/9499913" rel="nofollow">https://ieeexplore.ieee.org/document/9499913</a></p>
]]></description><pubDate>Wed, 26 Jun 2024 19:08:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=40803401</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=40803401</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40803401</guid></item><item><title><![CDATA[New comment by yarri in "Sohu – first specialized chip (ASIC) for transformer models"]]></title><description><![CDATA[
<p>Details from their technical memo at <a href="https://www.etched.com/announcing-etched" rel="nofollow">https://www.etched.com/announcing-etched</a><p>## How can we fit so much more compute on the silicon?<p>The NVIDIA H200 has 989 TFLOPS of FP16/BF16 compute without sparsity. This is state-of-the-art (more than even Google’s new Trillium chip), and the GB200 launching in 2025 has only 25% more compute (1,250 TFLOPS per die).<p>Since the vast majority of a GPU’s area is devoted to programmability, specializing on transformers lets you fit far more compute. You can prove this to yourself from first principles:<p>It takes 10,000 transistors to build a single FP16/BF16/FP8 multiply-add circuit, the building block for all matrix math. The H100 SXM has 528 tensor cores, and each has $4 \times 8 \times 16$ FMA circuits. Multiplying tells us the H100 has 2.7 billion transistors dedicated to tensor cores.<p>*But an H100 has 80 billion transistors! This means only 3.3% of the transistors on an H100 GPU are used for matrix multiplication!*<p>This is a deliberate design decision by NVIDIA and other flexible AI chips. If you want to support all kinds of models (CNNs, LSTMs, SSMs, and others), you can’t do much better than this.<p>By only running transformers, we can fit way more more FLOPS on our chip, without resorting to lower precisions or sparsity.<p>## Isn’t memory bandwidth the bottleneck on inference?<p>For modern models like Llama-3, no!</p>
]]></description><pubDate>Tue, 25 Jun 2024 18:11:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=40791636</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=40791636</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40791636</guid></item><item><title><![CDATA[New comment by yarri in "F-15s Scrambled from Portland Air National Guard Base"]]></title><description><![CDATA[
<p>Likely this one… <a href="https://en.wikipedia.org/wiki/Scrambling_(military)" rel="nofollow">https://en.wikipedia.org/wiki/Scrambling_(military)</a></p>
]]></description><pubDate>Sun, 12 Feb 2023 03:16:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=34759336</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=34759336</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34759336</guid></item><item><title><![CDATA[New comment by yarri in "Ask HN: What are your predictions for 2023?"]]></title><description><![CDATA[
<p>- Zuck will spin out FB & Instagram and merge with Twitter : TwiGramFace</p>
]]></description><pubDate>Sun, 25 Dec 2022 21:16:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=34131374</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=34131374</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34131374</guid></item><item><title><![CDATA[New comment by yarri in "Ask HN: Anyone else feel trapped in FANG? How did you get out?"]]></title><description><![CDATA[
<p>Take small bets, explain the value you are attempting to deliver and basically learn how to sell. Especially to skip levels. Be willing to fail.</p>
]]></description><pubDate>Sun, 21 Aug 2022 19:07:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=32543361</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=32543361</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=32543361</guid></item><item><title><![CDATA[New comment by yarri in "Apple's director of machine learning resigns due to return to office work"]]></title><description><![CDATA[
<p>We all still do a lot of lab work</p>
]]></description><pubDate>Sun, 08 May 2022 00:12:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=31299905</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=31299905</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=31299905</guid></item><item><title><![CDATA[New comment by yarri in "Welcome Yari: MDN Web Docs has a new platform"]]></title><description><![CDATA[
<p>The name is cool. That logo, not so much.</p>
]]></description><pubDate>Tue, 15 Dec 2020 18:13:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=25433076</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=25433076</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=25433076</guid></item><item><title><![CDATA[New comment by yarri in "Welcome Yari: MDN Web Docs has a new platform"]]></title><description><![CDATA[
<p>The name choice is so close...</p>
]]></description><pubDate>Tue, 15 Dec 2020 18:11:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=25433052</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=25433052</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=25433052</guid></item><item><title><![CDATA[New comment by yarri in "How not to create traffic jams: Don’t let people park for free"]]></title><description><![CDATA[
<p>The rise of punitive solutions is real. I was involved with discussions with local municipalities placing (private) local schools under restrictions for <i>not</i> providing sufficient carpool coverage -- levy fines based on percent of families carpooling.<p>Would the inverse of these punitive solutions, ie., encouraging carpool / ridesharing, not also work? It always amazes me how relatively unutilized the HOV lanes are.</p>
]]></description><pubDate>Sat, 08 Apr 2017 17:24:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=14067940</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=14067940</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=14067940</guid></item><item><title><![CDATA[New comment by yarri in "A right to repair: Why Nebraska farmers are taking on John Deere and Apple"]]></title><description><![CDATA[
<p>Depends on the industry. The auto insurance industry had facilitated this, but was then reprimanded for using generic parts to repair damaged cars. There was some irony in that US parts manufacturers claimed 3rd parties were importing "foreign" parts, but many of the US manufactures also subcontracted overseas. Thorny issue.</p>
]]></description><pubDate>Mon, 06 Mar 2017 18:13:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=13804680</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=13804680</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=13804680</guid></item><item><title><![CDATA[New comment by yarri in "Japanese village creates field-sized 3D paintings made of coloured rice shoots"]]></title><description><![CDATA[
<p>Civic engagement in Japanese agriculture was a challenge for rural communities when I was living in Japan, this village's activity seems similar to 4H clubs in the US.</p>
]]></description><pubDate>Mon, 06 Mar 2017 00:06:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=13799720</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=13799720</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=13799720</guid></item><item><title><![CDATA[New comment by yarri in "Show HN: SnapRides – Carpool scheduling with a panic button"]]></title><description><![CDATA[
<p>Thanks, Kevin, for the time & the detailed feedback.<p>> the problems with your flow are clearly posting calls to actions on the page
Sigh, agreed. I'd like to move away from a wizard-based flow and try either a) using a calendar to drag & drop the schedule, or b) create a WYSIWYG "signup page builder" flow.<p>>I’d just ask for where is the ultimate destination, maybe a date and a list of emails.<p>So try to get quicker to the step where a signup page is shared with users, right?<p>> If I have time, I’ll play around with this on my phone later. 
Thanks!</p>
]]></description><pubDate>Fri, 19 Jun 2015 18:09:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=9746582</link><dc:creator>yarri</dc:creator><comments>https://news.ycombinator.com/item?id=9746582</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=9746582</guid></item></channel></rss>