<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: gdiamos</title><link>https://news.ycombinator.com/user?id=gdiamos</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 04 Sep 2026 08:47:18 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=gdiamos" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by gdiamos in "Muse Spark 1.3"]]></title><description><![CDATA[
<p>I think it means that we should be aiming further ahead</p>
]]></description><pubDate>Wed, 02 Sep 2026 23:53:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49544205</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49544205</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49544205</guid></item><item><title><![CDATA[New comment by gdiamos in "How to build a diffusion language model"]]></title><description><![CDATA[
<p>I’d like to see more of these models.<p>I’ve been using diffusion Gemma and it is very fast on GPUs in output token/sec.<p>In the diffusion Gemma whitepaper, they say they could have done better with more time and compute.<p>Even with those caveats, it is very uses-able as a local model.</p>
]]></description><pubDate>Mon, 31 Aug 2026 05:21:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49505910</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49505910</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49505910</guid></item><item><title><![CDATA[New comment by gdiamos in "Nvidia agrees to acquire Hugging Face for $13B"]]></title><description><![CDATA[
<p>Best case scenario</p>
]]></description><pubDate>Thu, 27 Aug 2026 03:39:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=49459361</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49459361</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49459361</guid></item><item><title><![CDATA[New comment by gdiamos in "Choose Boring Technology (2015)"]]></title><description><![CDATA[
<p>There's certainly a place for enterprise and not breaking what's working.<p>Shouldn't that be 0 innovation tokens though?</p>
]]></description><pubDate>Fri, 14 Aug 2026 09:54:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49296617</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49296617</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49296617</guid></item><item><title><![CDATA[New comment by gdiamos in "Choose Boring Technology (2015)"]]></title><description><![CDATA[
<p>In hindsight I disagree.<p>Instead I like “only work on impossible problems”<p>Most of them turn out to be impossible, but some of them turn out to be possible.<p>I’ve never met anyone who could pick 3 and be confident in getting even one right. Tokens are a terrible analogy for innovation or research.<p>In hindsight I’ve had to sift through hundreds or more to fine one that worked.<p>I thought this post was helpful when I first started thinking about startups.<p>After more time, I think boring tech isn’t worth thinking about.</p>
]]></description><pubDate>Fri, 14 Aug 2026 08:02:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49295866</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49295866</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49295866</guid></item><item><title><![CDATA[New comment by gdiamos in "Mark Zuckerberg attacks 'closed' AI rivals as Meta returns to open models"]]></title><description><![CDATA[
<p>How big is the open model? 30B?</p>
]]></description><pubDate>Mon, 10 Aug 2026 19:18:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49248374</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49248374</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49248374</guid></item><item><title><![CDATA[New comment by gdiamos in "The Claudyssey: A line-for-line translation of Homer's Odyssey by Claude Fable 5"]]></title><description><![CDATA[
<p>Christopher Nolan beat you to it</p>
]]></description><pubDate>Fri, 07 Aug 2026 20:21:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49215749</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49215749</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49215749</guid></item><item><title><![CDATA[New comment by gdiamos in "Inside vLLM: Anatomy of a High-Throughput LLM Inference System (2025)"]]></title><description><![CDATA[
<p>vLLM is originally marketed as paged attention, but in hindsight, separating the web server and GPU process, continuous batching, kv caching / chunking, and a huge model library including low precision mattered more.<p>I wonder how much it would cost to vibe code the whole thing from scatch?<p>I wonder how much better models need to get before such a thing wouldn't look like code vomit?</p>
]]></description><pubDate>Fri, 07 Aug 2026 02:51:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49205405</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49205405</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49205405</guid></item><item><title><![CDATA[New comment by gdiamos in "Truth is not a direction: a Tarski attack on LLM probes"]]></title><description><![CDATA[
<p>I wish I could get a model to state its assumptions.</p>
]]></description><pubDate>Wed, 29 Jul 2026 05:48:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49093835</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49093835</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49093835</guid></item><item><title><![CDATA[New comment by gdiamos in "Startup founders urge Trump not to shut off Chinese open weight AI"]]></title><description><![CDATA[
<p>How do you ban melted sand?</p>
]]></description><pubDate>Thu, 23 Jul 2026 15:48:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49023539</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49023539</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49023539</guid></item><item><title><![CDATA[New comment by gdiamos in "Startup founders urge U.S. government not to shut off Chinese open weight AI"]]></title><description><![CDATA[
<p>I think we should shut it off. It would force US companies to build open models.</p>
]]></description><pubDate>Thu, 23 Jul 2026 15:45:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49023475</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49023475</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49023475</guid></item><item><title><![CDATA[New comment by gdiamos in "Counting ArXiv Delays"]]></title><description><![CDATA[
<p>I want a hosted paper to be archived.<p>That means that 10 years from now I don’t want think about making sure the hosting server is up.<p>I also want it to have a standard format for bibliography, DOI, and authors.<p>I agree it isn’t much, but it’s more than I get from a regular web hosting service and it is a standard format for papers so I don’t think it makes sense for every author to roll their own.</p>
]]></description><pubDate>Wed, 22 Jul 2026 16:37:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49009544</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49009544</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49009544</guid></item><item><title><![CDATA[New comment by gdiamos in "Gemini last models: temperature, top_p, and top_k are deprecated and ignored"]]></title><description><![CDATA[
<p>thank god, these parameters are so confusing</p>
]]></description><pubDate>Wed, 22 Jul 2026 02:31:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49001122</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=49001122</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49001122</guid></item><item><title><![CDATA[New comment by gdiamos in "Detecting LLM-Generated Texts with “Classical” Machine Learning"]]></title><description><![CDATA[
<p>as soon as you release a way of measuring it, you give LLMs a signal to optimize</p>
]]></description><pubDate>Fri, 17 Jul 2026 04:22:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=48943306</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48943306</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48943306</guid></item><item><title><![CDATA[New comment by gdiamos in "Demis Hassabis has a plan to harness AI safely"]]></title><description><![CDATA[
<p>Being on the review board comes with a promise to not be evil right?</p>
]]></description><pubDate>Tue, 14 Jul 2026 22:35:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=48913792</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48913792</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48913792</guid></item><item><title><![CDATA[New comment by gdiamos in "Counting ArXiv Delays"]]></title><description><![CDATA[
<p>No, I want arxiv to host the paper, not to review the paper.<p>I wouldn't want my google drive to start telling me my paper was too sloppy. I just want a link.</p>
]]></description><pubDate>Mon, 13 Jul 2026 19:12:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48897350</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48897350</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48897350</guid></item><item><title><![CDATA[New comment by gdiamos in "Costco is the anti-Amazon"]]></title><description><![CDATA[
<p>I wonder if Amazon eventually gets cut out by 3D printing/replicators for imitable objects.</p>
]]></description><pubDate>Fri, 03 Jul 2026 21:57:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=48780501</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48780501</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48780501</guid></item><item><title><![CDATA[New comment by gdiamos in "Scaling Laws, Carefully"]]></title><description><![CDATA[
<p>Scaling laws assume the error metric and data distribution.<p>There is a lot of follow on work that explains what happens as you change them, e.g. Scaling Laws for Transfer - <a href="https://arxiv.org/pdf/2102.01293" rel="nofollow">https://arxiv.org/pdf/2102.01293</a><p>I think it’s fortunate that transfer works in a similar way.<p>Common crawl (and Reddit, stack overflow, etc but not 4chan) was much easier to get access to at the time than using mechanical Turk.<p>There is certainly room for more work. There were many papers on scaling laws in NeurIPS this year.</p>
]]></description><pubDate>Wed, 01 Jul 2026 05:18:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=48742554</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48742554</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48742554</guid></item><item><title><![CDATA[New comment by gdiamos in "Scaling Laws, Carefully"]]></title><description><![CDATA[
<p>When I first saw scaling laws in that deep speech experiment notebook, I didn’t believe it could be real.  I was worried for months that we made a mistake, or that it only worked for that one dataset.<p>I started to believe it after we (Joel Hestness in particular) reproduced it in so many experiments in “scaling is predictable empirically”.<p>The OpenAI work replicated it in a completely different environment, and at that point I was sure it was real.<p>Sometimes people ask me why I was so surprised by it. Prior work like Banko and Brill and the unreasonable effectiveness of data argued for more data. ML theory had similar models for toy problems, eg coin flips.<p>At the time I thought deep learning was supposed to be complex. Speech and language datasets seemed much more complex than toy problems. Optimization of deep transformers was complex.<p>The idea that it was possible for the whole thing to be governed by a 3 term equation seemed too simple. The implication was that it was simple to manufacture intelligence.<p>Ten years later, I still think it is still the most interesting observation I have seen. We are still learning what it looks like to live in a world where it is possible to manufacture intelligence.</p>
]]></description><pubDate>Wed, 01 Jul 2026 03:03:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=48741817</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48741817</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48741817</guid></item><item><title><![CDATA[New comment by gdiamos in "We can still stop California's 3D printer surveillance scheme"]]></title><description><![CDATA[
<p>He started with tinkercad and thingiverse.<p>I tried basic elegoo and bambu printers.<p>He can’t read very well but he likes dragging shapes around on a tablet.<p>He would ask me to find shapes using the search engines then he mixes them together or reshapes them.<p>I would add them to his history.<p>This is why I was surprised to hear about 3D printed guns. I was quite sure there wasn’t anything like that in the history.<p>It was a good discussion topic about why adults get so bothered by things that look like guns.</p>
]]></description><pubDate>Sat, 27 Jun 2026 04:22:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=48695102</link><dc:creator>gdiamos</dc:creator><comments>https://news.ycombinator.com/item?id=48695102</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48695102</guid></item></channel></rss>