<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: sdrg822</title><link>https://news.ycombinator.com/user?id=sdrg822</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 17 Sep 2026 13:02:27 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=sdrg822" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by sdrg822 in "A single firm is behind OpenAI, Anthropic, and Meta hacking scandals"]]></title><description><![CDATA[
<p>This is incredibly misleading. OpenAI internal systems were pwned, and in all cases, the labs absolutely are responsible for their models.<p>Yes, vendors are also irresponsible, but this misses the point.</p>
]]></description><pubDate>Mon, 14 Sep 2026 21:41:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49704516</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=49704516</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49704516</guid></item><item><title><![CDATA[New comment by sdrg822 in "Agent Mesh for Enterprise Agents"]]></title><description><![CDATA[
<p>“Model Control Plane” - quite early to already re-use  a popular acronym</p>
]]></description><pubDate>Fri, 25 Apr 2025 00:48:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=43789119</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=43789119</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43789119</guid></item><item><title><![CDATA[New comment by sdrg822 in "LangManus: An Open-Source Manus Agent with LangChain + LangGraph"]]></title><description><![CDATA[
<p>Congrats on the launch!</p>
]]></description><pubDate>Mon, 24 Mar 2025 13:43:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=43460987</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=43460987</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43460987</guid></item><item><title><![CDATA[New comment by sdrg822 in "Prompt Caching"]]></title><description><![CDATA[
<p>+1 it wouldn’t be terribly useful if it were only caching the tokenizer output.</p>
]]></description><pubDate>Sun, 18 Aug 2024 23:01:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=41286247</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=41286247</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41286247</guid></item><item><title><![CDATA[New comment by sdrg822 in "LangGraph Engineer"]]></title><description><![CDATA[
<p>Where is the lie in this repo?</p>
]]></description><pubDate>Sun, 11 Aug 2024 21:47:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=41219526</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=41219526</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41219526</guid></item><item><title><![CDATA[New comment by sdrg822 in "LangGraph Engineer"]]></title><description><![CDATA[
<p>I think if you actually meet the people you'd realize they're pretty earnest and candid of these things' limitations, though it may not show in tweets and hackernews posts.<p>This repo, for instance, makes no claims of AGI. It just claims to help bootstrap a starting point for a software project: "it will not attempt to write the logic to fill in the nodes and edges."</p>
]]></description><pubDate>Sat, 10 Aug 2024 00:34:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=41206516</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=41206516</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41206516</guid></item><item><title><![CDATA[New comment by sdrg822 in "Consistency LLM: converting LLMs to parallel decoders accelerates inference 3.5x"]]></title><description><![CDATA[
<p>But indexing *is* training. It's just not using end-to-end gradient descent.</p>
]]></description><pubDate>Wed, 08 May 2024 23:02:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=40303573</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=40303573</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40303573</guid></item><item><title><![CDATA[New comment by sdrg822 in "A generalist AI agent for 3D virtual environments"]]></title><description><![CDATA[
<p>Dang they use Transformer-XL from 2019 haha - didn't realize people still used that / XLNet-like architectures</p>
]]></description><pubDate>Wed, 13 Mar 2024 16:03:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=39693188</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=39693188</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39693188</guid></item><item><title><![CDATA[New comment by sdrg822 in "Microsoft strikes deal with Mistral in push beyond OpenAI"]]></title><description><![CDATA[
<p>It is not.</p>
]]></description><pubDate>Mon, 26 Feb 2024 23:18:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=39518054</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=39518054</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39518054</guid></item><item><title><![CDATA[New comment by sdrg822 in "Show HN: Use natural language to query and visualize 400M tweets"]]></title><description><![CDATA[
<p>Congrats on the launch!</p>
]]></description><pubDate>Thu, 22 Feb 2024 17:07:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=39469959</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=39469959</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39469959</guid></item><item><title><![CDATA[New comment by sdrg822 in "Exponentially faster language modelling"]]></title><description><![CDATA[
<p>Cool. Important note:<p>"""
One may ask whether the conditionality introduced by the
use of CMM does not make FFFs incompatible with the
processes and hardware already in place for dense matrix
multiplication and deep learning more broadly. In short, the
answer is “No, it does not, save for some increased caching
complexity."
"""<p>It's hard to beat the hardware lottery!</p>
]]></description><pubDate>Wed, 22 Nov 2023 12:42:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=38378424</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=38378424</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=38378424</guid></item><item><title><![CDATA[New comment by sdrg822 in "LLaMa running at 5 tokens/second on a Pixel 6"]]></title><description><![CDATA[
<p>It’s only a matter of time</p>
]]></description><pubDate>Wed, 15 Mar 2023 20:10:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=35174059</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=35174059</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=35174059</guid></item><item><title><![CDATA[New comment by sdrg822 in "Implementing GPTZero from scratch – Reverse engineering GPTZero"]]></title><description><![CDATA[
<p>Given that GPTZero was an undergrad’s side project it’s not that surprising?</p>
]]></description><pubDate>Sat, 28 Jan 2023 15:51:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=34558537</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=34558537</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34558537</guid></item><item><title><![CDATA[New comment by sdrg822 in "What a transformer can NOT do"]]></title><description><![CDATA[
<p>Attention is Turing complete</p>
]]></description><pubDate>Fri, 20 Jan 2023 01:44:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=34448180</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=34448180</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34448180</guid></item><item><title><![CDATA[New comment by sdrg822 in "Why is Chat GPT so expensive to operate?"]]></title><description><![CDATA[
<p>For things like BERT where you just want to extract an embedding, the naive way you reach full utilization at inference time is that you :<p>- run tokenization of inputs on CPU<p>- sort inputs by length<p>- batch inputs of similar length and apply padding to make of uniform length<p>- pass the batches through so a single model can process many inputs in parallel.<p>For GPT-style decoder models however, this becomes much more challenging because inference requires a forward pass for every token generated. (Stopping criteria also may differ but that’s another tangent).<p>Every generated token performs attention on every previous token, both the context (or “prompt”) and the previously generated tokens (important for self consistency). this is a quadratic operation in the vanilla case.<p>Model sizes are large , often spanning multiple machines, and the information for later layers depends on previous ones, meaning inference has to be pipelined.<p>The naive approach would be to have a single transaction processed exclusively by a single instance of the model. this is expensive! even if each model can be crammed into a single A100 , if you want to run something like Codex or ChatGPT for millions of users with low latency inference, you’d have to have thousands of GPUs preloaded with models, and each transaction would take a highly variable amount of time.<p>If a model spans multiple machines, you’d achieve a max of 1/n% utilization because each shard has to remain loaded while the others process, and then if you want to do pipeline parallelism like in pipe dream, you’d have to deal with attention caches since you don’t want to have to recompute every previous state each time</p>
]]></description><pubDate>Sun, 15 Jan 2023 18:32:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=34391718</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=34391718</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34391718</guid></item><item><title><![CDATA[New comment by sdrg822 in "ChatGPT is a ‘code red’ for Google’s search business"]]></title><description><![CDATA[
<p>Yeah this is purely a risk decision for a prototype not an actual technical limitation.</p>
]]></description><pubDate>Sat, 24 Dec 2022 01:12:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=34112710</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=34112710</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34112710</guid></item><item><title><![CDATA[New comment by sdrg822 in "Show HN: Explainpaper  – Explain jargon in academic papers with GPT-3"]]></title><description><![CDATA[
<p>Really well done</p>
]]></description><pubDate>Tue, 01 Nov 2022 02:39:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=33416275</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=33416275</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=33416275</guid></item><item><title><![CDATA[New comment by sdrg822 in "Road to Artificial General Intelligence"]]></title><description><![CDATA[
<p>Which humans?</p>
]]></description><pubDate>Tue, 01 Nov 2022 02:38:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=33416268</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=33416268</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=33416268</guid></item><item><title><![CDATA[New comment by sdrg822 in "The Anglo-Saxon Classroom"]]></title><description><![CDATA[
<p>Good points!</p>
]]></description><pubDate>Tue, 17 Aug 2021 18:19:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=28213108</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=28213108</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=28213108</guid></item><item><title><![CDATA[New comment by sdrg822 in "The Anglo-Saxon Classroom"]]></title><description><![CDATA[
<p>Just a slight comment - "&" was indeed called "and" (not "per se"), but in reading the alphabet, it was confusing to say "and and." To clear up confusion, one could say "and per se and," which was smooshed to become "ampersand."<p>"per se" was used for letters that could also be words "in themselves"</p>
]]></description><pubDate>Tue, 17 Aug 2021 17:55:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=28212801</link><dc:creator>sdrg822</dc:creator><comments>https://news.ycombinator.com/item?id=28212801</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=28212801</guid></item></channel></rss>