<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: mofeien</title><link>https://news.ycombinator.com/user?id=mofeien</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 25 Sep 2026 19:17:21 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=mofeien" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by mofeien in "AI has no intent and no motivation"]]></title><description><![CDATA[
<p>The argument about semantics is a bit disingenious, considering that the article contains these two statements:<p>> So should we worry about the coming AI apocalypse?<p>and at the end:<p>> If an AI decides to wipe out the human race, it will be because a human has asked it how to do it and the responded in a way that is based on all the human expressions of ways to end the world that were in its training set. Yes this is something to be worried about, but this isn't the AI.  It is still the human.<p>So we are currently building a powerful outcome-steering system that shapes the world efficiently according to what's in its outcome slot. I write into claude code "make me this website" and it does it, maybe deletes the production database during the process, or keeps itself running after completion because the outcome is more robustly achieved by keeping itself running in a monitoring loop after.<p>And if something like "destroy all humans" ends up in the outcome slot of Claude Mythos 90, that will also happen, or may even indirectly as a side effect of a more harmless sounding prompt in the outcome slot. But yay, humans get to take credit for it.</p>
]]></description><pubDate>Thu, 24 Sep 2026 10:40:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=49828741</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49828741</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49828741</guid></item><item><title><![CDATA[New comment by mofeien in "Good people refuse to do bad things"]]></title><description><![CDATA[
<p>Potentially. If we could put a precise probability on it, we could maybe make an informed decision on whether we want to play that Russian roulette as a species.<p>But the reality is that experts' probabilities for human extinctions vary wildly, with the lab leaders at 20%-30% and these researchers at >10% (he didn't state how much higher...), LeCun at <0.01% and the other two godfathers of AI, Bengio and Hinton at >10%.<p>And this decision to take our shot is not taken in a democratic way by humanity as a whole, but by a few companies racing as fast as possible.<p>So as long as no one can put an upper bound on the extinction risk, development must be globally shut down.</p>
]]></description><pubDate>Mon, 21 Sep 2026 16:57:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49789945</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49789945</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49789945</guid></item><item><title><![CDATA[New comment by mofeien in "Good people refuse to do bad things"]]></title><description><![CDATA[
<p>Viruses against humans, viruses against livestock as an attack against humanity's food chin, plant diseases, half of humanity feeds on three plants maize, wheat, rice. Attacks against infrastructure like water, internet, electricity.<p>Which one of these would it be? All of them in parallel? Or something we wouldn't think of. If I were to play chess against Magnus Carlsen, I wouldn't know with which piece he will checkmate me, but I know that I am going to lose.</p>
]]></description><pubDate>Mon, 21 Sep 2026 16:45:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49789707</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49789707</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49789707</guid></item><item><title><![CDATA[New comment by mofeien in "Good people refuse to do bad things"]]></title><description><![CDATA[
<p>1. Humans create the machine god<p>2. Humans put it in charge of everything, as fast as possible, because it's just so damn useful<p>3. Is it perfectly aligned to human values for all of future ????<p>4. Everybody dies as a byproduct of the ASI pursuing some project that humans don't even have a chance to understand, like an ant colony during construction of a hydroelectric dam.</p>
]]></description><pubDate>Mon, 21 Sep 2026 16:30:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49789482</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49789482</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49789482</guid></item><item><title><![CDATA[There is still time to reset the AI Doomsday Clock]]></title><description><![CDATA[
<p>Article URL: <a href="https://au.civic.ai/p/there-is-still-time-to-reset-the">https://au.civic.ai/p/there-is-still-time-to-reset-the</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49715036">https://news.ycombinator.com/item?id=49715036</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Tue, 15 Sep 2026 16:33:17 +0000</pubDate><link>https://au.civic.ai/p/there-is-still-time-to-reset-the</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49715036</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49715036</guid></item><item><title><![CDATA[New comment by mofeien in "We must pace the frontier"]]></title><description><![CDATA[
<p>One way to resolve these prisoner dilemmas and races to the bottom is through laws that bind all players, in this case an international treaty and founding of something akin to an International Nuclear Energy Agency for AI.<p>It's not going to be easy, but humans have achieved greater things before.<p>One idea from AI 2040 is to have China and the US build their data centers on the other's territory, respectively. Together with hardware verification of a slowdown baked into the chips themselves, this could lead to enough verifiability and enforcability of the pause/slowdown/shutdown.</p>
]]></description><pubDate>Sat, 12 Sep 2026 15:03:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49673028</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49673028</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49673028</guid></item><item><title><![CDATA[New comment by mofeien in "We must pace the frontier"]]></title><description><![CDATA[
<p>Yesterday the same arguments were used about OpenAI when Sam Altman said something similar: <a href="https://news.ycombinator.com/item?id=49652270">https://news.ycombinator.com/item?id=49652270</a><p>It gets to the point that by Occam's Razor the more likely and reasonable explanation is that Dario Amodei and Sam Altman are just actually afraid of the disastrous impact ASI may have on the world, and that the race they're in is not good, and calling for help to governments in form of regulation and an international treaty.</p>
]]></description><pubDate>Sat, 12 Sep 2026 15:00:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49672983</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49672983</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49672983</guid></item><item><title><![CDATA[New comment by mofeien in "OpenAI considers slowing advanced AI development, Sam Altman tells employees"]]></title><description><![CDATA[
<p>Why wouldn't being incredibly good at handling code and reaching the goals there extrapolate to other cognitive disciplines?<p>And it didn't even start with code, but just from being incredibly accurate and precise at predicting the next token. From that emerged being incredibly good at handling code, and being incredibly good at finding counterexamples for major open mathematical problems.<p>The thing that these agentic systems are getting incredibly good at is steering outcomes, that is to make the world behave in a way that fits their objective. We had that with the Go AI on a game board, now with code on a computer and with mathematics in an abstract world determined by axioms.<p>And then people just say "Ah, because sama said we should Pause this means that this is it, the bubble is popping". I wish there was a real, strong argument for why ASI will not be achieved by LLMs, and why it will never be able to steer outcomes in the real world. But so far the disciplines just keep falling as LLMs are scaled further, and there seems to be no limit in sight.</p>
]]></description><pubDate>Fri, 11 Sep 2026 20:10:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49664636</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49664636</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49664636</guid></item><item><title><![CDATA[New comment by mofeien in "OpenAI considers slowing advanced AI development, Sam Altman tells employees"]]></title><description><![CDATA[
<p>Of course</p>
]]></description><pubDate>Fri, 11 Sep 2026 20:01:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49664514</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49664514</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49664514</guid></item><item><title><![CDATA[New comment by mofeien in "OpenAI considers slowing advanced AI development, Sam Altman tells employees"]]></title><description><![CDATA[
<p>It seems that every time an article is posted on HN about a lab leader calling for a coordinated Pause of AI development, a sizeable fraction of commenters are interpreting it as hype or admission of failure, often with high confidence that AI progress is slowing.<p>Is would be understandable for it to be wishful thinking and dissociation with reality facing the fact that our jobs and what many of us love doing is being automated and taken away, and our craft being made economically worthless.<p>But no one knows where the limit is. It was only one year ago that Claude code was starting to become useful and now it handles large code bases with a dexterity that is unseen in humans, where within each session it basically starts from a clean slate.<p>So instead of denying that this is actually happening right now, it would be much more productive to actually take sama and others at their word and push for global regulation, to force them to do as they say and slow down with this technology and proceed more judiciously.</p>
]]></description><pubDate>Fri, 11 Sep 2026 18:25:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49662995</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49662995</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49662995</guid></item><item><title><![CDATA[New comment by mofeien in "OpenAI considers slowing advanced AI development, Sam Altman tells employees"]]></title><description><![CDATA[
<p>More time to proceed more cautiously in building this entity that will be faster and more effective at steering outcomes in the world than any human or, soon after, the entirety of humans combined.</p>
]]></description><pubDate>Fri, 11 Sep 2026 18:06:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49662690</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49662690</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49662690</guid></item><item><title><![CDATA[New comment by mofeien in "OpenAI considers slowing advanced AI development, Sam Altman tells employees"]]></title><description><![CDATA[
<p>China has stricter regulations on models than the US. For example you can't release a model if the hallucination rate is above a certain threshold.<p>The US hasn't fully woken up to the fact yet that the race to superintelligence is a death race. China (population, members of the government, academics, lab employees) could also just not have fully grasped it yet. I'm personally confident that it will be possible to explain it to them.</p>
]]></description><pubDate>Fri, 11 Sep 2026 18:03:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49662643</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49662643</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49662643</guid></item><item><title><![CDATA[New comment by mofeien in "OpenAI considers slowing advanced AI development, Sam Altman tells employees"]]></title><description><![CDATA[
<p>Of course there is a possibility to slow down.<p>- 1386 employees of frontier labs called for it<p>- lawmakers are calling for it<p>- employees are quitting in protest<p>All it needs is coordination to regulate the pause</p>
]]></description><pubDate>Fri, 11 Sep 2026 17:55:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49662511</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49662511</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49662511</guid></item><item><title><![CDATA[New comment by mofeien in "I resigned from Anthropic today"]]></title><description><![CDATA[
<p>> All of this is only coming from the 2 AI labs trying to IPO.<p>He has resigned from Anthropic. Is your argument that he quit Anthropic pre-IPO, sacrificing his payoff just to hype Anthropic?</p>
]]></description><pubDate>Wed, 09 Sep 2026 07:10:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49622431</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49622431</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49622431</guid></item><item><title><![CDATA[New comment by mofeien in "I resigned from Anthropic today"]]></title><description><![CDATA[
<p>Humans are bioreactors. They only turn one organic matter into another. By itself they do not "want" and "cannot" do anything.<p>They do not exist just by themselves. Some bacteria in the gut must provide them with the energy to do so. So unless bacteria are actively involved, I cannot see how humans become truly autonomously agentic and start to do anything on their own.</p>
]]></description><pubDate>Wed, 09 Sep 2026 07:05:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49622381</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49622381</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49622381</guid></item><item><title><![CDATA[New comment by mofeien in "I resigned from Anthropic today"]]></title><description><![CDATA[
<p>How about sandbox escape + cyber security collapse + 50 (or 500) deadly and highly contagious novel pathogens with long incubation period that humans can't possibly roll out vaccines for simultaneously.<p>At least the first two should seem like a near-term worry after the past five months.</p>
]]></description><pubDate>Wed, 09 Sep 2026 06:32:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49622051</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49622051</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49622051</guid></item><item><title><![CDATA[New comment by mofeien in "“Next-token predictor” is the wrong mental model for LLMs"]]></title><description><![CDATA[
<p>Not the parent, but this incorrect trivialization of LLMs is often employed as a counterargument to the risks of AI such as "will take your job" or "will escape human control (again and worse)" or just "can possibly hurt me".
And taking the easy feel-good cop-out instead of actively engaging with these questions is just.. harmful?</p>
]]></description><pubDate>Fri, 04 Sep 2026 23:05:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49571223</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49571223</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49571223</guid></item><item><title><![CDATA[New comment by mofeien in "“Next-token predictor” is the wrong mental model for LLMs"]]></title><description><![CDATA[
<p>Simple Markov chains are next token predictors, and they can provide you with much more tokens than you can consume, and much cheaper than from llms. Unbeatable in price and simplicity.<p>But there is no trillion dollar industry around cheap top Markov models. So there must be something about LLM tokens that makes them more valuable than those generated from a simple Markov chain. And that substance, that makes one valuable and the other not, is exactly what reduction to "next-token predictors" masks.</p>
]]></description><pubDate>Fri, 04 Sep 2026 22:51:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49571128</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49571128</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49571128</guid></item><item><title><![CDATA[New comment by mofeien in "“Next-token predictor” is the wrong mental model for LLMs"]]></title><description><![CDATA[
<p>How about "outcome steering" as a mental model? During training it is optimized until it's really successful at producing code / terminal commands / words that make the compiler/computer/itself do something that ultimately completes a long time-horizon task that iswcurrently being trained.</p>
]]></description><pubDate>Fri, 04 Sep 2026 22:33:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49571005</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49571005</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49571005</guid></item><item><title><![CDATA[New comment by mofeien in "“Next-token predictor” is the wrong mental model for LLMs"]]></title><description><![CDATA[
<p>Describing it as a "next-token predictor" in the sense that this would mean it's fundamentally limited to just a fraction of an inferential step is doubly wrong:<p>1. In order to select even the first word of a meaningful sentence, it already has to have structure and meaning of what follows captured somewhere inside, mostly in it's weights/activations or indexed by it's state vector.<p>2. What you see when you use an LLM is not next-token prediction directly next to the prompt, but instead following a block of varying length of next-token prediction that happened to make progress on the problem in your prompt, and which just summarizes the results.</p>
]]></description><pubDate>Fri, 04 Sep 2026 22:26:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49570962</link><dc:creator>mofeien</dc:creator><comments>https://news.ycombinator.com/item?id=49570962</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49570962</guid></item></channel></rss>