<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: Bjorkbat</title><link>https://news.ycombinator.com/user?id=Bjorkbat</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 04 Sep 2026 09:07:18 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=Bjorkbat" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by Bjorkbat in "GPT-6 Astra"]]></title><description><![CDATA[
<p>The most straightforward answer is that despite efforts to design a benchmark that, in theory, is supposed to measure generalizable intelligence, performance on ARC-AGI-3 can't be reliably correlated to performance anywhere else.  I kind of lost faith in it after o1 or o3, I can't remember which, absolutely crushed ARC-AGI-1.<p>And, you know, maybe also some funny business.  I think it's good to be a little suspicious of a model that happens to shoot upwards in performance on a specific benchmark while also kind of keeping up with the pack on a bunch of other benchmarks.</p>
]]></description><pubDate>Fri, 04 Sep 2026 00:19:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49558900</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=49558900</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49558900</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Asus Bike Booster"]]></title><description><![CDATA[
<p>For my needs, sure.  One caveat is that if I’m doing a longer trip I might save power by turning it off when the ground is flat/downhill</p>
]]></description><pubDate>Mon, 17 Aug 2026 19:58:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49336700</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=49336700</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49336700</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Ask HN: Alternatives to GitHub"]]></title><description><![CDATA[
<p>Arguably the "social coding" angle is the reason why Tangled is the only actual alternative to GitHub.  Otherwise there's actually plenty of other alternatives to GitHub, but none of them have GitHub's punchcard, which I'm ashamed to admit is the primary draw for me.</p>
]]></description><pubDate>Mon, 17 Aug 2026 17:58:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49335057</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=49335057</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49335057</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Asus Bike Booster"]]></title><description><![CDATA[
<p>I actually bought one of these and love it.  My road bike is indistinguishable from any other normal road bike with the exception of it being much easier to get around now.<p>From experience I can say that it seems like a lot more common sense to simply replace the front wheel of your bike with a wheel containing an electric motor, rather than attaching whatever Asus is selling to your bike.</p>
]]></description><pubDate>Sun, 16 Aug 2026 05:03:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49317059</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=49317059</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49317059</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs"]]></title><description><![CDATA[
<p>One theory I've been entertaining is that whenever GPT-3.5 came out a lot of people were talking about the "bitter lesson" and how scale was all we really needed to get to AGI.  No need for any fancy tricks, just release a larger model trained on more data, by the time we released a hypothetical "GPT-5 sized" model we'd have AGI.<p>Anyway, the actual theory is that Google and Meta have fallen behind because they've been playing by this playbook of focusing on scale and training data, whereas OpenAI and Anthropic have done so well because they are likely doing much more interesting things to improve their models over time.  It makes sense when you realize that one of Google's key strengths, besides talent, is that they have an incredible amount of data they can use for training due to being both the world's leading search engine as well as having all that video data from YouTube.  Scaling the training data makes more sense to them than it does to Anthropic and OpenAI, who are both relatively data-disadvantaged.<p>You can kind of see this when you look at the Gemini 3 scorecard when it came out (<a href="https://blog.google/products-and-platforms/products/gemini/gemini-3/" rel="nofollow">https://blog.google/products-and-platforms/products/gemini/g...</a>) and notice that while it wasn't as good as Claude And GPT at coding, it scored higher on a bunch of other non-coding benchmarks, and I think the reason why is simply because of Google's data advantage.<p>If true, I feel even more vindicated for believing that the "scale is all we need" narrative was bullshit.</p>
]]></description><pubDate>Wed, 05 Aug 2026 17:18:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49185905</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=49185905</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49185905</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Sierra Raises $950M at $15B Valuation"]]></title><description><![CDATA[
<p>I was about to post a snarky comment along the lines of "Sierra?  The publishers of Homeworld and Homeworld 2?"</p>
]]></description><pubDate>Mon, 04 May 2026 17:11:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=48011616</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=48011616</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48011616</guid></item><item><title><![CDATA[New comment by Bjorkbat in "The Sideprocalypse"]]></title><description><![CDATA[
<p>I always got the impression that the only way a relatively average software developer could make a “successful” SaaS is if they built something weird and niche that appeals to an audience of <1000 paying customers.  In that sense, it doesn’t really matter if your competition is better at SEO since your competition never even thought to build something like this, never cared, and not to mention that the market for this thing is so small that SEO is arguably a wasted skill.  You’ll need to acquire these people by finding them directly or through word of mouth.<p>This blog post seems to fundamentally misunderstand the nature of solo-developer SaaS, but then again arguably mostly software developers also fundamentally misunderstand it.</p>
]]></description><pubDate>Mon, 16 Feb 2026 20:57:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=47040237</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=47040237</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47040237</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Cache Monet"]]></title><description><![CDATA[
<p>To me demoscene is kind of synonymous with the Amiga, so I would argue that peak demoscene lines up with the rise and fall of the Amiga brand.<p>So I think that's maybe the other differentiator between web experiment and art, because demoscene has a very distinct but difficult to describe cultural element that makes me identify it as art.</p>
]]></description><pubDate>Sat, 14 Feb 2026 05:05:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=47011754</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=47011754</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47011754</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Cache Monet"]]></title><description><![CDATA[
<p>I'm keenly aware, I have a pretty extensive collection of Hacker News bookmarks.  It's hard to articulate why I think these are different, but I think the best way to put it is that cachemonet feels a lot more avant garde, and perhaps also a reflection of a very particular form of "web culture" that has no clear successors.<p>People are experimenting with what you can do on the web, but the experiments aren't very "aesthetically inspiring".  For that reason I'm kind of lukewarm on neal.fun.<p>EDIT: so I think a better way to describe it is that when artists experiment with technology, you get something like cachemonet.  When developers experiment with technology, you get a web experiment that challenges conventional notions of what you can do with the web, but with varying degrees of creativity.  I think terra.layoutit.com is best appreciated by other web devs who can appreciate the sheer amount of work required to figure out how to render a terrain map in CSS, but otherwise it's basically just a tool to generate terrain height maps, and not a particularly good one.  Generating terrain maps in CSS is not a feature, but a handicap.</p>
]]></description><pubDate>Fri, 13 Feb 2026 17:14:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=47005109</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=47005109</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47005109</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Cache Monet"]]></title><description><![CDATA[
<p>Finding out that this is over 10 years old has made me profoundly sad.  Despite the age of LLMs arguably unlocking massive amounts of productivity and agency for developers and non-developers alike, it feels as though we are living in a dark age of creativity on the web, maybe even a dark age for computer culture in general.</p>
]]></description><pubDate>Fri, 13 Feb 2026 16:31:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=47004566</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=47004566</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47004566</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Project Genie: Experimenting with infinite, interactive worlds"]]></title><description><![CDATA[
<p>Ironically the physics are kind of my biggest criticism.  They call these "world models", but I think it's more accurate to call them "video game models" because they employ "video game physics" rather than real world physics, among other things<p>This is most evident in the way things collide.</p>
]]></description><pubDate>Thu, 29 Jan 2026 20:27:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=46816085</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=46816085</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46816085</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Experts explore new mushroom which causes fairytale-like hallucinations"]]></title><description><![CDATA[
<p>I’m sure there’s some boring neuro-chemical explanation for this, and I won’t doubt or deny the neuro-chemical explanation, but the fact that there’s a mushroom that consistently brings about hallucinations of tiny people is so bizarre that I kind of want to indulge in equally bizarre explanations.  Maybe it’s not a hallucination and this mushroom simply allows us to see the tiny people all around us.  Maybe mushrooms are intelligent and are intentionally making us hallucinate tiny people.<p>It’s a little bit crazy, I know, but it’s odd to me that evolutionary forces would produce a mushroom that makes you have some specific hallucinations, rather than simply make things swirl together or simply produce intense feelings of euphoria or dread.  I mean, marijuana just gets you high and that’s that.</p>
]]></description><pubDate>Fri, 26 Dec 2025 22:57:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=46397279</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=46397279</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46397279</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Nano Banana Pro"]]></title><description><![CDATA[
<p>Oh yeah, funny enough even though I’m a bit of an AI art hater I actually thought very early Midjourney looked good because of all had an impressionistic, dreamy quality.</p>
]]></description><pubDate>Thu, 20 Nov 2025 23:45:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=45999493</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45999493</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45999493</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Nano Banana Pro"]]></title><description><![CDATA[
<p>Funny enough that had crossed my mind with the woodchuck example, because at a glance I can't see any weird artifacts, but I felt confident I could tell it was AI generated immediately if I saw it in the wild, and I couldn't really explain why.  My immediate guess was "well, who the hell would actually bother to make something like this?"</p>
]]></description><pubDate>Thu, 20 Nov 2025 18:54:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=45996236</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45996236</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45996236</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Nano Banana Pro"]]></title><description><![CDATA[
<p>Something I find weird about AI image generation models is that even though they no longer produce weird "artifacts" that give away that the fact that it was AI generated, you can still recognize that it's AI due to stylistic choices.<p>Not all examples they gave were like this.  The example they gave of the word "Typography" would have fooled me as human-made.  The infographics stood out though.  I would have immediately noticed that the String of Turtles infographic was AI generated because of the stylistic choices.  Same for the guide on how to make chai.  I would be "suspicious" of the example they gave of the weather forecast but wouldn't immediately flag at as AI generated.<p>Similar note, earlier I was able to tell if something was AI generated right off the bat by noticing that it had a "Deviant Art" quality to it.  My immediate guess is that certain sources of training data are over-represented.</p>
]]></description><pubDate>Thu, 20 Nov 2025 18:00:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=45995553</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45995553</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45995553</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Doomsday scoreboard"]]></title><description><![CDATA[
<p>Reminds me of the old gem of the Web 1.0 internet that was Exit Mundi</p>
]]></description><pubDate>Tue, 21 Oct 2025 22:20:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=45662453</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45662453</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45662453</guid></item><item><title><![CDATA[New comment by Bjorkbat in "America's future could hinge on whether AI slightly disappoints"]]></title><description><![CDATA[
<p>One of my most frustrating things regarding the potential of an AI bubble was some very smart and intelligent researcher being incredibly bullish on AI on Twitter because if you extrapolate graphs measuring AI's ability to complete long-duration tasks (<a href="https://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/" rel="nofollow">https://metr.org/blog/2025-03-19-measuring-ai-ability-to-com...</a>) or other benchmarks then by 2026 or 2027 then you've basically invented AGI.<p>I'm going to take his statements at face value and assume that he really does have faith in his own predictions and isn't trying to fleece us.<p>My gripe with this statement is that this prediction is based on proxies for capability that aren't particularly reliable.  To elaborate, the latest frontier models score something like 65% on SWE-bench, but I don't think they're as capable as a human that also scored 65%.  That isn't to say that they're incapable, but just that they aren't <i>as</i> capable as an equivalent human.  I think there's a very real chance that a model absolutely crushes the SWE-bench benchmark but still isn't quite ready to function as an independent software engineering agent.<p>So a lot of this bullishness basically hinges on the idea that if you extrapolate some line on a graph into the future, then by next year or the year after all white-collar work can be automated.  Terrifying as that is, this all hinges on the idea that these graphs, these benchmarks, are good proxies.<p>And if they aren't, oh wow.</p>
]]></description><pubDate>Tue, 14 Oct 2025 05:30:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=45576578</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45576578</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45576578</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Vibe engineering"]]></title><description><![CDATA[
<p>I definitely agree with you there.  I contracted with a company that had some older engineers who were in largely managerial roles who really liked using AI for personal projects, and honestly, I kind of get it.  Their work flow was basically prompt, get results, prompt again with modifications, rinse and repeat, it's low effort and has a nice REPL-like loop.  Paraphrasing a bit, but it basically re-kindled the joy of programming for them.<p>Haven't gotten the chance to ask, but I imagine managing a team of AI agents would feel a little too much like their day job, and consequently, suck the fun out of it.<p>That said, looking back, I think the reason why generative AI is so fun for so many coders is because programming has become unnecessarily complex.  I have to admit, programming nowadays for me feels like a bit of a slog at times because of the sheer effort it can sometimes take to implement the simplest things.  Doesn't have to be that way, but I think LLM copy-paste machines are probably the wrong direction.</p>
]]></description><pubDate>Wed, 08 Oct 2025 16:29:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=45517955</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45517955</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45517955</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Vibe engineering"]]></title><description><![CDATA[
<p>I think people underestimate the degree to which fun matters when it comes to productivity.  If something isn’t fun then I’ll likely put it off.  A 15 minute task can become hours, maybe days long, because I’m going to procrastinate on doing it.<p>If managing a bunch of AI agents is a very un-fun way to spend time, then I don’t think it’s the future.  If the new way of doing this is more work and more tedium, then why the hell have we collectively decided this is the new way to work when historically the approach has been to automate and abstract tedium so we can focus on what matters?<p>The people selling you the future of work don’t necessarily know better than you.</p>
]]></description><pubDate>Wed, 08 Oct 2025 14:37:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=45516677</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45516677</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45516677</guid></item><item><title><![CDATA[New comment by Bjorkbat in "Claude Sonnet 4.5"]]></title><description><![CDATA[
<p>Yeah, I thought about that after I looked at the SWE-bench results.  It doesn't make sense that the SWE results are barely an improvement yet somehow the model is a more significant improvement when it comes to long tasks.  You'd expect a huge gain in one to translate to the other.<p>Unless the main area of improvement was tools and scaffolding rather than the model itself.</p>
]]></description><pubDate>Tue, 30 Sep 2025 06:41:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=45422608</link><dc:creator>Bjorkbat</dc:creator><comments>https://news.ycombinator.com/item?id=45422608</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45422608</guid></item></channel></rss>