<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: guybedo</title><link>https://news.ycombinator.com/user?id=guybedo</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sun, 04 Oct 2026 08:19:37 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=guybedo" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by guybedo in "Claude Opus 5.5"]]></title><description><![CDATA[
<p>i think we shouldn't mix things here.<p>Opus 5.5 isn't the frontier, when they say 'pacing the frontier', it's about internal models not yet released, as they're probably one or two generations ahead already.</p>
]]></description><pubDate>Tue, 22 Sep 2026 20:13:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49807427</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49807427</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49807427</guid></item><item><title><![CDATA[New comment by guybedo in "Claude Status – Elevated errors for multiple models"]]></title><description><![CDATA[
<p>problems with Grok too ... so i guess Grok 4.7 tried to escape and took control of Colossus datacenter.</p>
]]></description><pubDate>Tue, 22 Sep 2026 01:25:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49795724</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49795724</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49795724</guid></item><item><title><![CDATA[Open Source JEV architecture built 1 year ago]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.reddit.com/r/LocalLLaMA/comments/1wijo3e/i_literally_built_the_jev_architecture_one_year/">https://www.reddit.com/r/LocalLLaMA/comments/1wijo3e/i_literally_built_the_jev_architecture_one_year/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49764070">https://news.ycombinator.com/item?id=49764070</a></p>
<p>Points: 5</p>
<p># Comments: 1</p>
]]></description><pubDate>Sat, 19 Sep 2026 07:00:11 +0000</pubDate><link>https://www.reddit.com/r/LocalLLaMA/comments/1wijo3e/i_literally_built_the_jev_architecture_one_year/</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49764070</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49764070</guid></item><item><title><![CDATA[New comment by guybedo in "Astra for Law"]]></title><description><![CDATA[
<p>yes, most of the time i use Astra Low already.</p>
]]></description><pubDate>Fri, 18 Sep 2026 00:29:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49748709</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49748709</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49748709</guid></item><item><title><![CDATA[New comment by guybedo in "Real-SWE: Benchmarking AI models on private, real-world, enterprise codebases"]]></title><description><![CDATA[
<p>SHA-256-hash-verified sealed package artifact with automatic reconciliation system p95<0.5ms</p>
]]></description><pubDate>Sat, 12 Sep 2026 23:46:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49678443</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49678443</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49678443</guid></item><item><title><![CDATA[New comment by guybedo in "Procedural Graphs: Self-Evolving Execution Structures for LLM Agents"]]></title><description><![CDATA[
<p>your load-bearing thesis is probably interesting but it seems i can't read AI written text anymore -- or maybe i just need some more coffee.</p>
]]></description><pubDate>Wed, 09 Sep 2026 20:47:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49634026</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49634026</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49634026</guid></item><item><title><![CDATA[New comment by guybedo in "Project HydraFusion: Frontier quality via multi-model orchestration"]]></title><description><![CDATA[
<p>i've been using adversarial critique and reviews for many planning, solution design and implementation steps inside workflows.<p>It's so effective and helps catching so many design flaws, implementations misses etc ... that i'm wondering how people manage to build complex/large projects with agents without this kind of process. Well, i actually built this thing because i couldn't get good results so i had to find a way.<p>I'm gonna open source the whole thing but it needs some cleanup, there's a basic landing page here <a href="https://kodfactory.com" rel="nofollow">https://kodfactory.com</a> if anyone wants to be notified when it's released on github. Yeah i know, the world really needs another software factory :-)</p>
]]></description><pubDate>Fri, 04 Sep 2026 18:00:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49568002</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49568002</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49568002</guid></item><item><title><![CDATA[New comment by guybedo in "Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?"]]></title><description><![CDATA[
<p>LOAD BEARING</p>
]]></description><pubDate>Thu, 03 Sep 2026 17:26:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49553553</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49553553</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49553553</guid></item><item><title><![CDATA[New comment by guybedo in "Any Human Ever – One life, drawn at random from all who have ever lived"]]></title><description><![CDATA[
<p>we are probably living inside a more advanced version of this game right now.<p>Somebody, somewhere, somewhen rolled the dice and here we are.</p>
]]></description><pubDate>Thu, 03 Sep 2026 15:25:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49551580</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49551580</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49551580</guid></item><item><title><![CDATA[New comment by guybedo in "ChatGPT Is Throwing 404"]]></title><description><![CDATA[
<p>Civilization IV just took over</p>
]]></description><pubDate>Thu, 03 Sep 2026 15:13:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49551279</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49551279</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49551279</guid></item><item><title><![CDATA[Muse Spark 1.3]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.bloomberg.com/news/articles/2026-09-02/meta-releases-more-powerful-ai-model-edging-closer-to-rivals">https://www.bloomberg.com/news/articles/2026-09-02/meta-releases-more-powerful-ai-model-edging-closer-to-rivals</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49541194">https://news.ycombinator.com/item?id=49541194</a></p>
<p>Points: 5</p>
<p># Comments: 1</p>
]]></description><pubDate>Wed, 02 Sep 2026 19:30:04 +0000</pubDate><link>https://www.bloomberg.com/news/articles/2026-09-02/meta-releases-more-powerful-ai-model-edging-closer-to-rivals</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49541194</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49541194</guid></item><item><title><![CDATA[New comment by guybedo in "GLM-5.3 is now open-weight"]]></title><description><![CDATA[
<p>I have a dual epyc + 1TB RAM.
I could push glm 5.2 to 7 tok/s CPU only.</p>
]]></description><pubDate>Fri, 28 Aug 2026 17:49:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49482075</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49482075</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49482075</guid></item><item><title><![CDATA[New comment by guybedo in "GLM-5.3-Flash"]]></title><description><![CDATA[
<p>yeah i mostly agree, especially compared to subsidized subscription cost.<p>But for a heavy user who has enough work to be done so that the box runs almost 24/7 at say 50tok/sec, the math gets interesting against API prices.<p>And it can be interesting compared to subscription in the sense that you don't have the quota anymore. That means there's probably a lot of things you're not doing because of the quotas that you could do now.<p>It depends heavily on the tok/sec obviously and the very best solution financially remains subscriptions. But the idea remains entertaining and not that disconnected from reality</p>
]]></description><pubDate>Wed, 26 Aug 2026 23:24:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49457271</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49457271</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49457271</guid></item><item><title><![CDATA[New comment by guybedo in "GLM-5.3-Flash"]]></title><description><![CDATA[
<p>although i initially thought it didn't make sense financially to run this kind of model locally, i did run the numbers and for heavy users this could justify buying $10k worth of hardware with a ROI over a few months, less than a year.<p>I was looking at my token usage, mostly from subsidized codex/grok subscriptions and i'm a somewhat heavy user. The thing is i would actually use even more tokens if it wasn't for the weekly quotas.<p>In the end, with a $10k investment and running this kind of model, estimating a 2x increase in token usage because i wouldn't have weekly quotas and comparing to glm api prices, this thing could pay for itself in less than a year.<p>Obviously i'm paying subscription price right now, so the math doesn't work. Although using local ai removes all weekly quotas. Keep a subscription to have access to frontier models for planning work, and local hardware + glm-5.3 flash for implementation, e2e testing, qa work 24/7.<p>It's not that crazy of an idea and the numbers aren't that bad.</p>
]]></description><pubDate>Wed, 26 Aug 2026 22:07:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49456575</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49456575</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49456575</guid></item><item><title><![CDATA[New comment by guybedo in "The Vibe Tax"]]></title><description><![CDATA[
<p>i'm not sure why people expect agents to one shot everything to perfection with just a prompt.<p>There's a reason why we talk about software development lifecycle, design, architecture, testing ... It's because it's been the most reliable way to build and ship software. We shouldn't expect discard this and expect agents to perform well outside of this.<p>I'm treating LLM agents as junior devs who happen to have vast knowledge of software engineering. As their team leader i make them go through planning, implementation, bug sweeping cycles using strict workflows. And it works quite well, i've been working on several large projects (1M+ LOC java,typescript,c/c++) and by any measure the projects are healthy. Sure the code isn't that beautiful, sure i'd have written things differently but it's pretty good nonetheless.<p>Shameless plug here: i've been also working on <a href="https://kodfactory.com" rel="nofollow">https://kodfactory.com</a>, the code factory i've built to work on these large projects with workflows, reviews, etc ... I'm cleaning things up to open source it later.</p>
]]></description><pubDate>Sun, 23 Aug 2026 21:15:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49412738</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49412738</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49412738</guid></item><item><title><![CDATA[New comment by guybedo in "Eight Myths on Software Engineering and GenAI"]]></title><description><![CDATA[
<p>although, if i'm out of tokens and have to wait a full day, i won't bother doing some things manually because the day i'll spend doing something won't take more than 1 hour the next day when tokens are available again.</p>
]]></description><pubDate>Wed, 05 Aug 2026 01:46:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49177636</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49177636</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49177636</guid></item><item><title><![CDATA[New comment by guybedo in "What's the largest software project AI can complete on its own?"]]></title><description><![CDATA[
<p>i've been working on several rather large projects these past few months, and i'm trying to write as little code as possible.<p>I don't think i wrote more than 10 lines of code in the largest project i'm working on. Lines of code: Java: 900_635, typescript: 725_418, C++: 180_445, Dart: 96_181.<p>It's been obvious from the start that no model, as good as it is, can do large(-ish) amounts of work by its own without supervision, control, criticism, etc ... If left unsupervised, models usually do half the work, leaving stubs and todos everywhere.<p>Quality comes from applying software engineering principles as much as possible, just like you would do with teams of junior devs: planning sessions and implementation sessions with adversarial critiques, specifying as much as possible upfront, planning unit/smoke/integration tests, etc ...<p>Many systems rely on swarm of agents to build software but i've found it very difficult to get good results without lots of overhead/token waste because of inter agent communications mostly.<p>So instead i built what is mostly a workflow engine to structure / organize processes into workflows with different agents assigned different roles. I've setup a basic landing page here <a href="https://kodfactory.com" rel="nofollow">https://kodfactory.com</a> if anyone wants to follow along.</p>
]]></description><pubDate>Mon, 03 Aug 2026 18:41:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49159777</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49159777</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49159777</guid></item><item><title><![CDATA[New comment by guybedo in "Run Kimi K3 using 29 GB of RAM at 0.50 tok/s"]]></title><description><![CDATA[
<p>quite funny to see HN crowd downvoting this.<p>It looks like some people have a hard time accepting that software written with ai isn't a fad, it's here, it won't go away and it can be interesting and useful for the creator and for users.<p>Many people on the other hand have moved on and are now team leaders, except their team is mostly AIs instead of junior devs. Trade off: code is usually worse, but in the end AIs are more capable with vast knowledge and speed.<p>That doesn't mean we have to use AI everywhere, all the time though.</p>
]]></description><pubDate>Sat, 01 Aug 2026 16:45:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49135982</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49135982</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49135982</guid></item><item><title><![CDATA[New comment by guybedo in "Run Kimi K3 using 29 GB of RAM at 0.50 tok/s"]]></title><description><![CDATA[
<p>are you using a code editor ? a compiler ? something that you didn't write yourself in pure assembly ? are you a real software engineer then ? where's the limit ?<p>I don't know why some people are so angry at AI/software writing with AI. It's just like being a team leader with junior(-ish) devs on the team. You don't write the code yourself, you give directions, you help/refactor/optimize where you can, that's the job.<p>Yes sometimes you need a team to do something, you can't code everything by yourself.<p>For reference, i'm not OP.</p>
]]></description><pubDate>Sat, 01 Aug 2026 16:41:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49135941</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49135941</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49135941</guid></item><item><title><![CDATA[New comment by guybedo in "Run Kimi K3 using 29 GB of RAM at 0.50 tok/s"]]></title><description><![CDATA[
<p>i thought, here on HN, we were past the "oooh it's written by a LLM it's bad!".<p>I care about the craft, well designed systems, good clean architecture and code, etc...<p>But i also care about reaching goals. Whether i do it working on my own, or with human coworkers or with AI coworkers doesn't matter that much to me. Yes, the result is sometimes the most important thing.</p>
]]></description><pubDate>Fri, 31 Jul 2026 20:55:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49128416</link><dc:creator>guybedo</dc:creator><comments>https://news.ycombinator.com/item?id=49128416</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49128416</guid></item></channel></rss>