<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: jakswa</title><link>https://news.ycombinator.com/user?id=jakswa</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 04 Sep 2026 08:26:00 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=jakswa" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by jakswa in "Qwen 3.8 27B available on Cerebras at 1500 tokens/s"]]></title><description><![CDATA[
<p>dang only for certain nvidia GPUs, had my hopes up</p>
]]></description><pubDate>Thu, 03 Sep 2026 23:33:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49558583</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49558583</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49558583</guid></item><item><title><![CDATA[New comment by jakswa in "OpenAI begins rolling out GPT-6 Astra"]]></title><description><![CDATA[
<p>don't see it in my AWS bedrock model list yet, but boy has bedrock mantle been annoying today with the errors/downtimes, with NO status page entries >_<</p>
]]></description><pubDate>Thu, 03 Sep 2026 20:55:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49556833</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49556833</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49556833</guid></item><item><title><![CDATA[New comment by jakswa in "Show HN: We built open OpenRouter that turns usage into a better model"]]></title><description><![CDATA[
<p>Oh. The UI screenshot on github is... not actually in the github repo? It's platform/hosted only? There's my first awkward discovery, but makes sense in retrospect.</p>
]]></description><pubDate>Fri, 28 Aug 2026 20:01:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=49483508</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49483508</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49483508</guid></item><item><title><![CDATA[New comment by jakswa in "Show HN: We built open OpenRouter that turns usage into a better model"]]></title><description><![CDATA[
<p>I think I have to look at setting up one of these AI gateways for work, since AWS bedrock is such a PITA to hook a harness up to over IAM roles. Also I still can't believe bedrock hasn't released any open models in months (so there's paranoia that I'll want to swap in another provider).<p>Really though I'm hoping anthropic fixes the oppressive claude verbosity. I saw someone refer to being "clauderboarded" and my brain cannot let go of this as Claude's tokens bombard me.</p>
]]></description><pubDate>Fri, 28 Aug 2026 13:20:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49478084</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49478084</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49478084</guid></item><item><title><![CDATA[New comment by jakswa in "Ornith-1.5: From Self-Scaffolding to Self-Improvement"]]></title><description><![CDATA[
<p>ended up disabling ornith 9B. Oddly Ling 3 Tiny is pretty dang capable if its thinking is unleashed (tons of output tokens, maybe 5X the tokens but it's so fast it's maybe only twice as slow as a smarter model). This is a really interesting space if these super small models keep improving. SO FAST! Want to use!</p>
]]></description><pubDate>Thu, 20 Aug 2026 18:26:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49378312</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49378312</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49378312</guid></item><item><title><![CDATA[New comment by jakswa in "Ornith-1.5: From Self-Scaffolding to Self-Improvement"]]></title><description><![CDATA[
<p>I had to go down to UD-Q3_K_XL for Qwen 3.8 27B to get it to fit in VRAM and be usable, but I worry I'm gutting its intelligence somewhat. I too am interested in faster + more-usable alternative that can exchange blows with the Q3-dumbed 27B.</p>
]]></description><pubDate>Wed, 19 Aug 2026 17:50:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49364791</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49364791</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49364791</guid></item><item><title><![CDATA[New comment by jakswa in "Ornith-1.5: From Self-Scaffolding to Self-Improvement"]]></title><description><![CDATA[
<p>I'll be comparing the 9B vs Ling 3 Tiny (8B-A1B) as a scout model. Ling tiny is so fast but can be a little too dumb. Hope the 9B strikes a good middleground even if dense/slower.</p>
]]></description><pubDate>Wed, 19 Aug 2026 17:39:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49364652</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49364652</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49364652</guid></item><item><title><![CDATA[New comment by jakswa in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>I've been waiting on this model to show up on the Deep SWE benchmark results and treat its absence/delay as an indication of how slow and unusable it is for good results. I bet it thinks to the moon on some of those complex challenges.</p>
]]></description><pubDate>Mon, 17 Aug 2026 22:37:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49338619</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49338619</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49338619</guid></item><item><title><![CDATA[New comment by jakswa in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>Thanks for mentioning Ling 3 Tiny. This model has completely bypassed me and seems promising for how small it is.</p>
]]></description><pubDate>Mon, 17 Aug 2026 21:39:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=49338039</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49338039</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49338039</guid></item><item><title><![CDATA[New comment by jakswa in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>I'm waiting for this comparison too. I was impressed by a 1-shot GLM 5.3 did for me the other day.</p>
]]></description><pubDate>Mon, 17 Aug 2026 19:29:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49336364</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49336364</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49336364</guid></item><item><title><![CDATA[New comment by jakswa in "Launch HN: Speko (YC S26) – OpenRouter for Voice AI"]]></title><description><![CDATA[
<p>Gemma 4 (both E4B + 12B) performed really well as ears+brains. I mostly comment because I too am always scouting for a nice local all-in-one model.</p>
]]></description><pubDate>Mon, 17 Aug 2026 18:19:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49335347</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49335347</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49335347</guid></item><item><title><![CDATA[New comment by jakswa in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>I went back to Glimmer 30b for my 20GB of VRAM. Just a better experience fit-wise and speed-wise and tone-/voice-wise.</p>
]]></description><pubDate>Mon, 17 Aug 2026 02:55:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49326077</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49326077</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49326077</guid></item><item><title><![CDATA[New comment by jakswa in "Gemini 3.7 Flash"]]></title><description><![CDATA[
<p>thank you I'm still on ie8</p>
]]></description><pubDate>Sat, 15 Aug 2026 23:13:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49315176</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49315176</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49315176</guid></item><item><title><![CDATA[New comment by jakswa in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>I'm in the exact same boat with a 7900 XT and a good Glimmer 30B experience. I was really hoping qwen 3.8 would bring some memory/space efficiency savings along the lines of whatever is going on with Glimmer 30B. I have been surprised that a 30 billion model fits and runs better (at higher unsloth quantization! UD-Q4_K_XL fits!) than a 27 billion model.</p>
]]></description><pubDate>Sat, 15 Aug 2026 00:20:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49306222</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49306222</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49306222</guid></item><item><title><![CDATA[New comment by jakswa in "Qwen 3.8 27B"]]></title><description><![CDATA[
<p>anecdotes: 35B-A3B does want more memory, bigger model. But if you get it running it will be faster and more enjoyable to use -- text will fly by -- due to only 3B params being active, in my experience at least.</p>
]]></description><pubDate>Fri, 14 Aug 2026 18:26:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49302714</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49302714</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49302714</guid></item><item><title><![CDATA[New comment by jakswa in "How Compaction Works in Pi"]]></title><description><![CDATA[
<p>I dunno, I didn't read in-depth. Hopefully you don't gotta zoom in with human eyeballs.</p>
]]></description><pubDate>Fri, 14 Aug 2026 07:19:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49295596</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49295596</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49295596</guid></item><item><title><![CDATA[New comment by jakswa in "How Compaction Works in Pi"]]></title><description><![CDATA[
<p>OMP changed the default compaction to <i>images</i>! Kinda nuts to read about. Saves the generation cost of the traditional compaction step and writes the context as tiny text to an image, if I was following correctly.</p>
]]></description><pubDate>Thu, 13 Aug 2026 22:56:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49292778</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49292778</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49292778</guid></item><item><title><![CDATA[New comment by jakswa in "Gemini 3.7 Flash"]]></title><description><![CDATA[
<p>my whole world is shifting. have I been seeing _different pelicans_ from everyone else?!</p>
]]></description><pubDate>Thu, 13 Aug 2026 19:14:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49290624</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49290624</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49290624</guid></item><item><title><![CDATA[New comment by jakswa in "Gemini 3.7 Flash"]]></title><description><![CDATA[
<p>This pelican gave me a good laugh, because there's enough reasoning that the render is out of sight initially. The buildup!</p>
]]></description><pubDate>Thu, 13 Aug 2026 18:34:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49290134</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49290134</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49290134</guid></item><item><title><![CDATA[New comment by jakswa in "Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows"]]></title><description><![CDATA[
<p>I'll back up your smaller claim, but be specific that it's UD-Q4_K_XL size:<p>- muse glimmer: 15.9GB<p>- qwen 3.6 27B: 17.6GB<p>My video card is so close to its limit that these GB thresholds are mattering too much for me :D</p>
]]></description><pubDate>Mon, 10 Aug 2026 16:58:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49246456</link><dc:creator>jakswa</dc:creator><comments>https://news.ycombinator.com/item?id=49246456</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49246456</guid></item></channel></rss>