<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: aliljet</title><link>https://news.ycombinator.com/user?id=aliljet</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 04 Sep 2026 09:20:34 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=aliljet" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by aliljet in "The largest electric aircraft just flew [video]"]]></title><description><![CDATA[
<p>To be clear, this is a hybrid aircraft, but it's still an awesome step forward!</p>
]]></description><pubDate>Thu, 03 Sep 2026 21:51:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49557617</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49557617</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49557617</guid></item><item><title><![CDATA[New comment by aliljet in "GPT-6 Astra"]]></title><description><![CDATA[
<p>The ARC-AGI-3 score is ridiculously high. Is this benchmaxxing or something way different? It's really hard to discern how we're approaching breakthroughs...</p>
]]></description><pubDate>Thu, 03 Sep 2026 19:48:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49555743</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49555743</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49555743</guid></item><item><title><![CDATA[New comment by aliljet in "GPT-6 Astra"]]></title><description><![CDATA[
<p>The ARCC-AGI-3 performance is absolutely incredible. The magnitude of change here is so high that I'm almost incredulous. Is this real? Did the benchmark get gamed?</p>
]]></description><pubDate>Thu, 03 Sep 2026 19:15:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49555205</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49555205</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49555205</guid></item><item><title><![CDATA[New comment by aliljet in "ChatGPT Is Throwing 404"]]></title><description><![CDATA[
<p>This is probably Astra getting ready for release.</p>
]]></description><pubDate>Thu, 03 Sep 2026 15:08:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49551161</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49551161</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49551161</guid></item><item><title><![CDATA[New comment by aliljet in "At-home test for infected ticks could improve Lyme Disease diagnosis"]]></title><description><![CDATA[
<p>Does diagnosis offer a better prognosis for those that are infected?</p>
]]></description><pubDate>Sat, 15 Aug 2026 16:19:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49311821</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49311821</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49311821</guid></item><item><title><![CDATA[New comment by aliljet in "GLM-5.3: Frontier coding with emergent cyber capabilities"]]></title><description><![CDATA[
<p>This is absolutely still shy of Sol and Fable, but only just by a hair. Ridiculous results. There's still not a compelling economic reason to drop OpenAI courtesy of the ludicrous reset addiction that's taken place, but it feels like we're on the precipice.<p>How are you all toying with running this kind of thing in a mega quantized way locally? Two weeks out from released weights, but this is still just GLM 5.2 with post-training magic.</p>
]]></description><pubDate>Fri, 14 Aug 2026 05:31:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49295043</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49295043</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49295043</guid></item><item><title><![CDATA[New comment by aliljet in "Mistral OCR 4.1"]]></title><description><![CDATA[
<p>Accuracy is truly what people die for in the OCR game. Price isn't the primary function here.. it's an equation of price, accuracy, speed, and in mayn cases regulation.</p>
]]></description><pubDate>Thu, 13 Aug 2026 18:52:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49290363</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49290363</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49290363</guid></item><item><title><![CDATA[New comment by aliljet in "Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index"]]></title><description><![CDATA[
<p>Can you explain what you mean? These days courtesy of an addictive reset game OpenAI is playing, I can't find anything with frontier intelligence that's more cost efficient...</p>
]]></description><pubDate>Wed, 12 Aug 2026 17:11:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49275618</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49275618</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49275618</guid></item><item><title><![CDATA[New comment by aliljet in "Pixel 11 Pro Fold"]]></title><description><![CDATA[
<p>What genuinely disappointing result. Long time pixel user here and I've been routinely buying these phones with the argument that you're getting the most value of any modern smart phone. Now? I'm just waiting for the pixel 10 family to drop in price. Happy to just wait.</p>
]]></description><pubDate>Wed, 12 Aug 2026 16:56:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49275416</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49275416</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49275416</guid></item><item><title><![CDATA[New comment by aliljet in "Qwen3.8 Max now ranked as the best overall model by agentic index"]]></title><description><![CDATA[
<p>Is there a path to distill this model to do very specific things? Like a RAG strategy for a small (or even large) corpus?</p>
]]></description><pubDate>Thu, 06 Aug 2026 19:18:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49201063</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49201063</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49201063</guid></item><item><title><![CDATA[New comment by aliljet in "Beating GPT-5.6 Sol on retrieval with 100x cheaper open models"]]></title><description><![CDATA[
<p>There is a more serious question in here that's not being answered. How effective is the retrieval in finding buried needles in larger and larger haystacks. And there's a correlary question, how effective could you be in finding paired needles in that haystack where you need to hold a needle to unlock finding another needle.</p>
]]></description><pubDate>Wed, 05 Aug 2026 19:13:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49187553</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49187553</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49187553</guid></item><item><title><![CDATA[New comment by aliljet in "Qwen3.8-Max: A New Bar for Coding and Cowork"]]></title><description><![CDATA[
<p>I'm trying and failing to find value running a potential Qwen 3.8 27b dense model on a 16 core, 128 GB of ram, 2080ti box. Yes, the GPU yells for help, but the problem is that no math works to upgrade this machine even when pouring $200 in rent every month into the large model providers...<p>How are you all justifying economical use of these local models right now? What's the cost efficient way to do this and do better (even with models evolving over time and losing now vs later) than the big labs?</p>
]]></description><pubDate>Mon, 03 Aug 2026 04:33:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49151236</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49151236</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49151236</guid></item><item><title><![CDATA[New comment by aliljet in "RTX 2080 Ti Memory Upgrade to 22 GB"]]></title><description><![CDATA[
<p>Honestly, I have a 2080ti that I use to play and I can tell you the math isn't there to upgrade it. It's much easier to just find a 3090/4090/5090 and keep pace with the software and hardware simultaneously.</p>
]]></description><pubDate>Tue, 28 Jul 2026 04:05:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49079207</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=49079207</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49079207</guid></item><item><title><![CDATA[New comment by aliljet in "Show HN: Getting GLM 5.2 running on my slow computer"]]></title><description><![CDATA[
<p>I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...</p>
]]></description><pubDate>Fri, 10 Jul 2026 00:48:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=48854424</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48854424</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48854424</guid></item><item><title><![CDATA[New comment by aliljet in "Qwen-AgentWorld: Language World Models for General Agents"]]></title><description><![CDATA[
<p>The benchmarks here are confusing at best. Am I reading correctly that this model is essentially as good or better than all frontier models right now?</p>
]]></description><pubDate>Wed, 24 Jun 2026 07:24:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=48656399</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48656399</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48656399</guid></item><item><title><![CDATA[New comment by aliljet in "Mistral OCR 4"]]></title><description><![CDATA[
<p>I was just using infinity parser 2 (flash, to be fair) for pennies self-hosted to run through thousands of pages of documents with remarkable confidence. I decided to use <a href="https://huggingface.co/datasets/allenai/olmOCR-bench" rel="nofollow">https://huggingface.co/datasets/allenai/olmOCR-bench</a> to determine what was the best OCR tool, yesterday, but I've got no idea what the best is now. What is the dominant OCR eval right now? Between Baidu and Mistral this morning, I wonder if there's a new tool to switch to..</p>
]]></description><pubDate>Tue, 23 Jun 2026 15:26:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=48646583</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48646583</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48646583</guid></item><item><title><![CDATA[New comment by aliljet in "Unlimited OCR: One-Shot Long-Horizon Parsing"]]></title><description><![CDATA[
<p>I'm curious about this. What models/tools have you been using?</p>
]]></description><pubDate>Tue, 23 Jun 2026 15:10:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=48646269</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48646269</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48646269</guid></item><item><title><![CDATA[New comment by aliljet in "Unlimited OCR: One-shot long-horizon parsing"]]></title><description><![CDATA[
<p>How does this compare with infinty parser 2 which seemed to be running the table on every other OCR tool (<a href="https://huggingface.co/datasets/allenai/olmOCR-bench" rel="nofollow">https://huggingface.co/datasets/allenai/olmOCR-bench</a>). To be fair, there's no single winning OCR benchmark and this isn't showing up anywhere yet..</p>
]]></description><pubDate>Tue, 23 Jun 2026 15:02:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=48646125</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48646125</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48646125</guid></item><item><title><![CDATA[New comment by aliljet in "Qwen-Robot Suite: A Foundation Model Suite for Physical World Intelligence"]]></title><description><![CDATA[
<p>This sounds incredible. Have these models effectively solved the problem of trying to use a fast-processing network to predict the world's state ahead? For example, to catch a ball?</p>
]]></description><pubDate>Tue, 16 Jun 2026 20:07:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=48561258</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48561258</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48561258</guid></item><item><title><![CDATA[New comment by aliljet in "Running local models is good now"]]></title><description><![CDATA[
<p>The problem here is always the cost-benefit. For $200/mo, you're receiving subsidized best of breed access. There's no model competing for that price anywhere. If a 27B param model is what you choose, show me your hardware! I would love to be wrong...</p>
]]></description><pubDate>Tue, 16 Jun 2026 15:44:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=48557062</link><dc:creator>aliljet</dc:creator><comments>https://news.ycombinator.com/item?id=48557062</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48557062</guid></item></channel></rss>