<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: ismael_rr</title><link>https://news.ycombinator.com/user?id=ismael_rr</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 08 Oct 2026 06:07:26 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=ismael_rr" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by ismael_rr in "GPT‑6 and Intelligent UI for everyone"]]></title><description><![CDATA[
<p>I suspect that 6 sol was a better version of 5.6 terra (and note that in the 6 sol and luna release, they took out terra, and price 6 sol at 5.6 terra pricing), then the backlash from lesser capabilities made them roll out 6.1 sol as the actual 5.6 sol - size modee.</p>
]]></description><pubDate>Wed, 07 Oct 2026 19:21:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49997548</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49997548</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49997548</guid></item><item><title><![CDATA[New comment by ismael_rr in "Mistral Large 4"]]></title><description><![CDATA[
<p>Same with the US legal system protecting Anthropic from copyright laws from all the books and other stuff used for pretraining...</p>
]]></description><pubDate>Tue, 06 Oct 2026 14:54:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49979436</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49979436</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49979436</guid></item><item><title><![CDATA[New comment by ismael_rr in "GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price"]]></title><description><![CDATA[
<p>Distillation is also a broad term - I think most specifically, it refers to training a smaller model on a larger/better model's full output token distribution rather than normal pretraining, which only can access the next token in the data that was actually used.<p>It's also used to describe the SFT bootstrapping for posttraining, which is what people generally refer to as Chinese labs "distilling".<p>I would almost guarantee that smaller US frontier models (ex Luna/Sonnet) are distilled from their respective large models.</p>
]]></description><pubDate>Wed, 30 Sep 2026 01:26:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49903262</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49903262</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49903262</guid></item><item><title><![CDATA[New comment by ismael_rr in "Show HN: Share your AI Setup, Learn from others"]]></title><description><![CDATA[
<p>haha was also going to mention this at the risk of pedantry</p>
]]></description><pubDate>Thu, 17 Sep 2026 20:45:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49746309</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49746309</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49746309</guid></item><item><title><![CDATA[New comment by ismael_rr in ">10x More Efficient Pretraining"]]></title><description><![CDATA[
<p>Super awesome. Wish they would release the paper about what they did to achieve this. I remember nous released the token superposition paper which improved pretraining FLOPs some, but not 50x: <a href="https://nousresearch.com/token-superposition" rel="nofollow">https://nousresearch.com/token-superposition</a>. Wondering if they also found some cool tokenization strategiesa</p>
]]></description><pubDate>Thu, 10 Sep 2026 15:52:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49645831</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49645831</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49645831</guid></item><item><title><![CDATA[I gave frontier LLMs a canvas and told them to draw a self-portrait]]></title><description><![CDATA[
<p>Article URL: <a href="https://ismaelroblesrazzaq.github.io/blog/llm-self-portraits/">https://ismaelroblesrazzaq.github.io/blog/llm-self-portraits/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49362780">https://news.ycombinator.com/item?id=49362780</a></p>
<p>Points: 3</p>
<p># Comments: 3</p>
]]></description><pubDate>Wed, 19 Aug 2026 15:19:22 +0000</pubDate><link>https://ismaelroblesrazzaq.github.io/blog/llm-self-portraits/</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49362780</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49362780</guid></item><item><title><![CDATA[New comment by ismael_rr in "Models Are Getting Dumber on Purpose"]]></title><description><![CDATA[
<p>In my experience, it feels like this problem has been mostly solved. When I interact with a chatbot, I see it often uses web search to answer my questions and hallucinates much less then they used to. Furthermore, I suspect a lot of gains are merely from the system prompt instructing the chatbot to use web search to verify all facts it says.</p>
]]></description><pubDate>Sun, 16 Aug 2026 20:03:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49323146</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49323146</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49323146</guid></item><item><title><![CDATA[New comment by ismael_rr in "Models Are Getting Dumber on Purpose"]]></title><description><![CDATA[
<p>I agree that most AI use in terms of users may not be for coding or agentic use (everyday people are asking chatgpt for something or looking at google ai summary), but with respect to AI usage, I speculate that the vast majority of usage is coding and agentic because they're super token hungry.<p>In terms of the value proposition of AI replacing knowledge workers, all value is in coding agents (coding agents as general agents).</p>
]]></description><pubDate>Sun, 16 Aug 2026 20:00:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49323131</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49323131</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49323131</guid></item><item><title><![CDATA[New comment by ismael_rr in "Ask HN: What are you working on? (August 2026)"]]></title><description><![CDATA[
<p>for you tennis players out there, I'm working on RallyClip, an app for free tennis match segmentation. It's a free and open source alternative to SwingVision:<p>github: <a href="https://github.com/iroblesrazzaq/RallyClip" rel="nofollow">https://github.com/iroblesrazzaq/RallyClip</a>
website: rallyclip.app<p>I'm working on model performance and an iOS app right now; would love feedback on the macOS desktop app.</p>
]]></description><pubDate>Sun, 09 Aug 2026 23:35:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49237401</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=49237401</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49237401</guid></item><item><title><![CDATA[New comment by ismael_rr in "Show HN: Zanagrams"]]></title><description><![CDATA[
<p>It was fun! I like how unlike the nyt games one, you can get additional information from the edges' existence. Would advise removing having both plural/singular forms of a word. Great work!</p>
]]></description><pubDate>Sun, 28 Jun 2026 22:14:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=48712277</link><dc:creator>ismael_rr</dc:creator><comments>https://news.ycombinator.com/item?id=48712277</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48712277</guid></item></channel></rss>