<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: joshheitzman</title><link>https://news.ycombinator.com/user?id=joshheitzman</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 17 Aug 2026 03:57:36 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=joshheitzman" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by joshheitzman in "Claude Seems Down"]]></title><description><![CDATA[
<p>Personally I don't care about common benchmarks as I don't find actual coding agent performance correlates strongly with them.  One reason to use open-weight models is that they don't hide the reasoning, so you can do very aggressive context management in your harness to use significantly less tokens.  Smaller prompts are faster since KV is N^2 plus it can be dramatically cheaper (if you balance your aggressive context management with maintaining the prefix cache as much as possible).  Even paying for API prices directly and using agents as much as I want I spend less per month than the $200 I spent on a Claude MAX sub when I had it.</p>
]]></description><pubDate>Mon, 17 Aug 2026 00:48:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=49325381</link><dc:creator>joshheitzman</dc:creator><comments>https://news.ycombinator.com/item?id=49325381</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49325381</guid></item><item><title><![CDATA[New comment by joshheitzman in "Claude Seems Down"]]></title><description><![CDATA[
<p>Dozens of providers of open-weight models.  I have one session going with synthetic.new and another going with novita.ai right now.</p>
]]></description><pubDate>Sun, 16 Aug 2026 22:18:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49324306</link><dc:creator>joshheitzman</dc:creator><comments>https://news.ycombinator.com/item?id=49324306</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49324306</guid></item><item><title><![CDATA[New comment by joshheitzman in "Models Are Getting Dumber on Purpose"]]></title><description><![CDATA[
<p>As RAM prices are so nuts, I picked up a used gaming laptop early this year with an RTX 3070 (i.e. 8GB) to match the specs for my gaming tower that's run everything I'm interested in just fine.  That includes recent Unreal 5.x games (Satisfactory, Fortnite, etc.) with high graphics settings on a 4k TV (60hz).  There aren't a ton of games that require more than 8GB of GPU ram.</p>
]]></description><pubDate>Sun, 16 Aug 2026 21:15:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=49323753</link><dc:creator>joshheitzman</dc:creator><comments>https://news.ycombinator.com/item?id=49323753</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49323753</guid></item></channel></rss>