<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: chisleu</title><link>https://news.ycombinator.com/user?id=chisleu</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 03 Sep 2026 10:52:08 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=chisleu" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by chisleu in "Great resources for self-hosting AI hardware"]]></title><description><![CDATA[
<p>Many are investing in RTX6000pro hardware but hitting deployment walls. This guide covers everything from simple 4x builds to scaling up to 16 cards with PCIe switches.</p>
]]></description><pubDate>Thu, 09 Apr 2026 16:21:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=47705633</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=47705633</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47705633</guid></item><item><title><![CDATA[Great resources for self-hosting AI hardware]]></title><description><![CDATA[
<p>Article URL: <a href="https://github.com/voipmonitor/rtx6kpro">https://github.com/voipmonitor/rtx6kpro</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=47705632">https://news.ycombinator.com/item?id=47705632</a></p>
<p>Points: 2</p>
<p># Comments: 1</p>
]]></description><pubDate>Thu, 09 Apr 2026 16:21:18 +0000</pubDate><link>https://github.com/voipmonitor/rtx6kpro</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=47705632</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47705632</guid></item><item><title><![CDATA[New comment by chisleu in "Google ADK-Go"]]></title><description><![CDATA[
<p>An open-source, code-first Go toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.</p>
]]></description><pubDate>Sun, 09 Nov 2025 22:23:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=45869794</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45869794</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45869794</guid></item><item><title><![CDATA[Google ADK-Go]]></title><description><![CDATA[
<p>Article URL: <a href="https://github.com/google/adk-go">https://github.com/google/adk-go</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=45869793">https://news.ycombinator.com/item?id=45869793</a></p>
<p>Points: 2</p>
<p># Comments: 1</p>
]]></description><pubDate>Sun, 09 Nov 2025 22:23:46 +0000</pubDate><link>https://github.com/google/adk-go</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45869793</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45869793</guid></item><item><title><![CDATA[New comment by chisleu in "Hacking India's largest automaker: Tata Motors"]]></title><description><![CDATA[
<p>Total tangent, but I got to ride in some of these on a recent trip to India and I was really impressed with the build quality and utilitarian usefulness of the design.</p>
]]></description><pubDate>Sat, 01 Nov 2025 14:10:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=45781760</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45781760</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45781760</guid></item><item><title><![CDATA[New comment by chisleu in "Qwen3-Omni: Native Omni AI model for text, image and video"]]></title><description><![CDATA[
<p>Here is the demo video on it. The video w/ sound input -> sound output while doing translation from the video to another language was the most impressive display I've seen yet.<p><a href="https://www.youtube.com/watch?v=_zdOrPju4_g" rel="nofollow">https://www.youtube.com/watch?v=_zdOrPju4_g</a></p>
]]></description><pubDate>Mon, 22 Sep 2025 18:45:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=45337748</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45337748</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45337748</guid></item><item><title><![CDATA[New comment by chisleu in "iPhone Air"]]></title><description><![CDATA[
<p>Because of the prompt processing speed, small models like Qwen 3 coder 30b a3b are the sweet spot for mac platform right now. Which means a 32 or 64GB mac is all you need to use Cline or your favorite agent locally.</p>
]]></description><pubDate>Wed, 10 Sep 2025 18:01:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=45201441</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45201441</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45201441</guid></item><item><title><![CDATA[New comment by chisleu in "GLM 4.5 with Claude Code"]]></title><description><![CDATA[
<p>I've been using GLM 4.5 and GLM 4.5 Air for a while now. The Air model is light enough to run on a macbook pro and is useful for Cline. I can run the full GLM model on my Mac Studio, but the TPS is so slow that it's only useful for chatting. So I hooked up with openrouter to try but didn't have the same success. Any of the open weight models I try with open router give sub standard results. I get better results from Qwen 3 coder 30b a3b locally than I get from Qwen 3 Coder 480b through open router.<p>I'm really concerned that some of the providers are using quantized versions of the models so they can run more models per card and larger batches of inference.</p>
]]></description><pubDate>Sat, 06 Sep 2025 02:07:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=45145959</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45145959</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45145959</guid></item><item><title><![CDATA[New comment by chisleu in "Updates to Consumer Terms and Privacy Policy"]]></title><description><![CDATA[
<p>This is going to improve the quality of LLM responses for users. I'm for this.</p>
]]></description><pubDate>Fri, 29 Aug 2025 13:11:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=45063658</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45063658</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45063658</guid></item><item><title><![CDATA[New comment by chisleu in "AWS in 2025: Stuff you think you know that's now wrong"]]></title><description><![CDATA[
<p>> but eventually they did it<p>They can do this with manual partitioning indeed. I've done it before, but it's not ideal because the auto partitioner will scale beyond almost anything AWS will give you with manual partitioning unless you have 24/7 workloads.<p>> you can be throttled for more than a day before it kicks in<p>I expect that this would depends on your use case. If you are dropping content you need to scale out to tons of readers, that is absolutely the case. If you are dropping tons of content with well distributed reads, then the auto partitioner is The Way.</p>
]]></description><pubDate>Sun, 24 Aug 2025 19:49:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=45007149</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45007149</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45007149</guid></item><item><title><![CDATA[New comment by chisleu in "AWS in 2025: Stuff you think you know that's now wrong"]]></title><description><![CDATA[
<p>and indeed the bucket is not separate from the object key. the API separates it logically "for humans" but it's all one big string</p>
]]></description><pubDate>Sun, 24 Aug 2025 19:43:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=45007096</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=45007096</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45007096</guid></item><item><title><![CDATA[New comment by chisleu in "AWS in 2025: Stuff you think you know that's now wrong"]]></title><description><![CDATA[
<p>> You don’t have to randomize the first part of your object keys to ensure they get spread around and avoid hotspots.<p>As of when? According to internal support, this is still required as of 1.5 years ago.</p>
]]></description><pubDate>Wed, 20 Aug 2025 17:03:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=44963713</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44963713</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44963713</guid></item><item><title><![CDATA[New comment by chisleu in "What could have been"]]></title><description><![CDATA[
<p>/agree<p>We are in the infancy of LLM technology.</p>
]]></description><pubDate>Mon, 18 Aug 2025 22:55:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=44946187</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44946187</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44946187</guid></item><item><title><![CDATA[New comment by chisleu in "GPT-5 vs. Sonnet: Complex Agentic Coding"]]></title><description><![CDATA[
<p>How was he doing "complex agentic coding" when the APIs have such extreme context and throughput limitations?</p>
]]></description><pubDate>Fri, 08 Aug 2025 16:41:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=44839029</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44839029</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44839029</guid></item><item><title><![CDATA[New comment by chisleu in "Leonardo Chiariglione – Co-founder of MPEG"]]></title><description><![CDATA[
<p>holy shit it does. The scene with him inventing the new compression algorithm basically foreshadowed the gooning to follow local LLM availability.</p>
]]></description><pubDate>Thu, 07 Aug 2025 12:32:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=44823633</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44823633</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44823633</guid></item><item><title><![CDATA[New comment by chisleu in "Claude Opus 4.1"]]></title><description><![CDATA[
<p>I use opus or gemini 2.5 pro for plan mode and sonnet for act mode in Cline. <a href="https://cline.bot" rel="nofollow">https://cline.bot</a><p>It's my experience that Opus is better at solving architectural challenges where sonnet struggles.</p>
]]></description><pubDate>Wed, 06 Aug 2025 06:27:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=44808410</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44808410</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44808410</guid></item><item><title><![CDATA[New comment by chisleu in "Kimi-K2 Tech Report [pdf]"]]></title><description><![CDATA[
<p>It looks like qwen3-coder is going to steal K2's thunder in terms of agentic coding use.</p>
]]></description><pubDate>Wed, 23 Jul 2025 23:24:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=44665119</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44665119</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44665119</guid></item><item><title><![CDATA[New comment by chisleu in "Qwen3-Coder: Agentic coding in the world"]]></title><description><![CDATA[
<p>It's 480B params, not 480GB. The 4 bit version of this is 270GB. I believe it's trained at bf16, so you need over a TB of memory to operate the model at bf16. No one should be trying to replace claude with a quantized 8 bit or 4 bit model. It's simply not possible. Also, this model isn't going to be as versed as Claude at certain libraries and languages. I have something written entirely my claude which uses the Fyne library extensively in golang for UI. Claude knows it inside and out as it's all vibe coded, but the 4 bit Qwen3 coder just hallucinated functions and parameters that don't exist because it wasn't willing to admit it didn't know what it was doing. Definitely don't judge a model by it's quant is all I'm saying.</p>
]]></description><pubDate>Wed, 23 Jul 2025 17:12:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=44661573</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44661573</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44661573</guid></item><item><title><![CDATA[New comment by chisleu in "Qwen3-Coder: Agentic coding in the world"]]></title><description><![CDATA[
<p>A Mac Studio 512GB can run it in 4bit quantization. I'm excited to see unsloth dynamic quants for this today.</p>
]]></description><pubDate>Wed, 23 Jul 2025 15:38:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=44660457</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44660457</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44660457</guid></item><item><title><![CDATA[New comment by chisleu in "Qwen3-Coder: Agentic coding in the world"]]></title><description><![CDATA[
<p>I tried using the "fp8" model through hyperbolic but I question if it was even that model. It was basically useless through hyperbolic.<p>I downloaded the 4bit quant to my mac studio 512GB. 7-8 minutes until first tokens with a big Cline prompt for it to chew on. Performance is exceptional. It nailed all the tool calls, loaded my memory bank, and reasoned about a golang code base well enough to write a blog post on the topic: <a href="https://convergence.ninja/post/blogs/000016-ForeverFantasyFreshFoundation.md" rel="nofollow">https://convergence.ninja/post/blogs/000016-ForeverFantasyFr...</a><p>Writing blog posts is one of the tests I use for these models. It is a very involved process including a Q&A phase, drafting phase, approval, and deployment. The filenames follow a certain pattern. The file has to be uploaded to s3 in a certain location to trigger the deployment. It's a complex custom task that I automated.<p>Even the 4bit model was capable of this, but was incapable of actually working on my code, prefering to halucinate methods that would be convenient rather than admitting it didn't know what it was doing. This is the 4 bit "lobotomized" model though. I'm excited to see how it performs at full power.</p>
]]></description><pubDate>Wed, 23 Jul 2025 13:15:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=44658899</link><dc:creator>chisleu</dc:creator><comments>https://news.ycombinator.com/item?id=44658899</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44658899</guid></item></channel></rss>