<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: trouve_search</title><link>https://news.ycombinator.com/user?id=trouve_search</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sun, 04 Oct 2026 00:04:17 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=trouve_search" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by trouve_search in "ADHD, autism or complex trauma? [pdf]"]]></title><description><![CDATA[
<p>It's an article with no clear conclusion it's normal to feel confused.<p>My read of the article is balancing the fact that there's a lot of overlap between CPTSD/ADHD/ASD in the symptoms (emotional dysregulation, hyperarousal, etc) and in hereditary factors (undiagnosed parents causing trauma more often on average). Also that traumatic  childhood experiences are more likely to stick around as CPTSD in adulthood if there's also neurodivergence.<p>The author says there can be incredible relief to be correctly diagnosed with ADHD/ASD, so obviously she says it's helpful.<p>But she also warns that wrong treatment can often happen (eg. Giving stimulants to perpetual fight or flight PTSD brains), or inneffective therapy.</p>
]]></description><pubDate>Sat, 03 Oct 2026 20:09:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=49947327</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49947327</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49947327</guid></item><item><title><![CDATA[New comment by trouve_search in "I can't stop thinking about Papua New Guinea"]]></title><description><![CDATA[
<p>It's important to note that languages evolve much faster without a writing tradition to anchor the language down over time.</p>
]]></description><pubDate>Wed, 16 Sep 2026 00:55:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49720883</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49720883</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49720883</guid></item><item><title><![CDATA[New comment by trouve_search in "Your car is selling your data"]]></title><description><![CDATA[
<p>Couldn't you put some sort of faraday cage around the antenna?</p>
]]></description><pubDate>Sun, 13 Sep 2026 15:27:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49685078</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49685078</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49685078</guid></item><item><title><![CDATA[New comment by trouve_search in "ChatGPT Is Throwing 404"]]></title><description><![CDATA[
<p>pi.dev + anything else</p>
]]></description><pubDate>Thu, 03 Sep 2026 15:06:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49551076</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49551076</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49551076</guid></item><item><title><![CDATA[New comment by trouve_search in "DiffusionGemma Technical Report"]]></title><description><![CDATA[
<p>That's AMD's fault.<p>RDNA4 is pretty similar to CDNA4, yet over a year after the release of "pro AI" cards like the r9700, they had basic kernels lacking in vllm (like w4a16 int4 kernels) while they were implemented in the datacenter CDNA4 cards.<p>AMD hardware runs well on llama.cpp because basically anything runs on llama.cpp, especially with vulkan. It's not high praise of AMD's software team to say llama.cpp runs well on their hardware</p>
]]></description><pubDate>Thu, 20 Aug 2026 22:08:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49380921</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49380921</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49380921</guid></item><item><title><![CDATA[New comment by trouve_search in "DiffusionGemma Technical Report"]]></title><description><![CDATA[
<p>Yes, it runs in vllm happily. It gets >900TPS output reliably on a single 5090 with the nvfp4 model.<p>It's clearly worse than vanilla 26B-A4B, and lacks some things like structured outputs, and gets some tool calls wrong.<p>So you have to find a usecase or a hand rolled harness that leverages the cerebras-level TPS while not going off track during (even short) tasks.</p>
]]></description><pubDate>Thu, 20 Aug 2026 19:18:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49378908</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49378908</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49378908</guid></item><item><title><![CDATA[New comment by trouve_search in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>thanks for posting your setup! I think it's smart to set the reasoning effort default to something saner in the base config.<p>Here's a VLLM command for 3.6 (I'll update to 3.8 today) to test out:<p>```<p>PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True  \<p>vllm serve Qwen/Qwen3.6-27B-FP8 \<p>--dtype auto \<p>--kv-cache-dtype fp8 \<p>--enable-chunked-prefill \<p>--enable-prefix-caching \<p>--trust-remote-code \<p>--enable-auto-tool-choice \<p>--reasoning-parser qwen3 \<p>--tool-call-parser qwen3_coder \<p>--speculative-config '{"method":"qwen3_next_mtp","num_speculative_tokens":3}' \<p>--default-chat-template-kwargs '{<p><pre><code>    "enable_thinking": true, 

    "reasoning_effort":"medium"

 }' \

 --tensor-parallel-size 2 \

 --max-model-len 250000 \

 --gpu-memory-utilization 0.9 \

 --max-num-batched 12000 \

 --max-num-seqs 24
</code></pre>
```<p>I took the liberty of adding your reasoning effort chat template to my setup. You can play around with the last few parameters. In generall VLLM will be better in higher concurrency scenarios, so if you only use it for a personal vibe coding assistant and less as a general home model for task execution llama.cpp may be better.</p>
]]></description><pubDate>Tue, 18 Aug 2026 15:33:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49347233</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49347233</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49347233</guid></item><item><title><![CDATA[New comment by trouve_search in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>Yes, I mentioned the setup, but on vllm you can only use TP with speculative decoding or pipeline parallelism without, so there's tradeoff to both.<p>I gave general numbers of what I'm getting above, the performance ratios seemed similar regardless of setup (eg. getting a AWQ-in4 quant on a single GPU vs PP without speculative decoding vs TP with speculative decoding).<p>Overall single GPU is fastest, and TP+speculative decoding is still faster than PP, but for fp8 models you need dual GPUs whether you want it or not.</p>
]]></description><pubDate>Tue, 18 Aug 2026 15:26:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49347094</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49347094</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49347094</guid></item><item><title><![CDATA[New comment by trouve_search in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>I think it's a vllm vs llama_cpp performance thing, will pay more into it.<p>One note I had between the two is that gemma has a much higher prefix cache hit rate in general.</p>
]]></description><pubDate>Tue, 18 Aug 2026 15:24:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49347057</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49347057</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49347057</guid></item><item><title><![CDATA[New comment by trouve_search in "Qwen3.8 27B scores 52 on Artificial Analysis"]]></title><description><![CDATA[
<p>What configuration are you using? On both vllm and llama-cpp, I get significantly higher speeds from gemma4 than qwen3.6 (with their respective speculative decoding methods).<p>Output TPS in vllm for instance:<p>- Gemma4 26B-A4B: 200-300TPS<p>- Qwen3.6 35B-A3B: 120-180TPS<p>- Gemma4 31B: 80-120TPS<p>- Qwen3.6 27B: 60-80TPS<p>This is for a first request on a dual 5090 setup, with their respective speculative decoding methods.</p>
]]></description><pubDate>Mon, 17 Aug 2026 19:18:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49336209</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49336209</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49336209</guid></item><item><title><![CDATA[New comment by trouve_search in "llama.cpp"]]></title><description><![CDATA[
<p>Using 98.css would still leave you with the AI slop text wording.<p>The core problem is that some people don't even seem to notice / care.</p>
]]></description><pubDate>Wed, 12 Aug 2026 14:15:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49272739</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49272739</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49272739</guid></item><item><title><![CDATA[New comment by trouve_search in "Nvidia Nemotron 3.5 Lightning and NeMo Switchyard"]]></title><description><![CDATA[
<p>Laguna XS is MoE, however.</p>
]]></description><pubDate>Wed, 12 Aug 2026 01:04:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49266612</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49266612</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49266612</guid></item><item><title><![CDATA[New comment by trouve_search in "Manus will return to operating as an independent company"]]></title><description><![CDATA[
<p>Interesting, what does your tasks & workflow look like with them?<p>I generally found the quality decent (say, similar to other competitors), but the speed of task completion was very slow. I think because they would depend too much on Sonnet as a core backend, and relied on big/expensive models more than other harnesses.</p>
]]></description><pubDate>Tue, 11 Aug 2026 16:38:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49260854</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49260854</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49260854</guid></item><item><title><![CDATA[New comment by trouve_search in "Manus will return to operating as an independent company"]]></title><description><![CDATA[
<p>Does anyone here use Manus actively?<p>I found it worse than alternatives in the similar space (Claude, genspark, kagi research, etc.) and slightly baffling they were being acquired at a $2B valuation in the first place.</p>
]]></description><pubDate>Tue, 11 Aug 2026 15:59:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49260310</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49260310</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49260310</guid></item><item><title><![CDATA[New comment by trouve_search in "DeepSeek costs OpenCode Go user $1.14/day; dual DGX breaks even in 24 years"]]></title><description><![CDATA[
<p>The value prop really depends on what you're doing.<p>If you're just vibe coding with giant frontier models, yes, the value will be worse. Especially now, where GPU prices have spiked another 20% last month.<p>For some tasks where owning the setup and full kv cache matters, the payoff calculation is ridiculously in favor of running your own deployment.<p>For instance for some batch classifications jobs where the prefix cache hit rate will be >95%.<p>The calculus also changes if you just use AI as a light tool while coding and don't need the giant models; qwen3 27B runs at 80TPS on a 5090 properly deployed.</p>
]]></description><pubDate>Mon, 10 Aug 2026 13:13:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49243189</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49243189</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49243189</guid></item><item><title><![CDATA[New comment by trouve_search in "The Absurdity of Albert Camus"]]></title><description><![CDATA[
<p>I love these, thanks!</p>
]]></description><pubDate>Sat, 01 Aug 2026 01:48:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49130357</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49130357</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49130357</guid></item><item><title><![CDATA[New comment by trouve_search in "The Absurdity of Albert Camus"]]></title><description><![CDATA[
<p>Everyone has to find an answer to the meaning if life. For the non religious/escapist, you have to stare into the void and find an answer at some point.<p>Existentialism (Sartre) says you have to find your own meaning. Camus says there cannot be meaning; life is inherently "absurd".<p>Any of the existentialist philosophies will be adjacent to angsty teenager stuff; they dance closely with nihilism and cynicism.</p>
]]></description><pubDate>Sat, 01 Aug 2026 01:43:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49130334</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49130334</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49130334</guid></item><item><title><![CDATA[New comment by trouve_search in "Read this before you buy that TV streaming stick"]]></title><description><![CDATA[
<p>The nvidia shield is pretty damn good as well, even if old at this point.</p>
]]></description><pubDate>Thu, 30 Jul 2026 20:50:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49115589</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49115589</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49115589</guid></item><item><title><![CDATA[New comment by trouve_search in "Be skeptical of OpenAI's rogue hacker agent story"]]></title><description><![CDATA[
<p>From my reading, the sandbox escape came from the JS packages in the harness still having an internet connection (somehow!), the agent having access to the source of those packages, reading it and executing code from them to access the internet.</p>
]]></description><pubDate>Fri, 24 Jul 2026 20:25:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=49041170</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=49041170</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49041170</guid></item><item><title><![CDATA[New comment by trouve_search in "I co-founded Wikipedia, but an anonymous mob runs the show – and now I'm banned"]]></title><description><![CDATA[
<p>The fact that he could only get this story published in the Washington examiner of all places should be a signal that more reputable places don't want to attach their name to this</p>
]]></description><pubDate>Thu, 09 Jul 2026 08:41:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=48842789</link><dc:creator>trouve_search</dc:creator><comments>https://news.ycombinator.com/item?id=48842789</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48842789</guid></item></channel></rss>