<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: p1necone</title><link>https://news.ycombinator.com/user?id=p1necone</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 19 Aug 2026 13:03:08 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=p1necone" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by p1necone in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>This has not been my experience. Generally I do pin to 1 provider, or 1 provider with a couple fallbacks (especially with deepseek - most providers are 10x the cached token price compared to deepseek themselves), but even when I don't I still usually see 99%+ cache hit percentage. Specifically using pi with various ad-hoc customisations (that I was careful not to break prompt caching with).</p>
]]></description><pubDate>Thu, 13 Aug 2026 02:23:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49281132</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49281132</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49281132</guid></item><item><title><![CDATA[New comment by p1necone in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>It's interesting that all three of those used roughly the same amount of tokens, and almost entirely output. Feels like the thinking level lever didn't alter cost at all for this specific task, even though it did change the output.</p>
]]></description><pubDate>Thu, 13 Aug 2026 00:27:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280352</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49280352</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280352</guid></item><item><title><![CDATA[New comment by p1necone in "DeepSeek V4 Pro 0813"]]></title><description><![CDATA[
<p>50% cache hit is really low - in a standard agentic loop you should expect like 99%+ cache hit percentage (which should also lower that $12.50 to like a couple of $ for the same amount of tokens).<p>If you're using a customised harness you should make sure you don't have something that's e.g. changing your system prompt on some requests or rewriting history - it can be tempting to do stuff like strip old thinking tokens or compact tool call results to reduce context size but it's a trap - you want to <i>never</i> change history because of how cheap cache is, even more so with deepseek because their cache hit pricing is so low compared to most other models.</p>
]]></description><pubDate>Thu, 13 Aug 2026 00:21:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49280319</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49280319</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49280319</guid></item><item><title><![CDATA[New comment by p1necone in "Show HN: iPhone app takes simultaneous images from 2 lenses, fuses into 1 photo"]]></title><description><![CDATA[
<p>Apple customers are just built different.</p>
]]></description><pubDate>Tue, 11 Aug 2026 20:46:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=49264234</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49264234</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49264234</guid></item><item><title><![CDATA[New comment by p1necone in "Show HN: iPhone app takes simultaneous images from 2 lenses, fuses into 1 photo"]]></title><description><![CDATA[
<p>I just assumed this was what all phones with multiple rear cameras were doing, is it not? What are the multiple cameras for other than that?</p>
]]></description><pubDate>Tue, 11 Aug 2026 20:45:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49264228</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49264228</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49264228</guid></item><item><title><![CDATA[New comment by p1necone in "DeepSeek V4 Flash 0731"]]></title><description><![CDATA[
<p>I read that and thought "ah that's going to confuse people, but I can tell they mean 30k-250k loc", so thank you for the clarification.</p>
]]></description><pubDate>Sat, 08 Aug 2026 09:34:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49220190</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49220190</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49220190</guid></item><item><title><![CDATA[New comment by p1necone in "2027 memory capacity is reportedly sold out"]]></title><description><![CDATA[
<p>The capital investment needed to bring new chip fabs online and to staff them is likely orders of magnitude higher than that needed to buy land to grow corn on. And then the ratio of investment to sell price on that land + infrastructure is probably significantly worse for chip fabs that potentially aren't needed to satisfy demand anymore a few years from now.<p>"If the price is high enough" is of course technically true, but the <i>scale</i> of what high means in this context is important.</p>
]]></description><pubDate>Sat, 08 Aug 2026 00:59:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49217913</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49217913</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49217913</guid></item><item><title><![CDATA[New comment by p1necone in "Pi's Minimalism Is Its Advantage"]]></title><description><![CDATA[
<p>I love hacking away at pi extensions.<p>I realised how use case dependent harness behaviour is when I tried to use my customised-for-a-side-project pi config at work and realised I needed to tweak it significantly to be useful - I would not be surprised if tools like Claude Code needing to be all things for all people is hurting their peak usefulness.</p>
]]></description><pubDate>Tue, 04 Aug 2026 23:35:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49176734</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49176734</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49176734</guid></item><item><title><![CDATA[New comment by p1necone in "Developers are attached to tools because tools encode trust"]]></title><description><![CDATA[
<p>I'm speaking in the context of building scaffolding for e2e testing, not the tests themselves.<p>When I say "everything" I mean the frontend, backend, database, seeded data, optimizing runtime etc, I'm not talking about loc coverage. I would not recommend using llms to generate the tests themselves without heavy guidance and review, and you're right that 100% coverage is often a counterproductive goal if the code under test is anything less than pure business logic.</p>
]]></description><pubDate>Mon, 03 Aug 2026 19:31:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49160349</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49160349</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49160349</guid></item><item><title><![CDATA[New comment by p1necone in "Show HN: Shitty – fast terminal. Memory-unsafe and faster than yours"]]></title><description><![CDATA[
<p>This is cool, but I gotta say - I care <i>much</i> more about keypress-to-screen <i>latency</i> on my terminals than throughput - would love to see some numbers on that.</p>
]]></description><pubDate>Mon, 03 Aug 2026 01:03:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49150057</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49150057</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49150057</guid></item><item><title><![CDATA[New comment by p1necone in "Developers are attached to tools because tools encode trust"]]></title><description><![CDATA[
<p>I've found the process of building good reliable CI that thoroughly covers <i>everything</i> has been greatly improved by LLMs. There's <i>so much</i> tedious plumbing and grunt work involved in building CI and automated testing infrastructure for bespoke products that they can handle just fine while you concentrate on the important bits - I would say it's an area where agentic workflows are even more suited than regular product code.</p>
]]></description><pubDate>Sun, 02 Aug 2026 21:03:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=49148278</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49148278</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49148278</guid></item><item><title><![CDATA[New comment by p1necone in "The iPhone Upgrade Program is being replaced by Apple Upgrade"]]></title><description><![CDATA[
<p>Yeah I'm on the same page - I spend a couple hundred bucks on a low-end-of-midrange android phone every few years (usually because I broke the old one by doing something stupid like walking into the ocean with it in my pocket). I can't fathom why I would want to spend 5-10x on a flagship to do... what exactly?<p>My current phone (Motorola g34) runs all the games fine, and that's probably the heaviest thing people use phones for. I do take photos, but like I'm not running a photography business, and social media compresses them to shit anyway so the quality is perfectly fine.</p>
]]></description><pubDate>Wed, 29 Jul 2026 00:53:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49092061</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49092061</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49092061</guid></item><item><title><![CDATA[New comment by p1necone in "Our position on open-weights models"]]></title><description><![CDATA[
<p>Imo this amounts to caring about the wrong thing. The <i>only</i> thing an AI model can do is take in text/images/audio as input and spit out text/images/audio as output.<p>If you're going to analyse the safety of anything it should be the security controls in the harnesses we wrap around the models that take that output and treat it as instructions to actually do things.</p>
]]></description><pubDate>Tue, 28 Jul 2026 01:39:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49078193</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49078193</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49078193</guid></item><item><title><![CDATA[New comment by p1necone in "DeepSeek pause fundraise after comments on compute gap to US leaked (transcript) [pdf]"]]></title><description><![CDATA[
<p>I feel like if you showed current frontier models to someone 10 years ago, they'd probably call it AGI. Does AGI have a clear definition or is it just a pair of goalposts on wheels?</p>
]]></description><pubDate>Sun, 26 Jul 2026 10:58:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49056788</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49056788</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49056788</guid></item><item><title><![CDATA[New comment by p1necone in "So Reddit has decided that plain HTML is unsafe"]]></title><description><![CDATA[
<p>old.reddit.com works to bypass that on mobile.</p>
]]></description><pubDate>Thu, 23 Jul 2026 04:21:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49016896</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49016896</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49016896</guid></item><item><title><![CDATA[New comment by p1necone in "John C. Dvorak has died"]]></title><description><![CDATA[
<p>TIL the Dvorak layout does not derive its name the same way as the qwerty layout - I guess I never looked at a Dvorak keyboard before.</p>
]]></description><pubDate>Thu, 23 Jul 2026 03:14:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49016443</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49016443</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49016443</guid></item><item><title><![CDATA[New comment by p1necone in "Terence Tao's ChatGPT conversation about the Jacobian Conjecture counterexample"]]></title><description><![CDATA[
<p>I've noticed GPT specifically has more of a tendency to stop partway through things than many other models do. Although my most recent experience with it was 4.X I believe.<p>Hearing "here's what I've done, here's the completely unambiguous next steps, I'll wait for you to send a pointless message before I continue" over and over again is a real pain.</p>
]]></description><pubDate>Thu, 23 Jul 2026 00:21:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49015294</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49015294</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49015294</guid></item><item><title><![CDATA[New comment by p1necone in "Show HN: Read the Tape – Wordle for daytrading, five blind S&P 500 charts a day"]]></title><description><![CDATA[
<p>If a big enough proportion of the market is <i>also</i> following the same technical analysis strategies as you, it will predict the market. There doesn't have to be anything <i>actually</i> correct about the analysis.<p>In reality I don't think a big enough proportion of investors are reading tea leaves for this to be true.</p>
]]></description><pubDate>Wed, 22 Jul 2026 01:13:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49000578</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=49000578</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49000578</guid></item><item><title><![CDATA[New comment by p1necone in "Qwen 3.8"]]></title><description><![CDATA[
<p>I use GLM-5.2 as an orchestrator model which delegates to Deepseek subagents (v4 flash or pro depending on complexity) and it works pretty well for a quite complex compiler codebase when I do deep enough up front planning for features/fixes.<p>If you believe the benchmarks Deepseek v4 is pretty shitty at long running work in large codebases, but <i>really really good</i> at self contained algorithmic/math reasoning - which is basically ideal for a compiler for a language with a relatively complex type system. And with its cache pricing it's very cheap.<p>GLM-5.2 is too expensive at api pricing for my taste though even when it's not producing the bulk of the output tokens - I pay for the mid-tier subscription and switch the orchestrator over to other models via open router when I run out - Minimax M3 feels okay but definitely a step down from GLM-5.2.</p>
]]></description><pubDate>Sun, 19 Jul 2026 23:12:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48972560</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=48972560</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48972560</guid></item><item><title><![CDATA[New comment by p1necone in "More Bounce to the Ounce"]]></title><description><![CDATA[
<p>This worked for Snow Crash - felt like I went on a theme park ride, got to the end of it, and hopped off still buzzing. Doesn't work so well for something less tongue in cheek.</p>
]]></description><pubDate>Sat, 18 Jul 2026 01:19:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48954182</link><dc:creator>p1necone</dc:creator><comments>https://news.ycombinator.com/item?id=48954182</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48954182</guid></item></channel></rss>