<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: vlmutolo</title><link>https://news.ycombinator.com/user?id=vlmutolo</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 29 Sep 2026 10:09:23 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=vlmutolo" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by vlmutolo in "The case against JPEG XL"]]></title><description><![CDATA[
<p>What do you think the best use cases for jxl are? Where does it still have an advantage over other formats?</p>
]]></description><pubDate>Mon, 14 Sep 2026 01:39:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49690782</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=49690782</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49690782</guid></item><item><title><![CDATA[New comment by vlmutolo in "GPT-6 Astra"]]></title><description><![CDATA[
<p>They said 5.6 Sol got something around 40% with the corrected harness.</p>
]]></description><pubDate>Fri, 04 Sep 2026 14:52:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49565541</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=49565541</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49565541</guid></item><item><title><![CDATA[New comment by vlmutolo in "GPT-6 Astra"]]></title><description><![CDATA[
<p>The ARC-AGI-3 harness was throwing away reasoning tokens between turns. This is very bad harness design.<p>The models are designed to keep the reasoning tokens separate from the output and only publicly emit tool calls and the sometimes a summary of the reasoning tokens. The models are trained to depend on those private reasoning tokens. You can’t just delete them.<p><a href="https://openai.com/index/how-two-settings-tripled-our-arc-agi-3-scores/" rel="nofollow">https://openai.com/index/how-two-settings-tripled-our-arc-ag...</a></p>
]]></description><pubDate>Thu, 03 Sep 2026 21:40:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49557473</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=49557473</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49557473</guid></item><item><title><![CDATA[New comment by vlmutolo in "SQRL wan't wrong, it was early"]]></title><description><![CDATA[
<p>Another perspective on this is that coordination is the hardest problem to solve in deploying technology that depends on cooperation between different parties. It takes a few of the big centralized players to agree to put their weight behind it.<p>The tech is actually easier.</p>
]]></description><pubDate>Tue, 11 Aug 2026 15:11:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49259644</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=49259644</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49259644</guid></item><item><title><![CDATA[New comment by vlmutolo in "Google fixed more Chrome bugs in June than over the past two years, thanks to AI"]]></title><description><![CDATA[
<p>It’s a very soft skill that changes as the capabilities of AI change.<p>But one thing that doesn’t change is the need to specify an end goal correctly and precisely. I think the emphasis of knowledge/information-processing work is going to increasingly be placed on verification mechanisms. This is practically equivalent to precisely defining an end goal.<p>Spend time deeply thinking about what it means for a solution to be correct. What properties will a correct solution have? What of those properties are testable? Write those things down and tell the agent.<p>As models get better, agents will be able to target more and more difficult end goals. The strategy just becomes more useful. (It’s useful for people as well.)<p>For example, if you want an agent to write a photo editor, think about what end properties the editor should have. There are reference images for color space and rendering transformations. That’s a good start.<p>Sometimes the goal will be fuzzier. “I want a feature like CaptureOne where I provide a reference image and it makes my image look like that.” Well, time to think really hard about what that means. Iterate with AI on how to test for that precisely. Come up with some good metrics/heuristics. Maybe it means local contrast should match. Maybe the overall distribution of colors. Maybe something more complicated.<p>Then you have a target and you can let an implementation agent work against that. If it fails, it’s either because the agent is bad or your target was incorrect or incomplete. As models get better, the limiting factor becomes your ability to correctly define a problem.</p>
]]></description><pubDate>Fri, 31 Jul 2026 14:33:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49123673</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=49123673</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49123673</guid></item><item><title><![CDATA[New comment by vlmutolo in "Silurus/ooxml: Pixel-faithful Office documents, rendered in the browser"]]></title><description><![CDATA[
<p>Pretty cool, rendering PowerPoint files to an image is probably the only way for LLMs to make sense of them.<p>Does this work in Cloudflare’s workerd environment? Would be nice to have a cheap serverless render -> LLM (GLM-OCR / PaddleOCR) -> Markdown pipeline for the various MS Office formats.</p>
]]></description><pubDate>Sun, 07 Jun 2026 19:03:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=48437627</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=48437627</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48437627</guid></item><item><title><![CDATA[New comment by vlmutolo in "Gemini 3.5 Flash"]]></title><description><![CDATA[
<p>> While OpenAI originally pioneered Codex (which went on to power GitHub Copilot), Google’s direct answer for dedicated, native code completion and natural-language-to-code generation is CodeGemma.<p><a href="https://g.co/gemini/share/33e7a589a161" rel="nofollow">https://g.co/gemini/share/33e7a589a161</a></p>
]]></description><pubDate>Wed, 20 May 2026 03:17:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=48202703</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=48202703</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48202703</guid></item><item><title><![CDATA[New comment by vlmutolo in "We replaced RAG with a virtual filesystem for our AI documentation assistant"]]></title><description><![CDATA[
<p>If you give every agent an isolated container to use, you’re going to be paying for the reserved memory while the container is active, even if the agent isn’t doing anything.</p>
]]></description><pubDate>Sat, 04 Apr 2026 15:48:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=47640083</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47640083</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47640083</guid></item><item><title><![CDATA[New comment by vlmutolo in "Gemini 3.1 Flash-Lite: Built for intelligence at scale"]]></title><description><![CDATA[
<p>Wow, that’s very interesting. I wish more benchmarks were reported along with the total cost of running that benchmark. Dollars per token is kind of useless for the reasons you mentioned.</p>
]]></description><pubDate>Wed, 04 Mar 2026 01:02:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=47241588</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47241588</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47241588</guid></item><item><title><![CDATA[New comment by vlmutolo in "Gemini 3.1 Flash-Lite: Built for intelligence at scale"]]></title><description><![CDATA[
<p>Lots of comments about the price change, but Artifical Analysis reports that 3.1 Flash-Lite (reasoning) used fewer than half of the tokens of 2.5 Flash-Lite (reasoning).<p>This will likely bring the cost below 2.5 flash-lite for many tasks (depends on the ratio of input to output tokens).<p>That said, AA also reports that 3.1 FL was 20% more expensive to run for their complete Intelligence index benchmark.<p>The overall point is that cost is extremely task-dependent, and it doesn’t work to just measure token cost because reasoning can burn so many tokens, reasoning token usage varies by both task and model, and similarly the input/output ratios vary by task.</p>
]]></description><pubDate>Tue, 03 Mar 2026 18:03:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=47236248</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47236248</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47236248</guid></item><item><title><![CDATA[New comment by vlmutolo in "SynthID: A tool to watermark and identify content generated through AI"]]></title><description><![CDATA[
<p>C2PA has lots of problems.<p><a href="https://www.hackerfactor.com/blog/index.php?%2Farchives%2F1073-What-C2PA-Provides.html" rel="nofollow">https://www.hackerfactor.com/blog/index.php?%2Farchives%2F10...</a></p>
]]></description><pubDate>Thu, 26 Feb 2026 21:25:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=47172195</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47172195</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47172195</guid></item><item><title><![CDATA[New comment by vlmutolo in "Two Bits Are Better Than One: making bloom filters 2x more accurate"]]></title><description><![CDATA[
<p>Very interesting blog post. I’d never seen that method for quickly computing the patterns. I thought I had done a lot of research on bloom filters, too!</p>
]]></description><pubDate>Sun, 22 Feb 2026 17:56:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=47113077</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47113077</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47113077</guid></item><item><title><![CDATA[New comment by vlmutolo in "Two Bits Are Better Than One: making bloom filters 2x more accurate"]]></title><description><![CDATA[
<p>Yeah, I agree with this. I think there are open addressing hash tables like Swiss Table that do something similar. IIRC, they have buckets with a portion at the beginning with lossy “fingerprints” of items, which kind of serve a similar purpose as a bloom filter.</p>
]]></description><pubDate>Sun, 22 Feb 2026 17:45:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=47112975</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47112975</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47112975</guid></item><item><title><![CDATA[New comment by vlmutolo in "Two Bits Are Better Than One: making bloom filters 2x more accurate"]]></title><description><![CDATA[
<p>This article is a little confusing. I think this is a roundabout way to invent the blocked bloom filter with k=2 bits inserted per element.<p>It seems like the authors wanted to use a single hash for performance (?). Maybe they correctly determined that naive Bloom filters have poor cache locality and reinvented block bloom filters from there.<p>Overall, I think block bloom filters should be the default most people reach for. They completely solve the cache locality issues (single cache miss per element lookup), and they sacrifice only like 10–15% space increase to do it. I had a simple implementation running at something like 20ns per query with maybe k=9. It would be about 9x that for native Bloom filters.<p>There’s some discussion in the article about using a single hash to come up with various indexing locations, but it’s simpler to just think of block bloom filters as:<p>1. Hash-0 gets you the block index<p>2. Hash-1 through hash-k get you the bits inside the block<p>If your implementation slices up a single hash to divide it into multiple smaller hashes, that’s fine.</p>
]]></description><pubDate>Sun, 22 Feb 2026 04:40:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=47108228</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47108228</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47108228</guid></item><item><title><![CDATA[New comment by vlmutolo in "Cord: Coordinating Trees of AI Agents"]]></title><description><![CDATA[
<p>I wonder if the “spawn” API is ever preferable over “fork”. Do we really want to remove context if we can help it? There will certainly be situations where we have to, but then what you want is good compaction for the subagent. “Clean-slate” compaction seems like it would always be suboptimal.</p>
]]></description><pubDate>Sat, 21 Feb 2026 03:01:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=47097056</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=47097056</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47097056</guid></item><item><title><![CDATA[New comment by vlmutolo in "Show HN: DevicePrint – device fingerprinting without cookies"]]></title><description><![CDATA[
<p>AmIUnique.org has a good collection of non-cookie tracking techniques.<p><a href="https://amiunique.org/fingerprint" rel="nofollow">https://amiunique.org/fingerprint</a></p>
]]></description><pubDate>Mon, 12 Jan 2026 12:59:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=46587833</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=46587833</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46587833</guid></item><item><title><![CDATA[New comment by vlmutolo in "Roc Camera"]]></title><description><![CDATA[
<p>I wonder how this compares to similar initiatives by e.g. Sony [0] and Leica [1].<p>[0]: <a href="https://authenticity.sony.net/camera/en-us/" rel="nofollow">https://authenticity.sony.net/camera/en-us/</a><p>[1]: <a href="https://petapixel.com/2023/10/26/leica-m11-p-review-as-authenticated-as-they-come/" rel="nofollow">https://petapixel.com/2023/10/26/leica-m11-p-review-as-authe...</a></p>
]]></description><pubDate>Fri, 24 Oct 2025 04:07:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=45690658</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=45690658</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45690658</guid></item><item><title><![CDATA[New comment by vlmutolo in "Mathematicians have found a hidden 'reset button' for undoing rotation"]]></title><description><![CDATA[
<p>"Scientists unscramble egg proteins"<p><a href="https://www.science.org/content/article/scientists-unscramble-egg-proteins" rel="nofollow">https://www.science.org/content/article/scientists-unscrambl...</a></p>
]]></description><pubDate>Tue, 21 Oct 2025 19:43:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=45660698</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=45660698</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45660698</guid></item><item><title><![CDATA[New comment by vlmutolo in "Ultrasonic Chef's Knife"]]></title><description><![CDATA[
<p>It’s vibrating at a microscopic level. You won’t be able to feel it at all.</p>
]]></description><pubDate>Mon, 22 Sep 2025 21:19:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=45339571</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=45339571</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45339571</guid></item><item><title><![CDATA[New comment by vlmutolo in "Bevy's Fifth Birthday"]]></title><description><![CDATA[
<p>What are your thoughts on how Bevy's developing UI toolkit compares (in terms of goals and use cases) to some of the other Rust efforts in the space (egui, xilem, iced, etc.)? Do you expect it will be specialized/limited to scene development for games?</p>
]]></description><pubDate>Mon, 11 Aug 2025 15:00:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=44864909</link><dc:creator>vlmutolo</dc:creator><comments>https://news.ycombinator.com/item?id=44864909</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44864909</guid></item></channel></rss>