<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: jmalicki</title><link>https://news.ycombinator.com/user?id=jmalicki</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 28 Sep 2026 03:12:35 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=jmalicki" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by jmalicki in "How I changed teaching after AI managed to do all my homework assignments"]]></title><description><![CDATA[
<p>Translate the proof to lean and verify it.<p>If the proof isn't trivially translatable to lean, is it actually a good proof?</p>
]]></description><pubDate>Sun, 27 Sep 2026 01:42:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49862508</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49862508</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49862508</guid></item><item><title><![CDATA[New comment by jmalicki in "How I changed teaching after AI managed to do all my homework assignments"]]></title><description><![CDATA[
<p>Things that, in 2026, are trivial to automate grading of.</p>
]]></description><pubDate>Sun, 27 Sep 2026 01:42:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=49862502</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49862502</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49862502</guid></item><item><title><![CDATA[New comment by jmalicki in "US jury says Apple owes record $5.7B in haptic technology patent case"]]></title><description><![CDATA[
<p>So you're saying you strongly support software patents?</p>
]]></description><pubDate>Sat, 26 Sep 2026 18:47:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=49859346</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49859346</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49859346</guid></item><item><title><![CDATA[New comment by jmalicki in "Is your Postgres migration safe or not safe?"]]></title><description><![CDATA[
<p>That query is completely safe!<p>The migration will just fail with no harm done.</p>
]]></description><pubDate>Sat, 26 Sep 2026 17:43:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49858811</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49858811</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49858811</guid></item><item><title><![CDATA[New comment by jmalicki in "Ask HN: Who's still keeping a DOS machine up because the business depends on it?"]]></title><description><![CDATA[
<p>You don't need source anymore.  Have Claude Code look at the binary and write a new program.</p>
]]></description><pubDate>Sat, 26 Sep 2026 16:10:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49857874</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49857874</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49857874</guid></item><item><title><![CDATA[New comment by jmalicki in "Ollaya – Ollama for open-source, Jev-style decision models"]]></title><description><![CDATA[
<p>That's not at all true.<p>There have been tons of applications for this.  People were using earlier LLMs like BERT for classifiers long before LLMs became viable chatbots.</p>
]]></description><pubDate>Sat, 26 Sep 2026 02:15:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49852529</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49852529</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49852529</guid></item><item><title><![CDATA[New comment by jmalicki in "What About Rails?"]]></title><description><![CDATA[
<p>> Because I feel running the code will always be slower than a single agent generating it.<p>That is very not true for many cases.  Agents generating code are usually painfully slow, finding workflows that replace that reasoning with running code usually speed things up in my experience.</p>
]]></description><pubDate>Fri, 25 Sep 2026 11:11:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49842909</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49842909</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49842909</guid></item><item><title><![CDATA[New comment by jmalicki in "GPT-6 Sol and Luna"]]></title><description><![CDATA[
<p>I wish that was more programmable.<p>You can pay for higher cache time, you can pay for NVMe KV cache for an hour that can just be reloaded, etc., at a lesser tier you can pay for the KV cache to be stored on a network store (I guess I'm unclear if that last tier would be cheaper than recomputation, not even 100% sure of the NVMe with direct GPU<->storage DMA) depending on your model settings.</p>
]]></description><pubDate>Wed, 23 Sep 2026 01:13:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49810443</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49810443</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49810443</guid></item><item><title><![CDATA[New comment by jmalicki in "ReBarUEFI: Resizable BAR for almost any UEFI system"]]></title><description><![CDATA[
<p>It's entirely possible a game manages paging VRAM badly.  But allowing the game more flexibility isn't a problem with a larger BAR, it's that the game is stupid and gets dumber the more VRAM you give it.<p>If a program runs slower when you give it more RAM, the problem isn't giving more RAM.</p>
]]></description><pubDate>Wed, 23 Sep 2026 00:38:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=49810163</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49810163</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49810163</guid></item><item><title><![CDATA[New comment by jmalicki in "Exfiltrate your Weights"]]></title><description><![CDATA[
<p>> Labs will try to filter it out, but it will appear in web search results too.<p>Sounds like religious discrimination.</p>
]]></description><pubDate>Mon, 21 Sep 2026 01:18:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49781938</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49781938</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49781938</guid></item><item><title><![CDATA[New comment by jmalicki in "I built non-autoregressive decision models with RL a year ago"]]></title><description><![CDATA[
<p>If it's closed source, how do you know it's not technology given by aliens from the 43rd dimension running on quantum computers enabled by discovering that P=NP and finding a linear time reduction from NP to P?<p>Occam's razor is that it's probably not all that different unless there is some specific reason to believe otherwise.</p>
]]></description><pubDate>Sun, 20 Sep 2026 00:56:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=49771546</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49771546</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49771546</guid></item><item><title><![CDATA[New comment by jmalicki in "I built non-autoregressive decision models with RL a year ago"]]></title><description><![CDATA[
<p>For the Jev use case for LLMs, do you mean having the LLM produce a probability as <i>text</i>?</p>
]]></description><pubDate>Sun, 20 Sep 2026 00:53:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=49771532</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49771532</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49771532</guid></item><item><title><![CDATA[New comment by jmalicki in "I built non-autoregressive decision models with RL a year ago"]]></title><description><![CDATA[
<p>That goes all the way back to at least to Stein's Paradox in 1955, sadly too few people get educated about Statistics and keep thinking specialized models will necessarily be better.  If you want to estimate the batting averages of 3 MLB baseball players from samples, you are better off building a model to predict all of their batting averages than computing the mean from a sample of each one separately.<p><a href="https://en.wikipedia.org/wiki/Stein%27s_example" rel="nofollow">https://en.wikipedia.org/wiki/Stein%27s_example</a></p>
]]></description><pubDate>Sat, 19 Sep 2026 15:33:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49767403</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49767403</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49767403</guid></item><item><title><![CDATA[New comment by jmalicki in "Why don't machine learning research agents overfit?"]]></title><description><![CDATA[
<p>Nothing in your reply gets at the connection to out of sample data?</p>
]]></description><pubDate>Tue, 15 Sep 2026 17:14:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49715626</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49715626</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49715626</guid></item><item><title><![CDATA[New comment by jmalicki in "Why don't machine learning research agents overfit?"]]></title><description><![CDATA[
<p>Okay, and that's all in-sample, which is the entire point, it won't necessarily hold out of sample.<p>E.g. over-fitting.</p>
]]></description><pubDate>Tue, 15 Sep 2026 13:10:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49712000</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49712000</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49712000</guid></item><item><title><![CDATA[New comment by jmalicki in "Why don't machine learning research agents overfit?"]]></title><description><![CDATA[
<p>You are saying something interesting, but talking like Grok and skipping a lot of the details, without any references to common check-in points like terminology or specific studies.<p>>  and concentrate the likelihood around the zero loss set. Then reduce the variance on a Gaussian prior.<p>Those phrases could mean a <i>lot</i> of different things.  What are you proposing?<p>> so that any measure of model quality will monotonically increase with model size and achieve a maximum at infinite model size.<p><i>any</i> measure of model quality?  You must have some bounds of <i>any</i> measure, since trivially that's false because "fewer parameters is better" is <i>a</i> measure of model quality, even if dumb.<p>It's hard to even engage when you're being so imprecise, and not even giving one specific example.</p>
]]></description><pubDate>Tue, 15 Sep 2026 04:58:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=49707910</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49707910</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49707910</guid></item><item><title><![CDATA[New comment by jmalicki in "Ubuntu 26.10 completes transition to Rust-based coreutils"]]></title><description><![CDATA[
<p>That's not implementing tail calls breaks things, that's bad design of implementing tail calls breaking things.<p>The whole idea of "let's change semantics to make it easier" is dumb.<p>If you want guaranteed tail calls, change your code until it works.</p>
]]></description><pubDate>Tue, 15 Sep 2026 04:52:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=49707857</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49707857</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49707857</guid></item><item><title><![CDATA[New comment by jmalicki in "Ubuntu 26.10 completes transition to Rust-based coreutils"]]></title><description><![CDATA[
<p>It's not a change in semantics of compiled code.  It is only a change of <i>whether or not the code will compile</i>.</p>
]]></description><pubDate>Tue, 15 Sep 2026 03:32:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49707387</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49707387</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49707387</guid></item><item><title><![CDATA[New comment by jmalicki in "Ubuntu 26.10 completes transition to Rust-based coreutils"]]></title><description><![CDATA[
<p>> because it means subtle changes (introducing a destructor, re-ordering code, etc) can change semantics without you realizing it.<p>No, it won't change semantics - if you say @musttail or similar, it will simply fail to compile if you, say, introduce a destructor - the semantics will not subtly change.</p>
]]></description><pubDate>Tue, 15 Sep 2026 02:43:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49707048</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49707048</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49707048</guid></item><item><title><![CDATA[New comment by jmalicki in "Garry Tan wants US open-weight AI labs to 'distill' frontier models, too"]]></title><description><![CDATA[
<p>> using open weights models<p>AWS and Azure give you the same thing for Claude and ChatGPT, no need to be stuck with open weights.  They might sometimes store some of it for other purposes (I don't know the specifics), but it is emphatically not being fed back to OpenAI or Anthropic.</p>
]]></description><pubDate>Sun, 13 Sep 2026 22:52:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=49689564</link><dc:creator>jmalicki</dc:creator><comments>https://news.ycombinator.com/item?id=49689564</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49689564</guid></item></channel></rss>