<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: markasoftware</title><link>https://news.ycombinator.com/user?id=markasoftware</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 02 Sep 2026 08:28:58 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=markasoftware" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by markasoftware in "Claude Fable 5.1 and Claude Mythos 5.1"]]></title><description><![CDATA[
<p>You do not get it.<p>The llm never deterministically picks a shade of red. It's a probability distribution over shades of colors, with certain shades of red being more likely than others. Without fingerprinting, it randomly samples from the distribution using a certain pseudorandom RNG. With fingerprinting, it also selects from the distribution using a pseudorandom RNG. My understanding is that the fingerprinted prng is still a strong RNG. Neither output is more correct than the other.<p>If a certain token is far more likely than any other, it's usually chosen even in the fingerprinted output.</p>
]]></description><pubDate>Wed, 02 Sep 2026 07:24:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49532906</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49532906</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49532906</guid></item><item><title><![CDATA[New comment by markasoftware in "Claude Fable 5.1 and Claude Mythos 5.1"]]></title><description><![CDATA[
<p>Its essentially swapping out the psuedo random number generated with a differently seeded one iirc.<p>It has an effect on the output, but not the <i>output quality</i></p>
]]></description><pubDate>Tue, 01 Sep 2026 22:31:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49529135</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49529135</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49529135</guid></item><item><title><![CDATA[New comment by markasoftware in "One Nix flake to rule them all"]]></title><description><![CDATA[
<p>Please be satire</p>
]]></description><pubDate>Sun, 30 Aug 2026 17:15:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=49500621</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49500621</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49500621</guid></item><item><title><![CDATA[New comment by markasoftware in "Ox Alpha"]]></title><description><![CDATA[
<p>Anonymous unreleased models are made available on arena.ai all the time, it's not really news that one is on openrouter...</p>
]]></description><pubDate>Fri, 21 Aug 2026 04:30:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49383801</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49383801</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49383801</guid></item><item><title><![CDATA[New comment by markasoftware in "GLM-5.3 Artificial Analysis Benchmarks"]]></title><description><![CDATA[
<p>Aa also benchmarked k3 at max</p>
]]></description><pubDate>Tue, 18 Aug 2026 23:24:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49354328</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49354328</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49354328</guid></item><item><title><![CDATA[New comment by markasoftware in "GLM-5.3 Artificial Analysis Benchmarks"]]></title><description><![CDATA[
<p>Very impressive score for the size, though token use is higher than k3 and far higher than proprietary models, and its price to performance isn't all that far ahead of k3 as a result</p>
]]></description><pubDate>Tue, 18 Aug 2026 23:09:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49354143</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49354143</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49354143</guid></item><item><title><![CDATA[New comment by markasoftware in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>Openai has been focusing a lot on cutting down overthinking is the feel I get. If you look at the artificial analysis tokens per task benchmark Sol especially at lower effort uses far less tokens than the competition.</p>
]]></description><pubDate>Mon, 17 Aug 2026 04:32:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=49326576</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49326576</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49326576</guid></item><item><title><![CDATA[New comment by markasoftware in "Models Are Getting Dumber on Purpose"]]></title><description><![CDATA[
<p>It could be that the rotation in the helix manifold whatever is a low level representation of the logical steps (carry the 2, add the next column,...)  it's describing. The point stands that the explanation it generates doesn't necessarily in all cases reflect what it "actually did" but your counterexample doesn't hold.</p>
]]></description><pubDate>Sun, 16 Aug 2026 23:58:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49325056</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49325056</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49325056</guid></item><item><title><![CDATA[New comment by markasoftware in "Gemini 3.7 Flash"]]></title><description><![CDATA[
<p>Sol high is almost the same speed if you take into account drastically lower token use. Look at the artificial analysis speed vs token use. Gemini is 7x faster but 5x more tokens. And that's with Sol high being a substantially better model.<p>Edit: and Sol medium actually has the same AA intelligence score as Gemini 3.7, and has >7x fewer tokens, actually making it faster</p>
]]></description><pubDate>Thu, 13 Aug 2026 17:57:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49289650</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49289650</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49289650</guid></item><item><title><![CDATA[New comment by markasoftware in "Qwen3.8-2.4T"]]></title><description><![CDATA[
<p>There's no rhyme or reason to it. Quants aren't benchmarked much. Generally 4bit better than smaller model 8bit</p>
]]></description><pubDate>Wed, 12 Aug 2026 16:35:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49275089</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49275089</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49275089</guid></item><item><title><![CDATA[New comment by markasoftware in "llama.cpp"]]></title><description><![CDATA[
<p>possibly hitting front page because this website is fairly new? For me, it's certainly the first time I've seen a one-liner curl|bash installer for llama.cpp, which was basically the only reason to use ollama.</p>
]]></description><pubDate>Wed, 12 Aug 2026 07:11:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49268836</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49268836</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49268836</guid></item><item><title><![CDATA[New comment by markasoftware in "Nvidia Nemotron 3.5 Lightning and NeMo Switchyard"]]></title><description><![CDATA[
<p>yep, the person you're responding to created the benchmark and is using HN comments as advertisement.</p>
]]></description><pubDate>Tue, 11 Aug 2026 21:20:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49264670</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49264670</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49264670</guid></item><item><title><![CDATA[New comment by markasoftware in "Qwen3.8 Max now ranked as the best overall model by agentic index"]]></title><description><![CDATA[
<p>Learn about MoE models and also just look at the benchmarks of the two models. For example <a href="https://artificialanalysis.ai/models/comparisons/qwen3-6-27b-vs-qwen3-6-35b-a3b" rel="nofollow">https://artificialanalysis.ai/models/comparisons/qwen3-6-27b...</a> clearly shows both the intelligence and speed differences. It's a bit degenerate but you can get some useful info from e.g. the /r/localllama subreddit</p>
]]></description><pubDate>Sat, 08 Aug 2026 23:42:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49226953</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49226953</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49226953</guid></item><item><title><![CDATA[New comment by markasoftware in "Timeline of the OpenAI accidental attack against Hugging Face"]]></title><description><![CDATA[
<p>Certain traits simply cannot exist in a sufficiently intelligent mind. E.g., any "mind" of any type that's sufficiently intelligent will not tell you that 1+1=3 unless it's roleplaying, etc. It doesn't matter if it was trained via gradient descent or any other method. The comments you are responding to, and the original quote from the paper, are suggesting that absolute loyalty / subservience is similarly fundamentally incompatible with intelligence, not just a certain training algorithm or mind architecture. Of course, we have no actual evidence either way.</p>
]]></description><pubDate>Sat, 08 Aug 2026 23:06:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49226751</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49226751</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49226751</guid></item><item><title><![CDATA[New comment by markasoftware in "Qwen3.8 Max now ranked as the best overall model by agentic index"]]></title><description><![CDATA[
<p>It's well known 35b is much faster (on any hardware) and quite a bit dumber</p>
]]></description><pubDate>Thu, 06 Aug 2026 19:54:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49201548</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49201548</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49201548</guid></item><item><title><![CDATA[New comment by markasoftware in "Muse Code and Muse Spark 1.2"]]></title><description><![CDATA[
<p>There is no official dsv4 flash free tier api afaik. Did you get it from open router or something? Likely quantized.<p>The paid official dsv4 api already shares data with deepseek for training</p>
]]></description><pubDate>Thu, 06 Aug 2026 14:21:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=49197074</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49197074</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49197074</guid></item><item><title><![CDATA[New comment by markasoftware in "Qwen3.8-Max: A New Bar for Coding and Cowork"]]></title><description><![CDATA[
<p>Moonshot is printing money on k3. It likely costs the same to serve as qwen3.8. The license requires all major inference providers to sign an extra (secret) licensing agreement with moonshot that almost certainly requires them to agree to this price and pay royalties to moonshot. Watch as the k3 price plummets over the next 1-2 weeks.</p>
]]></description><pubDate>Mon, 03 Aug 2026 07:10:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49152239</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49152239</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49152239</guid></item><item><title><![CDATA[New comment by markasoftware in "Postmortem for Kernel Soundness Bug #14576"]]></title><description><![CDATA[
<p>good point, I didn't think about these cases.</p>
]]></description><pubDate>Sun, 02 Aug 2026 08:15:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49142234</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49142234</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49142234</guid></item><item><title><![CDATA[New comment by markasoftware in "Postmortem for Kernel Soundness Bug #14576"]]></title><description><![CDATA[
<p>Let a and b be as you describe (hash collision), and suppose that collisions are extremely rare. We have a theorem that a=b => a+1=b+1. But in this case, a=b according to our hash-equality, but a+1!=b+1, which contradicts the theorem we've already proved.<p>for real problems with my statement, see your sibling comment.</p>
]]></description><pubDate>Sun, 02 Aug 2026 08:14:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=49142229</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49142229</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49142229</guid></item><item><title><![CDATA[New comment by markasoftware in "OpenAI's claimed disproof of Connes' Rigidity Conjecture is invalid [pdf]"]]></title><description><![CDATA[
<p>Author is a crackpot. She does not meaningfully engage with anyone who points out the key flaw in her counterargument. See the thread here <a href="https://x.com/AcerFur/status/2083649346294382803" rel="nofollow">https://x.com/AcerFur/status/2083649346294382803</a></p>
]]></description><pubDate>Sun, 02 Aug 2026 04:58:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49141261</link><dc:creator>markasoftware</dc:creator><comments>https://news.ycombinator.com/item?id=49141261</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49141261</guid></item></channel></rss>