<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: _KnighT_</title><link>https://news.ycombinator.com/user?id=_KnighT_</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 17 Sep 2026 05:10:51 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=_KnighT_" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by _KnighT_ in "Efficient Reasoning with Hidden Thinking"]]></title><description><![CDATA[
<p>I'm new to this topic. Can someone help me understand this sentence?<p>"Meanwhile, through the next-token prediction constraint, the explicit
textual symbols of the hidden representations for Heima
Encoder are aligned to the text of the corresponding special
tokens {<CoT>(k)} in vocabulary, while the hidden representations contained in hidden states of thinking tokens remain distinct and variable depending on the inputs"<p>I understand that they have fine-tuned the MLLM to produce, in response to each query and image input, the CoT "thinking tokens" in addition to the answer.<p>How does that establish an association between the thinking tokens and the original plain-English CoT statements?<p>The second clause seems to say that the thinking tokens encode information that is "distinct and variable depending on the inputs." Is my interpretation correct?</p>
]]></description><pubDate>Mon, 03 Feb 2025 21:10:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=42923188</link><dc:creator>_KnighT_</dc:creator><comments>https://news.ycombinator.com/item?id=42923188</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42923188</guid></item></channel></rss>