<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: luulinh90s</title><link>https://news.ycombinator.com/user?id=luulinh90s</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 04 Aug 2026 11:26:13 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=luulinh90s" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by luulinh90s in "Steering interpretable language models with concept algebra"]]></title><description><![CDATA[
<p>We haven’t benchmarked our steering for scaffolding function-calling in an agent loop yet (and the model we are using is just a base model), so I can’t give a quantitative claim. But concept-based steering should be a good fit for keeping the agent on task and enforcing behavioral guardrails around tool use.<p>In practice, you can treat concepts as soft/hard constraints to bias the agent toward: (1) calling tools only when needed, (2) selecting the right tool/function, or (3) using the correct argument schema.</p>
]]></description><pubDate>Fri, 27 Feb 2026 08:17:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=47177990</link><dc:creator>luulinh90s</dc:creator><comments>https://news.ycombinator.com/item?id=47177990</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47177990</guid></item><item><title><![CDATA[New comment by luulinh90s in "Steering interpretable language models with concept algebra"]]></title><description><![CDATA[
<p>Hi! Thanks for checking.<p>We haven’t published the concept dictionary yet.<p>We plan to release it in soon with other important artifacts.</p>
]]></description><pubDate>Fri, 27 Feb 2026 08:05:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=47177899</link><dc:creator>luulinh90s</dc:creator><comments>https://news.ycombinator.com/item?id=47177899</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47177899</guid></item><item><title><![CDATA[Steering interpretable language models with concept algebra]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.guidelabs.ai/post/steerling-steering-8b/">https://www.guidelabs.ai/post/steerling-steering-8b/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=47159833">https://news.ycombinator.com/item?id=47159833</a></p>
<p>Points: 77</p>
<p># Comments: 8</p>
]]></description><pubDate>Wed, 25 Feb 2026 23:55:34 +0000</pubDate><link>https://www.guidelabs.ai/post/steerling-steering-8b/</link><dc:creator>luulinh90s</dc:creator><comments>https://news.ycombinator.com/item?id=47159833</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47159833</guid></item><item><title><![CDATA[New comment by luulinh90s in "Show HN: Steerling-8B, a language model that can explain any token it generates"]]></title><description><![CDATA[
<p>in the "Performance" section of the post: <a href="https://www.guidelabs.ai/post/steerling-8b-base-model-release/">https://www.guidelabs.ai/post/steerling-8b-base-model-releas...</a>, the authors show the model lags behind llama 8b but worth noting that llama 8b trained on > 2x more computes (see the FLOPs axis)</p>
]]></description><pubDate>Tue, 24 Feb 2026 05:43:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=47133293</link><dc:creator>luulinh90s</dc:creator><comments>https://news.ycombinator.com/item?id=47133293</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47133293</guid></item></channel></rss>