<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: neversupervised</title><link>https://news.ycombinator.com/user?id=neversupervised</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sat, 01 Aug 2026 02:32:31 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=neversupervised" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by neversupervised in "Ford AI hiccups push carmaker to rehire ‘gray beard’ inspectors"]]></title><description><![CDATA[
<p>This just feeds a certain narrative and allows people to take exactly the wrong conclusion. Just because there’s some uncertainty at the edge, it doesn’t change where things are going.</p>
]]></description><pubDate>Thu, 25 Jun 2026 15:36:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48675023</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48675023</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48675023</guid></item><item><title><![CDATA[New comment by neversupervised in "Roughly a quarter of American professionals hit a wall in their careers"]]></title><description><![CDATA[
<p>Years of experience don’t correlate to output in all careers. Surgeons and engineers get better over time. This might not be true for all jobs. Meanwhile, management is naturally capped because every manager necessarily needs people to manage under them, so there’ll be 1/N^y managers at the yth level of the org. Unless loyalty ought to be reward for its own sake, it’s not clear why 100% of workers should get promoted indefinitely.</p>
]]></description><pubDate>Mon, 01 Jun 2026 14:43:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=48357514</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48357514</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48357514</guid></item><item><title><![CDATA[Ask HN: Can someone help me understand the AI vibes on HN?]]></title><description><![CDATA[
<p>I’m just as worried as the next guy about the negative implications of AI. While I’m a long-term optimist, I have high confidence that we’re about to go through a rough period, with unpleasant labor markets, class clashes, enshittification of the consumer experience, and so on. However, I’m not in denial. I can clearly see why this will happen. It’s a combination of a technology that is incredibly capable and improving fast, with a world economy trying to absorb and adapt to this change.<p>Meanwhile, every AI post on HN seems to display a lot of denial. Many here seem convinced that AI won’t compete with their current jobs. Arguments are made about current capability gaps, instead of rate of improvement. People talk about how they don’t like AI, or don’t like working with AI, or how employers are mean and selfish for pushing AI or doing layoffs.<p>I find that viewpoint shortsighted. Regardless of how one feels, it’s better to have a more accurate understanding of the future than a less accurate one. I wish people were more inclined to understand why and how things are changing, and make personal decisions, including voting or going to work on AI safety, instead of just being upset about it.<p>Am I reading the room wrong? I don’t get the same vibes on other tech communities.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48321859">https://news.ycombinator.com/item?id=48321859</a></p>
<p>Points: 2</p>
<p># Comments: 1</p>
]]></description><pubDate>Fri, 29 May 2026 11:39:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=48321859</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48321859</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48321859</guid></item><item><title><![CDATA[New comment by neversupervised in "Five frontier LLMs disagree on 67% of 1k real-world fact-check claims"]]></title><description><![CDATA[
<p>This is not how people use LLMs. If you ask one of these questions you’d get a longer answer, often grounded on the internet. I speculate that conditional on a smart human operator interpreting the results, such interpretations across vendors converge more often than this report makes it seem.</p>
]]></description><pubDate>Thu, 28 May 2026 14:23:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=48309372</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48309372</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48309372</guid></item><item><title><![CDATA[New comment by neversupervised in "I'm Tired of Talking to AI"]]></title><description><![CDATA[
<p>This is just a glitch in time. It’ll be agents talking to agents. We won’t be able to keep up.</p>
]]></description><pubDate>Wed, 27 May 2026 13:33:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48294120</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48294120</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48294120</guid></item><item><title><![CDATA[New comment by neversupervised in "At least 25 Flock cameras have been destroyed in five states since April 2025"]]></title><description><![CDATA[
<p>Check out Pangram</p>
]]></description><pubDate>Sun, 17 May 2026 18:18:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=48171629</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48171629</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48171629</guid></item><item><title><![CDATA[New comment by neversupervised in "I don't think AI will make your processes go faster"]]></title><description><![CDATA[
<p>It’s completely wild to me that lifelong programmers come into contact with agentic coding and come to the conclusion that their jobs are safe for one reason or another. AI will definitely be able to write entire software, inclusive of figuring out requirements and asking the right questions. It’s not that far already. Why is it that everyone looks at weaknesses of a technology that didn’t exist a couple years ago instead of appreciating the incredible rate of improvement? I know why, because it’s inconvenient to the narrative of what makes us valuable. But still, our job is to turn ideas into a sequence of logical steps. Why can’t we do the same when forecasting the impact of AI on our jobs?</p>
]]></description><pubDate>Sun, 17 May 2026 14:34:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=48169285</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48169285</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48169285</guid></item><item><title><![CDATA[New comment by neversupervised in "The Terminal Bench 3.0 community is looking for task contributors"]]></title><description><![CDATA[
<p>This is a great way to become well-versed in benchmarking.</p>
]]></description><pubDate>Sun, 03 May 2026 21:19:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=48001612</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=48001612</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48001612</guid></item><item><title><![CDATA[New comment by neversupervised in "SWE-bench Verified no longer measures frontier coding capabilities"]]></title><description><![CDATA[
<p>Terminal Bench is the future</p>
]]></description><pubDate>Sun, 26 Apr 2026 15:04:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=47910915</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47910915</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47910915</guid></item><item><title><![CDATA[New comment by neversupervised in "SWE-bench Verified no longer measures frontier coding capabilities"]]></title><description><![CDATA[
<p>But this is the good kind of goalpost moving</p>
]]></description><pubDate>Sun, 26 Apr 2026 15:04:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=47910912</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47910912</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47910912</guid></item><item><title><![CDATA[New comment by neversupervised in "Show HN: Terminal-Wrench, a dataset of 331 realistic hackable environments"]]></title><description><![CDATA[
<p>That paper focuses on breaking the harness, the same hack applies to all tasks. Here we are breaking tasks individually. If these were put on a different, more secure harness, most of the exploits would still work.</p>
]]></description><pubDate>Wed, 15 Apr 2026 15:19:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=47780355</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47780355</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47780355</guid></item><item><title><![CDATA[Show HN: Terminal-Wrench, a dataset of 331 realistic hackable environments]]></title><description><![CDATA[
<p>I want to share a new dataset of 331 reward-hackable environments. These are real environments used in Terminal Bench and adjacent benchmarks. I first got interested in this because, as a reviewer of Terminal Bench, I noticed a lot of our tasks were hackable. I also noticed that many contributors to the benchmark do so because it provides credibility when selling environments to labs. Hence, TBench tasks are, in my opinion, held to a higher quality standard than those being used today for RL. No one is spending hours manually reviewing the $1B in tasks being purchased by major labs. As far as I understand, while everyone knows environments are hackable, nobody has released hundreds of "realistic" environments.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=47773298">https://news.ycombinator.com/item?id=47773298</a></p>
<p>Points: 6</p>
<p># Comments: 2</p>
]]></description><pubDate>Wed, 15 Apr 2026 00:42:30 +0000</pubDate><link>https://github.com/few-sh/terminal-wrench</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47773298</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47773298</guid></item><item><title><![CDATA[New comment by neversupervised in "90% of CEOs Say AI Changed Nothing. The Other 10% Have a PR Team"]]></title><description><![CDATA[
<p>This is nonsense. I’m sorry. AI will completely upend the workplace and the economy. Whether that’s self evident today in the numbers in the way that we track those numbers, which is based on how things have historically worked, is not relevant. First principles thinking is enough.<p>C’mon. Stop wishing for a future that feels convenient. This is not the world in which we live. Everything will change. Let’s help people accept and react to that.Let’s stop with the comfort talking and false hope.</p>
]]></description><pubDate>Tue, 14 Apr 2026 15:31:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=47766974</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47766974</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47766974</guid></item><item><title><![CDATA[New comment by neversupervised in "How to Make a Good Terminal Bench Task"]]></title><description><![CDATA[
<p>I've been a contributor and reviewer for terminal bench since last August, and this post is about what I've learned designing and reviewing tasks. The guidance is broadly applicable to anyone building an agentic benchmark.I would love feedback from the HN community.</p>
]]></description><pubDate>Mon, 23 Mar 2026 18:41:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=47493463</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47493463</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47493463</guid></item><item><title><![CDATA[How to Make a Good Terminal Bench Task]]></title><description><![CDATA[
<p>Article URL: <a href="https://twitter.com/neversupervised/status/2035455298417430911">https://twitter.com/neversupervised/status/2035455298417430911</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=47493447">https://news.ycombinator.com/item?id=47493447</a></p>
<p>Points: 3</p>
<p># Comments: 1</p>
]]></description><pubDate>Mon, 23 Mar 2026 18:39:47 +0000</pubDate><link>https://twitter.com/neversupervised/status/2035455298417430911</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47493447</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47493447</guid></item><item><title><![CDATA[New comment by neversupervised in "Reports of code's death are greatly exaggerated"]]></title><description><![CDATA[
<p>Can you explain what you think will happen, actually? People at OpenAi and Anthropic aren’t longer coding by hand. Are you saying everyone changes their mind and goes back? Not gonna happen. You have to work around this new constrain.</p>
]]></description><pubDate>Mon, 23 Mar 2026 14:34:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=47490129</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47490129</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47490129</guid></item><item><title><![CDATA[New comment by neversupervised in "Reports of code's death are greatly exaggerated"]]></title><description><![CDATA[
<p>The author’s intuition is still backward calibrated, even though he talks about the future. He doesn’t have an intuition for the future. All code will be AI generated. There’s no way to compete with the AI. And whatever new downsides this brings will be solved in ways we aren’t fully anticipating. But the solution is not to walk back vibecoding. You have to be blind to believe not most code will be vibecoded very soon.</p>
]]></description><pubDate>Sun, 22 Mar 2026 20:35:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=47481853</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47481853</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47481853</guid></item><item><title><![CDATA[New comment by neversupervised in "Atlassian says it had right to fire engineer for suggesting CEO is 'rich jerk'"]]></title><description><![CDATA[
<p>There’s no reason a company should put up with enemies within. In rare instances a disgruntled employee might be able to make a positive contribution. In most cases, even if the employee has valid reasons, by the time they are disgruntled there’s no coming back. It’s best for everyone to move on.</p>
]]></description><pubDate>Sun, 22 Mar 2026 16:19:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=47479072</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47479072</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47479072</guid></item><item><title><![CDATA[New comment by neversupervised in "I'm OK being left behind, thanks"]]></title><description><![CDATA[
<p>The mistake is that 1 every N waves of hype are in fact monumental shifts and it makes sense to embrace as soon as possible. Also being early to the right thing can have massive implications in appreciating the shift before the general public, which is upstream from making smart resource allocation (investments, career choices). I have a friend that went to super early OpenAI as a designer. He has more equity than most AI researchers there and made a 0.001% amount of wealth. Being early does very much matter in the right conditions.</p>
]]></description><pubDate>Sat, 21 Mar 2026 14:27:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=47467347</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47467347</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47467347</guid></item><item><title><![CDATA[New comment by neversupervised in "Warranty Void If Regenerated"]]></title><description><![CDATA[
<p>I don't oppose reading AI generated content in principle, but because it's free to generate, I always am less likely to read super long prose that is AI generated. So the question is whether someone has taken the time to keep it as long as necessary but not longer. Or if there are ways to make it easier for me to commit to the experience, with a sort of TLDR</p>
]]></description><pubDate>Wed, 18 Mar 2026 22:35:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=47432254</link><dc:creator>neversupervised</dc:creator><comments>https://news.ycombinator.com/item?id=47432254</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47432254</guid></item></channel></rss>