<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: heresalexandria</title><link>https://news.ycombinator.com/user?id=heresalexandria</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 22 Jul 2026 23:23:17 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=heresalexandria" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by heresalexandria in "NYC Roam: 3D world with real transit and building data"]]></title><description><![CDATA[
<p>An explorable Manhattan built from real map & transit data. Ride any subway, bus, or bike along its true routes & stops, or take helicopter mode and fly above the city to explore.<p>Walk up to any building's address plaque for its story Wikipedia info & link to historical photos from Old NYC.</p>
]]></description><pubDate>Mon, 20 Jul 2026 12:59:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=48978277</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48978277</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48978277</guid></item><item><title><![CDATA[NYC Roam: 3D world with real transit and building data]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.nycroam.com/">https://www.nycroam.com/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48978276">https://news.ycombinator.com/item?id=48978276</a></p>
<p>Points: 5</p>
<p># Comments: 2</p>
]]></description><pubDate>Mon, 20 Jul 2026 12:59:16 +0000</pubDate><link>https://www.nycroam.com/</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48978276</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48978276</guid></item><item><title><![CDATA[New comment by heresalexandria in "Airport Simulator"]]></title><description><![CDATA[
<p>This is super fun! Could be neat to add keyboard controls for auto routing (i.e. select plane number n and auto route it to strip x or have it fly a go-around to wait). Would also be sick if you added models of real world airports to play.<p>That said I love this as is and will definitely be playing & sharing it.</p>
]]></description><pubDate>Mon, 20 Jul 2026 12:30:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=48977921</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48977921</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48977921</guid></item><item><title><![CDATA[New comment by heresalexandria in "Show HN: FableCut – A browser video editor AI agents can drive (zero deps)"]]></title><description><![CDATA[
<p>My impression was that they made this editor with Fable, and its JSON project structure would only serve well for manipulation by lesser models.</p>
]]></description><pubDate>Thu, 09 Jul 2026 14:33:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=48846544</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48846544</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48846544</guid></item><item><title><![CDATA[New comment by heresalexandria in "Show HN: FableCut – A browser video editor AI agents can drive (zero deps)"]]></title><description><![CDATA[
<p>Cool concept, will try it out! I've had decent results with computer use operating conventional editing tools, but being able to directly edit JSON project files is a solid optimization and opens up a lot of opportunities with things like modular templating.</p>
]]></description><pubDate>Thu, 09 Jul 2026 14:31:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=48846511</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48846511</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48846511</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>For sure, I appreciate your comment - this is a tough crowd, but it's their loss.</p>
]]></description><pubDate>Fri, 03 Jul 2026 17:33:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=48777570</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48777570</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48777570</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>That's fair, I do agree that you don't need a harness or ultra-high thinking mode for many problems. Many folks evaluate without those things on a task that would benefit from them leading to the sort of attitudes in this article and its comments section, which is where my comment was coming from.<p>If you're just saying different tools are best suited for different problems, apologies - that's my take as well.</p>
]]></description><pubDate>Fri, 03 Jul 2026 16:35:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776939</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776939</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776939</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>Never said I was bad at math, but I am aware of the fact that computers can do math better and faster than me - and with our powers combined...</p>
]]></description><pubDate>Fri, 03 Jul 2026 16:28:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776864</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776864</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776864</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>This was pretty cool, knocking out a problem that the best minds in maths couldn't for 80 years: <a href="https://openai.com/index/model-disproves-discrete-geometry-conjecture/" rel="nofollow">https://openai.com/index/model-disproves-discrete-geometry-c...</a><p>Also this is a remarkable (and realistic) evaluation of where these systems are for general work which speaks to both the room to grow as well as the pace: <a href="https://www.remotelabor.ai/" rel="nofollow">https://www.remotelabor.ai/</a><p>For some practical examples of what the leading consumer grade AI can do, Ethan Mollick consistently has great writeups with demos: <a href="https://www.oneusefulthing.org/p/what-it-feels-like-to-work-with-mythos" rel="nofollow">https://www.oneusefulthing.org/p/what-it-feels-like-to-work-...</a></p>
]]></description><pubDate>Fri, 03 Jul 2026 15:52:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776479</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776479</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776479</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>Offline models are becoming increasingly more capable - merely a few years ago it would've been unthinkable to run the LLM I have on my phone even on my MacBook Pro.<p>Are you suggesting that losing electricity in the modern age (entirely absent AI) doesn't upend one's world?<p>You seem to be saying "we should avoid this thing because we'll become dependent on it," but we're highly dependent on all manners of technology for all sorts of things and would seem to be better for it.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:44:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776380</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776380</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776380</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>It shouldn't be a surprise that the baseline for "best" shifts as better tech comes out, but that doesn't make dated models any less capable than they were when they came out.<p>Skeptics continue to move the goalposts on what constitutes this mattering, but the fact that frontier systems are making novel maths & sciences discoveries and I can run an LLM on my phone for simple tasks that would've been unthinkable a few years ago are testaments to the directionality of the tech.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:37:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776299</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776299</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776299</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>That's exactly what the people in my orbit and whom I'm watching are doing, and some of their outputs are fueling the excitement.<p>If you aren't seeing remarkable things being done with this tech, I'd argue you aren't looking hard enough. I understand there's a lot of noise obscuring the signal, but that's always the case with a "big thing."</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:33:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776237</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776237</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776237</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>Qwen is a lightweight locally hosted model that's many months behind the SoTA available from the big three - while the crowd here (myself included) is excited for locally hosted models to catch up to the usable baseline, regardless of what benchmarks you based that selection on they aren't there yet.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:29:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776193</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776193</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776193</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>This sounds like you may be using subpar models and/or tools - have you had this experience using Codex with GPT-5.5 on at least "high" reasoning or on Claude Code using Opus 4.8 (both with ability to browse web and sufficient context for your project)?</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:27:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776168</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776168</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776168</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>Totally agree that the lack of a common base of evaluation is terrible for the discussion, and benchmaxing only contributes to this.<p>The only way to get a sense for these systems is to use them on things you know well, and everyone knows different things at different levels.<p>People also tend to underestimate how fast this is moving and base their take on dated and subpar systems for a variety of reasons, a key one being that the firehose is too big for any one person to have a proper focus on all of it.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:24:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776146</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776146</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776146</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>Did you try providing it documentation for the respective formats (via browsing/tool use or input to the prompt)? And were you using a modern thinking model from Anthropic or OpenAI?<p>The crucial breakdown here sounds like either lack of proper context/harness or insufficiently capable model (there's a huge gulf between GPT-5.5/Opus 4.8/Fable class models and anything not from the big three) or both.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:20:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776109</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776109</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776109</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>Something a lot of folks struggling with these systems don't get is that the instruction and management of them is often quite important - just because they're capable doesn't mean they're mind readers.<p>Most of the skepticism I encounter on this front is due to lack of proper direction, process involving planning and review before execution, and appropriate attention given to evaluation and feedback loops.<p>If you asked the smartest person in the world to YOLO a task with the sort of instruction the average denier uses to evaluate an LLM, you'd likely find they wouldn't get back what they were expecting either - and if you're evaluating on subpar models/tools, you shouldn't be surprised to get subpar results.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:16:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776057</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776057</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776057</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>They're literally doing novel research. The smartest mathematicians in the world couldn't solve Erdős' planar unit distance problem for 80 years, and OpenAI's models knocked that out a couple months ago.<p>This stuff is moving fast, and if you aren't evaluating SoTA on at least a quarterly basis, you're going to have a bad time.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:12:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=48776023</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48776023</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48776023</guid></item><item><title><![CDATA[New comment by heresalexandria in "Please stop the AI confidence theater"]]></title><description><![CDATA[
<p>The same attitude has been directed at points through history for people "who depend on the internet," "who depend on computers," and "who depend on machines."<p>I was told growing up "you won't always have a calculator in your pocket" and yet now my phone has an offline LLM on it.</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:09:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=48775985</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48775985</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48775985</guid></item><item><title><![CDATA[New comment by heresalexandria in "Quake in 13 Kilobytes (2021)"]]></title><description><![CDATA[
<p>This is really remarkable, and frankly more playable than a number of niche ports like this that I've tried - great work!</p>
]]></description><pubDate>Fri, 03 Jul 2026 15:04:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=48775930</link><dc:creator>heresalexandria</dc:creator><comments>https://news.ycombinator.com/item?id=48775930</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48775930</guid></item></channel></rss>