<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: jermaustin1</title><link>https://news.ycombinator.com/user?id=jermaustin1</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Mon, 28 Sep 2026 07:27:45 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=jermaustin1" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by jermaustin1 in "A single function Jev-like wrapper for LLMs, including vision models"]]></title><description><![CDATA[
<p>A is significant in it's own right. But most models follow instructions well enough when you prompt it, "Choose one of the following answers:", it will follow that 90% of the time. Use temperature tuning, and a LoRA, and you are 99% of the way there, just without the speed that Jev has.<p>Yesterday, just to prove to a friend that Jev isn't that "revolutionary" I extracted some image classification code that Claude had written for my private image organizer that used Qwen3-VL, and stopped the output at a single token, then used the probabilities. Input processing on my GPU was somewhere around 1000ms per image, so not too fast, but each question used the prompt cache, so followups were 100ms-ish.<p>That was my baseline of an untrained, non-optimized single pass classification. It would connect to my llama-server, and use the logprobs for the choices.<p>After that, I had Claude remove llama-server from the solution, and write it directly to the transformers, then I kept prompting it to profile and find more speed. Eventually my "Decision Engine" running locally on a trained 1B model got  to just under 85% accuracy across my 500 validation prompts (images and text) not used or derived from the training set, and an 8MB image, with 10 questions with 5 choices per question, got down to just under 500ms. Pure text prompts and questions are below 100ms for 300tokens + 10 questions + 5 choices per question (average).<p>It did better on text than images, just because my training set included 90% text. I'll do more training and validation for images when I get home, but for now, I'm more than happy that I can get a local "decision engine" running in 4GB of VRAM and responding in under 30ms for most use cases I've had.</p>
]]></description><pubDate>Sat, 26 Sep 2026 16:14:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49857928</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49857928</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49857928</guid></item><item><title><![CDATA[New comment by jermaustin1 in "A single function Jev-like wrapper for LLMs, including vision models"]]></title><description><![CDATA[
<p>I don’t know how this wrapper works, but if it is like any of the classifiers I’ve had Claude build off an LLM in the past, it grabs the probabilities of the tokens you are looking for, and then computes their relative probs against each other.<p>Even if the LLM thinks it’s made up D is the highest probability, that isn’t part of the set.<p>You never actually generate the prose, only the first pass, and grab the probabilities. It couldn’t ask for more details even if it wants to. It gets stopped before the first token renders.</p>
]]></description><pubDate>Sat, 26 Sep 2026 11:48:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49855635</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49855635</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49855635</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Meta VR Glasses"]]></title><description><![CDATA[
<p>I bought my Quest 2 when it first came out, and I played it for a while, about 2-3 hours a week, peaking at about an hour a day when I found that Blade and Sorcery is a really good game for "exercise".<p>I bought the Quest 3 as an upgrade because I had noticed my Quest 2 started having intermittent input lag issues. I never pinpointed what was causing it, but my Golf clubs would fly out of my hand and I'd have to wait a few seconds for them to eventually find their way back to me. My controllers would drift out of view when navigating menus. I uninstalled everything, reinstalled it again. When that didn't work, I upgraded.<p>The issues persisted in the new version as well. So for about the last 20-ish months, I've averaged under an hour a month, most months I don't even pick it up. I might do some homemade PC VR "experiments" (still have access to mouse and keyboard for the game, and not the controllers) on some random claude code experiments. I have a "cyberpunk" mobius strip space station walking sim I've been working on, that will eventually hook into a treadmill.</p>
]]></description><pubDate>Thu, 24 Sep 2026 13:30:53 +0000</pubDate><link>https://news.ycombinator.com/item?id=49830328</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49830328</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49830328</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Microsoft killed FoxPro in 2007. Anyway, here's FoxPro revived"]]></title><description><![CDATA[
<p>My first real job was 2006, turning a Visual FoxPro application into a Web Application using ASP.Net 1.1 Web Forms and VB.Net.<p>With that said, I've never actually used FoxPro. I only had a database as the contract for what my web app was supposed to do.</p>
]]></description><pubDate>Tue, 22 Sep 2026 21:40:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808530</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49808530</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808530</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Ask HN: Can we please limit the AI news flood?"]]></title><description><![CDATA[
<p>Not in my experience. Multiple submissions (non Show HN) of mine have been second chanced. I can't be certain, but I don't think any of my Shows have been... I guess that says more about the content I create than the content I share.</p>
]]></description><pubDate>Fri, 11 Sep 2026 13:42:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=49658310</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49658310</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49658310</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Ask HN: Can we please limit the AI news flood?"]]></title><description><![CDATA[
<p>They already do this. It's called the "second chance" pile. You will get an email saying that your submission didn't get the love they felt it should, and will add it to a list of submissions that get submitted to the front page, to see if the community agrees.</p>
]]></description><pubDate>Fri, 11 Sep 2026 13:37:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49658222</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49658222</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49658222</guid></item><item><title><![CDATA[New comment by jermaustin1 in "GPT-6 Astra on OpenRouter"]]></title><description><![CDATA[
<p>A bespoke design like that a year ago would have cost $200-2000+ between design and development. Hell, it probably still does unless the person wanting the design is already a developer who knows how to prompt.<p>Everything costs more now than it did a year ago... except for THIS, and we are still complaining that a 90-99% reduction in cost is STILL too expensive. And a 50-75% reduction in time is STILL too long.<p>We used to have to wait for weeks for a design like that when I worked at a consultancy, and that is a week of salary. For the design, then it got handed off to a front end developer to slice it and get built so the back end developer can hook it up to a CRM. We are talking a month turn around with design, revisions, development, testing, and bug fixing.<p>It can now be done in a couple of hours for less than a single hour's cost. If it were 10x slower and 10x more expensive, it would STILL be "good deal".</p>
]]></description><pubDate>Sat, 05 Sep 2026 17:02:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=49578428</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49578428</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49578428</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Invisible Companies"]]></title><description><![CDATA[
<p>I worked for a different Steve Ross (Dolphins owner, Related Companies, etc), and we did this in house.<p>From 2008-2017, “we” (I was a contractor 08-09, and employee 15-17) basically turned every spreadsheet into a web application internally.<p>We added authentication and authorization, used a LOT of ETL-type processing to move data around. So many things could have been packaged and sold, but that was the “secret sauce” that kept the company so profitable.<p>Then the end in 2017 when a new CIO came in, killed off IT (laid off all non-managers over the course of a year), and replaced everyone with South African consultants to turn IT from a cost center to a profit center.<p>He lasted another couple years then left.</p>
]]></description><pubDate>Thu, 03 Sep 2026 13:30:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49549694</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49549694</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49549694</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Muse Spark 1.3"]]></title><description><![CDATA[
<p>Years ago, I lived in NYC, and my roommate was a director of photography for National Geographic, and various other nature documentaries. I loved photography (still do, but much less time for it as a late 30s adult than a mid 20s adult), and she was kind enough to answer any question I had regarding film/photo.<p>She told me that "left to right" denoted progression in the story, "right to left" told the viewer the subject was "exiting" the current scene.<p>She didn't go into the details of WHY, and I probably didn't probe deeper, but it stuck with me, and I notice it all the time in film and television.</p>
]]></description><pubDate>Thu, 03 Sep 2026 11:26:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=49548634</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49548634</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49548634</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Debian votes to allow "responsible use of generative AI""]]></title><description><![CDATA[
<p>Opposite policy at one of my clients (kind of). I am responsible for the code that upper management's Claude produces. Some Mondays, I will start work with a half dozen emails with attachments of Claude generated code for something I don't even know what the point is, with the task of "integrate this and make sure it works." without any context to go along with it, so I have to read the code, usually hundreds of lines and understand WHY manager wanted it, before I can start to code it myself, because it is 1) in the wrong language, 2) doesn't understand our codebase, 3) is using libraries we can't license, etc.<p>My job has been less watching Claude Code, and more watching Managers Claude Code.<p>I don't know which I hate more as a programmer.</p>
]]></description><pubDate>Sat, 29 Aug 2026 16:38:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49491297</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49491297</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49491297</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Show HN: Galaxium, an experimental WebGPU space explorer"]]></title><description><![CDATA[
<p>Wow! I'm utterly speechless.<p>I love space "browsers" - I lose myself in them, just clicking through all the entities.<p>Also TIL that there are satellite galaxies, and I've never noticed those on other toys/tools like this. I always though Andromeda was the closest galaxy, but it is FAR away compared to `Sgr dSph` which is within 50kly from the center of the Milky Way.</p>
]]></description><pubDate>Sat, 29 Aug 2026 12:57:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=49489541</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49489541</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49489541</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>I've done pretty decent local prose->json extraction using Qwen and Phi and Gemma.<p>I'm sure most of it comes down to prompts, and all of them run over 100tps on a 3090. Smaller cards will likely be slower, but Qwen3.5 9B is small enough to fit on most consumer cards.</p>
]]></description><pubDate>Fri, 28 Aug 2026 15:38:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=49480170</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49480170</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49480170</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>They cost more to run than hosted anyway. But that isn't the point of having them. They are a playground, a backup when the internet is down, or claude is down. They can render Blender scenes pretty well. They play any game I want.<p>You can do each of those at various hosts and own nothing. Or own a couple "over priced" cards and do it all at home on battery power for a few hours while the power is out.</p>
]]></description><pubDate>Thu, 27 Aug 2026 18:57:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49469576</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49469576</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49469576</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>I don't think there is tunnel vision. I'm just saying that I have a couple 3090s I invested in a handful of years ago, and they are still going strong today as multiple GPU-needing technologies emerged.<p>I'm not saying everyone has to run local LLMs, because the APIs are in a race to the bottom, and my $10 of OpenRouter credits I bought months ago is down to $8.94 because most models give you MILLIONS of tokens for a US Quarter.</p>
]]></description><pubDate>Thu, 27 Aug 2026 18:54:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49469534</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49469534</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49469534</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>Having multiple 6 year old cards doesn't seem like it's that big of burden for local LLMs.<p>I get that a lot of people don't have them. And a single one can be VERY performant. And the smaller models like a 7B can run on much smaller hardware like a mid-range [3|4|5]060.<p>My entire AI Dev Box cost $4500 in parts. 128GB RAM, i7-10700, 1TB and 2TB SSD, and 2x 3090s. Today's prices and inflation have definitely made that price tag seem a lot better than it was, but it was an investment in all things GPU that were happening in 2020 (crypto, blender, image gen), then LLMs exploded.</p>
]]></description><pubDate>Thu, 27 Aug 2026 17:41:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49468441</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49468441</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49468441</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Small Models Have Arrived"]]></title><description><![CDATA[
<p>To me, most local models work just fine for anything you can be patient for. If I want something quicker, I will go to a SOTA model via API, but with multiple 3090s, I have never really needed a hosted model for a lot of my experiments.<p>For code, they are great, but for creativity for NPC controllers, they leave something to be desired, but work well enough for testing, so I don't burn tokens until I'm actually playing my games.<p>But nothing one-shots a prototype better than Fable 5. I can have a prototype built in 30 minutes, hooked up to my local LLMs and Claude Code is very good at testing the interactions and even tuning the prompts of the NPCs for better experiences.</p>
]]></description><pubDate>Thu, 27 Aug 2026 17:21:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=49468141</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49468141</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49468141</guid></item><item><title><![CDATA[New comment by jermaustin1 in "CEO fired developers to make room for AI. Developers create open source AI CEO"]]></title><description><![CDATA[
<p>I know this is mostly true at huge companies, but having spent my career at various "small" corporations, with maybe a few thousand employees total (and only ~50 of which were in IT). I was always surprised when the CEO knew names of people like me, the lowly developer 5 layers below them on the org chart, but beyond that, remembered the last conversation we had, or an interest we might have shared (golf, hiking, travel, etc.).<p>Even as a temporary consultant, at a company for only a couple of months, I remember seeing the CEO walk the hallway, greeting every single person on the way between a meeting room and the bathroom. I can barely remember my family member's names, or the last thing we talked about, let alone have the memory space for all that information to seem kind and courteous to "low level" employees.<p>In my entire career, I've never worked at a behemoth-sized company, though, so my experience has been much more "intimate" environments, and primarily in consulting firms or satellite offices away from the main HQ (even by only a couple blocks some times). So everyone in the office knew each other.</p>
]]></description><pubDate>Thu, 27 Aug 2026 13:10:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49464358</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49464358</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49464358</guid></item><item><title><![CDATA[New comment by jermaustin1 in "New Mac Studio with M5 Max and M5 Ultra"]]></title><description><![CDATA[
<p>I just wait for a cycle or two where the leaps and bounds are more like hops and steps. So if the M7 Ultra improves inference by 2x over the M5, but the M9 Ultra only improves by 1.2x over the M7, that's my signal to buy. Unfortunately they haven't slowed down yet.</p>
]]></description><pubDate>Tue, 25 Aug 2026 14:19:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49434635</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49434635</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49434635</guid></item><item><title><![CDATA[New comment by jermaustin1 in "New Mac Studio with M5 Max and M5 Ultra"]]></title><description><![CDATA[
<p>I've been holding out, because I think my next purchase will be a Studio with an Ultra Chip in it. I'm wanting it to be a "forever" server, so I'm holding out while I can.</p>
]]></description><pubDate>Tue, 25 Aug 2026 13:28:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49433693</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49433693</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49433693</guid></item><item><title><![CDATA[New comment by jermaustin1 in "Show HN: Module for LLM Homeostasis (PoC)"]]></title><description><![CDATA[
<p>Have you actually read your README?<p>There are so many parts that are just "words" thrown together to sound technical. It reminds me of books you pick up in a science fiction game.</p>
]]></description><pubDate>Mon, 24 Aug 2026 21:53:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=49426346</link><dc:creator>jermaustin1</dc:creator><comments>https://news.ycombinator.com/item?id=49426346</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49426346</guid></item></channel></rss>