<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: taylorhou</title><link>https://news.ycombinator.com/user?id=taylorhou</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 02 Sep 2026 09:39:48 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=taylorhou" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by taylorhou in "My local model setup on an M4 Pro Mac Mini"]]></title><description><![CDATA[
<p>i have a 512gb ram m3 ultra mac studio setup with a gas city that runs one of my companies. today was the first time ever that a local model (GLM5.3 8-bit) was able to match fable5 in our tests.<p>GLM-5.3-Flash at true 8-bit: 341 GB on disk, 328 GB resident, 288 experts across 46 layers, loads in 65 seconds.
• 18.7 tokens/s generation, 35 tokens/s prompt, on a desk, on a $0 per-token bill.
• Runs beside our whole agent city on one box with ~130 GB to spare.
• Review test: caught 6 of 6 planted P1 defects, zero false positives, same score as the frontier model we pay for.
• CRM test: 11 of 11 required records extracted, zero wrong writes, 45 minutes, first local model to clear the bar.
• Serving a 131k-token window today; the model itself supports 1,048,576. Widened to 4 concurrent slots and still have 50gb+ of excess ram.<p>granted my cto still isn't moving all of our inference to glm5.3 but we've identified 40%+ that is currently handled by fable that we're routing locally instead and will do concurrent requests to verify/compare responses for a while.</p>
]]></description><pubDate>Wed, 02 Sep 2026 06:47:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=49532605</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=49532605</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49532605</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: What are you working on? (June 2026)"]]></title><description><![CDATA[
<p>teale.com - distributed ai inference using networked devices
essentially folding@home but sharing underutilized ram (when you're asleep, someone else in the world is awake)<p>would really appreciate testers but also any companies thinking about distributed inference powered by their own company devices on a private network. my own company has 200+ 16gb ram machines that we're using for inference.</p>
]]></description><pubDate>Sun, 14 Jun 2026 21:03:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=48532680</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=48532680</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48532680</guid></item><item><title><![CDATA[New comment by taylorhou in "Open source AI must win"]]></title><description><![CDATA[
<p>Let's collab. I'm one guy too but I built distributed inference network (teale.com) banging away for about a month with opus/gpt</p>
]]></description><pubDate>Sat, 13 Jun 2026 19:21:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=48520517</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=48520517</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48520517</guid></item><item><title><![CDATA[New comment by taylorhou in "Open source AI must win"]]></title><description><![CDATA[
<p>I built Teale.com and opensourced it. My domain contribution to society. It powers fully distributed inference on Mac, windows, Linux, android, iOS, hell even harmonyOS.<p>Opensource/weight models will get better and better and eventually we will have mythos level running on smartphone/eyeglass hardware.<p>It is stupidly tedious currently to match supply with demand though because physical hardware like a 16gb ram MacBook doesn't mean there's truly 16gb available let alone matching models and all of their settings (kvcache, context limit, temperature, etc) to demand.<p>Would appreciate any help cus we need ai inference by the people for the people.</p>
]]></description><pubDate>Sat, 13 Jun 2026 19:18:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=48520485</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=48520485</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48520485</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: Who is hiring? (June 2026)"]]></title><description><![CDATA[
<p>APMHelp.com / PM-in-a-box.org || AI native growth marketer | Remote | Contract/Full-time/part-time<p>We're an outsourced accounting firm specializing in property management (appfolio, buildium, rentvine) that's building an open source ai harness/orchestrator for property managers.<p>$7M ARR, cash flow positive, looking to re-envision marketing/growth ai native.<p>email taylor at the apmhelp domain or comment and i'll reach out. =)</p>
]]></description><pubDate>Wed, 03 Jun 2026 15:52:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=48385726</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=48385726</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48385726</guid></item><item><title><![CDATA[New comment by taylorhou in "Show HN: Needle: We Distilled Gemini Tool Calling into a 26M Model"]]></title><description><![CDATA[
<p>this is awesome. i'm the founder/maintainer at teale.com (open source distributed inference app) and one of the biggest challenges has been an actually usable/reliable local model that can run on 8gb macbook air's, 6gb android smartphones, etc... will get some of our test machines serving needle</p>
]]></description><pubDate>Thu, 14 May 2026 15:49:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=48137173</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=48137173</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48137173</guid></item><item><title><![CDATA[New comment by taylorhou in "Show HN: Rapid-MLX – Run local LLMs on Mac, 2-3x faster than alternatives"]]></title><description><![CDATA[
<p>is this available for other open source projects? i'm stealing tokens from my employer effectively and keep hitting my limits re: codex tokens o_O<p>i'm working on github.com/teale-ai (distributed inference)</p>
]]></description><pubDate>Tue, 05 May 2026 19:05:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=48027023</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=48027023</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48027023</guid></item><item><title><![CDATA[New comment by taylorhou in "Several Mac mini and Mac Studio configs are now out of stock at Apple"]]></title><description><![CDATA[
<p>250+ employees on the lowest paid subscriptions of claude or openai is $5k/month. their usage of AI is chatbot most of the time which local models/inference can easily handle. so just being able to cancel those subscriptions with local hardware makes my break even on a single $10k mac studio 2 months and honestly the mac studio for their use is overkill.</p>
]]></description><pubDate>Mon, 13 Apr 2026 17:05:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=47755011</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=47755011</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47755011</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: What Are You Working On? (April 2026)"]]></title><description><![CDATA[
<p>Distributed ai inference pool for any Mac/iOS device where devices are paid for contributing unused ram. To help with the demand, also doing multiplayer AI. $0.05/million tokens - teale.com</p>
]]></description><pubDate>Mon, 13 Apr 2026 11:59:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=47750765</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=47750765</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47750765</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: What Are You Working On? (April 2026)"]]></title><description><![CDATA[
<p>Booked a time! We built senior smartphone assistance without humans</p>
]]></description><pubDate>Mon, 13 Apr 2026 11:53:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=47750717</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=47750717</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47750717</guid></item><item><title><![CDATA[New comment by taylorhou in "Several Mac mini and Mac Studio configs are now out of stock at Apple"]]></title><description><![CDATA[
<p>We combo frontier coding models with last frontier for more admin stuff.</p>
]]></description><pubDate>Sun, 12 Apr 2026 22:50:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=47745356</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=47745356</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47745356</guid></item><item><title><![CDATA[New comment by taylorhou in "Several Mac mini and Mac Studio configs are now out of stock at Apple"]]></title><description><![CDATA[
<p>Annecdotal here but I've been buying studios with as much ram as possible. I probably got in one of the last orders for the m3 ultra 512gb ram officially from apple late February right before they took that config down. I know this because the very next day I went back to order more and couldn't.<p>There's a ton of spam on ebay currently with configs of the 512gb ram variant going for sub $4k. Those are all scams/spam. You can tell instantly because the seller's account is/was created this month 2026.<p>The $20k+ listings aren't selling but what's interesting is technically, $20k for 512gb ram is still less than $40/gb ram on a single device. Compare that to a Nvidia spark with 128gb ram selling at $4k+ or if you're lucky, you snagged one for ~$3500. That's $31.25/gb<p>The retail on the 512gb was $9500 and after tax if you don't ship to a sales tax free state makes it only $20/gb ram.<p>I'm curious what price apple's newest studio variant will be priced at per gb ram but I recently bought a "refurb" from a reputable seller I've purchased many laptops from in the past at $13,750 which still put the 512gb at a $27/gb ram price point.<p>I'm bullish on local models especially due to gemma4. Good luck out there!</p>
]]></description><pubDate>Sat, 11 Apr 2026 22:37:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=47734619</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=47734619</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47734619</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: What Are You Working On? (March 2026)"]]></title><description><![CDATA[
<p>A bunch of ideas that have had domains but never enough engineers. Now there isn't enough time it seems except when I've hit my LLM subscription limits and they need to cool down.<p>Already launched biz-in-a-box.org and a life-in-a-box.org spinoff as frameworks to replace every entity's QuickBooks. I'm using them myself for every project my agents are spinning up.<p>Stealth project is related to classpass but for another category of need that won't go away even in the age of AI that really is only possible with critical mass of supply to meet existing demand. Super excited cus there's no better time to build with unlimited agents that scale without people problems.<p>Lastly, can't wait to run local LLMs so no longer limited by tokens/money.</p>
]]></description><pubDate>Mon, 09 Mar 2026 04:17:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=47304836</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=47304836</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47304836</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: What Are You Working On? (December 2025)"]]></title><description><![CDATA[
<p>Deploying robots within the next 6 months, not some 6+ years from now.. if anyone is interested in joinin9, DM.</p>
]]></description><pubDate>Mon, 15 Dec 2025 06:10:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=46270999</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=46270999</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46270999</guid></item><item><title><![CDATA[New comment by taylorhou in "Is America's jobs market nearing a cliff?"]]></title><description><![CDATA[
<p>996 - it's a global marketplace of talent and very few in America are willing to work 996. If you are, you either are the founder type or young and unshackled.</p>
]]></description><pubDate>Mon, 01 Dec 2025 11:10:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=46106028</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=46106028</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46106028</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: Who is hiring? (November 2025)"]]></title><description><![CDATA[
<p>Fyxed.com | Fintech Engineers | Fully Remote<p>Fyxed deploys capital to landlords and their property managers for cash needs like repairs, turnovers, and improvements on rental properties.<p>We're at an inflection point where we have the deployment capital ($50M) to really take a big swing at this opportunity.<p>The perfect candidate has prior fintech (lending) experience, loves wrangling the complexities of reconciling/tracking money, and bonus points if you've built on top of Column.com's API.<p>Email me directly: taylor at fyxed dot com</p>
]]></description><pubDate>Mon, 03 Nov 2025 18:45:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=45802697</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=45802697</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45802697</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: Who is hiring? (September 2025)"]]></title><description><![CDATA[
<p>Houmanoids | <a href="https://houmanoids.com" rel="nofollow">https://houmanoids.com</a> | Lots of roles<p>We're integrating robots in and around buildings (focused in America for now). Example usecases:<p>- security guards, doormen, receptionists (low mobility using humanoids)<p>- security patrol around perimeter of buildings (high mobility using dogs)<p>- moving things/packages within a building (last step package delivery)<p>We are building on top of Unitree's hardware and there's lots to do.<p>- abstraction layer similar to Swift for iOS and Kotlin for Android<p>- appstore<p>- example reference apps (think mail, calculator, camera, maps)<p>996 work environment due to collaboration/communication requirements with folks in China and insatiable demand for deploying robots.<p>deploy at houmanoids.com goes to myself (founder) and cto/cofounder</p>
]]></description><pubDate>Tue, 02 Sep 2025 16:24:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=45105207</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=45105207</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45105207</guid></item><item><title><![CDATA[New comment by taylorhou in "Show HN: An API for human-powered browser tasks"]]></title><description><![CDATA[
<p>to add to the post, our humans are all over the world but we have folks in the States, Latam, Philippines, and other SE Asian countries. they primarily do accounting/bookkeeping tasks today (for us at APM Help which I'm the founder).</p>
]]></description><pubDate>Mon, 21 Jul 2025 17:04:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=44637576</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=44637576</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44637576</guid></item><item><title><![CDATA[New comment by taylorhou in "Ask HN: Who is hiring? (February 2025)"]]></title><description><![CDATA[
<p>APM Help | Remote | Full-time Software Engineers (engineering manager, security manager, go developer, banking & fintech engineers<p>We help manage 1 million+ residential rentals across America (primarily accounting, banking & finance) and we're likely to double our engineering team from 8 > 16 this year (broader co is 250+ FTE).<p>We have 3 engineering & product teams (US, Brazil, South Asia) and likely will prioritize folks in those regions.<p>Specific call outs for:
- product folks with banking/fintech experience
- financial/data analyst on hundreds of thousands of financials
- crypto/dao lawyer<p>Please don't just send a resume but help us get excited about you by sending an email to "hn at apmhelp.com" the group is monitored by founder (me), CIO & CTO</p>
]]></description><pubDate>Mon, 03 Feb 2025 17:19:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=42920490</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=42920490</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42920490</guid></item><item><title><![CDATA[New comment by taylorhou in "Horrible Airbnb Experience (With Receipts)"]]></title><description><![CDATA[
<p>sigh. hate having to do this but i've exhausted all other efforts. really hope airbnb can get back to its glory days as my experience is nothing short of embarrassing.</p>
]]></description><pubDate>Mon, 11 Nov 2024 18:14:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=42109205</link><dc:creator>taylorhou</dc:creator><comments>https://news.ycombinator.com/item?id=42109205</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42109205</guid></item></channel></rss>