<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: TechExpert2910</title><link>https://news.ycombinator.com/user?id=TechExpert2910</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 06 Oct 2026 00:09:16 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=TechExpert2910" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by TechExpert2910 in "Why does Opus 5 feel worse to work with?"]]></title><description><![CDATA[
<p>i think it's because they've fine-tuned an old opus 4.6 base model more and more with code<p>to the point where it's really great at coding, but lost its knowledge on how to write well!<p>if instead opus 5 was a full fresh training run, i don't think it'd be this trash at writing.<p>really great coding perf, but now the weights are adapted to so much code fine tuning that it writes like a weird engineer that's trying to sound smart.</p>
]]></description><pubDate>Sun, 23 Aug 2026 02:54:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49405755</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=49405755</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49405755</guid></item><item><title><![CDATA[New comment by TechExpert2910 in "Show HN: Sentient OS – On-device intelligence layer for your entire digital life"]]></title><description><![CDATA[
<p>Hey! Really great question. My plan is to charge a ~$2 a month subscription if you'd like to analyze more than the last 6 months’ worth of your data (I'll also allow a one-time lifetime license option!).<p>And I can let you use it on this 6-month window for free because it costs me nothing per user! Your own device does all the processing :D</p>
]]></description><pubDate>Sat, 02 May 2026 18:57:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=47989310</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=47989310</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47989310</guid></item><item><title><![CDATA[Show HN: Sentient OS – On-device intelligence layer for your entire digital life]]></title><description><![CDATA[
<p>Hi HN :D I'm 20 and I spent a year building something that shouldn't be possible: a custom on-device vision LLM that processes your entire digital life overnight on a phone.<p>We all have thousands of buried screenshots, notes, files, bookmarks, saved posts, etc we'll never find again. The only way to make AI understand all of it is to upload everything to the cloud -- privacy nightmare, and way too expensive at scale. And it shouldn't be possible on-device either: small models are too dumb, and phones are too slow for thousands of LLM inference runs.<p>So I spent a year deeply optimizing <i>every</i> layer of the on-device inference stack to make it possible anyway.<p>Sentient OS runs a custom multimodal vision LLM on your phone and laptop while they charge overnight. It understands your entire digital life -- every screenshot, note, file, email, bookmark, plus integrations for external services -- with nothing ever leaving your device.<p>This gives you three things that weren't possible before:<p>-> <i>Talk to your entire digital life in natural language:</i> "what was that wine I liked?" / "who did I wanna meet next week?" [on-device RAG]. And with MCP, your existing LLM (ChatGPT, Claude, etc.) can talk to your digital life too -- so it actually understands you.<p>-> <i>Proactive reminders surfaced from your own data:</i> "that tax return in your Downloads is due next week" / "tickets for that concert you screenshotted open tomorrow"<p>-> <i>Knowledge graphs of your entire digital life:</i> tap any node to find what you buried!<p>Here's what I had to build to make this possible:<p><i>Inference speed:</i><p>- KV cache reuse: the system prompt + few-shot examples are identical across all 3,000 analysis calls. I run inference on that prefix once, cache the KV state, and reuse it for every image. Prefill drops to just processing the image itself.<p>- Thermal-aware scheduling: I throttle the moment iOS reports thermal state > fair. I have all night, so I trade speed for not cooking the device.<p>- iOS jetsam awareness: iOS kills apps above a specific RAM threshold. I profiled that threshold across different iPhones and push right up to the edge.<p><i>Model quality at small size:</i><p>- Vision transplant: a 2B Qwen model has terrible vision. I transplanted Qwen 3.5 9B's multimodal projector onto the 2B base. Same architecture family makes this possible.<p>- Selective quantization on MLX: MLX doesn't support k-quant style mixed precision. I built it manually: less quantization on first/last layers and high-activation layers, more aggressive on the rest.<p>The alpha processes ~3,000 screenshots entirely on-device on a 6 year old iPhone. Coming to Mac and iPhone!<p>Previously I researched Apple's neural accelerators: <a href="https://www.reddit.com/r/LocalLLaMA/comments/1ohrn20/" rel="nofollow">https://www.reddit.com/r/LocalLLaMA/comments/1ohrn20/</a><p>And I love OSS! I built <a href="https://github.com/theJayTea/WritingTools" rel="nofollow">https://github.com/theJayTea/WritingTools</a> (2K+ stars, ~30 press features). I'm considering making Sentient OS OSS under AGPL (so no one else can profit off of my work haha).<p>I think this is one of the coolest consumer usecases to take advantage of on-device LLMs. I'd love to hear what you all think, and happy to answer any questions (I love geeking out about the deep work that's gone into optimizing models and inference!) :D</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=47989158">https://news.ycombinator.com/item?id=47989158</a></p>
<p>Points: 3</p>
<p># Comments: 2</p>
]]></description><pubDate>Sat, 02 May 2026 18:40:54 +0000</pubDate><link>https://sentient-os.ai</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=47989158</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47989158</guid></item><item><title><![CDATA[New comment by TechExpert2910 in "Show HN: Mercury – No-code orchestration for human and agent teams"]]></title><description><![CDATA[
<p>looks interesting... where do you draw the line between human in the loop and agents going all in?</p>
]]></description><pubDate>Mon, 13 Apr 2026 22:19:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=47758673</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=47758673</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47758673</guid></item><item><title><![CDATA[Hacked iPhone running iPadOS + a Mac-like experience on an external monitor]]></title><description><![CDATA[
<p>Article URL: <a href="https://old.reddit.com/r/iphone/comments/1p3e2bf/my_hacked_iphone_running_ipados_and_running_a/">https://old.reddit.com/r/iphone/comments/1p3e2bf/my_hacked_iphone_running_ipados_and_running_a/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=46011367">https://news.ycombinator.com/item?id=46011367</a></p>
<p>Points: 7</p>
<p># Comments: 2</p>
]]></description><pubDate>Sat, 22 Nov 2025 01:57:54 +0000</pubDate><link>https://old.reddit.com/r/iphone/comments/1p3e2bf/my_hacked_iphone_running_ipados_and_running_a/</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=46011367</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46011367</guid></item><item><title><![CDATA[Show HN: Open-Source Apple Intelligence Writing Tools for Windows and Linux]]></title><description><![CDATA[
<p>I posted this previously, but there's been a major update.<p>As a high school student, I love LLMs and the execution of Apple's Appel Intelligence Writing Tools, but was disappointed that nothing like that existed for Windows. I thus created Writing Tools, a better than Apple Intelligence open-source alternative that works system-wide on Windows.
It's much better than the tiny 2B parameter Apple Intelligence model!<p>Features (works system-wide):
1. Select text in a textbox, press the hotkey, and optimise your text in a click (proofread, rewrite, tone changes, custom instructions...). You can also select code to fix/refactor it with the custom instructions.
2. Press the hotkey without selecting any text to bring up a chat UI with the LLM.
3. Select all text on a webpage, press the hotkey, and use the Summarise button to get summary of the webpage/text!<p>It can use the free Gemini API, or a multitude of local LLMs via Ollama, llama.cpp, KoboldCPP, TabbyAPI, vLLM, etc.<p>It's completely free & open source, and is heavily privacy-focused (local LLM support, your API key stays on-device, no logging, no tracking, etc.)<p>GitHub (with a demo video): <a href="https://github.com/theJayTea/WritingTools">https://github.com/theJayTea/WritingTools</a><p>Writing Tools has also been featured on XDA, Beebom, Windows Central, Neowin, and more! :D<p>P.S. This is my first major Windows coding project, so I'm especially keen on advice for best practices and potential improvements! Thank you so much for your time :)</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=42193478">https://news.ycombinator.com/item?id=42193478</a></p>
<p>Points: 6</p>
<p># Comments: 0</p>
]]></description><pubDate>Wed, 20 Nov 2024 12:50:32 +0000</pubDate><link>https://github.com/theJayTea/WritingTools/releases/tag/v5</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=42193478</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42193478</guid></item><item><title><![CDATA[Show HN: I built open-source Apple Intelligence-like Writing Tools for Windows]]></title><description><![CDATA[
<p>I posted this a while back, but there's been a major update — it now supports local LLMs, multiple cloud LLMs, code editing, chat mode, themes, dark mode, and more!<p>As a high school student, I love LLMs and the execution of Apple's Appel Intelligence Writing Tools, but was disappointed that nothing like that existed for Windows.
So, I created Writing Tools, a better than Apple Intelligence open-source alternative that works system-wide on Windows!
It can use the free Gemini API, or a multitude of local LLMs via Ollama, llama.cpp, KoboldCPP, TabbyAPI, vLLM, etc.
It's much better than the tiny 2B parameter Apple Intelligence model!
It works in any application with a customizable hotkey - Proofreads, rewrites, summarizes, and more - Free and privacy-focused (your API key stays local, no logging, no tracking, local model options, etc.)<p>It's built with Python and PySide6, and I've made it easy to install with a pre-compiled exe. For the technically inclined, the well-documented source is available to run or modify.<p>I'd love feedback from the HN community on both the concept and the implementation. Are there features you'd like to see? Any thoughts on making it more robust or user-friendly?<p>GitHub: <a href="https://github.com/theJayTea/WritingTools">https://github.com/theJayTea/WritingTools</a><p>P.S. This is my first major Windows coding project, so I'm especially keen on advice for best practices and potential improvements! Thanks so much for your time :)</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=41896518">https://news.ycombinator.com/item?id=41896518</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Sun, 20 Oct 2024 16:32:02 +0000</pubDate><link>https://github.com/theJayTea/WritingTools/releases/tag/v3</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=41896518</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41896518</guid></item><item><title><![CDATA[Show HN: I built an open-source Apple Intelligence-like Writing Tool for Windows]]></title><description><![CDATA[
<p>As a high school student, I love LLMs and the execution of Apple's Appel Intelligence Writing Tools, but was disappointed that nothing like that existed for Windows. So, I created Writing Tools, a better than Apple Intelligence open-source alternative that works system-wide on Windows!<p>Key features:
- Uses Google's Gemini 1.5 Flash model for superior AI-assisted writing (much better than the tiny 2B parameter Apple Intelligence model)
- Works in any application with a customizable hotkey
- Proofreads, rewrites, summarizes, and more
- Free and privacy-focused (your API key stays local, no logging, no tracking, etc.)<p>It's built with Python and PySide6, and I've made it easy to install with a pre-compiled exe. For the technically inclined, the well-documented source is available to run or modify.<p>I'd love feedback from the HN community on both the concept and the implementation. Are there features you'd like to see? Any thoughts on making it more robust or user-friendly?<p>GitHub: <a href="https://github.com/theJayTea/WritingTools">https://github.com/theJayTea/WritingTools</a><p>P.S. This is my first major Windows coding project, so I'm especially keen on advice for best practices and potential improvements!
Thanks so much for your time :)</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=41864096">https://news.ycombinator.com/item?id=41864096</a></p>
<p>Points: 9</p>
<p># Comments: 1</p>
]]></description><pubDate>Wed, 16 Oct 2024 21:38:13 +0000</pubDate><link>https://github.com/theJayTea/WritingTools/blob/main/README.md</link><dc:creator>TechExpert2910</dc:creator><comments>https://news.ycombinator.com/item?id=41864096</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41864096</guid></item></channel></rss>