<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: 3Sophons</title><link>https://news.ycombinator.com/user?id=3Sophons</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 01 Sep 2026 07:49:13 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=3Sophons" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by 3Sophons in "Qwen 3.8 27B is excellent, but it defaults to overthinking things"]]></title><description><![CDATA[
<p>powerful enough models are becoming a reality on personal hardware. Olares is an open-source personal AI cloud OS, supporting local AI, Ollama, Open WebUI, Hermes etc</p>
]]></description><pubDate>Mon, 17 Aug 2026 12:32:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=49329811</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=49329811</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49329811</guid></item><item><title><![CDATA[New comment by 3Sophons in "How do I permanently disable random Google Photos popup to backup photos? (2024)"]]></title><description><![CDATA[
<p>Here is a open source OS to allow you to self host immich with 1 click <a href="https://www.olares.com/docs/use-cases/immich" rel="nofollow">https://www.olares.com/docs/use-cases/immich</a></p>
]]></description><pubDate>Mon, 17 Aug 2026 12:06:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=49329542</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=49329542</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49329542</guid></item><item><title><![CDATA[New comment by 3Sophons in "Ask HN: Is there any AI OS that has local AI and all are integrated to the OS?"]]></title><description><![CDATA[
<p>Not so sure about the shadowban policy so i won't paste the github link but you can search Olares. It is an AI OS for one click AI deployment and more.<p>try if out if you wanted AI woven into the OS itself: the model runs on your machine, your files and chats never leave your walls, and the agent layer talks to the system natively instead of through a browser tab.</p>
]]></description><pubDate>Fri, 24 Jul 2026 12:56:59 +0000</pubDate><link>https://news.ycombinator.com/item?id=49034905</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=49034905</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49034905</guid></item><item><title><![CDATA[New comment by 3Sophons in "Ask HN: AI Agent and harness containerization/security recommendations"]]></title><description><![CDATA[
<p>isolated container per run, no mounted secrets, vault at runtime, egress off. Open Source Olares (biased) runs agents as isolated workloads on your own hardware.</p>
]]></description><pubDate>Wed, 15 Jul 2026 05:19:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=48916564</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=48916564</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48916564</guid></item><item><title><![CDATA[New comment by 3Sophons in "Can I run AI locally?"]]></title><description><![CDATA[
<p>a lighter-weight alternative of docker and python is the Rust+Wasm stack <a href="https://github.com/LlamaEdge/LlamaEdge" rel="nofollow">https://github.com/LlamaEdge/LlamaEdge</a></p>
]]></description><pubDate>Sat, 14 Mar 2026 00:23:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=47371883</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=47371883</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47371883</guid></item><item><title><![CDATA[New comment by 3Sophons in "Running a full voice stack (ASR –> LLM –> TTS) locally with Docker"]]></title><description><![CDATA[
<p>covers how to wire up the components to run a real-time voice agent that you can actually talk to via an ESP32 client, keeping the heavy lifting on your local machine.</p>
]]></description><pubDate>Thu, 18 Dec 2025 14:25:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=46313045</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=46313045</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46313045</guid></item><item><title><![CDATA[Running a full voice stack (ASR –> LLM –> TTS) locally with Docker]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.docker.com/blog/develop-deploy-voice-ai-apps/">https://www.docker.com/blog/develop-deploy-voice-ai-apps/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=46313044">https://news.ycombinator.com/item?id=46313044</a></p>
<p>Points: 2</p>
<p># Comments: 1</p>
]]></description><pubDate>Thu, 18 Dec 2025 14:25:30 +0000</pubDate><link>https://www.docker.com/blog/develop-deploy-voice-ai-apps/</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=46313044</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46313044</guid></item><item><title><![CDATA[New comment by 3Sophons in "So – – – Is AI a Bubble?"]]></title><description><![CDATA[
<p>huh</p>
]]></description><pubDate>Tue, 09 Dec 2025 12:59:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=46204470</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=46204470</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46204470</guid></item><item><title><![CDATA[So – – – Is AI a Bubble?]]></title><description><![CDATA[
<p>Article URL: <a href="https://blog.andrewyang.com/p/so-is-ai-a-bubble">https://blog.andrewyang.com/p/so-is-ai-a-bubble</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=46204464">https://news.ycombinator.com/item?id=46204464</a></p>
<p>Points: 1</p>
<p># Comments: 1</p>
]]></description><pubDate>Tue, 09 Dec 2025 12:58:09 +0000</pubDate><link>https://blog.andrewyang.com/p/so-is-ai-a-bubble</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=46204464</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46204464</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>This post summarizes WasmEdge's recent talks at OSSummit Korea and KubeCon NA 2025, focusing on its new role as a premier AI runtime. Key technical highlights include expanded multi-GPU support (NVIDIA, AMD, Apple), multi-backend inference (TensorRT, OpenVINO), and the use of Embedded Rust/Wasm for building real-time voice AI agents (EchoKit project).</p>
]]></description><pubDate>Fri, 28 Nov 2025 12:03:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=46077910</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=46077910</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46077910</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>customize both hardware and software. Their voice cloning lets you recreate your own voice for truly personal AI experiences.</p>
]]></description><pubDate>Thu, 20 Nov 2025 17:50:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=45995439</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45995439</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45995439</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>Voice is the next frontier of conversational AI. It is the most natural modality for people to chat and interact with another intelligent being.</p>
]]></description><pubDate>Mon, 27 Oct 2025 12:04:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=45720019</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45720019</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45720019</guid></item><item><title><![CDATA[New comment by 3Sophons in "LLMs can get "brain rot""]]></title><description><![CDATA[
<p>why they don't prompt ai to avoid using dashes. and bullet points etc</p>
]]></description><pubDate>Fri, 24 Oct 2025 04:37:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=45690821</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45690821</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45690821</guid></item><item><title><![CDATA[Show HN: EchoKit – An open-source, ESP32-based AI voice agent with a Rust server]]></title><description><![CDATA[
<p>Have you wanted a voice AI agent that you could fully control, customize, and understand from the ground up—not just a black box.<p>EchoKit is a DIY, open-source voice agent running on an ESP32-S3. The fun part is the server backend, which I wrote entirely in Rust to handle the AI pipeline (ASR, LLM, TTS).<p>The stack is:<p>Hardware: EchoKit board (ESP32-S3)<p>Firmware: ESP-IDF<p>Server: Rust (Actix Web/Tungstenite)<p>AI: Customizable pipeline. The tutorial uses Groq for the Whisper, Llama 3, and TTS models, which makes the response time incredibly fast (usually just a few seconds for the full ASR->LLM->TTS roundtrip).<p>It's designed to be easy for makers, students, or anyone curious about AI to build in just a few minutes. You can modify the system prompts, swap out models, or even add custom actions (Step 6 in the guide).<p>The tutorial (linked) walks through assembly, flashing, and setting up the server. The server code is on GitHub (also linked).<p>Happy to answer any questions. Would love to hear your thoughts and what you think we could build with this!<p>Server Repo:<a href="https://github.com/second-state/echokit_server" rel="nofollow">https://github.com/second-state/echokit_server</a></p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=45644113">https://news.ycombinator.com/item?id=45644113</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 20 Oct 2025 14:09:45 +0000</pubDate><link>https://www.instructables.com/Create-Your-Own-AI-Voice-Agent-Using-EchoKit-ESP32/</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45644113</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45644113</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>The talk explores the idea that while human-centric languages like Python are popular for their ease of use, they are suboptimal for AI-driven development. Rust, with its strict compiler and strong type system, provides an excellent reward function and a tight feedback loop for LLMs, forcing them to generate correct, efficient code. This could make Rust the de facto language in a future where most code is written by AI. The video demonstrates this with a voice AI agent built with Echokit and Rust. the orignal tak is there <a href="https://www.youtube.com/watch?v=bbq0b_FpYEY" rel="nofollow">https://www.youtube.com/watch?v=bbq0b_FpYEY</a></p>
]]></description><pubDate>Fri, 17 Oct 2025 11:02:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=45615282</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45615282</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45615282</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>made the anime with Veo 3</p>
]]></description><pubDate>Thu, 16 Oct 2025 08:59:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=45603035</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45603035</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45603035</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>We Built an Open-Source Voice AI Kit (Hardware + AI Server + Docs) for Education & Tinkering</p>
]]></description><pubDate>Mon, 08 Sep 2025 14:42:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=45168934</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=45168934</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45168934</guid></item><item><title><![CDATA[Show HN: Videolangua – end-to-end video translate and subtitle/ dub]]></title><description><![CDATA[
<p>What it is
A practical pipeline to go from video → transcripts → bilingual subtitles (A/B at once) or voice dub (choose female or male). It targets lectures, tutorials, tech talks etc<p>Current constraints (honest version)<p>No speaker diarization yet → one TTS voice for the entire video<p>Voice choice: female / male<p>Bilingual subtitles (e.g., EN+JP, EN+ZH, EN+KR) with length-aware line breaking<p>You can add customized terms and translations in front of the listening model and the translation model to get the best accuracy<p>Long videos: we re-anchor timestamps periodically to fight drift<p>Pipeline (condensed)<p>ASR with word-level timestamps<p>Segment cleanup + merge tiny fragments<p>MT twice → produce A/B subtitle tracks; constrain length to reduce overflow; have multiple language model to supervise/double check the translation quality<p>TTS → single voice (female/male) for the full track<p>Mixback → keep ambience, duck original, SRT (mono or dual-lang) + dubbed MP4<p>Why post this now
It’s not “magic studio” quality, but it’s dependable for many real-world cases: course videos, onboarding, webinars. We found being explicit about limits (no diarization) actually speeds teams up.<p>What we’d love feedback on<p>The current User experience<p>Acceptable subtitle overflow rate on 30–60 min content<p>TTS pacing rules that feel most natural for multilingual reading speed<p>(If there’s interest, we’ll extract a minimal CLI with the exact steps and corner cases called out.)</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=44893553">https://news.ycombinator.com/item?id=44893553</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Wed, 13 Aug 2025 20:38:20 +0000</pubDate><link>https://videolangua.com/</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=44893553</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44893553</guid></item><item><title><![CDATA[Show HN: EchoKit –Fully Open Source AI Voice Agent Hardware and Server]]></title><description><![CDATA[
<p>It is powered by Rust
most “open” AI devices are actually closed on the backend, limiting customization and privacy.
EchoKit is fully open—from ESP32-S3 hardware to the Rust-based server. You can customize ASR, TTS (including voice cloning), LLM prompts, and even add your own MCP actions.
We’d love your feedback, ideas, and contributions. Try it out, check the code, and let us know what you think!</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=44596016">https://news.ycombinator.com/item?id=44596016</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Thu, 17 Jul 2025 17:45:04 +0000</pubDate><link>https://echokit.dev/</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=44596016</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44596016</guid></item><item><title><![CDATA[New comment by 3Sophons in "[dead]"]]></title><description><![CDATA[
<p>Edge LLM Osmosis-Structure-0.6B for structured data</p>
]]></description><pubDate>Tue, 03 Jun 2025 15:37:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=44171183</link><dc:creator>3Sophons</dc:creator><comments>https://news.ycombinator.com/item?id=44171183</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44171183</guid></item></channel></rss>