<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: trueforma</title><link>https://news.ycombinator.com/user?id=trueforma</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 18 Sep 2026 07:49:54 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=trueforma" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by trueforma in "Show HN: Voice bots with 500ms response times"]]></title><description><![CDATA[
<p>I too am excited about voice inferencing. I wrote my own Websocket Faster whisper implementation before OpenAI's gpt4o release . They steamrolled my interview coach concept <a href="https://intervu.trueforma.ai" rel="nofollow">https://intervu.trueforma.ai</a> and <a href="https://sales.trueforma.ai" rel="nofollow">https://sales.trueforma.ai</a> - sales pitch coach implementations. I defaulted to Push to talk implementation as I couldn't get VAD to work reliably. I run it all on a panda Latte :) Was looking to implement Groq's hosted whisper. I love the idea of having Llama3 uncensored on Groq as the LLM as I'm tired of the boring corporate conversations. I hope to reduce my latency and learn from your examples - Kudos to your efforts. I wish I could try the demo - seems to be over subscribed as I can't get in to talk to the bot. I'm sure my latte Panda would melt if just 3 people try to inference at the same time :)</p>
]]></description><pubDate>Thu, 27 Jun 2024 21:21:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=40815267</link><dc:creator>trueforma</dc:creator><comments>https://news.ycombinator.com/item?id=40815267</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40815267</guid></item></channel></rss>