<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: karimf</title><link>https://news.ycombinator.com/user?id=karimf</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Sun, 11 Oct 2026 07:39:02 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=karimf" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by karimf in "Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost"]]></title><description><![CDATA[
<p>This is awesome. Thanks for pushing the audio pareto frontier forward.<p>Probably far fetched for now, but I think the next big evolution is building the pareto/much cheaper alternative to GPT-Live-1.<p>The STT/TTS market is quite saturated, while today, there's almost no cheap/open source alternative to GPT-Live-1.</p>
]]></description><pubDate>Tue, 15 Sep 2026 02:50:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49707098</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49707098</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49707098</guid></item><item><title><![CDATA[New comment by karimf in "DeepSeek v4.1 Flash"]]></title><description><![CDATA[
<p>While this is very impressive benchmark-wise, GPT-6 Astra showed us that benchmarks don't always correlate 1:1 to intelligence of a model.<p>When Astra launched, I think Artifical Analysis showed that it was on par with GPT-5.6 Sol and lower than Opus or something like that? Then, they updated the scoring.<p>I hope that more open source models, including this model, to be "as good to use" as Astra.</p>
]]></description><pubDate>Thu, 10 Sep 2026 08:36:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=49640373</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49640373</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49640373</guid></item><item><title><![CDATA[K2 Horizon: A connected fleet of six open models]]></title><description><![CDATA[
<p>Article URL: <a href="https://ifm.ai/blog/k2/">https://ifm.ai/blog/k2/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49551760">https://news.ycombinator.com/item?id=49551760</a></p>
<p>Points: 335</p>
<p># Comments: 131</p>
]]></description><pubDate>Thu, 03 Sep 2026 15:36:43 +0000</pubDate><link>https://ifm.ai/blog/k2/</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49551760</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49551760</guid></item><item><title><![CDATA[New comment by karimf in "Show HN: Free Inference Engineer and Model Training Roadmap"]]></title><description><![CDATA[
<p>I think a curriculum like this is neat and might help with interviews since you go wide and have a checklist of things that you need to learn.<p>I'm on a totally different path for learning inference engineering. I self-host a voice AI app that has ~2000 monthly active users on my own GPU box.<p>This forces me to learn about production serving, KV cache, quantization, inference engine, observability and economics, prefill optimization since I'm optimizing for TTFT instead of decode speed, and many more.<p>It's fun since every optimization you do directly translate to a better user experience or allow you to serve more users using the same hardware.</p>
]]></description><pubDate>Mon, 24 Aug 2026 17:03:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=49422749</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49422749</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49422749</guid></item><item><title><![CDATA[New comment by karimf in "GLM-5.3 Artificial Analysis Benchmarks"]]></title><description><![CDATA[
<p>Yes. Please seriously try other models. See relevant thread here: <a href="https://news.ycombinator.com/item?id=49296740">https://news.ycombinator.com/item?id=49296740</a></p>
]]></description><pubDate>Wed, 19 Aug 2026 05:12:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=49357120</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49357120</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49357120</guid></item><item><title><![CDATA[New comment by karimf in "Why does Opus 5 feel worse to work with?"]]></title><description><![CDATA[
<p>This 100%. I was Anthropic-pilled. I had a $200/mo subscription and I only used Anthropic models. I was frustrated by the verbose output and the writing style. I tried ASD-STE-100, it helped a bit, but it's still too verbose for my taste.<p>Then I tried GPT 5.6 Sol. It's night and day.<p>I think Anthropic just RL too hard on coding capabilities and never calibrated or benchmarked the writing styles.</p>
]]></description><pubDate>Fri, 14 Aug 2026 10:42:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=49296949</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49296949</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49296949</guid></item><item><title><![CDATA[New comment by karimf in "llama.cpp"]]></title><description><![CDATA[
<p>Not sure why it's on the front page now, but I highly recommend using llama.cpp for running AI model locally vs using other inference framework, unless you have a very specific requirement.<p>ggerganov and the team have done a stellar job maintaining the quality while still being fast to implement new models/improvements.</p>
]]></description><pubDate>Wed, 12 Aug 2026 06:22:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=49268500</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49268500</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49268500</guid></item><item><title><![CDATA[We built a realtime system for responsive voice AI in six months]]></title><description><![CDATA[
<p>Article URL: <a href="https://openai.com/index/continuous-voice-interaction-with-gpt-live/">https://openai.com/index/continuous-voice-interaction-with-gpt-live/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49256018">https://news.ycombinator.com/item?id=49256018</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Tue, 11 Aug 2026 10:36:51 +0000</pubDate><link>https://openai.com/index/continuous-voice-interaction-with-gpt-live/</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49256018</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49256018</guid></item><item><title><![CDATA[New comment by karimf in "Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows"]]></title><description><![CDATA[
<p>Yes, and also waiting for the next iteration of Gemma. Muse or Qwen are optimized for coding, while IMO Gemma is still better for non-coding tasks.<p><a href="https://x.com/osanseviero/status/2086107547535122767" rel="nofollow">https://x.com/osanseviero/status/2086107547535122767</a></p>
]]></description><pubDate>Mon, 10 Aug 2026 11:32:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=49242278</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49242278</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49242278</guid></item><item><title><![CDATA[New comment by karimf in "Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows"]]></title><description><![CDATA[
<p>Practically ~20GB with KV cache<p>> We quantize weights to ~4-bit, bringing the LM under 20 GB. We validated minimal to no degradation on agentic tasks under compression.<p><a href="https://www.reddit.com/r/LocalLLaMA/comments/1vkgsum/introducing_muse_glimmer_an_openweight_model/" rel="nofollow">https://www.reddit.com/r/LocalLLaMA/comments/1vkgsum/introdu...</a></p>
]]></description><pubDate>Mon, 10 Aug 2026 11:29:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=49242260</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49242260</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49242260</guid></item><item><title><![CDATA[Nvidia NemotronLabs VoiceChat 11B – real-time full duplex with tool calling]]></title><description><![CDATA[
<p>Article URL: <a href="https://huggingface.co/nvidia/NVIDIA-NemotronLabs-VoiceChat-11B">https://huggingface.co/nvidia/NVIDIA-NemotronLabs-VoiceChat-11B</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49163542">https://news.ycombinator.com/item?id=49163542</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Tue, 04 Aug 2026 01:50:05 +0000</pubDate><link>https://huggingface.co/nvidia/NVIDIA-NemotronLabs-VoiceChat-11B</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49163542</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49163542</guid></item><item><title><![CDATA[New comment by karimf in "Turn And Face The Strange"]]></title><description><![CDATA[
<p>Most people are going under identity crisis right now because of recent LLM advancements. This post is a good example that shows that it's not only happening at the individual level, but also on the company/organization level.<p>Is it still worth building products or companies that can be one-shotted by AI? Probably not.<p>One interesting consequence is that this force everyone to be more ambitious and do something bigger that's impossible before.<p>I hope more people are working on something that can always bring net positive to humanity even if there are hundreds of people working on the same thing, like clean energy.</p>
]]></description><pubDate>Sat, 25 Jul 2026 22:45:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49052511</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49052511</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49052511</guid></item><item><title><![CDATA[Turn and Face the Strange]]></title><description><![CDATA[
<p>Article URL: <a href="https://fly.io/blog/kurt-scott-money-sprites/">https://fly.io/blog/kurt-scott-money-sprites/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49043871">https://news.ycombinator.com/item?id=49043871</a></p>
<p>Points: 4</p>
<p># Comments: 0</p>
]]></description><pubDate>Sat, 25 Jul 2026 02:14:50 +0000</pubDate><link>https://fly.io/blog/kurt-scott-money-sprites/</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=49043871</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49043871</guid></item><item><title><![CDATA[New comment by karimf in "Local, CPU-Friendly, High-Quality TTS (Text-to-Speech) with Kokoro"]]></title><description><![CDATA[
<p>This repo is a good starting point for comparing TTS models <a href="https://github.com/5uck1ess/tts-bench" rel="nofollow">https://github.com/5uck1ess/tts-bench</a><p>Kokoro is a really good model, considered it’s released 1.5 years ago. It’s punching above its weight <a href="https://5uck1ess.github.io/tts-bench/scores.html" rel="nofollow">https://5uck1ess.github.io/tts-bench/scores.html</a></p>
]]></description><pubDate>Tue, 07 Jul 2026 23:21:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=48825367</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=48825367</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48825367</guid></item><item><title><![CDATA[Why WebRTC beats WebSockets for realtime voice AI]]></title><description><![CDATA[
<p>Article URL: <a href="https://livekit.com/blog/why-webrtc-beats-websockets-for-voice-ai-agents">https://livekit.com/blog/why-webrtc-beats-websockets-for-voice-ai-agents</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48616386">https://news.ycombinator.com/item?id=48616386</a></p>
<p>Points: 5</p>
<p># Comments: 0</p>
]]></description><pubDate>Sun, 21 Jun 2026 07:05:20 +0000</pubDate><link>https://livekit.com/blog/why-webrtc-beats-websockets-for-voice-ai-agents</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=48616386</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48616386</guid></item><item><title><![CDATA[New comment by karimf in "1-Click GitHub Token Stealing via a VSCode Bug"]]></title><description><![CDATA[
<p>I've been using Zed for a few weeks now and these two are also my main complaints as well.</p>
]]></description><pubDate>Wed, 03 Jun 2026 12:07:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=48382915</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=48382915</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48382915</guid></item><item><title><![CDATA[Unsloth Joins PyTorch Ecosystem]]></title><description><![CDATA[
<p>Article URL: <a href="https://unsloth.ai/blog/pytorch">https://unsloth.ai/blog/pytorch</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48095906">https://news.ycombinator.com/item?id=48095906</a></p>
<p>Points: 8</p>
<p># Comments: 2</p>
]]></description><pubDate>Mon, 11 May 2026 14:59:53 +0000</pubDate><link>https://unsloth.ai/blog/pytorch</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=48095906</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48095906</guid></item><item><title><![CDATA[Denial of Service Vulnerability in React Server Components]]></title><description><![CDATA[
<p>Article URL: <a href="https://github.com/facebook/react/security/advisories/GHSA-rv78-f8rc-xrxh">https://github.com/facebook/react/security/advisories/GHSA-rv78-f8rc-xrxh</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48058448">https://news.ycombinator.com/item?id=48058448</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Fri, 08 May 2026 04:09:58 +0000</pubDate><link>https://github.com/facebook/react/security/advisories/GHSA-rv78-f8rc-xrxh</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=48058448</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48058448</guid></item><item><title><![CDATA[New comment by karimf in "Cloudflare Email Service"]]></title><description><![CDATA[
<p>Oh yeah for sure. At that point, using SES is probably a better option compared to running a VPS just for SMTP. I posted that to let them know that SMTP support is a requirement for some developers.</p>
]]></description><pubDate>Thu, 16 Apr 2026 19:17:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=47798167</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=47798167</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47798167</guid></item><item><title><![CDATA[New comment by karimf in "Cloudflare Email Service"]]></title><description><![CDATA[
<p>Ok I just tried the service since I want to migrate from Resend.<p>Seems like you can only send email via the worker or REST API for now?<p>Can I send via SMTP? I'm using Supabase and it needs the SMTP credentials.<p>I can't find anything on the dashboard or on the docs, even though last year they said it supports SMTP [0]<p>[0] <a href="https://blog.cloudflare.com/email-service/" rel="nofollow">https://blog.cloudflare.com/email-service/</a></p>
]]></description><pubDate>Thu, 16 Apr 2026 16:01:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=47795371</link><dc:creator>karimf</dc:creator><comments>https://news.ycombinator.com/item?id=47795371</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47795371</guid></item></channel></rss>