<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: pranshuchittora</title><link>https://news.ycombinator.com/user?id=pranshuchittora</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 24 Jul 2026 02:39:42 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=pranshuchittora" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[Claude Opus 5]]></title><description><![CDATA[
<p>Article URL: <a href="https://artificialanalysis.ai/models/claude-opus-5">https://artificialanalysis.ai/models/claude-opus-5</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49025676">https://news.ycombinator.com/item?id=49025676</a></p>
<p>Points: 4</p>
<p># Comments: 1</p>
]]></description><pubDate>Thu, 23 Jul 2026 18:01:03 +0000</pubDate><link>https://artificialanalysis.ai/models/claude-opus-5</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=49025676</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49025676</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Ask HN: What Are You Working On? (July 2026)"]]></title><description><![CDATA[
<p>I am working on building an self-improving QA agent for software teams. Free and open-source.
It is an agentic testing harness with batteries, includes the test runner infra (web & mobile), memory (vector store, local embedding model powered by transformers.js, self improving loop, issue reporting.
More details - <a href="https://vostride.ai" rel="nofollow">https://vostride.ai</a> | 
Code - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Mon, 13 Jul 2026 11:23:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=48891005</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48891005</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48891005</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Ask HN: How are you using AI agents for testings features?"]]></title><description><![CDATA[
<p>GitHub - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Tue, 30 Jun 2026 17:24:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=48736053</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48736053</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48736053</guid></item><item><title><![CDATA[Ask HN: How are you using AI agents for testings features?]]></title><description><![CDATA[
<p>Hey, I am working on building agent-qa https://github.com/vostride/agent-qa which is a self-improving QA agent for software teams.<p>Processed over 500M+ tokens and thousands of test runs, would love to understand your thoughts on this and how you are approaching testing today.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48736050">https://news.ycombinator.com/item?id=48736050</a></p>
<p>Points: 1</p>
<p># Comments: 1</p>
]]></description><pubDate>Tue, 30 Jun 2026 17:24:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=48736050</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48736050</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48736050</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Ask HN: Is Codex with GPT 5.5 Extra High being dumbed down?"]]></title><description><![CDATA[
<p>Yes, I feel so. I started happening from june first weekish.
I have shifted to claude code for planning things and codex for execution (as it is faster, though dumb)</p>
]]></description><pubDate>Tue, 30 Jun 2026 17:22:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=48736023</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48736023</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48736023</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Ask HN: How do you handle QA at a startup with no QA team? Genuinely curious"]]></title><description><![CDATA[
<p>Yes - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Mon, 29 Jun 2026 20:59:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=48725153</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48725153</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48725153</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Ask HN: How do you handle QA at a startup with no QA team? Genuinely curious"]]></title><description><![CDATA[
<p>Use <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a> you don't need a QA team just ask coding harness to write tests and you can review the test runs.</p>
]]></description><pubDate>Mon, 29 Jun 2026 19:14:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=48723787</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48723787</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48723787</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Yes a lot. So with the acceleration of AI in the software engineering. Features are being shipped faster but causes regressions. The only way to verify is either you write tests with AI and spend hours reviewing them or you do manual QA. agent-qa aims to solve the later. First your product should work for the end user, later you can write clean test etc.</p>
]]></description><pubDate>Fri, 19 Jun 2026 12:25:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=48597786</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48597786</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48597786</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Try the OSS alternative - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Fri, 19 Jun 2026 07:50:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=48595984</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48595984</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48595984</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Try the OSS alternative - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Fri, 19 Jun 2026 07:50:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=48595983</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48595983</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48595983</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Try the OSS alternative - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Fri, 19 Jun 2026 07:50:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=48595977</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48595977</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48595977</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Try the OSS alternative - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Fri, 19 Jun 2026 07:49:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=48595970</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48595970</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48595970</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>I did. Checkout the OSS alternative - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Fri, 19 Jun 2026 07:46:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=48595945</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48595945</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48595945</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Try the OSS alternative - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a></p>
]]></description><pubDate>Fri, 19 Jun 2026 07:46:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=48595944</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48595944</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48595944</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Some digging 
FAST_MODEL        = "google/gemini-3-flash"   (fast mode primary)
DEEP_MODEL        = "openai/gpt-5.4"           (deep mode primary)
VISION_CLICK_MODEL= "openai/gpt-5.4"           (the visual grounder)<p>fast: gemini-3-flash, falls back to gpt-5.4, 15-min run timeout, max 2 visual calls/step.
deep: gpt-5.4, 15-min timeout, max 3 visual calls/step.<p>Why such a hard timeout, and why not latest models?</p>
]]></description><pubDate>Thu, 18 Jun 2026 23:05:11 +0000</pubDate><link>https://news.ycombinator.com/item?id=48592802</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48592802</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48592802</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps"]]></title><description><![CDATA[
<p>Hey, I just gave it a try and ran a quick test on booking.com. It took ~3 mins for a basic test. Do you cache the test steps so that future runs are faster and they don't call LLMs for the subsequent runs?<p>Also your current pricing is $300 for 1K tests which means $0.3 for each test. We tried out playwright mcp and it easily consumes 1M+ tokens for a test with ~20 steps (including image input). So with this pricing are you guys default alive?<p>Also is there a benchmark which you ran to prove the efficacy of your testing agent? because in the current stage it is a trust me bro kinda thing.</p>
]]></description><pubDate>Thu, 18 Jun 2026 22:42:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=48592625</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48592625</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48592625</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Show HN: Id-agent – Token efficient UUID alternative for AI agents"]]></title><description><![CDATA[
<p>The word dictionary is curated with guardrails. Also the dictionary contains words which are 1 BPE token long. Under 5-6 characters</p>
]]></description><pubDate>Thu, 21 May 2026 14:24:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=48223170</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48223170</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48223170</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Show HN: Open-Source Agentic QA Harness with Memory"]]></title><description><![CDATA[
<p>Thanks for those kind words. The landing page's demo required lots of sculpting.
I would say that agent-qa is not only frontend focused. As you can run hooks in sandboxed env to test apis, so a better way to put it is with agent-qa you can test the product end-to-end not only UI.<p>But the issue with API testing / backend is that coding harnesses are really good at it. A product manager who writes user stories should be able to write tests for the product, and usually PMs don't care about the APIs.<p>Do give agent-qa a try, and consider giving it a star on GH <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a><p>Thanks!</p>
]]></description><pubDate>Wed, 20 May 2026 16:21:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=48210171</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48210171</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48210171</guid></item><item><title><![CDATA[Show HN: Open-Source Agentic QA Harness with Memory]]></title><description><![CDATA[
<p>GitHub - <a href="https://github.com/vostride/agent-qa" rel="nofollow">https://github.com/vostride/agent-qa</a>
Live Demos - <a href="https://vostride.com/demo/agent-qa" rel="nofollow">https://vostride.com/demo/agent-qa</a></p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48205901">https://news.ycombinator.com/item?id=48205901</a></p>
<p>Points: 18</p>
<p># Comments: 2</p>
]]></description><pubDate>Wed, 20 May 2026 11:10:47 +0000</pubDate><link>https://vostride.com/agent-qa</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48205901</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48205901</guid></item><item><title><![CDATA[New comment by pranshuchittora in "Show HN: Id-agent – Token efficient UUID alternative for AI agents"]]></title><description><![CDATA[
<p>It is being used in production at <a href="https://vostride.com/agent-qa" rel="nofollow">https://vostride.com/agent-qa</a>
The issues was agent-qa have many different kinds of files tests, memory etc and there's too much FK references which LLMs need to resolve. Using id-agent worked like a charm</p>
]]></description><pubDate>Tue, 19 May 2026 19:52:02 +0000</pubDate><link>https://news.ycombinator.com/item?id=48198537</link><dc:creator>pranshuchittora</dc:creator><comments>https://news.ycombinator.com/item?id=48198537</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48198537</guid></item></channel></rss>