<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: sethkim</title><link>https://news.ycombinator.com/user?id=sethkim</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 09 Oct 2026 02:39:00 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=sethkim" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[A classifier is not a classifier is not a classifier]]></title><description><![CDATA[
<p>Article URL: <a href="https://sutro.sh/blog/a-classifier-is-not-a-classifier-is-not-a-classifier">https://sutro.sh/blog/a-classifier-is-not-a-classifier-is-not-a-classifier</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49869159">https://news.ycombinator.com/item?id=49869159</a></p>
<p>Points: 3</p>
<p># Comments: 0</p>
]]></description><pubDate>Sun, 27 Sep 2026 18:04:15 +0000</pubDate><link>https://sutro.sh/blog/a-classifier-is-not-a-classifier-is-not-a-classifier</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=49869159</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49869159</guid></item><item><title><![CDATA[New comment by sethkim in "LLM Ass Bench"]]></title><description><![CDATA[
<p>Folks, we've reached the top.</p>
]]></description><pubDate>Tue, 22 Sep 2026 22:09:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=49808894</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=49808894</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49808894</guid></item><item><title><![CDATA[Show HN: Jev-align, a CLI to calibrate Jev to your judgement]]></title><description><![CDATA[
<p>Article URL: <a href="https://github.com/sutro-sh/jev-align">https://github.com/sutro-sh/jev-align</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49770872">https://news.ycombinator.com/item?id=49770872</a></p>
<p>Points: 3</p>
<p># Comments: 0</p>
]]></description><pubDate>Sat, 19 Sep 2026 23:05:55 +0000</pubDate><link>https://github.com/sutro-sh/jev-align</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=49770872</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49770872</guid></item><item><title><![CDATA[New comment by sethkim in "The Analytical AI Handbook"]]></title><description><![CDATA[
<p>I appreciate the classic HN sarcasm!</p>
]]></description><pubDate>Fri, 28 Aug 2026 20:22:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=49483748</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=49483748</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49483748</guid></item><item><title><![CDATA[The Analytical AI Handbook]]></title><description><![CDATA[
<p>Article URL: <a href="https://handbook.sutro.sh">https://handbook.sutro.sh</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49482925">https://news.ycombinator.com/item?id=49482925</a></p>
<p>Points: 49</p>
<p># Comments: 2</p>
]]></description><pubDate>Fri, 28 Aug 2026 19:01:47 +0000</pubDate><link>https://handbook.sutro.sh</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=49482925</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49482925</guid></item><item><title><![CDATA[Useful Black Boxes]]></title><description><![CDATA[
<p>Article URL: <a href="https://sethkim.me/l/thesolutionspace/?postnum=12">https://sethkim.me/l/thesolutionspace/?postnum=12</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=47753761">https://news.ycombinator.com/item?id=47753761</a></p>
<p>Points: 3</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 13 Apr 2026 15:48:26 +0000</pubDate><link>https://sethkim.me/l/thesolutionspace/?postnum=12</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=47753761</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47753761</guid></item><item><title><![CDATA[New comment by sethkim in "If DSPy is so great, why isn't anyone using it?"]]></title><description><![CDATA[
<p>This is extremely true. In fact, from what we see many/most of the problems to be solved with LLMs do not have ground-truth values; even hand-labeled data tends to be mostly subjective.</p>
]]></description><pubDate>Mon, 23 Mar 2026 16:44:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=47491914</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=47491914</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47491914</guid></item><item><title><![CDATA[New comment by sethkim in "If DSPy is so great, why isn't anyone using it?"]]></title><description><![CDATA[
<p>Feel free to shoot me a note at seth@sutro.sh if you want to check it out!</p>
]]></description><pubDate>Mon, 23 Mar 2026 16:15:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=47491524</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=47491524</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47491524</guid></item><item><title><![CDATA[New comment by sethkim in "If DSPy is so great, why isn't anyone using it?"]]></title><description><![CDATA[
<p>We build a product that's somewhat similar in spirit to DSPy, but people come to us for different reasons than the OP listed here.<p>1) It's slow: you first have to get acquainted with DSPY and then get hand-labeled data for prompt optimization. This can be a slow process so it's important to just label cases that are ambiguous, not obvious.<p>2) They know that manual prompt engineering is brittle, and want a prompt that's optimized and robust against a model they're invoking, which DSPy offers. However, it's really the optimizer (ex. GEPA) doing the heavy-lifting.<p>3) They don't actually want a model or prompt at all. They want a task completed, reliably, and they want that task to not regress in performance. Ideally, the task keeps improving in production.<p>Curious if folks in this thread feel more of these pains than the ones in the article.</p>
]]></description><pubDate>Mon, 23 Mar 2026 16:13:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=47491489</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=47491489</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47491489</guid></item><item><title><![CDATA[New comment by sethkim in "My trick for getting consistent classification from LLMs"]]></title><description><![CDATA[
<p>Under-discussed superpower of LLMs is open-set labeling, which I sort of consider to be inverse classification. Instead of using a static set of pre-determined labels, you're using the LLM to find the semantic clusters within a corpus of unstructured data. It feels like "data mining" in the truest sense.</p>
]]></description><pubDate>Mon, 20 Oct 2025 22:04:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=45650018</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=45650018</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45650018</guid></item><item><title><![CDATA[New comment by sethkim in "The Future (and Present) of AI Is Synthetic Data"]]></title><description><![CDATA[
<p>The models you called out at the beginning were all released this year. What do you think is the difference between this generation of models and previous ones?</p>
]]></description><pubDate>Fri, 26 Sep 2025 19:12:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=45389964</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=45389964</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45389964</guid></item><item><title><![CDATA[New comment by sethkim in "The End of Moore's Law for AI? Gemini Flash Offers a Warning"]]></title><description><![CDATA[
<p>Yes! Both Llama 3 and Gemma 3 have 128k context windows.</p>
]]></description><pubDate>Thu, 03 Jul 2025 18:53:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=44458100</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44458100</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44458100</guid></item><item><title><![CDATA[New comment by sethkim in "The End of Moore's Law for AI? Gemini Flash Offers a Warning"]]></title><description><![CDATA[
<p>Yes, we're a startup! And LLM inference is a major component of what we do - more importantly, we're working on making these models accessible as analytical processing tools, so we have a strong focus on making them cost-effective at scale.</p>
]]></description><pubDate>Thu, 03 Jul 2025 18:50:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=44458078</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44458078</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44458078</guid></item><item><title><![CDATA[New comment by sethkim in "The End of Moore's Law for AI? Gemini Flash Offers a Warning"]]></title><description><![CDATA[
<p>My two cents here is the classic answer - it depends. If you need general "reasoning" capabilities, I see this being a strong possibility. If you need specific, factual information baked into the weights themselves, you'll need something large enough to store that data.<p>I think the best of both worlds is a sufficiently capable reasoning model with access to external tools and data that can perform CPU-based lookups for information that it doesn't possess.</p>
]]></description><pubDate>Thu, 03 Jul 2025 18:47:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=44458059</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44458059</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44458059</guid></item><item><title><![CDATA[New comment by sethkim in "The End of Moore's Law for AI? Gemini Flash Offers a Warning"]]></title><description><![CDATA[
<p>No doubt prices will continue to drop! We just don't think it will be anything like the orders-of-magnitude YoY improvements we're used to seeing. Consequently, developers shouldn't expect the cost of building and scaling AI applications to be anything close to "free" in the near future as many suspect.</p>
]]></description><pubDate>Thu, 03 Jul 2025 18:40:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=44457994</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44457994</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44457994</guid></item><item><title><![CDATA[New comment by sethkim in "The End of Moore's Law for AI? Gemini Flash Offers a Warning"]]></title><description><![CDATA[
<p>Both great points, but more or less speak to the same root cause - customer usage patterns are becoming more of a driver for pricing than underlying technology improvements. If so, we likely have hit a "soft" floor for now on pricing. Do you not see it this way?</p>
]]></description><pubDate>Thu, 03 Jul 2025 18:29:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=44457905</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44457905</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44457905</guid></item><item><title><![CDATA[The End of Moore's Law for AI? Gemini Flash Offers a Warning]]></title><description><![CDATA[
<p>Article URL: <a href="https://sutro.sh/blog/the-end-of-moore-s-law-for-ai-gemini-flash-offers-a-warning">https://sutro.sh/blog/the-end-of-moore-s-law-for-ai-gemini-flash-offers-a-warning</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=44457371">https://news.ycombinator.com/item?id=44457371</a></p>
<p>Points: 113</p>
<p># Comments: 75</p>
]]></description><pubDate>Thu, 03 Jul 2025 17:34:05 +0000</pubDate><link>https://sutro.sh/blog/the-end-of-moore-s-law-for-ai-gemini-flash-offers-a-warning</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44457371</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44457371</guid></item><item><title><![CDATA[New comment by sethkim in "Making 2.5 Flash and 2.5 Pro GA, and introducing Gemini 2.5 Flash-Lite"]]></title><description><![CDATA[
<p>I run a batch inference/LLM data processing service and we do a lot of work around cost and performance profiling of (open-weight) models.<p>One odd disconnect that still exists in LLM pricing is the fact that providers charge linearly with respect to token consumption, but costs are actually quadratic with an increase in sequence length.<p>At this point, since a lot of models have converged around the same model architecture, inference algorithms, and hardware - the chosen costs are likely due to a historical, statistical analysis of the shape of customer requests. In other words, I'm not surprised to see costs increase as providers gather more data about real-world user consumption patterns.</p>
]]></description><pubDate>Tue, 17 Jun 2025 19:29:38 +0000</pubDate><link>https://news.ycombinator.com/item?id=44302932</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44302932</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44302932</guid></item><item><title><![CDATA[New comment by sethkim in "Ask HN: Who is hiring? (June 2025)"]]></title><description><![CDATA[
<p>Sutro.sh (fka Skysight) | Infrastructure/LLMs & Research Engineering | SF Bay Area | Full-time<p>We are building batch inference infrastructure and a great/user developer experience around it. We believe LLMs have not yet been meaningfully unlocked as data processing tools - we're changing that.<p>Our work involves interesting distributed systems and LLM research problems, newly-imagined user experiences, and a meaningful focus on mission and values.<p>Open Roles:<p>Infrastructure/LLM Engineer — <a href="https://jobs.skysight.inc/Member-of-Technical-Staff-Infrastructure-LLMs-1a32de87d04a80d583fdfabdb4fe9dba" rel="nofollow">https://jobs.skysight.inc/Member-of-Technical-Staff-Infrastr...</a><p>Research Engineer - <a href="https://jobs.skysight.inc/Member-of-Technical-Staff-Research-Engineering-1d62de87d04a80759341c70cce869fec" rel="nofollow">https://jobs.skysight.inc/Member-of-Technical-Staff-Research...</a><p>If you're interested in applying, please send an email to jobs@sutro.sh with a resume/LinkedIn Profile. For extra priority, please include [HN] in the subject line.</p>
]]></description><pubDate>Tue, 03 Jun 2025 00:00:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=44164721</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=44164721</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44164721</guid></item><item><title><![CDATA[New comment by sethkim in "Ask HN: Who is hiring? (May 2025)"]]></title><description><![CDATA[
<p>Skysight | Infrastructure/LLMs & Research Engineering | SF Bay Area | Full-time<p>We are building large-scale batch inference infrastructure and a great/user developer experience around it. We believe LLMs have not yet been meaningfully unlocked as data processing tools - we're changing that.<p>Our work involves interesting distributed systems and LLM research problems, newly-imagined user experiences, and a meaningful focus on mission and values.<p>Open Roles:<p>Infrastructure/LLM Engineer — <a href="https://jobs.skysight.inc/Member-of-Technical-Staff-Infrastructure-LLMs-1a32de87d04a80d583fdfabdb4fe9dba" rel="nofollow">https://jobs.skysight.inc/Member-of-Technical-Staff-Infrastr...</a><p>Research Engineer - <a href="https://jobs.skysight.inc/Member-of-Technical-Staff-Research-Engineering-1d62de87d04a80759341c70cce869fec" rel="nofollow">https://jobs.skysight.inc/Member-of-Technical-Staff-Research...</a><p>If you're interested in applying, please send an email to jobs@skysight.inc with a resume/LinkedIn Profile. For extra priority, please include [HN] in the subject line.</p>
]]></description><pubDate>Thu, 01 May 2025 22:42:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=43864168</link><dc:creator>sethkim</dc:creator><comments>https://news.ycombinator.com/item?id=43864168</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43864168</guid></item></channel></rss>