<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: feliixh</title><link>https://news.ycombinator.com/user?id=feliixh</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 11 Aug 2026 16:33:18 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=feliixh" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by feliixh in "Ask HN: What are you working on? (August 2026)"]]></title><description><![CDATA[
<p>I just hacked together <a href="https://www.firstbranch.ai" rel="nofollow">https://www.firstbranch.ai</a> in a few days, conversational intelligence for congress hearings.<p>I was talking about healthcare price transparency with a friend. He pointed me to a hearing in congress and I was surprised by how much I learned from one of the witnesses' testimonies so I built a "serendipity engine" for congress hearings.</p>
]]></description><pubDate>Sun, 09 Aug 2026 20:46:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=49235677</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=49235677</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49235677</guid></item><item><title><![CDATA[New comment by feliixh in "Ask HN: What are you working on? (July 2026)"]]></title><description><![CDATA[
<p>I have been working on <a href="https://www.accessmrf.com" rel="nofollow">https://www.accessmrf.com</a> - a catalog of all American health care prices. Since there's way too much data, I am building a taxonomy and deduplicating it extensively first, with the idea that I can shrink it by ~4 orders of magnitude and make the problem tractable. I have written a bit about how I do it here: <a href="https://www.felixhaba.com/writing/simplifying-healthcare-price-transparency-files-with-minhashing/" rel="nofollow">https://www.felixhaba.com/writing/simplifying-healthcare-pri...</a><p>My goal this week has been to index all insurers that publish machine readable files. For now, my project helps researchers and founders in the space access the data via API, it is not consumer-grade yet. But if you are interested in this data I would love to brainstorm ideas or share my knowledge. Shoot me a message.</p>
]]></description><pubDate>Thu, 02 Jul 2026 16:56:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=48764214</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=48764214</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48764214</guid></item><item><title><![CDATA[LLMs Put Style over Substance, You Should Put Substance over Style]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.felixhaba.com/writing/llms-put-style-over-substance/">https://www.felixhaba.com/writing/llms-put-style-over-substance/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48587531">https://news.ycombinator.com/item?id=48587531</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Thu, 18 Jun 2026 16:06:21 +0000</pubDate><link>https://www.felixhaba.com/writing/llms-put-style-over-substance/</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=48587531</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48587531</guid></item><item><title><![CDATA[New comment by feliixh in "Ask HN: What are you working on? (June 2026)"]]></title><description><![CDATA[
<p>I am working on <a href="https://www.accessmrf.com" rel="nofollow">https://www.accessmrf.com</a> a catalog for all health care prices published by insurers under the Transparency in Coverage rule.<p>I recently wrote a blog post using Min Hashing to estimate that at least 90% of the 1.17 petabyte dataset are duplicates and I keep investigating new ways of making this dataset manageable. <a href="https://www.felixhaba.com/writing/simplifying-healthcare-price-transparency-files-with-minhashing/" rel="nofollow">https://www.felixhaba.com/writing/simplifying-healthcare-pri...</a></p>
]]></description><pubDate>Wed, 17 Jun 2026 14:46:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=48571286</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=48571286</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48571286</guid></item><item><title><![CDATA[New comment by feliixh in "Ask HN: What are you working on? (May 21)"]]></title><description><![CDATA[
<p>I've been building out <a href="https://www.accessmrf.com/sources" rel="nofollow">https://www.accessmrf.com/sources</a> to index Transparency in Coverage files across insurers in a single place. Most recently I have been looking at the overlap across files using Min Hashing and found that there is a lot of duplicate data. Currently, I'm figuring out a way of making this dataset more digestible by identifying and eliminating the 90+% duplicate and ghost data.</p>
]]></description><pubDate>Thu, 21 May 2026 13:34:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=48222341</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=48222341</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48222341</guid></item><item><title><![CDATA[New comment by feliixh in "I’m Not a Robot"]]></title><description><![CDATA[
<p>2,3,8,9</p>
]]></description><pubDate>Sun, 21 Sep 2025 00:49:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=45319025</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=45319025</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45319025</guid></item><item><title><![CDATA[New comment by feliixh in "Ask HN: What are you working on? (Aug 2025)"]]></title><description><![CDATA[
<p>I'm building a catalog for health care price transparency data that aggregates the rates published by all insurers, to put everything in one place and make it easier for developers / researchers to access this data. <a href="https://www.accessmrf.com/" rel="nofollow">https://www.accessmrf.com/</a></p>
]]></description><pubDate>Wed, 03 Sep 2025 13:49:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=45115786</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=45115786</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45115786</guid></item><item><title><![CDATA[New comment by feliixh in "Spotting base64 encoded JSON, certificates, and private keys"]]></title><description><![CDATA[
<p>Very useful!</p>
]]></description><pubDate>Wed, 06 Aug 2025 04:38:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=44807692</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=44807692</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44807692</guid></item><item><title><![CDATA[New comment by feliixh in "Ask HN: What Are You Working On? (June 2025)"]]></title><description><![CDATA[
<p>I'm building a catalog for health care price transparency data that aggregates the rates published by all insurers, to put everything in one place and make it easier for developers / researchers to access this data. <a href="https://www.accessmrf.com/" rel="nofollow">https://www.accessmrf.com/</a></p>
]]></description><pubDate>Mon, 30 Jun 2025 03:07:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=44418889</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=44418889</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44418889</guid></item><item><title><![CDATA[New comment by feliixh in "Mapping almost every law, regulation and case in Australia"]]></title><description><![CDATA[
<p>Great job, I intend to reproduce this on a similar dataset I've been collecting!<p>I will say, it would be great to see the color labeling done on domain url alone, to see how much of the topography of the map is driven simply by the different formatting characteristics of the websites you're gathering data from.</p>
]]></description><pubDate>Fri, 22 Mar 2024 18:33:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=39793408</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=39793408</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39793408</guid></item><item><title><![CDATA[New comment by feliixh in "CatalaLang/catala: Programming language for law specification"]]></title><description><![CDATA[
<p>Def a psyop. Classic Catalonian move of seizing independence by writing self determination laws in code and carefully introducing a bug that they can then exploit to secede.</p>
]]></description><pubDate>Sun, 17 Sep 2023 21:36:48 +0000</pubDate><link>https://news.ycombinator.com/item?id=37549593</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=37549593</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=37549593</guid></item><item><title><![CDATA[New comment by feliixh in "Building LLM Applications for Production"]]></title><description><![CDATA[
<p>Haha, fair point, what I really meant is that LLMs will translate natural language to code, so building will be mostly in English while debugging will still happen in code.</p>
]]></description><pubDate>Sat, 15 Apr 2023 13:04:49 +0000</pubDate><link>https://news.ycombinator.com/item?id=35580278</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=35580278</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=35580278</guid></item><item><title><![CDATA[New comment by feliixh in "Building LLM Applications for Production"]]></title><description><![CDATA[
<p>One thing I think will dominate in the future is to write software documentation geared towards the easy understanding of it by LLMs, with documentation possibly including a fine-tunning dataset with which a model can be tested for proficiency in using that particular tool (like OpenAI Evals). Software will be written to be used by humans through LLMs because humans will code in natural language, and not in the language of your interface.</p>
]]></description><pubDate>Fri, 14 Apr 2023 16:45:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=35571604</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=35571604</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=35571604</guid></item><item><title><![CDATA[New comment by feliixh in "Substack Notes Launched"]]></title><description><![CDATA[
<p>+1</p>
]]></description><pubDate>Tue, 11 Apr 2023 19:14:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=35529909</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=35529909</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=35529909</guid></item><item><title><![CDATA[New comment by feliixh in "AI is in danger of being swallowed up by copyright law"]]></title><description><![CDATA[
<p>I don't think that's so clear. When you train a deep learning model you are making it extract the gist or insight of many works and then use that pattern to produce new works. While the NN does not experience the work like a human it is definitely not memorizing.<p>A silly example. Making GPT write a rap battle between Keynes and Mises goes beyond a performative remix, it is transformational work, nothing is copied explicitly. If a human were to write it that would not violate copyright.<p>I think that to tackle this we need a new lens other than copyright in the long term.</p>
]]></description><pubDate>Sun, 22 Jan 2023 15:44:20 +0000</pubDate><link>https://news.ycombinator.com/item?id=34478427</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=34478427</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34478427</guid></item><item><title><![CDATA[New comment by feliixh in "Ask HN: Books that teach programming by building a series of small projects?"]]></title><description><![CDATA[
<p>Automate the boring stuff with Python</p>
]]></description><pubDate>Tue, 17 Jan 2023 15:53:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=34413996</link><dc:creator>feliixh</dc:creator><comments>https://news.ycombinator.com/item?id=34413996</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=34413996</guid></item></channel></rss>