<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: __mharrison__</title><link>https://news.ycombinator.com/user?id=__mharrison__</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 22 Sep 2026 17:54:50 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=__mharrison__" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>I misspoke, you need to use .groupby/.transform to add a new filtering column:<p><pre><code>    (sales
      .assign(country_median=lambda df_: (
          df_.groupby("country")["amount"].transform("median")
      ))
      .query("amount <= country_median * 10")
      .assign(net=pd.col('amount') - pd.col('discount'))
      .groupby("country", as_index=False)
      .agg(total=("net", "sum"))
    )</code></pre></p>
]]></description><pubDate>Sat, 12 Sep 2026 14:34:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=49672734</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49672734</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49672734</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Would love to see examples.</p>
]]></description><pubDate>Sat, 12 Sep 2026 13:58:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=49672382</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49672382</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49672382</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>I'm confused, you can use filter after a groupby in pandas...<p>It's late here, I'm going to bed, perhaps I'll write the code tomorrow when I'm at my laptop and not on my phone.</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:47:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669265</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669265</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669265</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Those are generally rules of thumbs for data pipelines. (Plus duckdb can read Excel these days I think).</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:38:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669221</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669221</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669221</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>This is a common complaint I get all the time (heard it this week while teaching pandas). I compare this to whitespace indentation in Python.<p>Lots of folks complain about it before using it. After they use it it is a non issue.<p>If it really is an issue (and it generally isn't a cause of vectorization removal when used correctly) and you can't get over the syntactic noise if the lambda, pandas 3 introduced pd.col (that work in most (I filed a big about some exceptions) places when you'd use lambda).</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:37:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669209</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669209</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669209</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>I looked at your code very quickly, but it looks like you need to use .filter after a .groupby...</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:32:45 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669177</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669177</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669177</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Did you ever report these issues to polars developers?<p>I reported many bugs to both pandas and polars over the years and both teams tend to address relatively big ones. (Still have some outstanding pandas bugs that I think are a big deal but the devs disagree.)</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:25:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669131</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669131</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669131</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Most of the time it works, but there are still corner cases throughout the python ML ecosystem where polars fails.<p>This and the decision not to natively plot with matplotlib keep me in pandas for most tasks. (Plus there's is still a relatively large demand for pandas training.)</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:22:23 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669108</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669108</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669108</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Rewriting pandas to polars is relatively trivial for most tasks these days. Especially if you wrote your pandas code correctly.<p>I still prefer (and use) Pandas for EDA. I think matplotlib integration is a better choice for most viz.<p>Also, I'm probably in the top 3-5 worldwide for number of folks I've trained with pandas. I offer Polars training and there is little demand for it.</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:19:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669090</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669090</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669090</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Use pandas if you want advanced analytics, visualization, or ml.<p>Use SQL if you need to move data around.</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:15:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669065</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669065</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669065</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>If you learn to write pandas correctly you end up writing it very similar to polars (or tidyverse).<p>I agree that the API has warts, though typically it is more concise than polars.</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:14:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669056</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669056</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669056</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>Most folks just need to learn how to use pandas well and that will open enough doors. Then they can move to polars or duck if needed.</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:12:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669040</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669040</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669040</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pandas Should Go Extinct"]]></title><description><![CDATA[
<p>I'm in the middle of wrapping up the edits for Effective Pandas 3rd Edition. (I also wrote a Polars book and just wrapped up a weeklong training session on pandas this week.)<p>Pandas is not perfect, it has a bunch of warts. But it is good enough for most. (And many of those folks are using Excel or tableau or power bi... These were the types I was training this week).<p>If you have medium data, migrating from pyarrow backed pandas to duck or Polars is trivial.</p>
]]></description><pubDate>Sat, 12 Sep 2026 05:09:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=49669028</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49669028</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49669028</guid></item><item><title><![CDATA[New comment by __mharrison__ in "DaVinci Resolve 21.1"]]></title><description><![CDATA[
<p>I spent a few tokens last year trying to create a Python lib for native (undocumented) Resolve files. I was making progress, but it was slow (I ended up scripting Resolve in different ways). I imagine that current models could reverse-engineer the native file format relatively quickly.</p>
]]></description><pubDate>Tue, 08 Sep 2026 20:24:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49616509</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49616509</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49616509</guid></item><item><title><![CDATA[New comment by __mharrison__ in "DaVinci Resolve 21.1"]]></title><description><![CDATA[
<p>Highly recommend the studio version. The ability to edit by transcript is worth the price alone.</p>
]]></description><pubDate>Tue, 08 Sep 2026 20:17:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=49616402</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49616402</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49616402</guid></item><item><title><![CDATA[New comment by __mharrison__ in "DaVinci Resolve 21.1"]]></title><description><![CDATA[
<p>I thought DaVinci Resolve could automatically sync camera footage for a few years now. I did it about 4 years ago.</p>
]]></description><pubDate>Tue, 08 Sep 2026 20:15:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49616373</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49616373</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49616373</guid></item><item><title><![CDATA[New comment by __mharrison__ in "DaVinci Resolve 21.1"]]></title><description><![CDATA[
<p>I had AI write a program to edit with resolve. It does transcription, looks at the transcript and decides where to make cuts. If does a decent job and saves me a bunch of time.</p>
]]></description><pubDate>Tue, 08 Sep 2026 14:31:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=49610876</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49610876</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49610876</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Ask HN: Resources to get good at soldering?"]]></title><description><![CDATA[
<p>Build a keyboard. By the end you should be well aware with soldering (and how to remove parts that you put on the wrong side of the PCB).</p>
]]></description><pubDate>Sat, 05 Sep 2026 14:42:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=49577003</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49577003</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49577003</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Show HN: Open-Source eInk Bike Computer"]]></title><description><![CDATA[
<p>Very cool. I was messing around vibe-coding an HR monitor on a LilyGo T5.<p>My thought was I'd love to set an HR zone and then connect it to my e-MTB and have it change power to keep me in a zone.<p>The refresh rate on my device was horrible, and the models (this was a few months ago) struggled to create a historical line plot of my HR.</p>
]]></description><pubDate>Fri, 04 Sep 2026 18:41:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=49568497</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49568497</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49568497</guid></item><item><title><![CDATA[New comment by __mharrison__ in "Pre-Release of Polars 2.0"]]></title><description><![CDATA[
<p>This is the way. Favor keyword arguments to alias.</p>
]]></description><pubDate>Thu, 03 Sep 2026 13:38:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=49549793</link><dc:creator>__mharrison__</dc:creator><comments>https://news.ycombinator.com/item?id=49549793</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49549793</guid></item></channel></rss>