<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: wtf242</title><link>https://news.ycombinator.com/user?id=wtf242</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 03 Sep 2026 12:17:04 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=wtf242" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by wtf242 in "Show HN: I scraped 3B Goodreads reviews to train a better recommendation model"]]></title><description><![CDATA[
<p>That's super cool. I launched a book recommendations feature this year, which works vastly different. I ask users what their favorite books are(which you can rank), and then allow them to import their goodreads data which includes star reviews, and books they have read, then I determine your favorite style of books based on genres and subjects, then use opensearch to find similar books. It's a lot more complicated than that, but seems to work well. I'm always looking for ideas on how to improve this feature. Interested in what you are actually doing on the backend on the how-it-works page. thanks!<p>here's my recommendations feature: <a href="https://thegreatestbooks.org/recommendations" rel="nofollow">https://thegreatestbooks.org/recommendations</a><p>It's much more powerful if you're a member. you can restrict the results to certain genres or book lengths, as we as published date ranges, etc. If someone wants to try out the more powerful feature, DM me and i'll mark your account as a member.</p>
]]></description><pubDate>Fri, 07 Nov 2025 19:41:14 +0000</pubDate><link>https://news.ycombinator.com/item?id=45850177</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=45850177</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45850177</guid></item><item><title><![CDATA[New comment by wtf242 in "The year of peak might and magic"]]></title><description><![CDATA[
<p>One of my favorite games of all time. It's so simple yet can be so complex. You can legit spend 8+ hours playing the first couple levels easily.</p>
]]></description><pubDate>Fri, 18 Jul 2025 21:18:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=44609945</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=44609945</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44609945</guid></item><item><title><![CDATA[New comment by wtf242 in "EverQuest"]]></title><description><![CDATA[
<p>I failed out of college my senior year because I discovered EQ. So many fond memories. I created a guild on a new server and it ended up one of the most famous and best guilds on any of the servers. The amount of planning and management it took to lead a large guild was just ridiculous. It was a full time job. I even created one of the first "loot" web apps in php 3 and mysql just to keep track of player participation and make loot more fair.<p>Most of the team who created World of Warcraft were members in the guild.<p>some of my fondest memories:<p>- getting pretty far in the Plane of Air, which was an incomplete end game zone with almost impossible to beat bosses.
- defeating the Avatar of War, which was not supposed to be killable. We figured out we could charm his guards by having a huge number of enchanters and use the guards to tank him. We managed to beat him and they patched/fixed the guards and made them uncharmable shortly afterwards<p>The death penalty in the end game zones was originally very tough to work around. You needed a key to reach the zone, but if you died the key required to get into the zone was on your corpse inside the zone. So if everyone wiped/died getting everyone's corpse back was a multi hour event.</p>
]]></description><pubDate>Sat, 05 Jul 2025 12:44:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=44472439</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=44472439</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44472439</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: What Are You Working On? (June 2025)"]]></title><description><![CDATA[
<p>I think it's most useful for discovering new reads, especially with the advanced search and recommendations functionality. I do agree i could do a better job of non spoiler summaries. good idea<p>- Series have always been a problem. Some book lists will include the entire series, and then some will have individual books. If the series is sold as a single book I'll often just include that. Like Lord of the Rings. Sometimes I will include only the first book in the series on a list, to prevent always adding every single book in a series when a list mentions "harry potter series".<p>basically I don't have a perfect way of handling series'<p>for the last point, kind of. If you add a book to the default "My Favorite Books" user list, it gets aggregated and used for this book list which is included in the rankings. <a href="https://thegreatestbooks.org/lists/463" rel="nofollow">https://thegreatestbooks.org/lists/463</a></p>
]]></description><pubDate>Mon, 30 Jun 2025 21:39:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=44428194</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=44428194</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44428194</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: What Are You Working On? (June 2025)"]]></title><description><![CDATA[
<p>Still working on my books site <a href="https://thegreatestbooks.org" rel="nofollow">https://thegreatestbooks.org</a> that I started in 2008. It's been a 1 man team the entire time. I recently made some major algorithm changes that I think greatly improves the rankings. My algorithm code is open source <a href="https://github.com/ssherman/weighted_list_rank">https://github.com/ssherman/weighted_list_rank</a><p>I do plan on open sourcing more of the code over time. I also have started working on other sites using the same algorithm implementation (music, movies, video games)<p>This has just been a side project over the year generating passive income. I get around 250,000 page views a day, and with ads, memberships, and affiliate links I make around $2,500~ a month.<p>Tech stack is ruby on rails 8, postgresql 17, opensearch, redis, bootstrap 5.3 hosting on 3 servers on linode.</p>
]]></description><pubDate>Mon, 30 Jun 2025 07:22:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=44420422</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=44420422</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44420422</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: What are you working on? (May 2025)"]]></title><description><![CDATA[
<p>recently launched book recommendations feature for my books side project that I put a LOT of work into. I might be biased but I think it works well as long as you give it your favorite books.<p><a href="https://thegreatestbooks.org/recommendations?demo=tgb2025" rel="nofollow">https://thegreatestbooks.org/recommendations?demo=tgb2025</a><p>warning: account required, and the full featured version where you can specify book length, include/exclude genres/subjects, etc requires a membership. if you would like to test it though just e-mail me at contact@thegreatestbooks.org and I'll mark your account as paid.</p>
]]></description><pubDate>Sun, 25 May 2025 22:55:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=44092032</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=44092032</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44092032</guid></item><item><title><![CDATA[New comment by wtf242 in "Side projects I've built since 2009"]]></title><description><![CDATA[
<p>that's awesome! I've had many many side projects launched in the past 2 decades, but the only one still going is my books site <a href="https://thegreatestbooks.org" rel="nofollow">https://thegreatestbooks.org</a><p>I created it 17~ years ago mostly as just a tool for myself and now it gets roughly 8 million views a month.<p>The hardest part of any side project is actually launching it and making it somewhat production ready. I always spend the vast majority of my time dealing with devops/deployment issues/tasks</p>
]]></description><pubDate>Mon, 19 May 2025 15:14:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=44030747</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=44030747</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44030747</guid></item><item><title><![CDATA[New comment by wtf242 in "RubyLLM: A delightful Ruby way to work with AI"]]></title><description><![CDATA[
<p>been using <a href="https://github.com/alexrudall/ruby-openai" rel="nofollow">https://github.com/alexrudall/ruby-openai</a> for years with no issues which is a fine gem and works great.</p>
]]></description><pubDate>Sat, 15 Mar 2025 04:31:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=43369998</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=43369998</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=43369998</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: Those making $500/month on side projects in 2024 – Show and tell"]]></title><description><![CDATA[
<p>yeah the amazon product API has some severe limitations. I have 25,000~ books on my site, and I just don't have enough API calls in a day to keep the prices 100% updated. It's on my todo list to revisit this, but it's low on my priority list. I don't make that much from amazon refs(couple hundred a month)<p>I will say that I don't think they really defend their TOS too much from my experience. I used to have a cookbooks site for years that used the same affiliate tag i use for my greatest books site, and never had any issues.</p>
]]></description><pubDate>Wed, 11 Dec 2024 14:59:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=42388380</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=42388380</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42388380</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: Those making $500/month on side projects in 2024 – Show and tell"]]></title><description><![CDATA[
<p>I built it in 2008, and have rewritten it twice now since them. I have spent quite a bit of time adding new lists. It's definitely a labor of love and I do spend quite a bit of time on it.<p>I do use Gen AI now to generate genres, descriptions, and to grab other data. Previously years ago i would just scrape it or manually set it.<p>Amazon has a nice product API with up to date prices.<p>I am working on bookshop.org integration. I used to also do barnes & noble. The problem is neither of them have APIs to programmatically search for books, so i have to do complicated scraping. example: <a href="https://github.com/ssherman/bookshop-search">https://github.com/ssherman/bookshop-search</a></p>
]]></description><pubDate>Wed, 11 Dec 2024 07:11:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=42385566</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=42385566</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42385566</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: Those making $500/month on side projects in 2024 – Show and tell"]]></title><description><![CDATA[
<p>The Greatest Books <a href="https://thegreatestbooks.org" rel="nofollow">https://thegreatestbooks.org</a><p>I created it in 2008 and have maintained and improved it over the years. I am trying to figure out how to monetize it more. I currently make around $2k a month. I just use adsense and have a paid membership feature through buymeacoffee. I get massive traffic and I'm pretty much the #1 result for anything related to best/greatest books.<p>It's built with Rails and Postgresql and hosted on 3 linode servers. I get around 250k page visits a day.</p>
]]></description><pubDate>Wed, 11 Dec 2024 01:15:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=42383656</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=42383656</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42383656</guid></item><item><title><![CDATA[New comment by wtf242 in "Nearly 90% of our AI crawler traffic is from ByteDance"]]></title><description><![CDATA[
<p>I had the same issue with TikTok/ByteDance. They were using almost 100gb of my traffic per month.<p>I now block all ai crawlers at the cloudflare WAF level. On Monday I noticed a HUGE spike in traffic and my site was not handling it well. After a lot of troubleshooting and log parsing, I was getting millions of requests from China that were getting past cloudflare's bot protection.<p>I ended up having to force a CF managed challenge for the entire country of China to get my site back in a normal working state.<p>In the past 24 hours CF has blocked 1.66M bot requests. Good luck running a site without using CloudFlare or something similar.<p>AI crawlers are just out of control</p>
]]></description><pubDate>Thu, 31 Oct 2024 18:31:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=42009935</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=42009935</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42009935</guid></item><item><title><![CDATA[New comment by wtf242 in "Show HN: I mapped HN's favorite books with GPT-4o"]]></title><description><![CDATA[
<p>This is awesome! Do you mind if I add this list to my books site? (<a href="https://thegreatestbooks.org" rel="nofollow">https://thegreatestbooks.org</a>) I'll give you full credit and link back to your site.</p>
]]></description><pubDate>Mon, 09 Sep 2024 21:59:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=41494513</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=41494513</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41494513</guid></item><item><title><![CDATA[New comment by wtf242 in "Streaming every NFL game this season requires 7 different services, costs $2,500"]]></title><description><![CDATA[
<p>This is a bad article. The highest cost is network television/cable which is free if you buy an hd antenna</p>
]]></description><pubDate>Fri, 06 Sep 2024 13:48:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=41466175</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=41466175</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41466175</guid></item><item><title><![CDATA[New comment by wtf242 in "Ask HN: What are you working on (August 2024)?"]]></title><description><![CDATA[
<p>still working on <a href="https://thegreatestbooks.org" rel="nofollow">https://thegreatestbooks.org</a><p>been my main side project since like 2008. working on goodreads import right now. Always working on improving the algorithm. would love to collab with a data scientist on ways to improve my algorithm. <a href="https://github.com/ssherman/weighted_list_rank">https://github.com/ssherman/weighted_list_rank</a></p>
]]></description><pubDate>Sun, 25 Aug 2024 15:40:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=41348088</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=41348088</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41348088</guid></item><item><title><![CDATA[New comment by wtf242 in "Google is the only search engine that works on Reddit now, thanks to AI deal"]]></title><description><![CDATA[
<p>This problem is only going to get worse. for my thegreatestbooks.org site i used to just get indexed/scraped by google and bing. now it's like 50+ AI bots scraping my entire site just so they can train a LLM to answer questions my site answers without having a user ever visit my site. I just checked cloudflare and in the past 24 hours I've had 1.2 million bot/automated requests</p>
]]></description><pubDate>Wed, 24 Jul 2024 16:40:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=41058861</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=41058861</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41058861</guid></item><item><title><![CDATA[New comment by wtf242 in "Perplexity AI is lying about their user agent"]]></title><description><![CDATA[
<p>The amount of AI bots scraping/indexing content is just mind boggling. for my books site <a href="https://thegreatestbooks.org" rel="nofollow">https://thegreatestbooks.org</a>, without blocking any bots, I was probably getting 500,000~ requests a day from ONLY ai bots. Claudebot, amazon ai bot, bing ai bot, bytespider, openai. Endless ai bots just non-stop indexing/scraping my data.<p>Before i moved my dns to cloudflare and got on their pro plan, which offers robust bot blocking, they were severely hurting my performance to the point that I bought a new server to offload the traffic.</p>
]]></description><pubDate>Sat, 15 Jun 2024 19:38:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=40692183</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=40692183</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40692183</guid></item><item><title><![CDATA[New comment by wtf242 in "React 19 Breaks Async Composability"]]></title><description><![CDATA[
<p>It's insane how toxic the js environment is. it seems like if a project is over 6 months old, nothing will work. When I yarn install on an old project, i'm rolling the dice. I had a 2 year old next.js side project i was working on and the amount of work to make it work the latest version with just updating the dependencies and reading the upgrade docs were infinitely more complex than just starting over from scratch.<p>no thanks. I will stick with Rails.</p>
]]></description><pubDate>Fri, 14 Jun 2024 06:48:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=40678167</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=40678167</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40678167</guid></item><item><title><![CDATA[Show HN: The Global Literary Canon]]></title><description><![CDATA[
<p>The biggest complaint with aggregated book lists, is they are way too western focused. I made a page with some very specific country and author limits at an attempt to correct this.<p>My site aggregates 305~ book lists and uses a weighted min max normalization algorithm. This has been my little side project for almost 15 years now.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=40639923">https://news.ycombinator.com/item?id=40639923</a></p>
<p>Points: 4</p>
<p># Comments: 1</p>
]]></description><pubDate>Mon, 10 Jun 2024 22:25:08 +0000</pubDate><link>https://thegreatestbooks.org/global-canon</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=40639923</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40639923</guid></item><item><title><![CDATA[New comment by wtf242 in "OpenAI working on web search product"]]></title><description><![CDATA[
<p>"Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; GPTBot/1.0; +<a href="https://openai.com/gptbot" rel="nofollow">https://openai.com/gptbot</a>)" "52.230.152.37"</p>
]]></description><pubDate>Thu, 02 May 2024 14:00:08 +0000</pubDate><link>https://news.ycombinator.com/item?id=40236422</link><dc:creator>wtf242</dc:creator><comments>https://news.ycombinator.com/item?id=40236422</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=40236422</guid></item></channel></rss>