<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: jrussbowman</title><link>https://news.ycombinator.com/user?id=jrussbowman</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Thu, 30 Jul 2026 00:13:31 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=jrussbowman" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by jrussbowman in "YaCy, a distributed Web Search Engine, based on a peer-to-peer network"]]></title><description><![CDATA[
<p>And for my immature moment of the day, the above comment was comment #69</p>
]]></description><pubDate>Wed, 06 Mar 2024 17:31:33 +0000</pubDate><link>https://news.ycombinator.com/item?id=39618406</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=39618406</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39618406</guid></item><item><title><![CDATA[New comment by jrussbowman in "YaCy, a distributed Web Search Engine, based on a peer-to-peer network"]]></title><description><![CDATA[
<p>Nice to see search projects are still popping up. After a move, family life taking over and me getting more interested in Unreal Engine, my poor search engine is now more of an experiment in seeing how well it runs while basically on life-support maintenance updates I do. Starting to think I honestly should just take it down and save my $50 a month I spend maintaining it.<p>But I'll post it in a hacker news comment and maybe you all will give it enough traffic I can get excited about it again, lol<p><a href="https://www.unscatter.com" rel="nofollow">https://www.unscatter.com</a></p>
]]></description><pubDate>Wed, 06 Mar 2024 17:28:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=39618365</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=39618365</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=39618365</guid></item><item><title><![CDATA[New comment by jrussbowman in "Build a search engine, not a vector DB"]]></title><description><![CDATA[
<p>I just used postgres to build my search engine and it also helps with the last 2 questions. Keeping the content context consistent helps with the first. Unscatter.com for example is content shared only in the last 30 days. Helps with keeping my operating costs under $50 a month too.<p>I wish I had time to mess with it more. Job and life has taken over. My first goal with AI would be to use it to for key word and phrase extraction and also analyzing all the links I pull in hourly to see if there is a larger story I could make visible.</p>
]]></description><pubDate>Wed, 20 Dec 2023 11:56:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=38707647</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=38707647</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=38707647</guid></item><item><title><![CDATA[New comment by jrussbowman in "Prompt Engineering Is Hard"]]></title><description><![CDATA[
<p>I've seen some people use () and other syntax on the stable diffusion subreddit. I've been trying to find a guide for the syntax but haven't had much luck with Google. Is there a resource for this?</p>
]]></description><pubDate>Tue, 04 Oct 2022 12:08:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=33079199</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=33079199</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=33079199</guid></item><item><title><![CDATA[New comment by jrussbowman in "Yep: Google alternative that shares revenue with creators – by Ahrefs"]]></title><description><![CDATA[
<p>If we're adding our own, <a href="https://www.unscatter.com" rel="nofollow">https://www.unscatter.com</a> is mine. It's focused on the news vertical rather than general search.<p>It's not monetized but it's also only costing me less than $50 a month in hosting and cloudflare costs to support.</p>
]]></description><pubDate>Sat, 04 Jun 2022 01:00:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=31615734</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=31615734</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=31615734</guid></item><item><title><![CDATA[New comment by jrussbowman in "Almost all searches on my independent search engine are now from SEO spam bots"]]></title><description><![CDATA[
<p>"Want to become rich? Make a search engine which indexes the fresh relevant data from the big siloed websites, and ignores the general dead Internet."<p>Did that to some degree. Unscatter.com pulls from reddit and twitter to source links.<p>I found reddit only created an echo chamber bubble of obvious bias and twitter only diluted it a little.</p>
]]></description><pubDate>Mon, 16 May 2022 12:49:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=31396340</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=31396340</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=31396340</guid></item><item><title><![CDATA[New comment by jrussbowman in "Almost all searches on my independent search engine are now from SEO spam bots"]]></title><description><![CDATA[
<p>I do this all the time</p>
]]></description><pubDate>Mon, 16 May 2022 11:15:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=31395592</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=31395592</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=31395592</guid></item><item><title><![CDATA[New comment by jrussbowman in "Almost all searches on my independent search engine are now from SEO spam bots"]]></title><description><![CDATA[
<p>It's been the same for unscatter.com for years but I've always attributed to that to me not having a real marketing strategy or even sticking with the ones I've tried to start.</p>
]]></description><pubDate>Mon, 16 May 2022 11:13:57 +0000</pubDate><link>https://news.ycombinator.com/item?id=31395581</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=31395581</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=31395581</guid></item><item><title><![CDATA[New comment by jrussbowman in "Reddit can't build a better search engine"]]></title><description><![CDATA[
<p>So if anyone is following this, right now it appears that not many Kanye stories are bubbling to the top. It's simply not in the top 100 of the top 100. For the full 30 day index, checking just on Kanye in story titles, this is all I have in  my index.<p>Kanye West - Gold Digger
 Kanye wants Billie Eilish to say sorry or he'll pull out of Coachella
 Kanye West: ‘Stop Asking Me to Do NFTs... Ask Me Later’
 Kanye West Does Not Want to Get Involved With NFTs
 Kanye West Rejects NFTs, Tells Fans To Stop Asking: 'I Make Music In The Real World'<p>So long story short, I think I may need to consider increasing my pull. Going from the top 100 of the top 100 to maybe the top 1000 of the top 100 or the top 100 of the top 1000. I'll have to do some research and also validate my crawl can support it.</p>
]]></description><pubDate>Thu, 17 Feb 2022 23:43:28 +0000</pubDate><link>https://news.ycombinator.com/item?id=30380312</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=30380312</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=30380312</guid></item><item><title><![CDATA[New comment by jrussbowman in "Reddit can't build a better search engine"]]></title><description><![CDATA[
<p>I get the top 100 posts from the top 100 popular subreddits for the past hour (as defined by reddit). I then do some basic filtering on subreddits to exclude a few that are often in the top that I find are either mostly text or media content, I'm looking for links.<p>The lack of Kanye West content is interesting. I'm going to try to find some time this weekend to dive into that. The only thing I can think of is it's so well known less people are sharing links and more people are just making text posts on reddit about it.<p>However, it could be an indexing issue too. I'm using postgres for the index so maybe there is something there. I'll research that. Thanks for noticing and calling it out!</p>
]]></description><pubDate>Thu, 17 Feb 2022 17:54:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=30376299</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=30376299</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=30376299</guid></item><item><title><![CDATA[New comment by jrussbowman in "Reddit can't build a better search engine"]]></title><description><![CDATA[
<p>Manipulated or not, I do agree it's an echo chamber. I added Twitter in attempt of breaking the echo chamber aspect but wasn't as successful as I hoped. I do occasionally look for other sources to add but finding similar sites that can match the volume of those two is difficult. Which makes it more difficult to figure out how to weigh other sites results to those two.<p>This discussion did make me poke around some more. I may consider using the free tier of Bing's api just to pull in trending topics.</p>
]]></description><pubDate>Thu, 17 Feb 2022 15:01:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=30373759</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=30373759</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=30373759</guid></item><item><title><![CDATA[New comment by jrussbowman in "Reddit can't build a better search engine"]]></title><description><![CDATA[
<p>Right now the terms are available only via the word cloud. I have considered trying to put together an api and also keeping more metrics about each link. For example how many times it's popped up, reddit and twitter users that posted it and which subreddits the post is in. I just haven't gotten around to it.<p>The concern of it being an echo chamber is one of the major reasons I added Twitter and I still look for more sources. For most of Trump's presidency some form of his name was the top trend 24 hours a day. Crypto is another trend, that while it's great for me because I'm interested in it as well, I question if it's really reflective of what the world is talking about.<p>I have considered creating some sub-sites as well to try and dig more. Focusing on subreddits for specific categories, but the Twitter api (at least what I can afford which means free) isn't quite as flexible for doing that kind of thing while staying inside my api call limits.</p>
]]></description><pubDate>Thu, 17 Feb 2022 14:42:22 +0000</pubDate><link>https://news.ycombinator.com/item?id=30373518</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=30373518</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=30373518</guid></item><item><title><![CDATA[New comment by jrussbowman in "Reddit can't build a better search engine"]]></title><description><![CDATA[
<p>I built <a href="https://www.unscatter.com" rel="nofollow">https://www.unscatter.com</a> using Reddit to source links for the search index. Last year I added Twitter as another source.<p>I don't think it's a "better" search engine. It's a different lens through which to search. Reddit and Twitter are an information source for what people are talking about. This is why I limit my index to articles that have popped in my Reddit/Twitter input in the last 30 days, deleting anything older.<p>I've actually had it up for years now, just don't know what to do with it. Been focused on my career in IT rather than entrepreneurism because well, life. I can say just this morning I saw "Stanytsia Luhanska" pop as a trending term on the front page of Unscatter and at the time mainstream media has not picked up the story of the school being hit by Russian shells.<p>I think over all the quality results still come from Reddit. Twitter often gets gamed and I see content terms pop up in the trending list. However, Twitter overnight (my time, US East) gets a more international flavor with lots of Korean and other Asia Pac country content bubbling to the top during that time because of Twitter.</p>
]]></description><pubDate>Thu, 17 Feb 2022 13:52:44 +0000</pubDate><link>https://news.ycombinator.com/item?id=30373001</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=30373001</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=30373001</guid></item><item><title><![CDATA[New comment by jrussbowman in "Ask HN: Do people still use DeviantArt?"]]></title><description><![CDATA[
<p>My 14 year old daughter managed to find it by herself through her friends. I've only really ever used it to find nice wallpapers but my daughter is a pretty talented artist and she loves the site.</p>
]]></description><pubDate>Sat, 29 Jan 2022 16:41:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=30127728</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=30127728</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=30127728</guid></item><item><title><![CDATA[New comment by jrussbowman in "Postgres is a great pub/sub and job server (2019)"]]></title><description><![CDATA[
<p>Postgres is really just great for being able to build just about anything to get that first viable product built. It's basically the swiss army knife for anything data in my opinion. You got sql, nosql, job queues, full text indexing. It's great.<p>I use it as a sql database and full text search for little personal project I work on off and on and it works great. I haven't touched it except to check every few weeks for security updates for months since I got a promotion and it, the golang app server and python scripts have had no issue just churning along keeping a 30 day archive of links found via reddit and twitter. Postgres is great.</p>
]]></description><pubDate>Sat, 18 Dec 2021 03:12:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=29601228</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=29601228</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=29601228</guid></item><item><title><![CDATA[New comment by jrussbowman in "Show HN: A search engine that lets you refine your queries"]]></title><description><![CDATA[
<p>Thanks, I just got back to this thread and you already linked it for me :) I didn't want to risk being seen as hijacking a thread.</p>
]]></description><pubDate>Tue, 30 Nov 2021 00:29:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=29387250</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=29387250</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=29387250</guid></item><item><title><![CDATA[New comment by jrussbowman in "Show HN: A search engine that lets you refine your queries"]]></title><description><![CDATA[
<p>Good luck with the project. Out of curiosity are you crawling and indexing yourself or using a search api?<p>It's really fast, I like that. The amount of tags to choose from is a bit overwhelming.<p>Hope you get good responses here, the couple times I've posted my search engine I haven't gotten a lot of feedback.</p>
]]></description><pubDate>Sun, 28 Nov 2021 05:39:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=29366478</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=29366478</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=29366478</guid></item><item><title><![CDATA[Unscatter.com: Trump top trend for months, since election called it's been Biden]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.unscatter.com">https://www.unscatter.com</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=25037495">https://news.ycombinator.com/item?id=25037495</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 09 Nov 2020 17:55:34 +0000</pubDate><link>https://www.unscatter.com</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=25037495</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=25037495</guid></item><item><title><![CDATA[Show HN: Stuck at home, I rebuilt unscatter.com to be an index of trending links]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.unscatter.com">https://www.unscatter.com</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=22804854">https://news.ycombinator.com/item?id=22804854</a></p>
<p>Points: 3</p>
<p># Comments: 0</p>
]]></description><pubDate>Tue, 07 Apr 2020 16:46:52 +0000</pubDate><link>https://www.unscatter.com</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=22804854</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=22804854</guid></item><item><title><![CDATA[Show HN: Unscatter.com is now a searchable news index]]></title><description><![CDATA[
<p>Hi,<p>For the past several years I've been working on <a href="https://www.unscatter.com" rel="nofollow">https://www.unscatter.com</a> as a hobby. Used it first as a project to keep up to date with python and a few years ago rewrote it to learn go.<p>For the past several years it's primarily been not much more than a proxy for APIs for sites like Faroo, Reddit, Youtube and Tumblr.<p>I recently just started a rebuild of the site that's now live. Instead of just being a proxy for API's, I now use the reddit api to determine what news is trending. Mostly I'm pulling from r/all, r/worldnews and r/news. I'm then pulling and analyzing the stories and creating a searchable index of them.<p>My goal is to have it act a 1 month archive of what news has been trending. What's driving me is what I have perceived as a lack of trust of the press based on my interactions with people.  My goal with the Unscatter brand is to provide tools people can use to see if the press is giving them news or driving a message. This index is just the first tool I'm pushing out.<p>I am doing a Show HN in the hopes of getting some feed back. My goal is for it to be a 1 month archive, articles that haven't been seen by my fetcher in over a month will be deleted. I don't have the full 1 month archive and probably won't start deleting until June as I'm still tuning.<p>Another reason I'm doing the Show HN is to see if anyone has some suggestions on other sites that might have an API I can work with? I realize that pulling just from reddit I still risk getting a bias based off of the visitors of that site being their own small society. I'm hoping to find other sources to pull from in an attempt to try and avoid bias.<p>I should also note if you have a great tool you wish to sell me, at this time I don't have a real monetization strategy. I do serve adsense adds but I also have /search blocked in robots.txt so it's not like google can really tune the ads that well. This is 100% pure hobby, I'm only paying $20 a month for hosting.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=16792992">https://news.ycombinator.com/item?id=16792992</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 09 Apr 2018 13:59:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=16792992</link><dc:creator>jrussbowman</dc:creator><comments>https://news.ycombinator.com/item?id=16792992</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=16792992</guid></item></channel></rss>