<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: akshayubhat</title><link>https://news.ycombinator.com/user?id=akshayubhat</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 09 Oct 2026 05:53:26 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=akshayubhat" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by akshayubhat in "Delicious's Data Policy is Like Setting a Museum on Fire"]]></title><description><![CDATA[
<p>I can only think of using 80legs as a crawler, since its distributed enough to make sure that you don't run into any IP address based rate limitation. But it's just a guess.</p>
]]></description><pubDate>Fri, 17 Dec 2010 16:14:01 +0000</pubDate><link>https://news.ycombinator.com/item?id=2016606</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=2016606</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=2016606</guid></item><item><title><![CDATA[New comment by akshayubhat in "Bayesian Learning in Social Networks"]]></title><description><![CDATA[
<p>So many upvotes, I find information cascade happening here.</p>
]]></description><pubDate>Tue, 07 Dec 2010 13:01:30 +0000</pubDate><link>https://news.ycombinator.com/item?id=1978900</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1978900</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1978900</guid></item><item><title><![CDATA[New comment by akshayubhat in "Students: You Are Probably Not Mark Zuckerberg, So Stay In School"]]></title><description><![CDATA[
<p>a harvard (or UIUC, stanford, cornell, MIT, UCB) CS grad can easily make 70k$ a year right out of college and much more in future. 100k$ tuition might be an issue for liberal arts and communication majors, but surely not for a CS major (assuming he is good enough to have a successful startup in first place).</p>
]]></description><pubDate>Sun, 26 Sep 2010 06:02:42 +0000</pubDate><link>https://news.ycombinator.com/item?id=1728651</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1728651</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1728651</guid></item><item><title><![CDATA[New comment by akshayubhat in "Google Sibyl: A system for large scale machine learning [pdf]"]]></title><description><![CDATA[
<p>Interesting! There is an open source Apache Mahout project for doing machine learning via hadoop. Also have a look at Vowpal wabbit an open source framework for fast SGD for online learning.<p>Also another interesting point is their use of boosting, since i recently attended a tech talk by facebook engineers where they told us that bagged decision trees while predicting the friends for a user.</p>
]]></description><pubDate>Sun, 19 Sep 2010 19:33:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=1706739</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1706739</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1706739</guid></item><item><title><![CDATA[New comment by akshayubhat in "Bots playing the market make some bizarre patterns"]]></title><description><![CDATA[
<p>Interesting anyone knows which strategies they use to decide to buy/sell, once they have determined the expected price?<p>do they use physics based techniques, like black scholes?<p>or do they use statical learning based techniques?</p>
]]></description><pubDate>Sun, 19 Sep 2010 03:27:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=1705523</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1705523</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1705523</guid></item><item><title><![CDATA[New comment by akshayubhat in "Tell HN: I accidentally ran up a $1000 Heroku bill"]]></title><description><![CDATA[
<p>Something like this wont happen to Google App Engine, thankfully.</p>
]]></description><pubDate>Tue, 14 Sep 2010 03:00:00 +0000</pubDate><link>https://news.ycombinator.com/item?id=1689250</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1689250</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1689250</guid></item><item><title><![CDATA[New comment by akshayubhat in "NoSQL, Heroku, and You"]]></title><description><![CDATA[
<p>err HBASE is the database and not HDFS.</p>
]]></description><pubDate>Mon, 13 Sep 2010 21:16:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=1688359</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1688359</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1688359</guid></item><item><title><![CDATA[New comment by akshayubhat in "NoSQL, Heroku, and You"]]></title><description><![CDATA[
<p>I completely agree, I find it weird whenever someone mentions Hadoop as as NOSQL. Hadoop is a Distributed Computing system. While HDFS is the Database.<p>Hadoop is more of an Distributed O.S. to run process and store data across multiple machines.</p>
]]></description><pubDate>Mon, 13 Sep 2010 05:25:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=1685661</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1685661</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1685661</guid></item><item><title><![CDATA[New comment by akshayubhat in "The Psychology of Loners and Introverts"]]></title><description><![CDATA[
<p>Well he has Indianized English Grammatical style.</p>
]]></description><pubDate>Sun, 12 Sep 2010 06:44:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=1683198</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1683198</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1683198</guid></item><item><title><![CDATA[New comment by akshayubhat in "Phantom Android Apps: they don't copy your ideas, they steal your code"]]></title><description><![CDATA[
<p>On a side note, had this been a case with an MacOSX application or Windows application, would you have petitioned Apple or MSFT?</p>
]]></description><pubDate>Sat, 11 Sep 2010 11:09:56 +0000</pubDate><link>https://news.ycombinator.com/item?id=1681356</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1681356</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1681356</guid></item><item><title><![CDATA[New comment by akshayubhat in "Apple products are a mutant virus... says Acer founder"]]></title><description><![CDATA[
<p>oh are i-OS API and code cross platform compatible?
 Do any other companies except Apple create iOS devices?<p>in fact Apple used 3.3.1 to make sure that code was not cross platform compatible.</p>
]]></description><pubDate>Thu, 09 Sep 2010 18:31:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=1676340</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1676340</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1676340</guid></item><item><title><![CDATA[New comment by akshayubhat in "Apple products are a mutant virus... says Acer founder"]]></title><description><![CDATA[
<p>no because you can run programs on Acer netbook which will also run on say HP and Dell netbook's.</p>
]]></description><pubDate>Thu, 09 Sep 2010 17:34:26 +0000</pubDate><link>https://news.ycombinator.com/item?id=1676136</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1676136</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1676136</guid></item><item><title><![CDATA[New comment by akshayubhat in "Apple Relaxes iOS Restrictions On Development Tools, Publishes Review Guidelines"]]></title><description><![CDATA[
<p>no more 3.3.1 that's good. Finally sanity prevails over "interface decides programming paradigm/language argument".</p>
]]></description><pubDate>Thu, 09 Sep 2010 17:31:40 +0000</pubDate><link>https://news.ycombinator.com/item?id=1676123</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1676123</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1676123</guid></item><item><title><![CDATA[New comment by akshayubhat in "Scientists: Go ahead, kill all the mosquitoes."]]></title><description><![CDATA[
<p>Evolution, why DDT wouldn't work.
<a href="http://en.wikipedia.org/wiki/DDT#Mosquito_resistance" rel="nofollow">http://en.wikipedia.org/wiki/DDT#Mosquito_resistance</a></p>
]]></description><pubDate>Thu, 09 Sep 2010 09:05:41 +0000</pubDate><link>https://news.ycombinator.com/item?id=1674832</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1674832</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1674832</guid></item><item><title><![CDATA[New comment by akshayubhat in "How I Used Amazon’s Mechanical Turk to Validate my Startup Idea"]]></title><description><![CDATA[
<p>People who use AMT for earning money may not necessarily be the right audience you are looking for.</p>
]]></description><pubDate>Tue, 07 Sep 2010 12:11:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=1668664</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1668664</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1668664</guid></item><item><title><![CDATA[New comment by akshayubhat in "Once a Dynamo, the Tech Sector Is Slow to Hire"]]></title><description><![CDATA[
<p>Entrepreneurship isn't the cure-all pill, good Ideas are hard to come by, even if you are looking for them continuously. Second Large Businesses have huge efficiency advantage due to their scale.<p>While HN would seem extremely pro - entrepreneurship, in reality, starting your business is not always the smartest move, not from just personal perspective but from perspective of a nation as well. Sure you want Larry Page and Sergey Brin to create Google but before that they went to Stanford for  Ph.D. in CS. Entrepreneurship just for sake of it and without any good idea is no better than playing stock market with different strategies.<p><pre><code>     If 15% of the unemployed engineers in Corvallis could start companies that hired 5 full timers, they would (probably) have a booming economy all over again. Shit, maybe we could start exporting to China (I think they value entrepeneurship and industrialization there...).
</code></pre>
You highly overestimate success rate, also what would they sell? What if they compete with each other and end up cannibalizing the market share between each other. That would at most create re-distribution of wealth and not generation.<p>The only way to generate income is via means which are non zero sum games [at least non zero sum within that nation].</p>
]]></description><pubDate>Tue, 07 Sep 2010 06:31:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=1668264</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1668264</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1668264</guid></item><item><title><![CDATA[New comment by akshayubhat in "Once a Dynamo, the Tech Sector Is Slow to Hire"]]></title><description><![CDATA[
<p>Wow, Doctors are different case altogether<p>Their supply is limited and controlled, anyone can come and say that they are "web programmer" after reading HTML/CSS/Ruby, Even Universities increase class sizes arbitrarily. All this is not possible in medical profession. Doctors are few, they are trained extensively, not anyone reading anatomy book can go for a residency!</p>
]]></description><pubDate>Tue, 07 Sep 2010 06:14:50 +0000</pubDate><link>https://news.ycombinator.com/item?id=1668241</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1668241</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1668241</guid></item><item><title><![CDATA[New comment by akshayubhat in "Why a Do it Yourself Big Data Stack Is a Better Option"]]></title><description><![CDATA[
<p>I could not understand several points made in this article:<p>1: He mentions infochimps, but according to my knowledge its more of an ebay for datasets rather than support/provider for Big Data Stack, also is it successful? I am unsure how infochimps is related to the Big Data stack.<p>2: From what I remember reading about 80legs, is that it uses distributed grid computing to run the crawlers (something like SETI @ Home), I doubt Hadoop was ever designed for such applications. So this is surely isn't a Hadoop use case.<p>3: Quoting:<p><pre><code>       While the standard big data stack has made huge strides in making big data more accessible to everyone, it will always fall short against our stack when it comes to the cost of collecting data.  We actually don’t store that much data. Because 80legs users can filter their data on the nodes, they’re able to return the minimum amount of data in their result sets.  The processing (or reduction, pardon the pun) is done on the nodes.  Actual result sets are very small relative to the size of the input set.
</code></pre>
Again I am unsure how it is different from Hadoop? First Hadoop uses same principle "to move computation closer to data" hence a crawler implemented using Hadoop (something Hadoop is not intended to do) will also store data locally and not on some other node.<p>Also he mentions """ We have about 50,000 computers using their excess bandwidth.""" 50,000?? The biggest Hadoop cluster That I know (Yahoo) has ~10-20k nodes, and Hadoop was never meant to be used at 50K scale for crawling. So they had no option other than building their own system, even if they had to build it today.<p>4: Quoting<p><pre><code>      One advantage is optimization — an “off-the-shelf” system is going to have some generalities built into it that can’t be optimized to fit your needs. The opportunity cost of going “standard” is a slew of competitive advantages.
</code></pre>
The only issue I can think of regarding Hadoop is that its written in JAVA, otherwise its an extremely extensible piece of software. Unless you are designing a real time messaging system or distributed system for High Frequency Trading, Hadoop is good enough for most of the applications. Also what about cost of finding good enough programmers who are capable of building a system? Another advantage of Hadoop is that in case of a low load the remaining nodes can be used to do something else, maybe processing some data, with your own solution it would be harder to do it. Also your IP and your Secret Sauce isn't of much use, if you dont have solid Patents for them, otherwise they would mostly end up becoming a maintenance nightmare, after original engineers cash out. Also what if the the big company already has Hadoop cluster, it would be even difficult for them to integrate with your computing power.<p>While I seem to agree with Authors conclusion that a highly focused startup should make their proprietary solution, I cannot agree with his evidence behind that argument. A grid based crawler with 50K machines isn't something that Hadoop was ever designed to support.</p>
]]></description><pubDate>Sun, 05 Sep 2010 22:09:32 +0000</pubDate><link>https://news.ycombinator.com/item?id=1665293</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1665293</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1665293</guid></item><item><title><![CDATA[New comment by akshayubhat in "Mining social networks"]]></title><description><![CDATA[
<p>Yup he is young, he did a post doc for a year at Cornell [under Prof. Kleinberg] after his PhD, and directly became Asst. Prof at Stanford.<p>Well you can start with reading this book <a href="http://www.cs.cornell.edu/home/kleinber/networks-book/" rel="nofollow">http://www.cs.cornell.edu/home/kleinber/networks-book/</a> for overview, regarding application of these techniques are considered you can look for papers at recent WWW, NIPS,ICML conferences. The most popular and well studied areas are Link Prediction and Community Detection. The SNAP library comes with some good example code. You can also have a look at Divisi project at MIT Media Lab if you are interested in reasoning/ analogy over the networks.<p>For real datasets, there are a lot of encyclopedic datasets such as Wikipedia / DBPedia /Semantic Web/ Music Brainz, as well as social ones such as Twitter follower network dataset. If you are in a university, you can even get full Web Graph from Yahoo [for research use alone].</p>
]]></description><pubDate>Sun, 05 Sep 2010 15:41:37 +0000</pubDate><link>https://news.ycombinator.com/item?id=1664659</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1664659</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1664659</guid></item><item><title><![CDATA[New comment by akshayubhat in "IPad: the perfect computing device for children?"]]></title><description><![CDATA[
<p>iPad should come out with Disney specific version (in pink for girls and in red/blue for boys). Consider the opportunities, an inbuilt iTunes  store with all superhero / Cinderella type cartoons [they will learn to pay for the content from the childhood itself no need for RIAA / MPAA lawsuits], there is so much opportunity in this field, I can totally see kids harrowing their parents for one. Wonder how Steve jobs hasn't already began producing one.</p>
]]></description><pubDate>Sun, 05 Sep 2010 15:31:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=1664646</link><dc:creator>akshayubhat</dc:creator><comments>https://news.ycombinator.com/item?id=1664646</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=1664646</guid></item></channel></rss>