<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: ImageXav</title><link>https://news.ycombinator.com/user?id=ImageXav</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Wed, 19 Aug 2026 01:52:33 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=ImageXav" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by ImageXav in "GPT 5.6 Sol is the best "vision" model OpenAI ever released"]]></title><description><![CDATA[
<p>Gemini tops their vision evals [0] by a mile, with 4/5 top spots going to variants of it. Qwen is the only other contender, likely due to how good it is for object detection, where it crushes the competition [1].<p>[0] <a href="https://playground.roboflow.com/evals" rel="nofollow">https://playground.roboflow.com/evals</a><p>[1] <a href="https://playground.roboflow.com/evals/object-detection" rel="nofollow">https://playground.roboflow.com/evals/object-detection</a></p>
]]></description><pubDate>Mon, 17 Aug 2026 18:23:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=49335420</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=49335420</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49335420</guid></item><item><title><![CDATA[New comment by ImageXav in "GPT 5.6 Sol is the best "vision" model OpenAI ever released"]]></title><description><![CDATA[
<p>Me too. This is an interesting comparison but in my experience Qwen and Gemini have typically been the top contenders for image related tasks. For that reason it would be great to have the comparison here, as I'm not surprised by Gemini's dominance over the other models.</p>
]]></description><pubDate>Mon, 17 Aug 2026 14:36:55 +0000</pubDate><link>https://news.ycombinator.com/item?id=49331751</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=49331751</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49331751</guid></item><item><title><![CDATA[Measuring Autonomous AI Research]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.primeintellect.ai/blog/measuring-autonomous-research">https://www.primeintellect.ai/blog/measuring-autonomous-research</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=49330109">https://news.ycombinator.com/item?id=49330109</a></p>
<p>Points: 1</p>
<p># Comments: 0</p>
]]></description><pubDate>Mon, 17 Aug 2026 12:58:14 +0000</pubDate><link>https://www.primeintellect.ai/blog/measuring-autonomous-research</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=49330109</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49330109</guid></item><item><title><![CDATA[New comment by ImageXav in "Explorative modeling: Train on the best of K guesses"]]></title><description><![CDATA[
<p>This feels extremely close in nature to being a generalisation of discrete distribution networks.<p>Paper: <a href="https://arxiv.org/abs/2401.00036" rel="nofollow">https://arxiv.org/abs/2401.00036</a>
Project page: <a href="https://discrete-distribution-networks.github.io/" rel="nofollow">https://discrete-distribution-networks.github.io/</a><p>Given that these were published at ICML 2025, at which the author was a top reviewer as per their own website <a href="https://alexiglad.github.io/" rel="nofollow">https://alexiglad.github.io/</a>, I wonder how influenced they were to pursue this avenue of the back of it. They do have a related works section in appendix E, but it somehow seems to miss this. Which is odd, as the paper made a splash at the time at the conference and even made it close to the top of hacker news due to the novelty.<p>Of course, DDNs are a fundamentally new architecture, whereas this is more a generaliseable training strategy, but there is a very similar core.</p>
]]></description><pubDate>Sat, 01 Aug 2026 20:36:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=49138215</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=49138215</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=49138215</guid></item><item><title><![CDATA[Quoting Sam Altman]]></title><description><![CDATA[
<p>Article URL: <a href="https://simonwillison.net/2026/Jul/20/sam-altman/">https://simonwillison.net/2026/Jul/20/sam-altman/</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48994248">https://news.ycombinator.com/item?id=48994248</a></p>
<p>Points: 4</p>
<p># Comments: 1</p>
]]></description><pubDate>Tue, 21 Jul 2026 16:10:50 +0000</pubDate><link>https://simonwillison.net/2026/Jul/20/sam-altman/</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48994248</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48994248</guid></item><item><title><![CDATA[New comment by ImageXav in "Kimi K3: Open Frontier Intelligence"]]></title><description><![CDATA[
<p>I've been avidly using Fable since it was re-released and while it has been excellent at building the apps I want, the reasoning has been completely opaque.<p>Kim, however, has exposed the whole reasoning trace, or enough of it to matter. I'd almost forgotten how nice it is to see this. I've been able to see all of the weird twist and turns it takes and it is joyful. But also, far, far more informative and means I can debug ideas far more thoroughly. Also, at a first glance it seems to have gotten quite far on a niche hobby horse of mine that no LLM has been able to crack. I'll be testing this more for sure.</p>
]]></description><pubDate>Thu, 16 Jul 2026 19:46:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=48939333</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48939333</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48939333</guid></item><item><title><![CDATA[New comment by ImageXav in "Mistral's Robostral Navigate: a state of the art robotics navigation model"]]></title><description><![CDATA[
<p>Ok, this is really cool. The fact that the robot can use pointing to decide where to go is a great design decision, and robotics really is the next frontier. Definitely cheering on Mistral here!</p>
]]></description><pubDate>Wed, 08 Jul 2026 15:21:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=48833139</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48833139</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48833139</guid></item><item><title><![CDATA[New comment by ImageXav in "We're extending access to Fable 5 on all paid plans through July 12"]]></title><description><![CDATA[
<p>I think it depends on your workflow. I've had a great experience with the trial. I work in research, and have set up something similar to Kaparthy's auto research. I, with Fable, have managed to get an image generation model down from 80M parameters to 10M and keep the quality of the generated images on par (similar FID). And, importantly, every change was modular, explained by Fable, reviewed by myself, and understood and documented. If not understood, I read relevant docs until I did it didn't accept it as part of the plan. So it ended up being a simple composition of existing ideas which I had previously encountered, but stacked much more rapidly than I could have.<p>The structure of the code is easily readable as I enforce concenventions followed by good libraries. And I can easily plug in new datasets. It's pretty good frankly.</p>
]]></description><pubDate>Wed, 08 Jul 2026 08:15:12 +0000</pubDate><link>https://news.ycombinator.com/item?id=48829079</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48829079</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48829079</guid></item><item><title><![CDATA[New comment by ImageXav in "Is The Economist Always Wrong?"]]></title><description><![CDATA[
<p>I read the Economist for over a decade growing up. It was a great way to learn about the world, who was in power where, and the challenges facing economies at the time. I found their exposition to be pretty good given the fact they were restricted to a few pages for important events. However, their proposed solutions were always the same. More market freedom, etc.<p>I did feel with the change in editorial direction a while back that they lost some of their edge. I've since mostly just stuck to the Financial Times. It feels less worldly, but the content of the articles feels better.</p>
]]></description><pubDate>Wed, 08 Jul 2026 07:32:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=48828748</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48828748</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48828748</guid></item><item><title><![CDATA[S&P 500 Indices Consultation on Treatment of MegaCap Companies – Results [pdf]]]></title><description><![CDATA[
<p>Article URL: <a href="https://www.spglobal.com/spdji/en/documents/indexnews/announcements/20260604-1483731/1483731_spdji-us-indices-megacaps-results-20260604.pdf">https://www.spglobal.com/spdji/en/documents/indexnews/announcements/20260604-1483731/1483731_spdji-us-indices-megacaps-results-20260604.pdf</a></p>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48409535">https://news.ycombinator.com/item?id=48409535</a></p>
<p>Points: 5</p>
<p># Comments: 1</p>
]]></description><pubDate>Fri, 05 Jun 2026 08:17:39 +0000</pubDate><link>https://www.spglobal.com/spdji/en/documents/indexnews/announcements/20260604-1483731/1483731_spdji-us-indices-megacaps-results-20260604.pdf</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48409535</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48409535</guid></item><item><title><![CDATA[New comment by ImageXav in "Stack Overflow’s forum is dead but the company’s still kicking"]]></title><description><![CDATA[
<p>Agreed. Which is also odd, if you think about it. Surely with the amount of compute Anthropic and others have available, they could test each of the solutions in the SO data they surely have and rank them based on efficiency/elegance/other criteria and remove poor solutions from their training data.</p>
]]></description><pubDate>Tue, 26 May 2026 19:28:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=48284767</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48284767</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48284767</guid></item><item><title><![CDATA[How do you reduce LLM spam in PR reviews?]]></title><description><![CDATA[
<p>Title. I finally got annoyed enough at work with a colleague who posted an 11 point list they clearly hadn't read or reviewed as a comment on my PR that my reply started with 'Thanks Claude...'. No doubt in my mind that I spent far longer on my curt rebuttal than they did on the review. I'd like to hear from folks whose organisation uses LLMs for coding effectively and what kind of best practices they have put in place to avoid these situations.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=48193561">https://news.ycombinator.com/item?id=48193561</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Tue, 19 May 2026 14:11:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=48193561</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=48193561</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48193561</guid></item><item><title><![CDATA[New comment by ImageXav in "New iron nanomaterial wipes out cancer cells without harming healthy tissue"]]></title><description><![CDATA[
<p>It may feel that way due to the iterative nature of medical improvements, but over the past few decades there has been a consistent reduction in cancer mortality rates across most types of cancer [0]. Treatments really are getting better and more targeted. Immunotherapy has made huge breakthroughs. Combination treatments allow for significantly improved lifespans and better quality of life during treatments. There are a few cancers that remain hard to treat, but I have a lot of confidence that in the coming decades we will make strides in attacking them. That being said, I'm very sorry to hear about the pain you and your family must be going through. I've had a few close loved ones undergo cancer treatment and it was tough.<p>[0] <a href="https://acsjournals.onlinelibrary.wiley.com/doi/10.3322/caac.70043" rel="nofollow">https://acsjournals.onlinelibrary.wiley.com/doi/10.3322/caac...</a></p>
]]></description><pubDate>Sun, 01 Mar 2026 19:09:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=47209667</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=47209667</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=47209667</guid></item><item><title><![CDATA[New comment by ImageXav in "A tough labor market for white-collar workers has turned recruiting upside down"]]></title><description><![CDATA[
<p>This seems like such an easy way to create perverse incentives and profit off people who are already down on their luck. Imagine being told that the only way to get considered is to pay a fee. Then later on you get told to pay the gold fee for priority. Oh you're still not getting hired? Go for our platinum package that will definitely make the difference! Not enough money? No worries, we'll take 30% of your salary for the first few years. Or maybe we'll just give you some a fixed debt at a high interest rate. Aren't you glad you used us?</p>
]]></description><pubDate>Mon, 09 Feb 2026 10:58:03 +0000</pubDate><link>https://news.ycombinator.com/item?id=46943888</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=46943888</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46943888</guid></item><item><title><![CDATA[New comment by ImageXav in "The Smol Training Playbook: The Secrets to Building World-Class LLMs"]]></title><description><![CDATA[
<p>Agreed. One at a time testing (OAT) has been outdated for almost a century at this point. Factorial and fractional factorial experiments have been around for that long and give detailed insights into the effect of not just single changes but the interaction between changes, which means you can superpower your learnings as many variables in DL do in fact interact.<p>Or, more modern Bayesian methods if you're more interested in getting the best results for a given hyperparameter sweep.<p>However, that is not to detract from the excellent effort made here and the great science being investigated. Write ups like this offer so much gold to the community.</p>
]]></description><pubDate>Sun, 02 Nov 2025 12:05:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=45789710</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=45789710</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45789710</guid></item><item><title><![CDATA[New comment by ImageXav in "How Europe crushes innovation"]]></title><description><![CDATA[
<p>I guess that's the crux of it. From an individual perspective it makes sense to stay in a stable environment, especially if a family is involved. However, I think from a societal perspective it is desirable to have people who gamble on creating new products which can raise the bar in their given industries.<p>Also, just because the start up fails doesn't mean it was a waste of time. If you manage to provide employment for even just 3 or 4 people for a few years, help them and yourself develop, that is a valuable success.</p>
]]></description><pubDate>Tue, 07 Oct 2025 12:46:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=45502437</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=45502437</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45502437</guid></item><item><title><![CDATA[New comment by ImageXav in "How Europe crushes innovation"]]></title><description><![CDATA[
<p>I would add an aspect that is not covered here but is often ignored: the strong labour protection laws result in a mentality where if you get a good job you are much less likely to want to take risks e.g. start your own business. There was a post on the HENRY (high earner, not rich yet) UK subreddit the other day from someone who had a wealth of experience and had the opportunity to join a start up as a CTO. It honestly sounded like a great chance to initiate change. All of the comments were telling the poster that they had it good, that 99% of start ups fail, that the hours would be gruelling. I feel as though the conversation would have been quite different in a US subreddit.<p>A term they like to use is 'crabs in a bucket'.</p>
]]></description><pubDate>Mon, 06 Oct 2025 18:24:13 +0000</pubDate><link>https://news.ycombinator.com/item?id=45494498</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=45494498</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=45494498</guid></item><item><title><![CDATA[New comment by ImageXav in "Everything is correlated (2014–23)"]]></title><description><![CDATA[
<p>This is an interesting point. I've been trying to think about something similar recently but don't have much of an idea how to proceed. I'm gathering periodic time series data and am wondering how to factor in the frequency of my sampling for the statistical tests. I'm not sure how to assess the difference between 50Hz and 100Hz on the outcome, given that my periods are significantly longer. Would you have an idea of how to proceed? The person I'm working with currently just bins everything in hour long buckets and uses the mean for comparison between time series but this seems flawed to me.</p>
]]></description><pubDate>Fri, 22 Aug 2025 21:55:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=44990339</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=44990339</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44990339</guid></item><item><title><![CDATA[Best workflow for quick ideation with LLMs from phone]]></title><description><![CDATA[
<p>As per the title. I tend to travel a lot and don't often have comfortable access to my laptop. I would like to run quick, small research ideas quickly. My workflow would be:<p>- Ask an LLM to put together a PoC
- Get the code and send it to my home PC or some cloud offering
- Run and get results<p>Has anyone here put together a pipeline like this? I would be curious to hear thoughts on processes people have set up.</p>
<hr>
<p>Comments URL: <a href="https://news.ycombinator.com/item?id=44970222">https://news.ycombinator.com/item?id=44970222</a></p>
<p>Points: 2</p>
<p># Comments: 0</p>
]]></description><pubDate>Thu, 21 Aug 2025 07:50:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=44970222</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=44970222</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44970222</guid></item><item><title><![CDATA[New comment by ImageXav in "An engineer's perspective on hiring"]]></title><description><![CDATA[
<p>I've had the complete opposite experience, and feel the complete opposite way. What is there to learn from failing a leetcode? It feels like luck of the draw - I didn't study that specific problem type and so failed. Also, there is an up front cost of several months to cover and study a wide array of leetcode problems.<p>With a take home I can demonstrate how I would perform at work. I can sit on it, think things over in my head, come up with an attack plan and execute it. I can demonstrate how I think about problems and my own value more clearly. Using a take home as a test is indicative to me that a company cares a bit more about its hiring pipeline and is being careful not to put candidates under arbitrary pressures.</p>
]]></description><pubDate>Sun, 10 Aug 2025 05:31:39 +0000</pubDate><link>https://news.ycombinator.com/item?id=44852984</link><dc:creator>ImageXav</dc:creator><comments>https://news.ycombinator.com/item?id=44852984</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=44852984</guid></item></channel></rss>