<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: kkielhofner</title><link>https://news.ycombinator.com/user?id=kkielhofner</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Tue, 21 Jul 2026 01:53:56 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=kkielhofner" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by kkielhofner in "Understanding the BM25 full text search algorithm"]]></title><description><![CDATA[
<p>I've used it for hybrid search and it works quite well.<p>Overall I'm really happy to see Typesense mentioned here.<p>A lot of the smaller scale RAG projects, etc you see around would be well served by Typesense but it seems to be relatively unknown for whatever reasons. It's probably one of the easiest solutions to deploy, has reasonable defaults, good docs, easy clustering, etc while still be very capable, performant, and powerful if you need to dig in further.</p>
]]></description><pubDate>Wed, 20 Nov 2024 16:50:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=42195777</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42195777</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42195777</guid></item><item><title><![CDATA[New comment by kkielhofner in "Netflix buffering issues: Boxing fans complain about Jake Paul vs. Mike Tyson"]]></title><description><![CDATA[
<p>> I'm an engineering manager<p>How are you involved in the hiring process?<p>> Our engineers are fucking morons. And this guy was the dumbest of the bunch.<p>Very indicative of a toxic culture you seem to have been pulled in to and likely have contributed to by this point given your language and broad generalizations.<p>Describing a wide group of people you're also responsible for as "fucking morons" says more about you than them.</p>
]]></description><pubDate>Sat, 16 Nov 2024 04:35:10 +0000</pubDate><link>https://news.ycombinator.com/item?id=42154276</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42154276</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42154276</guid></item><item><title><![CDATA[New comment by kkielhofner in "Serving 15 terabytes of 4K video for $2.18"]]></title><description><![CDATA[
<p>They will.<p>- It's a cornerstone of their brand.<p>- R2 users are at least paying for something (storage).<p>- Their network is massively overbuilt to be able to absorb DDoS attacks.<p>- They offer free bandwidth with their CDN - including completely free users. These resources need to get fetched from origin and/or cached. R2 doesn't have to fetch from origin which eliminates the bandwidth required for the fetch.<p>Large providers typically pay very little/nothing for bandwidth other than equipment costs (ports, etc). As a large provider they have free peering to most of the "last mile" ISP/eyeball networks in the world. This benefits all parties because these ISPs don't have to pay transit providers and neither does Cloudflare and it's faster. Same goes for all of the big clouds.<p>People who think AWS/GCP/Azure/etc bandwidth pricing of $0.12/GB (or whatever) is fair/reasonable have no idea what bandwidth actually costs these operators. As noted it's effectively nothing and the big clouds capitalize on this ignorance by charging insane markups for bandwidth.</p>
]]></description><pubDate>Fri, 08 Nov 2024 18:59:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=42089517</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42089517</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42089517</guid></item><item><title><![CDATA[New comment by kkielhofner in "AMD outsells Intel in the datacenter space"]]></title><description><![CDATA[
<p>ICC, IPP, QAT, etc are definitely an edge.<p>In AI world they have OpenVINO, Intel Neural Compressor, and a slew of other implementations that typically offer dramatic performance improvements.<p>Like we see with AMD trying to compete with Nvidia software matters - a lot.</p>
]]></description><pubDate>Tue, 05 Nov 2024 23:39:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=42056180</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42056180</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42056180</guid></item><item><title><![CDATA[New comment by kkielhofner in "Colorado scrambles to change voting-system passwords after accidental leak"]]></title><description><![CDATA[
<p>> The US voting machines are just waiting to be hacked, just a matter of when, not if.<p>The US election system is very distributed and fragmented - there is virtually no standardization.<p>Even in the tightest margins for something like President you'd need to have seriously good data to figure out which random municipality voting system(s) you'd need to target to actually affect the outcome.</p>
]]></description><pubDate>Sat, 02 Nov 2024 22:24:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=42029647</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42029647</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42029647</guid></item><item><title><![CDATA[New comment by kkielhofner in "Embeddings are underrated"]]></title><description><![CDATA[
<p>> Do you mind sharing why you chose SPLADE-esque sparse embeddings?<p>I can provide what I can provide publicly. The first thing we ever do is develop benchmarks given the uniqueness of the nuclear energy space and our application. In this case it's FermiBench[0].<p>When working with operating nuclear power plants there are some fairly unique challenges:<p>1. Document collections tend to be in the billions of pages. When you have regulatory requirements to extensively document EVERYTHING and plants that have been operating for several decades you end up with a lot of data...<p>2. There are very strict security requirements - generally speaking everything is on-prem and hard air-gapped. We don't have the luxury of cloud elasticity. Sparse embeddings are very efficient especially in terms of RAM and storage. Especially important when factoring in budgetary requirements. We're already dropping in eight H100s (minimum) so it starts to creep up fast...<p>3. Existing document/record management systems in the nuclear space are keyword search based if they have search at all. This has led to substantial user conditioning - they're not exactly used to what we'd call "semantic search". Sparse embeddings in combination with other techniques bridge that well.<p>4. Interpretability. It's nice to be able to peek at the embedding and be able to get something out of it at a glance.<p>So it's basically a combination of efficiency, performance, and meeting users where they are. Our Fermi model series is still v1 but we've found performance (in every sense of the word) to be very good based on benchmarking and initial user testing.<p>I should also add that some aspects of this (like pretrained BERT) are fairly compute-intense to train. Fortunately we work with the Department of Energy Oak Ridge National Laboratory and developed all of this on Frontier[1] (for free).<p>[0] - <a href="https://huggingface.co/datasets/atomic-canyon/FermiBench" rel="nofollow">https://huggingface.co/datasets/atomic-canyon/FermiBench</a><p>[1] - <a href="https://en.wikipedia.org/wiki/Frontier_(supercomputer)" rel="nofollow">https://en.wikipedia.org/wiki/Frontier_(supercomputer)</a></p>
]]></description><pubDate>Fri, 01 Nov 2024 21:25:05 +0000</pubDate><link>https://news.ycombinator.com/item?id=42021703</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42021703</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42021703</guid></item><item><title><![CDATA[New comment by kkielhofner in "Embeddings are underrated"]]></title><description><![CDATA[
<p>> word Wood dominated the embedding values, but these were supposed to go into 2 different categories<p>When faced with a similar challenge we developed a custom tokenizer, pretrained BERT base model[0], and finally a SPLADE-esque sparse embedding model[1] on top of that.<p>[0] - <a href="https://huggingface.co/atomic-canyon/fermi-bert-1024" rel="nofollow">https://huggingface.co/atomic-canyon/fermi-bert-1024</a><p>[1] - <a href="https://huggingface.co/atomic-canyon/fermi-1024" rel="nofollow">https://huggingface.co/atomic-canyon/fermi-1024</a></p>
]]></description><pubDate>Fri, 01 Nov 2024 19:00:21 +0000</pubDate><link>https://news.ycombinator.com/item?id=42020375</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42020375</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42020375</guid></item><item><title><![CDATA[New comment by kkielhofner in "Embeddings are underrated"]]></title><description><![CDATA[
<p>> we don't want to hurt performance on other real-world tasks just to do well on MTEB<p>Nice!<p>Fortunately MTEB lets you sort by model parameter size because using 7B parameter LLMs for embeddings is just... Yuck.</p>
]]></description><pubDate>Fri, 01 Nov 2024 18:44:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=42020222</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42020222</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42020222</guid></item><item><title><![CDATA[New comment by kkielhofner in "Embeddings are underrated"]]></title><description><![CDATA[
<p>LLMs have nearly completely sucked the oxygen out of the room when it comes to machine learning or "AI".<p>I'm shocked at the number of startups, etc you see trying to do RAG, etc that basically have no idea what they are, how they actually work, etc.<p>The "R" in RAG stands for retrieval - as in the entire field of information retrieval. But let's ignore that and skip right to the "G" (generative)...<p>Garbage in, garbage out people!</p>
]]></description><pubDate>Fri, 01 Nov 2024 18:41:51 +0000</pubDate><link>https://news.ycombinator.com/item?id=42020196</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42020196</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42020196</guid></item><item><title><![CDATA[New comment by kkielhofner in "Embeddings are underrated"]]></title><description><![CDATA[
<p>My startup (Atomic Canyon) developed embedding models for the nuclear energy space[0].<p>Let's just say that if you think off-the-shelf embedding models are going to work well with this kind of highly specialized content you're going to have a rough time.<p>[0] - <a href="https://huggingface.co/atomic-canyon/fermi-1024" rel="nofollow">https://huggingface.co/atomic-canyon/fermi-1024</a></p>
]]></description><pubDate>Fri, 01 Nov 2024 18:33:16 +0000</pubDate><link>https://news.ycombinator.com/item?id=42020115</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42020115</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42020115</guid></item><item><title><![CDATA[New comment by kkielhofner in "Embeddings are underrated"]]></title><description><![CDATA[
<p>> they're not a complete replacement for simpler methods like BM25<p>There are embedding approaches that balance "semantic understanding" with BM25-ish.<p>They're still pretty obscure outside of the information retrieval space but sparse embeddings[0] are the "most" widely used.<p>[0] - <a href="https://zilliz.com/learn/sparse-and-dense-embeddings" rel="nofollow">https://zilliz.com/learn/sparse-and-dense-embeddings</a></p>
]]></description><pubDate>Fri, 01 Nov 2024 18:30:09 +0000</pubDate><link>https://news.ycombinator.com/item?id=42020090</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=42020090</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=42020090</guid></item><item><title><![CDATA[New comment by kkielhofner in "AMD Will Need Another Decade to Try to Pass Nvidia"]]></title><description><![CDATA[
<p>Jensen has said for years that 30% of their R&D spend is on software. Needless to say as they continue to crush it financially this number continues to completely race past AMD.<p>Turns out people don’t actually want GPUs, they want solutions that happen to run best on GPUs. Nvidia understands that, AMD doesn’t.<p>Lisa Su keeps talking about “chips chips chips” and MAYBE “Oh btw here’s a minor ROCm update”. Meanwhile, Nvidia continues to masterfully execute deeper and wider into overall solutions and ecosystems - a substantial portion of which is software.<p>Nvidia is at the point where they’re eating the entire stack. They do a lot of work on their own models and then package them up nice and tight for you with NIM and Nvidia AI Enterprise. On top of stuff like Metropolis, RIVA, countless things. They even have a ton of frameworks to ingest/handle data, finetune/train, and then deploy via NIM.<p>Enterprise customers can be 100% Nvidia for a solution. When Nvidia is the #1-#2 most valuable company in the world “no one ever got fired for buying Nvidia” hits hard.<p>The people who say “AMD and Nvidia are equal - it’s all PyTorch anyway” have no view of the larger picture.<p>With x86_64, day one you could take a drive out of an Intel system, put it in an AMD system, and it would boot and run perfectly. You can still do that today unless you build something for REALLY specific/obscure CPU instructions.<p>Needless to say that’s not the case with GPUs and a lot of people that make the AMD vs Intel comparison don’t seem to understand that.</p>
]]></description><pubDate>Wed, 30 Oct 2024 16:19:34 +0000</pubDate><link>https://news.ycombinator.com/item?id=41996912</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41996912</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41996912</guid></item><item><title><![CDATA[New comment by kkielhofner in "Strava was used to locate the most powerful people"]]></title><description><![CDATA[
<p>Shouldn't be much of a surprise, this made news back in 2018 when the same was realized with soldiers and secret military bases:<p><a href="https://www.theguardian.com/world/2018/jan/28/fitness-tracking-app-gives-away-location-of-secret-us-army-bases" rel="nofollow">https://www.theguardian.com/world/2018/jan/28/fitness-tracki...</a></p>
]]></description><pubDate>Tue, 29 Oct 2024 21:43:25 +0000</pubDate><link>https://news.ycombinator.com/item?id=41989710</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41989710</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41989710</guid></item><item><title><![CDATA[New comment by kkielhofner in "OpenAI builds first chip with Broadcom and TSMC, scales back foundry ambition"]]></title><description><![CDATA[
<p>For reference seven trillion dollars is 25% of US GDP.<p>Yeah, that's um, wild.</p>
]]></description><pubDate>Tue, 29 Oct 2024 21:38:46 +0000</pubDate><link>https://news.ycombinator.com/item?id=41989682</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41989682</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41989682</guid></item><item><title><![CDATA[New comment by kkielhofner in "People Are Sick and Tired of All Their Subscriptions"]]></title><description><![CDATA[
<p>Living in a climate (Wisconsin) that has extreme highs and lows my understanding is this is typically intended to smooth-out a gas bill (as one example) moving from $10/mo in the summer to $400/mo in the winter.<p>It’s a budgeting thing.</p>
]]></description><pubDate>Mon, 28 Oct 2024 19:59:35 +0000</pubDate><link>https://news.ycombinator.com/item?id=41975529</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41975529</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41975529</guid></item><item><title><![CDATA[New comment by kkielhofner in "GenAI is set to create a mountainous increase in e-waste"]]></title><description><![CDATA[
<p>As one example take a peek at /r/LocalLLaMA[0] (I suspect you know). These people are snapping up anything and everything they can get their hands on at a reasonable price.<p>To your point on the P40, it's an eight year old card but fortunately Nvidia has a history of long term support (especially for "datacenter" GPUs). The Pascal series is still fully supported by the latest Nvidia driver and CUDA releases, and projects like llama.cpp are still fairly regularly adding performance optimizations for even Maxwell series GPUs!<p>Current V100/A100/H100/etc hardware families are not going to end up as e-waste anytime soon. In fact, compare used pricing (and demand) of GPUs to CPUs, RAM, disk, motherboards, etc from eight years ago... That hardware ends up in the trash/at e-waste recyclers much, much sooner (even with /r/homelab).<p>[0] - <a href="https://old.reddit.com/r/LocalLLaMA/" rel="nofollow">https://old.reddit.com/r/LocalLLaMA/</a></p>
]]></description><pubDate>Mon, 28 Oct 2024 19:51:17 +0000</pubDate><link>https://news.ycombinator.com/item?id=41975412</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41975412</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41975412</guid></item><item><title><![CDATA[New comment by kkielhofner in "Dramatic drop in marijuana use among U.S. youth over a decade"]]></title><description><![CDATA[
<p>> If you get wasted on anything, or do anything silly, or act weird, someone will pull out a phone and video you.<p>Anecdotally this seems to be the key impact. With social media almost everyone now has a "brand" and that brand is typically not supported with a post of you out of it, sloppy, etc.<p>Along those lines, there also seems to be MUCH more emphasis on health - granted superficial health (looking good) but health nonetheless. Fortunately the standards for "ideal beauty" for women especially have shifted from the 90s/2000s no-such-thing-as-too-thin dangerous and extremely unhealthy to a physique that is well-muscled and actually healthy (while being inclusive of different body types).<p>When I'm at the gym and the high school/college kids show up I just can't believe their level of physical fitness and development. Self-selecting given it's the gym but when I was in high school (class of 2002) the most fit kid on the football, basketball, track, volleyball, etc teams would look out of shape next to what appears to be the "average" gym-goer of this generation. The numbers also seem to be quite a bit higher - there are A LOT of these kids hitting it really hard in the gym.<p>Needless to say this clearly obsessive-level focus and work is not supported by using drugs like marijuana and alcohol. If nothing else having a lot of followers is much more important and "cool".<p>If anything I'm more interested in usage statistics of steroids and other performance-enhancing drugs. Some of the physiques, performance, etc I see just don't seem possible to achieve naturally at 16-25.</p>
]]></description><pubDate>Mon, 28 Oct 2024 19:10:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=41974889</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41974889</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41974889</guid></item><item><title><![CDATA[New comment by kkielhofner in "Geico repatriates work from the cloud, continues ambitious infra overhaul"]]></title><description><![CDATA[
<p>> That so many people working in software don’t have deep hardware expertise or are not familiar with data centers plays to that hand.<p>I like to remind myself that AWS is 20 years old. That's an entire generation of people from devs to C-Suite that likely don't know anything else. For many of these people all they know about hardware is their laptop. All they know about bandwidth is what they pay their local ISP. All they know about storage is (maybe) USB flash drives and what Apple charges for 256GB vs 512GB.<p>This is not a criticism. More of a reality check to myself and others that at this point not a lot of people outside of bigger cloud and ISPs know what an Autonomous System is or what buying transit and peering costs. Nor have they bought a cabinet of bare metal or talked to a co-lo provider.<p>> Not a criticism, just an observation from my experiences.<p>Exactly!</p>
]]></description><pubDate>Sun, 27 Oct 2024 03:20:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=41959622</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41959622</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41959622</guid></item><item><title><![CDATA[New comment by kkielhofner in "Understanding Round Robin DNS"]]></title><description><![CDATA[
<p>TTL isn't universally respected. Consider the following path:<p>Your machine -> Local router -> Configured upstream DNS Server (ISP/CF/Quad8/etc) -> ? -> Authoritative DNS Server<p>Any one of those layers can override/mess with/cache in a variety of ways including TTL. This is why Cloudflare and a variety of other providers use IP anycast. They accepted DNS for what it is and worked around it.<p>Not only is the IP always the IP, the "global" BGP routing table actually universally and consistently updates much faster than DNS. Then whatever routers, machines, etc downstream from that don't matter.</p>
]]></description><pubDate>Sun, 27 Oct 2024 02:49:52 +0000</pubDate><link>https://news.ycombinator.com/item?id=41959482</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41959482</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41959482</guid></item><item><title><![CDATA[New comment by kkielhofner in "Geico repatriates work from the cloud, continues ambitious infra overhaul"]]></title><description><![CDATA[
<p>> Dell/EMC says "Hey, here is drive replacement." We do it, 2 hours later, the volume is knocked offline. Apparently, there was mismatch between backplane version, drive version and through some weird edge case, it knocked the volume offline. Yes, they fixed it, no it wasn't pretty since a bunch of applications had to be recovered.<p>Anecdotal (as is my position). I can theoretically understand this happening but not only have I never seen it, such an issue would need to be escalated. That's a "this is unacceptable" high-level phone call. A call you more than likely have a chance of someone in actual authority answering because IME unless you have SERIOUS spend with big cloud you'll be lucky to make it a rung or two up sales/support.<p>Plus backups and redundancies that should prevent even the failure of a chassis/storage/etc from being a significant critical issue.<p>> their failures tend to be you twiddling your thumbs vs hair on fire on phone with the vendor trying to get it resolved<p>As a Founder/CTO I have the opposite take - put me and my team in a position to /do something/ vs sitting around waiting for AWS to come back whenever it decides to and while they obscure comms, don't update the fake status dashboards, etc. Meanwhile you're telling your customer "Ummm, we don't know - Amazon has a problem. When it comes back I guess it's back".<p>Coming from a background of telecom, healthcare, and nuclear energy I can't believe that even flies.</p>
]]></description><pubDate>Sat, 26 Oct 2024 00:16:18 +0000</pubDate><link>https://news.ycombinator.com/item?id=41951380</link><dc:creator>kkielhofner</dc:creator><comments>https://news.ycombinator.com/item?id=41951380</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=41951380</guid></item></channel></rss>