<rss version="2.0" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Hacker News: Xyra</title><link>https://news.ycombinator.com/user?id=Xyra</link><description>Hacker News RSS</description><docs>https://hnrss.org/</docs><generator>hnrss v2.1.1</generator><lastBuildDate>Fri, 24 Jul 2026 01:33:49 +0000</lastBuildDate><atom:link href="https://hnrss.org/user?id=Xyra" rel="self" type="application/rss+xml"></atom:link><item><title><![CDATA[New comment by Xyra in "Ask HN: What Are You Working On? (July 2026)"]]></title><description><![CDATA[
<p>making the internet sql queryable, crawling, cleaning, indexing, embedding many sources into source-aware schemas, solving contention problem with free-floating pricing.<p>Currently crawling over 1M records/sec. software is still in alpha.<p>scry.io.<p>goal is a 10PB NVMe cluster online by November (need funding champions) as a public benefit project, so prosocial researchers and builders and their agents can have low-friction access to running analytical queries over the public internet.</p>
]]></description><pubDate>Mon, 13 Jul 2026 21:29:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=48899164</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=48899164</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=48899164</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Hetzner, Postgres, Rust, SvelteKit</p>
]]></description><pubDate>Sat, 03 Jan 2026 03:25:07 +0000</pubDate><link>https://news.ycombinator.com/item?id=46472511</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46472511</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46472511</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>What did you think?</p>
]]></description><pubDate>Sat, 03 Jan 2026 03:19:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=46472479</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46472479</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46472479</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>emailed you, and it's <a href="https://venmo.com/u/XyraSinclair" rel="nofollow">https://venmo.com/u/XyraSinclair</a>.</p>
]]></description><pubDate>Sat, 03 Jan 2026 03:17:58 +0000</pubDate><link>https://news.ycombinator.com/item?id=46472473</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46472473</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46472473</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>We can iterate fast with understanding useful paradigms of vector manipulation. Yesterday I added `debias_vector(axis, topic)` and l2_normalization guidance.</p>
]]></description><pubDate>Thu, 01 Jan 2026 20:28:54 +0000</pubDate><link>https://news.ycombinator.com/item?id=46457711</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46457711</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46457711</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Thank you! I got the idea December 3, and initially released it December 19.</p>
]]></description><pubDate>Thu, 01 Jan 2026 20:11:19 +0000</pubDate><link>https://news.ycombinator.com/item?id=46457566</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46457566</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46457566</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>I'm raising at least $175k and doing a serious startup.</p>
]]></description><pubDate>Thu, 01 Jan 2026 19:39:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=46457285</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46457285</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46457285</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Thanks, that's very interesting.</p>
]]></description><pubDate>Thu, 01 Jan 2026 19:30:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=46457199</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46457199</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46457199</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>~300 token chunks right now. Have other exciting embedding strategies in the works.</p>
]]></description><pubDate>Thu, 01 Jan 2026 19:25:06 +0000</pubDate><link>https://news.ycombinator.com/item?id=46457155</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46457155</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46457155</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>The scale is there. I'm scraping, cleaning, token efficientizing dozens of sources every single hour. The lack of monies for embedding everything was a temporary problem.</p>
]]></description><pubDate>Thu, 01 Jan 2026 03:11:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=46450941</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46450941</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46450941</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>in the direction of "empowering the public with new capabilities they didn't have before", Scry offers, with the copy and paste of a prompt and talking with an agent:<p>1) Full readonly-SQL + vector manipulation in a live public database. Most vector DB products expose a much narrower search API. Basically only a few enterprise level services let you run arbitrary SQL on remote machines. Google BigQuery gives users SQL power, but it mostly doesn't have embeddings, connect public corpora, have as good of indexes, and doesn't have support an agentic research experience. Beyond object-level research, Scry a good tool for exploring and acquiring intuitions about embedding-space.<p>2) An agent-native text-to-SQL + lexical + semantic deep research workflow. We have a prompt that's been heavily optimized for taking full advantage of our machine and Claude Code for exploration and answering nuanced questions. Claude fires off many exploratory queries and builds towards really big queries that lean on the SQL query planner. You can interrupt at any time. You have the compute limits to do lots of exhaustive exploration--often more epistemically powerful than finding a document often, is being confident than one doesn't exist.<p>3) dozens of public commons in one database, with embeddings.</p>
]]></description><pubDate>Thu, 01 Jan 2026 03:01:31 +0000</pubDate><link>https://news.ycombinator.com/item?id=46450880</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46450880</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46450880</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Thank you! I'll be getting millions more quality, embedded documents, it'll be here just getting more useful.</p>
]]></description><pubDate>Thu, 01 Jan 2026 01:27:27 +0000</pubDate><link>https://news.ycombinator.com/item?id=46450330</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46450330</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46450330</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Thank you!</p>
]]></description><pubDate>Thu, 01 Jan 2026 01:25:43 +0000</pubDate><link>https://news.ycombinator.com/item?id=46450315</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46450315</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46450315</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>You submit a SQL query to periodically run, we run it and store the results. As we ingest more documents (dozens of sources are being ingested every day), we run it again. If there's different outputs, you get an email.</p>
]]></description><pubDate>Thu, 01 Jan 2026 01:10:36 +0000</pubDate><link>https://news.ycombinator.com/item?id=46450191</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46450191</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46450191</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Maybe more actually, server costs and API credits for my agent-coordination research are expensive.</p>
]]></description><pubDate>Thu, 01 Jan 2026 01:08:04 +0000</pubDate><link>https://news.ycombinator.com/item?id=46450166</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46450166</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46450166</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Exactly, people want precision and control sometimes. Also it's very hard to beat SQL query planners when you have lots of material views and indexes. Like this is a lot more powerful for most use cases for exploring these documents than if you just had all these documents as json on your local machine and could write whatever python you wanted.<p>Yeah I've out a lot of care into rate-limiting and security. We do AST parsing and block certain joins, and Hacker News has not bricked or overloaded my machine yet--there's actually a lot more bandwidth for people to run expensive queries.<p>As for getting good semantic queries for different domains, one thing Claude can do besides use our embed endpoint to embed arbitrary text as a search vector, is use compositions of centroids (averages) of vectors in our database, as search vectors. Like it can effortlessly average every lesswrong chunk embedding over text mentioning "optimization" and search with that. You can actually ask Claude to run an experiment averaging the "optimization" vectors from different sources, and see what kind of different queries you get when using them on different sources. Then the fun challenge would be figuring out legible vectors that bridge the gap between these different platform's vectors. Maybe there's half the cosine distance when you average the lesswrong "optimization" vector with embed("convex/nonconvex optimization, SGD, loss landscapes, constrained optimization.")</p>
]]></description><pubDate>Thu, 01 Jan 2026 00:14:15 +0000</pubDate><link>https://news.ycombinator.com/item?id=46449765</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46449765</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46449765</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Thank you, I've started ingestion operations of pubmed.</p>
]]></description><pubDate>Wed, 31 Dec 2025 22:04:47 +0000</pubDate><link>https://news.ycombinator.com/item?id=46448840</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46448840</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46448840</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>What is hyperbole? We are collectively experiencing a software intelligence explosion (people are shipping good software at prolific rates now due to Opus 4.5 and GPT-5.2-Codex-xhigh). With Scry, you can run arbitrary SELECT SQL statements over a large corpus and have an easier time composing embedding vectors in whatever mathematical ways you want, than any other tool I've seen.</p>
]]></description><pubDate>Wed, 31 Dec 2025 22:02:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=46448826</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46448826</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46448826</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>I've since improved it, and also discovered a new method of vector composition I have added as a first-class primitive:<p>debias_vector(axis, topic) removes the projection of axis onto topic:
axis − topic * (dot(axis, topic) / dot(topic, topic))<p>That preserves the signal in axis while subtracting only the overlap with topic (not the whole topic). It’s strictly better than naive subtraction for “about X but not Y.”</p>
]]></description><pubDate>Wed, 31 Dec 2025 20:32:24 +0000</pubDate><link>https://news.ycombinator.com/item?id=46447981</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46447981</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46447981</guid></item><item><title><![CDATA[New comment by Xyra in "Show HN: Use Claude Code to Query 600 GB Indexes over Hacker News, ArXiv, etc."]]></title><description><![CDATA[
<p>Yes, thanks for explaining it.</p>
]]></description><pubDate>Wed, 31 Dec 2025 20:24:29 +0000</pubDate><link>https://news.ycombinator.com/item?id=46447918</link><dc:creator>Xyra</dc:creator><comments>https://news.ycombinator.com/item?id=46447918</comments><guid isPermaLink="false">https://news.ycombinator.com/item?id=46447918</guid></item></channel></rss>