{"author":"GeorgeCurtis","children":[{"author":"brene","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T16:09:52.000Z","created_at_i":1781107792,"id":48478475,"options":[],"parent_id":48478409,"points":null,"story_id":48478148,"text":"We see comparable results for vectors and FTS.<p>For vector search we have warm and cold p99s of approx 20ms and 400ms respectively.\nFor FTS, warm and cold query p99s of approx 15ms and 250ms respectively.<p>Both of these benchmarks were run on 1m docs.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:05:04.000Z","created_at_i":1781107504,"id":48478409,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"How does this compare vs. Turbopuffer?","title":null,"type":"comment","url":null},{"author":"mentioum","children":[{"author":"GeorgeCurtis","children":[{"author":"mentioum","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T19:06:23.000Z","created_at_i":1781118383,"id":48481120,"options":[],"parent_id":48479451,"points":null,"story_id":48478148,"text":"Sure! You can email me personally at george@helix-db.com","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T17:11:38.000Z","created_at_i":1781111498,"id":48479451,"options":[],"parent_id":48478429,"points":null,"story_id":48478148,"text":"Hmmm... I&#x27;ll get in touch.  Got an email i can reach out to, there doesn&#x27;t seem to be one listed on your website?<p>I&#x27;m more concerned about if the p99s stay consistent when things get spikey.<p>dgraph is fine otherwise...","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:06:27.000Z","created_at_i":1781107587,"id":48478429,"options":[],"parent_id":48478411,"points":null,"story_id":48478148,"text":"In prod we see p99\u2019s of &lt;10ms ms for warm queries and around 50ms per hop for cold queries.","title":null,"type":"comment","url":null},{"author":"zw17","children":[{"author":"GeorgeCurtis","children":[{"author":"jauntywundrkind","children":[{"author":"fouc","children":[],"created_at":"2026-06-11T07:25:18.000Z","created_at_i":1781162718,"id":48487367,"options":[],"parent_id":48480889,"points":null,"story_id":48478148,"text":"where is puppygraph&#x27;s source code?","title":null,"type":"comment","url":null},{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-11T12:41:05.000Z","created_at_i":1781181665,"id":48489564,"options":[],"parent_id":48480889,"points":null,"story_id":48478148,"text":"puppy graph is not open source","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T18:47:30.000Z","created_at_i":1781117250,"id":48480889,"options":[],"parent_id":48479246,"points":null,"story_id":48478148,"text":"And is open source.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:58:57.000Z","created_at_i":1781110737,"id":48479246,"options":[],"parent_id":48478801,"points":null,"story_id":48478148,"text":"PuppyGraph is a good fit for OLAP for sure.<p>We\u2019re just two young founders sharing what we\u2019ve been building, so I\u2019ll take the drive-by competitor plug as a compliment :)<p>Definitely a different focus though. Helix is OLTP, built for operational graph + vector workloads, especially apps&#x2F;agent memory where low-latency traversals and writes are concerned.","title":null,"type":"comment","url":null},{"author":"mentioum","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T17:13:59.000Z","created_at_i":1781111639,"id":48479486,"options":[],"parent_id":48479390,"points":null,"story_id":48478148,"text":"This sounds like a perfect usecase. Would love to learn more and see if we can help!<p>email us: founders@helix-db.com","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T17:08:11.000Z","created_at_i":1781111291,"id":48479390,"options":[],"parent_id":48478801,"points":null,"story_id":48478148,"text":"It&#x27;s not, its actually our prod db with direct user usage - we self host a large dgraph cluster.  We have a very large number of people manage their car and car histories with us and host a full replica of the UK MOT Database.<p>We&#x27;re fine with clickhouse and redshift for the OLAP work we do.  I&#x27;ve been looking at ParaQuery lately if I really want to speed that up.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:31:03.000Z","created_at_i":1781109063,"id":48478801,"options":[],"parent_id":48478411,"points":null,"story_id":48478148,"text":"If your use case is OLAP based, please check it out PuppyGraph. It\u2019s a graph query engine that sits on top of your Lakehouse (no ETL required). Our benchmark has shown consistently that 10-hop queries across billions of edges in &lt;2 seconds. Our customers including some most data demanding companies like Coinbase, Datadog, Palo Alto Network, Netskope, AMD, etc.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:05:08.000Z","created_at_i":1781107508,"id":48478411,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"We&#x27;ve been having some issues with intermittent performance on multi hop queries.<p>What&#x27;s your p99 like for multi hops?","title":null,"type":"comment","url":null},{"author":"maxrumpf","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T16:20:16.000Z","created_at_i":1781108416,"id":48478638,"options":[],"parent_id":48478581,"points":null,"story_id":48478148,"text":"Yes you can put vectors, full text data, secondary and range indexes on both nodes and edges.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:16:17.000Z","created_at_i":1781108177,"id":48478581,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"does it support fts&#x2F;vector on edges of the graph?","title":null,"type":"comment","url":null},{"author":"raufakdemir","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T16:49:50.000Z","created_at_i":1781110190,"id":48479097,"options":[],"parent_id":48478985,"points":null,"story_id":48478148,"text":"We don&#x27;t support cypher or gremlin. We can<p>You can query HelixDB using JSON or directly in your programming language of choice by using our Rust, TypeScript, Go or Python SDKs. \nWe\u2019ve found AI is very good at working with the SDKs and JSON itself to query, making the development experience much better than before: <a href=\"https:&#x2F;&#x2F;docs.helix-db.com&#x2F;database&#x2F;querying\">https:&#x2F;&#x2F;docs.helix-db.com&#x2F;database&#x2F;querying</a>","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T16:42:54.000Z","created_at_i":1781109774,"id":48478985,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"what language does this support? cypher&#x2F;gremlin?","title":null,"type":"comment","url":null},{"author":"Bnjoroge","children":[{"author":"GeorgeCurtis","children":[{"author":"Bnjoroge","children":[],"created_at":"2026-06-11T14:08:03.000Z","created_at_i":1781186883,"id":48490584,"options":[],"parent_id":48481028,"points":null,"story_id":48478148,"text":"it does, think I misunderstood your value prop. best of luck -definitely a real usecase.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T18:59:20.000Z","created_at_i":1781117960,"id":48481028,"options":[],"parent_id":48480409,"points":null,"story_id":48478148,"text":"tpuffer is a vector&#x2F;fts database. Surreal is a bit of an &quot;everything database&quot;.<p>We&#x27;re a graph database with vector and FTS capabilities. Our vector and FTS benchmarks are comparable with tpuffer, but you would primarily use us for building whole applications, knowledge graphs, or AI memory&#x2F;retrieval. Anything that is relationship intense.<p>Let me know if this properly answers your question","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T18:15:23.000Z","created_at_i":1781115323,"id":48480409,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"congrats! how does this compare to turbopuffer, surreal or other multi-model ones built on object storage or not","title":null,"type":"comment","url":null},{"author":"cjlm","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T18:47:14.000Z","created_at_i":1781117234,"id":48480887,"options":[],"parent_id":48480822,"points":null,"story_id":48478148,"text":"yooo this is awesome. Didn&#x27;t even realise :)","title":null,"type":"comment","url":null},{"author":"dig1","children":[{"author":"cjlm","children":[],"created_at":"2026-06-10T21:14:13.000Z","created_at_i":1781126053,"id":48482809,"options":[],"parent_id":48481901,"points":null,"story_id":48478148,"text":"The feature scores are from a paper cited on the site. The ranking features are proprietary to dissuade gaming but are a blend of open public data across a number of different sources.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T20:08:12.000Z","created_at_i":1781122092,"id":48481901,"options":[],"parent_id":48480822,"points":null,"story_id":48478148,"text":"How rankings are calculated? HelixDB has zero feature scores, unlike e.g. JanusGraph which is at #27 or Neo4j at #1.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T18:42:40.000Z","created_at_i":1781116960,"id":48480822,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"Currently #5 on gdb-engines.com - definitely worth a look.","title":null,"type":"comment","url":null},{"author":"rajit","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-10T19:04:00.000Z","created_at_i":1781118240,"id":48481089,"options":[],"parent_id":48480891,"points":null,"story_id":48478148,"text":"We plan on launching end of month.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T18:47:32.000Z","created_at_i":1781117252,"id":48480891,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"when will the graph memory layer be available?","title":null,"type":"comment","url":null},{"author":"caust1c","children":[{"author":"GeorgeCurtis","children":[{"author":"tao_oat","children":[],"created_at":"2026-06-11T08:37:13.000Z","created_at_i":1781167033,"id":48487822,"options":[],"parent_id":48481678,"points":null,"story_id":48478148,"text":"Unfortunately it&#x27;s not possible to read this without an X account.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T19:50:53.000Z","created_at_i":1781121053,"id":48481678,"options":[],"parent_id":48481629,"points":null,"story_id":48478148,"text":"This was a TEMPORARY decision we made, and I  wrote a bit about why we did this here: <a href=\"https:&#x2F;&#x2F;x.com&#x2F;georgecurtiss&#x2F;status&#x2F;2060043184059912470\" rel=\"nofollow\">https:&#x2F;&#x2F;x.com&#x2F;georgecurtiss&#x2F;status&#x2F;2060043184059912470</a><p>We\u2019re 100% committed to going back to open-source on an Apache 2.0 license as soon as possible. In the meantime, you can continue to deploy us completely for free, however you like, using the compiled docker container.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T19:46:21.000Z","created_at_i":1781120781,"id":48481629,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"Where&#x27;s the source code for the database itself?  Looks like the repo is just a client.<p>Congrats on the launch!","title":null,"type":"comment","url":null},{"author":"ymir_e","children":[],"created_at":"2026-06-10T19:53:14.000Z","created_at_i":1781121194,"id":48481696,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"Congrats on the launch George!<p>Looking forward to looking into the generalised AI memory layer when it comes out.","title":null,"type":"comment","url":null},{"author":"NexoraDev","children":[],"created_at":"2026-06-10T21:15:52.000Z","created_at_i":1781126152,"id":48482830,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"bellissimo","title":null,"type":"comment","url":null},{"author":"jesol","children":[],"created_at":"2026-06-10T21:29:59.000Z","created_at_i":1781126999,"id":48482985,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"I&#x27;ve been working on a graph database in Rust this year actually! I&#x27;d love to hear anything you can talk about wrt the query planner and&#x2F;or how you decided to do cardinality estimation. I decided to go with an EAV graph which makes CE pretty complex, and it&#x27;s been an interesting challenge to balance quality and speed and expressiveness in the query language","title":null,"type":"comment","url":null},{"author":"rgbrgb","children":[{"author":"GeorgeCurtis","children":[{"author":"Onawa","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-11T05:01:22.000Z","created_at_i":1781154082,"id":48486410,"options":[],"parent_id":48485979,"points":null,"story_id":48478148,"text":"Lapse of communication. For now as in you\u2019ll be able to host it without reading the source code.<p>Soon you\u2019ll be able to host it yourself AND have access to the source code","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T03:46:34.000Z","created_at_i":1781149594,"id":48485979,"options":[],"parent_id":48484498,"points":null,"story_id":48478148,"text":"Why do you say &quot;for now&quot;?","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T00:01:00.000Z","created_at_i":1781136060,"id":48484498,"options":[],"parent_id":48483348,"points":null,"story_id":48478148,"text":"You can definitely host it for free locally for now.<p>We aim to launch our GA cloud at the end of this month, which will be much more affordable.","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T22:04:37.000Z","created_at_i":1781129077,"id":48483348,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"congrats on the launch! site and docs look great.<p>can you host this yourself or do you need to use helix-cloud? the chat thing on the side seems to push me to helix-cloud but it looks like that starts at like $600&#x2F;mo which is above my experimentation budget.<p>looking for a db for an agent memory application and i&#x27;d probably start with something that&#x27;s just self-hosted &#x2F; freeish. postgres is working ok but I want to start ingesting server and chat logs.","title":null,"type":"comment","url":null},{"author":"thedreammachine","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-11T13:27:31.000Z","created_at_i":1781184451,"id":48490071,"options":[],"parent_id":48485416,"points":null,"story_id":48478148,"text":"OLAP queries, deep multi-hops where latency is a priority.<p>As long as the sub-graph you&#x27;re trying to hop is cached, then there&#x27;s no problem or latency issues. However, if you need to do a deep hop query, where all those nodes and edges are in cold storage, each hop costs ~50ms. So a 10-hop would take ~0.5 seconds.<p>Again though, we find most people are using us for agentic workloads, so even this worst case scenario the LLMs make up the majority of the latency.","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T02:05:43.000Z","created_at_i":1781143543,"id":48485416,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"What kinds of graph shapes or query patterns do you feel are the worst case for object storage?","title":null,"type":"comment","url":null},{"author":"let_rec","children":[{"author":"GeorgeCurtis","children":[{"author":"aitchnyu","children":[{"author":"GeorgeCurtis","children":[{"author":"aitchnyu","children":[{"author":"let_rec","children":[],"created_at":"2026-06-17T14:51:20.000Z","created_at_i":1781707880,"id":48571360,"options":[],"parent_id":48567550,"points":null,"story_id":48478148,"text":"Probably? Other clouds implement the S3 API","title":null,"type":"comment","url":null}],"created_at":"2026-06-17T08:44:55.000Z","created_at_i":1781685895,"id":48567550,"options":[],"parent_id":48556344,"points":null,"story_id":48478148,"text":"I mean object stores from other clouds.","title":null,"type":"comment","url":null}],"created_at":"2026-06-16T15:01:08.000Z","created_at_i":1781622068,"id":48556344,"options":[],"parent_id":48494192,"points":null,"story_id":48478148,"text":"No, you can run it in-memory or on-disk too","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T18:10:14.000Z","created_at_i":1781201414,"id":48494192,"options":[],"parent_id":48489527,"points":null,"story_id":48478148,"text":"Does it work only on S3?","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T12:38:05.000Z","created_at_i":1781181485,"id":48489527,"options":[],"parent_id":48487964,"points":null,"story_id":48478148,"text":"Can&#x27;t imagine why they&#x27;d be hesitant, Helix is awesome, we&#x27;ve never had any data loss issues, and are completely ACID.<p>I&#x27;d encourage them to start a local instance with claude&#x2F;codex to build a mini project and see what it&#x27;s like.","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T08:58:57.000Z","created_at_i":1781168337,"id":48487964,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"This seems like a great idea.<p>What reassurance can you offer devs that are hesitant to try a new data-store?","title":null,"type":"comment","url":null},{"author":"lennertjansen","children":[{"author":"GeorgeCurtis","children":[],"created_at":"2026-06-16T15:44:23.000Z","created_at_i":1781624663,"id":48557067,"options":[],"parent_id":48494231,"points":null,"story_id":48478148,"text":"yes, and yes.<p><a href=\"https:&#x2F;&#x2F;docs.helix-db.com&#x2F;database&#x2F;multi-tenancy\">https:&#x2F;&#x2F;docs.helix-db.com&#x2F;database&#x2F;multi-tenancy</a>","title":null,"type":"comment","url":null}],"created_at":"2026-06-11T18:12:35.000Z","created_at_i":1781201555,"id":48494231,"options":[],"parent_id":48478148,"points":null,"story_id":48478148,"text":"Are you ACID? And does this version have multi-tenancy?","title":null,"type":"comment","url":null}],"created_at":"2026-06-10T15:47:31.000Z","created_at_i":1781106451,"id":48478148,"options":[],"parent_id":null,"points":159,"story_id":48478148,"text":"Hey HN, it\u2019s been just over a year since we launched HelixDB (<a href=\"https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=43975423\">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=43975423</a>), a project a friend and I started in college. It\u2019s an OLTP graph database built on object-storage, with native vector search and full-text search (FTS).<p>Why graph, vector and FTS? Graph databases provide a natural cognitive model for data, vectors allow for a semantic understanding of the entities and relationships in the graph, and FTS provides more specific filtering. Many AI-driven applications attempt to combine all of these functionalities by stitching together multiple disconnected systems, but even then there\u2019s no native way to perform joins or queries that span all systems. You still need to handle this logic at the application level.<p>Helix started as a graph DB, but we moved to a hybrid graph&#x2F;vector approach after attempting to build an AI memory system, which led us down the GraphRAG and HybridRAG rabbit hole, where we would need separate graph and vector databases.<p>We knew scalability would be a challenge at each stage of our product&#x27;s development, however our initial focus this past year was to prove out the product through local deployments and was only meant to be run on a single node. Scaling graph DBs remained a difficult and expensive problem we\u2019d have to solve later.\nSome common ways other graph DBs solve scaling is by duplicating entire datasets across distributed machines (extremely expensive per node), or by sharding the data.<p>Sharding databases is effective and affordable, however, graph data doesn\u2019t have explicit partitions like relational databases do. For example, sharding a relational DB involves splitting up tables. When it comes to graph DBs, the edges can span across any of the partitions, and hopping across multiple machines when traversing nodes is ineffective and computationally expensive.<p>Replicating graph DBs for high availability and better throughput drastically increases the operational cost of the db and still has a limit of how big you can vertically scale. The workload that we\u2019re used for requires storing a huge amount of data for agents, where only a subset of that data is ever needed at any one time. So rather than having the whole thing in memory, we can store it all in object-storage and get the bits we need when they\u2019re needed.<p>Agents benefit from better context, which is achieved from more and better data (more relationships etc). By using S3 as the persistence&#x2F;data layer there is <i>no limit</i> to how big the graph can be or how many relationships you can have, and we can scale to serve throughput and requests by horizontally spinning up nodes and caching relevant subsets of the graph on each node. This way, you get extremely low latency for \u201chot\u201d data and a p99 of ~100ms for writes and ~50ms for reads from cold storage (S3). Plus you get the benefit of dirt cheap storage.<p>Workloads that HelixDB is currently supporting:\n- Huge amounts of data (TBs) from which the agents need to search and traverse over\n- Offering affordable graph storage for companies where cost of graph data is a bottleneck\n- Consolidating multiple databases, enabling AI agents to have autonomy over companies, helping them become more autonomous.\n- AI memory\n- Company brains<p>We\u2019re currently working on our own generalised AI memory layer which will use HelixDB under the hood and be completely open-source. Also, we\u2019re finishing up on pre-filtering for vector search which will allow you to pre-filter based on relationships in the graph, metadata, and sub-graphs. And lastly, GA cloud will be available in the coming weeks.<p>If you want to run Helix locally (either on-disk or in-memory), you can find more info on our github (<a href=\"https:&#x2F;&#x2F;github.com&#x2F;HelixDB&#x2F;helix-db\" rel=\"nofollow\">https:&#x2F;&#x2F;github.com&#x2F;HelixDB&#x2F;helix-db</a>) or via our docs (<a href=\"https:&#x2F;&#x2F;docs.helix-db.com&#x2F;database&#x2F;local-development\">https:&#x2F;&#x2F;docs.helix-db.com&#x2F;database&#x2F;local-development</a>). If you\u2019re interested in getting started with our distributed cloud, please email us founders@helix-db.com.<p>Many thanks! Comments and feedback welcome!","title":"Show HN: HelixDB \u2013 A graph database built on object storage","type":"story","url":"https://github.com/HelixDB/helix-db/tree/main"}
