6 ms·
Pinecone Open-Sources AWS Reference Architecture with Pulumi
- manojlds 3y agoSounds silly to go to a vector DB vendor to get reference architecture for my app where it will only be a small part of my app.
- CharlesW 3y agoI believe the point is that you can leverage the reference architecture for that small part of your app.
- zackproser 3y agoYes - this and to make it easier to understand how you'd use Pinecone at scale. We get a lot of feedback that our many open-source Jupyter notebooks (github.com/pinecone-io/examples) are very helpful for learning new techniques and understanding how patterns work - even for starting new applications. However, we're also often asked how to go from Notebook to prod. The Reference Architecture is a step in the direction of making this easier and more clear.
- whalesalad 3y agoThe massive (and contextually irrelevant) AI generated images every other paragraph remind me of setting line spacing to 2.2 in order to hit the minimum number of pages for your school essay.
- zackproser 3y agoThanks for the feedback!
- isoprophlex 3y agoThe massive paragraph heading size compared to the text itself doesn't do the readability a whole lot of good, either.
- cdchn 3y agoMost of the text felt AI generated as well.
- zackproser 3y agoInteresting! I can assure you none of the text was auto-generated. I've written before about how I don't like having text generated for me. I do use Grammarly for editing assistance when I'm writing something for work.
- _a_a_a_ 3y agoAppreciated, but 16+ MB of images - there was no need given ~4K of text. That's a 4,000 times bloat. Please, professional devs don't need this. Just give us the info.
- manvillej 3y agoI do want code examples with the pretty dark syntax highlighting though
- zackproser 3y agoI appreciate the feedback, but we may need to agree to disagree on this one :) As a professional dev, I enjoy articles with eye candy and AI-generated images, especially those with a pixel-art bent make me happy.
- whalesalad 3y agoImages like this can be used to break up content, but the ratio is way off. Also, the images are not full width to the content, so it creates a zig-zag noise issue when reading down the page. I would left align them, or make them full width. The images should be used to break up logical sections, but in this case they truly just look scattered randomly and just cause frustration. There is really not enough text to warrant this many (unrelated) images. The images themselves are cool though. Kinda reminds me of the way that baby snakes are more dangerous than older ones - they are not sure how to regulate their venom so they end up using too much and it is more lethal. Save some of these images for your next blog posts =) Don't shoot them all off at once.
- paulddraper 3y agoDon't forget to bump up those page margins.
- quickthrower2 3y agoThey are the new cat images, borat jokes and rick rolling.
- ryoshu 3y agoThe isometric architecture diagram is hard to read as well.
- candiddevmike 3y agoAs someone interested in kicking the tires of VectorDBs, where does Pinecone rank? Is there one that will be "future proof"?
- federationfive 3y agoRedis is probably better suited
- gk1 3y agoI work for Pinecone so my bias is obvious, but objectively: Pinecone is ranked as the most popular vector DB according to multiple sources [1][2][3][4], the best funded ($100M Series B), and is used in production by the likes of Shopify, Gong, Plaid, Zapier, Midjourney, CVS, and thousands of others. [1] https://retool.com/reports/state-of-ai-2023 https://retool.com/reports/state-of-ai-2023 [2] https://state-of-llm.streamlit.app/#third https://state-of-llm.streamlit.app/#third [3] https://db-engines.com/en/ranking/vector+dbms https://db-engines.com/en/ranking/vector+dbms [4] https://www.g2.com/categories/vector-database https://www.g2.com/categories/vector-database
- dmezzetti 3y agoCouple relatively recent HN threads that give a good overview of the vector database landscape. https://news.ycombinator.com/item?id=36943318 https://news.ycombinator.com/item?id=36943318 https://news.ycombinator.com/item?id=38416994 https://news.ycombinator.com/item?id=38416994 https://news.ycombinator.com/item?id=38420554 https://news.ycombinator.com/item?id=38420554
- fzliu 3y agoRegarding "future-proof-ness": we've been building production-grade vector search since 2018 and have a number of organizations running it at billion+ scale in production environments. It's all open source too. https://milvus.io https://milvus.io
- redwood 3y agoAnyone have good experience with Pulumi? The IaC space feels a bit crowded
- zackproser 3y agoI've worked with Terraform extensively, including doing a lot of similar work (Reference Architectures, modules, AWS) over the past several years. This was my first time using Pulumi, and it was very much a delight. There's clearly been a lot of thought and effort put into the plugins and overall developer experience. For example, you can have Pulumi handle the Docker image builds for you - as well as the ECR repository logins and pushes. This means that your end users don't need to manually build images, log into ECR and push them - they just run `pulumi up` instead.
- rtuin 3y agoMy setup extends on this: the pulumi stack creates ECR repo, IAM user+access token, adds credentials and ECR details to GitHub actions secrets, and GitHub actions builds tags and pushes the images to ECR when a release is (automatically) tagged. Pulumi is (mostly) a bliss!
- Jemaclus 3y agoI've used Pulumi for professional projects and personal projects, and I think it's fantastic. You can write your code in pretty much any major language (professionally, i use Python, personally I use Go), and you get the same output. Running the Pulumi program itself couldn't be simpler. My one critique is sort of weird in the sense that Pulumi supports way more than its public documentation says. I hope they up their documentation game in the future, because there are a lot of things that I had to go digging through source code to find out how to implement. But that's really a minor nitpick in the grand scheme of things. Big Pulumi fan here.
- quickthrower2 3y agoI agree! Looking at the source code is invaluable. Also if you use Azure, look at the generated ARM templates, for clues on how to set things up where the docs are short on details. It is nice they don't abstract over the ARM templates too much.
- next_xibalba 3y agoPinecone’s aggressive marketing here on HN is getting really old.
- sabareesh 3y agoOther day i was trying out Azure SQL for storing vector DB. After a POC i was able to get results on sub 1 second to get the results with 2 core serverless instance. It begs me to reconsider does dedicated Vector db really worth it ? Here is the article on how to store and query vector data https://devblogs.microsoft.com/azure-sql/vector-similarity-search-with-azure-sql-database-and-openai/ https://devblogs.microsoft.com/azure-sql/vector-similarity-s...
- marcinzm 3y agoThe article says you had 500ms latency while searching through only 25k records. Presumably at 1 query at a time? That's probably 2 orders of magnitude slower than a vector database.
- beoberha 3y agoAsk yourself the same question about an OLAP system. Once you’re at a large enough scale, it makes sense to use tools custom designed for your access patterns. I work on Azure SQL, but it’s pretty clear that this is a cool demo. It may scale for some usages, but I wouldn’t build a production system on it.
- sabareesh 3y agoWell tried it with ~ 34 million records , which is good enough sample for our usecase
- marcinzm 3y agoAt what number of queries per second?
- bob1029 3y ago> Azure SQL > large enough scale You mean something like the Hyperscale service tier? This is what we are building our new products on top of.
- 3y ago
- Dowwie 3y agoNow THAT is developer advocacy-- a reference architecture and a complete video series. I'm intrigued by the fact that Typescript was chosen for this work. Are you finding that clients are rolling out ML architectures based on Typescript microservices?
- zackproser 3y agoThanks so much for your support We are finding that JavaScript is an often overlooked language for working in AI: https://www.pinecone.io/learn/javascript-ai/ https://www.pinecone.io/learn/javascript-ai/ There's also a very nice synergy between using a language like JavaScript (many are familiar enough to read through and understand what's happening) and the type safety that TypeScript introduces. We are seeing a lot of interest in working with Typescript hence why many of our examples are TypeScript applications.
- dvfjsdhgfv 3y agoWhy is this on the main page? Just a few paragraphs of text with autogenerated images, and the contents feels like LLM-generated, too.
- zackproser 3y agoThanks for your question! The images were generated via DALL-E - so they're not technically "autogenerated". They were created using a generative model, however. The content itself definitely was not generated by an LLM - although that feels like a low-effort comment =/ What in particular felt LLM-generated to you? I think this is on the main page because being able to set 3 environment variables and run `pulumi up` to get to a production-ready system in your own AWS account, without having to purchase anything, is a massive time-saver for anyone with a high-scale use cases that wants to use Pinecone's vector database.