r/Rag
Migrate the embeddings model or the Database infrastructure
- upvotes
- 5
- comments
- 9
Post
Highlighted: the lines this signal was extracted from
Hey! Im usually active on these feeds but never comment but this problems made me stress out. I've built a rag across 2000 documents for my company right now and i was originally using the Text-embeddings-3-model from OpenAI and realized it wasn't able to gather nuanced contexts like images so i decided to migrate to the Gemini multi modal. Our current DB runs on supabase and uses pg vector as our vector db. Currently we use HNSW search + BM 25 in our search algorithm and hit a constraint during migration as Geminis vectors are bigger than the 2000 limit we get using postgreSQL. We can either use truncated vectors and accept some loss of information or migrate to a cohereV4 multimodal tech that fits in the vector constraints allowed. My coworker wants to migrate our entire DB to something like pinecone but something tells me migrating Databases for something like this isn't worth doing. (We aren't in production for other users yet and have a small blast radius). I would love to know your suggestions!