databendlabs/databend
quality grade B, 79 out of 100Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. — rebuilt from scratch. Unified architecture on your S3.
- stars
- 9.4k
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Storage and retrieval for embeddings: vector indexes, hybrid search and ANN libraries.
Signals: vector-database, vector-search, vector-store, similarity-search, approximate-nearest-neighbor, ann-search, faiss, hnsw
195 results
Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. — rebuilt from scratch. Unified architecture on your S3.
Semantic search over videos using Gemini Embedding 2 or Qwen3-VL.
🌌 A complete search engine and RAG pipeline in your browser, server or edge network with support for full-text, vector, and hybrid search in less than 2kb.
[MLsys2026]: RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal device.
Open Source alternative to Algolia + Pinecone and an Easier-to-Use alternative to ElasticSearch ⚡ 🔍 ✨ Fast, typo tolerant, in-memory fuzzy Search Engine for building delightful search experiences
A 7-layer memory operating system for Hermes Agent — persistent memory with Qdrant, structured facts, fabric recall, auto-curated wiki, and surgical context injection. Runs locally, any LLM provider.
Jupyter Notebooks to help you get hands-on with Pinecone vector databases
Lite & Super-fast re-ranking for your search & retrieval pipelines. Supports SoTA Listwise and Pairwise reranking based on LLMs and cross-encoders and more. Created by Prithivi Da, open for PRs & Collaborations.
A curated list of awesome works related to high dimensional structure/vector search & database
The codebase for the book "AI-Powered Search" (Manning Publications, 2025) and associated "AI-Powered Search: Modern Retrieval for Humans & Agents" Maven course
Optimized Agentic and LLM Bulk Processing Over Your Data
A multi-modal vector database that supports upserts and vector queries using unified SQL (MySQL-Compatible) on structured and unstructured data, while meeting the requirements of high concurrency and ultra-low latency.
A modern desktop application for exploring, managing, and analyzing vector databases
[CoRL25] GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data
FloTorch is an open-source tool for optimizing Generative AI workloads on AWS. It automates RAG proof-of-concept development with features like hyperparameter tuning, vector database optimization, and LLM integration. FloTorch streamlines experimentation, ensures security, and accelerates production with cost-efficient, validated workflows.
Cosdata: A cutting-edge AI data platform for next-gen search pipelines. Features semantic search, hybrid capabilities, real-time scalability, and ML integration. Designed for immutability and version control to enhance AI projects.
Picky is an easy to use and fast Ruby semantic search engine that helps your users find what they are looking for.
Local semantic search. Stupidly simple.
基于向量数据库与GPT3.5的通用本地知识库方案(A universal local knowledge base solution based on vector database and GPT3.5)
A hyper-fast local vector database for use with LLM Agents. Now accepting SAFEs at $135M cap.
Your filesystem as a vector database
A transactional, relational-graph-vector database that uses Datalog for query. The hippocampus for AI!
PostgreSQL vector database extension for building AI applications
fastest vector database made in numpy
24,523 repositories in the index in total.