Category Β· 18 repos
Vector Databases
Databases and indexes built for embeddings and similarity search. Ranked by star velocity over the last 24 hours.
Not another ChatGPT wrapper. This is a local-first tool that turns your notes, web pages, PDFs and audio/video into a personal knowledge graph β running entirely on your own machine.
The Postgres development platform. Supabase gives you a dedicated Postgres database to build your web, mobile, and AI applications.
π PageIndex: Document Index for Vectorless, Reasoning-based RAG
World's first open-source enterprise world model.
MinHash, LSH, LSH Forest, Weighted MinHash, HyperLogLog, HyperLogLog++, LSH Ensemble and HNSW
Dealing with all unstructured data, such as reverse image search, audio search, molecular search, video analysis, question and answer systems, NLP, etc.
All-in-one Zotero search plugin - parallel academic search across 13 sources, offline journal quality badges (IF, JCR/CAS quartiles, warning & predatory lists), one-click import with open-access PDFs, citation traversal, and in-library vector search. MIT.
A complete serverless research paper ingestion and search system built on AWS. The system fetches papers from arXiv, generates embeddings using Amazon Titan, stores them in PostgreSQL with pgvector, and provides semantic search with AI-powered chat responses using OpenAI models via AWS Bedrock.
Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search
Postgres with GPUs for ML/AI apps.
Local-first AI job intelligence workbench for scraping roles, ranking fit, and generating tailored application materials.
Build realtime AI voice agents using FastRTC for low-latency streaming, Superlinked for vector search, Twilio for live phone calls, and Runpod for scalable GPU deployment.
Official code for "Retrieval Over Training: Similarity search-based Model Selection for Time Series Anomaly Detection" (NeurIPS 2026). RAMSAD selects the best anomaly detector for a new series by retrieving similar series from a knowledge base, with no selector training.
Deeplake is AI Data Runtime for Agents. It provides serverless postgres with a multimodal datalake, enabling scalable retrieval and training.
A library for efficient similarity search and clustering of dense vectors.
One Postgres for your application data, full-text search, vector retrieval, and aggregations. Home of the pg_search extension.
Enterprise AI compliance platform for retrieval-augmented regulatory verification, human-in-the-loop review workflows, evidence management and audit-ready reporting built with FastAPI, PostgreSQL,and Qdrant.
Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory with small models for free