Vector Watcher is a cross-platform desktop GUI for exploring and managing LanceDB vector databases. Built with Tauri, React and TypeScript over a bundled Python/FastAPI sidecar, it offers connection management, schema inspection and similarity search across local or remote instances.
Open source. Open possibilities.
Discover quality open-source projects, submit projects anonymously, and claim and edit your own project.
A little curiosity. A world of open source.
THE FIRST COLLECTIONLess than 10 stars, 5+ commits in the last 30 days — projects made with care.
MCP Context Server is a FastMCP-based server providing persistent multimodal context storage for LLM agents, featuring full-text, semantic, and hybrid search with cross-encoder reranking, thread-based scoping, and support for SQLite and PostgreSQL backends.
This open-source collection features AI Agent books, tutorials, and code repositories from GitHub, covering LLM Agent resources, courses, and frameworks. It supports daily automatic Star count updates, sorting by popularity, and provides recommended reading paths for developers and learners.
Lemory is a local middleware that transforms markdown notes (e.g., Obsidian vaults) into a context database for AI agents. It supports keyless on-device hybrid search, Korean-specific processing, hierarchical token savings, and knowledge graph visualization, integrating with various AI tools via MCP.
by2kb converts video URLs from platforms like Bilibili and YouTube into durable Markdown knowledge base artifacts, including transcripts, abstracts, and study notes, supporting agent-first workflows and local/cloud ASR.
TalaDB is an embedded vector and document database designed for on-device AI, providing a unified TypeScript API across Browser, Node.js, and React Native.
Zotero CLI is an English-language command-line tool for managing Zotero libraries and supporting rigorous systematic literature reviews. It combines Zotero item and collection operations, multi-source import, interactive screening, audit trails, citation snowballing, data extraction, PRISMA reporting, and optional local RAG/LLM-backed semantic search.
Utility OS is a Next.js academic workspace for university students with Google Drive-synced notes, hybrid RAG AI search with grounded citations, a study planner, GPA calculator, SRS flashcards, Pomodoro timer, and a PWA client.
Web of Flaws is a structured, markdown-first catalog of web security vulnerabilities and secure code replacements designed for developers, security engineers, and AI coding agents.
Python SDK for the AskNews API that injects real-time news context into any LLM with a single line of code, offering endpoints for news search, event tracking, forecasting, analytics, deep research, knowledge graphs, and web search.
A curated collection of 107 foundational generative AI research papers with comprehensive summaries, learning roadmaps, glossaries, and decision guides—making cutting-edge AI research accessible to everyone.
OpenSelf is a local-first, open-source context layer for AI agents. It stores source-attributed, time-aware memories in SQLite and exposes them via MCP or JavaScript API, with optional encryption and privacy controls.
A community-maintained catalog of Bangla (Bengali) NLP resources — 712 papers, 63 datasets, 20 models, and 9 tools across 13 tasks — built as a static Astro site with strict data verification and shareable filters.
AI Engineering is an open-source e-book providing a systematic guide to building production-grade AI systems, from LLM fundamentals to advanced agent architectures. Maintained by the community and Hermes Agent, it features daily updates, traceable sources, and advanced search capabilities.
Starbase is a web-based database and toolkit for exploring Starship transposable elements in fungi, offering sequence browsing, BLAST/HMMER search, submission, and visualization features.
npm/PyPI packages and an MCP server for querying the Brazilian BNCC curriculum. Bundled data with 1,721 verified learnings from MEC/CNE, zero runtime dependencies. Integrates BNCC into TypeScript, Python, or AI agents via MCP.
An unofficial MCP server that gives AI assistants seven read-only tools to search, list and read any Fumadocs documentation site, working against a deployed URL or a local docs repo via a single npx command.
pyGecko is an open-source Python library for parsing, processing, and analyzing GC-MS and GC-FID raw data, supporting automated workflows, spectral comparison, and standardized exports like ORD.
Lexos is an event-driven AI document processing engine combining a Go gateway, Python workers, Redis queues and self-hosted models for offline RAG, summarization and speech transcription.
A personal repository of articles and experiments covering statistics, machine learning, economics, and mathematics, focusing on clear, practical explanations of complex topics.
AksharaMD is a local, parser-agnostic tool that grades how well a chosen document parser converted a source file into AI-friendly Markdown, returning a per-document readiness score, block-level provenance tags, named warnings, and an optional source-grounded ACCEPT/REJECT verdict.
A Python library and CLI tool for searching the SLB Energy Glossary in English and Spanish, featuring live site scraping, local SQLite caching, and semantic search capabilities.
Oneiron is an embedded, ACID-compliant retrieval engine for AI agents, combining vector, text, graph, temporal, and phonetic signals in a single LMDB environment.
MaluDB is a PostgreSQL 17 extension implementing a memory DBMS for long-term institutional memory, human-AI knowledge sharing, and contextual recall. It provides bitemporal data, graph traversal, hybrid search, and governed memory surfaces.
TESSERA is a text-first memory and evidence layer for AI agents: it turns Markdown project knowledge into structured, provenance-backed evidence that agents can query via a Python API, CLI or MCP, with stable identity and explainable retrieval (MIT, v0.0.1).
Onya is a knowledge‑graph data model and Markdown‑native serialization format with a Python library. It supports first‑class relationships, recursive assertions, shared vocabularies, and LLM‑friendly authoring for extracted knowledge and agent context.
AllSource is an AI-native event store built in Rust, featuring high-throughput ingestion, low-latency reads, and native MCP integration for AI agents. It supports durable event sourcing without Postgres, includes an agent memory engine, and offers community and enterprise editions.
Python library for on-device AI data infrastructure, offering quantized image/text embedding providers, batch indexing for images, videos and documents, incremental clustering, semantic search and few-shot classification.