CocoIndex is an open-source incremental indexing framework designed to provide AI agents and LLM applications with continuously fresh context by reprocessing only changed data.
Open source. Open possibilities.
Discover quality open-source projects, submit projects anonymously, and claim and edit your own project.
A little curiosity. A world of open source.
THE FIRST COLLECTIONA local, read‑only dashboard that visualizes Claude Code usage—token counts, cost estimates, sessions, tools, skills, projects, and activity patterns—by scanning ~/.claude/projects/ and serving a FastAPI‑based UI.
DB-GPT is an open-source agentic AI data assistant that connects to databases, spreadsheets and knowledge bases, autonomously writes SQL and Python code, runs reusable skills in sandboxed environments, and turns analysis into charts, dashboards and reports.
Apache SeaTunnel is a high-performance, distributed data integration tool supporting 160+ connectors, batch-stream integration, and multi-engine execution.
LightRAG is a lightweight graph-based retrieval-augmented generation framework from HKUDS, offering dual-level retrieval, knowledge-graph indexing, multimodal parsing, multiple storage backends and an API server with WebUI.
lore is a Bun-based command line tool that mirrors Claude Code's local transcripts into an archive, indexes them with SQLite FTS5, and serves a local web explorer for searching sessions and profiling token usage.
Chat2DB Community is a free, cross-platform, local-first database client and SQL workspace from OtterMind. It supports 40+ databases, data management, import/export, dashboards, and a bring-your-own-model AI assistant for SQL, with desktop, web, Docker, and CLI/MCP options.
An all-in-one AI framework for semantic search, LLM orchestration, and language model workflows.
A Multimarket Stock Intelligent Analysis System Based on Large Language Models, Covering A-Shares, Hong Kong Stocks, US Stocks, and More, Supporting Multi-Source Market Data Aggregation, AI-Generated Decision Reports, and Automatic Push to Multiple Channels, with Zero-Cost Timed Execution and Web Platform with Strategy Stock Query Function.
Deep Lake is a serverless multimodal database for AI that combines vector search with storage for embeddings, images, video, audio, and documents. It streams data directly from cloud storage to PyTorch/TensorFlow, supports versioning, and integrates with LangChain, LlamaIndex, and Weights & Biases.
PageIndex is a vectorless, reasoning-based RAG engine that builds hierarchical tree indexes for long documents and lets an LLM reason through the tree to retrieve relevant sections, avoiding vector databases and chunking.
LLM Router routes prompts to cost-effective Perplexity models, tracks tokens and costs in PostgreSQL, and provides analytics via API and React dashboard. Provider-agnostic design allows swapping APIs.
Python SDK for the AskNews API that injects real-time news context into any LLM with a single line of code, offering endpoints for news search, event tracking, forecasting, analytics, deep research, knowledge graphs, and web search.
An automated pipeline that scrapes GitHub Trending, enriches entries via the GitHub API, generates LLM-written summaries and trend analysis, then publishes Markdown reports plus JSON data through a Docusaurus site deployed by GitHub Actions.
LEANN is a lightweight vector database designed for personal RAG applications, reducing storage requirements by up to 97% through graph-based selective recomputation.