evidently: an open-source framework to evaluate, test, and monitor ML and LLM-powered systems
An open-source Python framework to evaluate, test, and monitor ML and LLM-powered systems, helping developers detect data drift and ensure output quality.
BentoML: a unified model serving framework for building and deploying production-ready AI inference APIs
A Python framework for building and deploying high-performance model inference APIs and multi-model serving systems for any AI/ML model.
agentset
Agentset is an open-source platform for building, evaluating, and shipping production-ready RAG and agentic applications, providing end-to-end tooling from ingestion to hosting.
llm-for-zotero
A research agent system for Zotero that allows users to chat with PDFs, summarize papers, and manage their library using LLMs with grounded citations.
nestia
A set of helper libraries for NestJS that provides high-performance validation, automatic SDK generation, and AI-powered development tools for typed API servers.
kernel-memory
Kernel Memory is a multi-modal AI service for efficient dataset indexing and Retrieval Augmented Generation (RAG), providing tools for data ingestion pipelines and natural language querying with citations.
vearch
Vearch is a cloud-native distributed vector database that enables efficient similarity search of embedding vectors for AI applications.
hamilton
A lightweight Python library for creating portable and expressive data transformation DAGs, used to structure ML workflows, ETL pipelines, and RAG systems.
seekdb
A MySQL-compatible state store for AI agents that provides high-performance streaming writes, hybrid vector/full-text search, and copy-on-write sandboxes for safe exploration.
autoflow
AutoFlow is an open-source Graph RAG knowledge base tool that enables users to create conversational search experiences using a built-in website crawler and TiDB Vector.
swirl-search
An open-source federated metasearch and RAG platform that enables unified search and AI summaries across enterprise data sources without moving or indexing the data.
rag-web-ui
An intelligent dialogue system that allows users to build custom Q&A services by combining their own document knowledge bases with LLMs using Retrieval-Augmented Generation.
fastembed
A lightweight, fast Python library for generating text, image, and multimodal embeddings using ONNX Runtime to avoid heavy PyTorch dependencies.
EvoAgentX
EvoAgentX is an open-source framework for building and automatically evolving LLM-based agents and workflows, featuring built-in evaluation and a rich library of tools.
AnyCrawl
AnyCrawl is a high-performance web crawling and scraping toolkit that enables the extraction of structured JSON data from websites using LLMs, making web content LLM-ready.
SimpleMem
A unified memory stack for LLM agents that uses semantically lossless compression to store and retrieve text and multimodal memories efficiently.
GenerativeAIExamples
A collection of reference implementations and tutorials for building generative AI systems, RAG pipelines, and agentic workflows using the NVIDIA software ecosystem and NIM microservices.
Lealone
Lealone is a high-performance, self-evolving general agent that enables full-stack application and enterprise AI service development through natural language and SQL-like commands.
holmesgpt
An open-source AI agent for SREs that automates production incident investigation and root cause analysis across any infrastructure stack.
ai
A type-safe, provider-agnostic TypeScript SDK for building streaming chat, tool-calling agents, and multimodal AI applications across multiple JS frameworks.
AdalFlow
AdalFlow is a PyTorch-like library for building and auto-optimizing LLM workflows, including chatbots, RAG, and agents, by replacing manual prompting with automated textual gradient descent.
chonkie
A lightweight, high-performance text chunking library for RAG pipelines that provides diverse splitting strategies and seamless integrations with vector databases and embedding providers.
OpenMemory
A cognitive memory engine for AI agents that provides long-term, multi-sector memory and temporal reasoning, moving beyond simple vector-based RAG.
llm-graph-builder
A tool that uses LLMs and LangChain to transform unstructured data from various sources into structured Knowledge Graphs stored in Neo4j.
LLM-Engineers-Handbook
A production-ready framework and codebase for building end-to-end LLM systems, covering training, RAG, and deployment on AWS.
openagent
OpenAgent is an open-source personal AI assistant platform that combines LLMs, RAG, and autonomous agent loops into a single self-hostable binary for browser, shell, and office automation.
ComfyUI-Copilot
An intelligent assistant for ComfyUI that automates workflow generation, debugging, and parameter tuning to streamline AI image generation development.
MineContext
MineContext is a proactive, context-aware AI partner that captures screen activity and digital content to automatically generate summaries, to-do lists, and insights.
PixelRAG
PixelRAG is a visual RAG system that renders documents as screenshots instead of parsing them to text, allowing AI to retrieve and reason over visual elements like tables and charts.
helix-db
HelixDB is a graph-vector database built in Rust that consolidates graph, vector, relational, and document data into a single platform for AI memory and knowledge graphs.
honcho
Honcho is memory infrastructure for stateful AI agents that allows them to maintain a persistent, evolving understanding of people and projects through background reasoning and peer-centric representations.
UltraRAG
A lightweight RAG development framework based on the Model Context Protocol (MCP) that enables low-code orchestration of complex workflows and rapid prototyping via a visual IDE.
trafilatura
A Python package and command-line tool for discovering and extracting clean, structured text and metadata from the web, removing HTML noise to create high-quality datasets.
airweave
An open-source context retrieval layer for AI agents and RAG systems that syncs data from 50+ integrations into a unified, LLM-friendly search interface.
vespa
Vespa is a high-performance platform for search, recommendation, and personalization that allows for real-time inferences and data organization using vectors, tensors, and text at scale.
Upsonic
A Python framework for building autonomous and traditional AI agents, featuring secure workspace execution, custom tool integration, and a unified OCR pipeline.
paper-qa
PaperQA2 is an agentic RAG system for scientific literature that provides high-accuracy, grounded answers with in-text citations from PDFs and other document formats.
garden-skills
A curated collection of production-ready skills for AI coding agents (Claude Code, Cursor, Codex) to perform specialized tasks in web design, video production, and image generation.
claude-context
An MCP plugin that adds semantic code search to AI coding agents, allowing them to efficiently retrieve relevant snippets from large codebases using a vector database.
turbovec
A Rust-based vector index with Python bindings that implements Google's TurboQuant algorithm to provide extreme memory compression and fast SIMD-accelerated vector search for RAG applications.
zvec
Zvec is an open-source, in-process vector database that provides low-latency similarity search and hybrid retrieval directly embedded within applications.
coze-studio
Coze Studio is an open-source, low-code visual development platform for creating, debugging, and deploying AI agents, workflows, and AI apps.
DeepTutor
DeepTutor is an agent-native personalized tutoring workspace that integrates tutoring, research, and mastery practice into a single system with shared memory and multi-engine RAG.
kotaemon
An open-source, customizable RAG UI for chatting with documents, featuring hybrid retrieval, multi-modal parsing, and advanced citations.
claude-mem
A persistent memory compression system for Claude Code and other AI CLIs that preserves project context and tool observations across sessions.
opendataloader-pdf
An open-source PDF parser and accessibility tool that extracts structured data (Markdown, JSON) for AI pipelines and automates the creation of Tagged PDFs for accessibility compliance.
ragflow
RAGFlow is an open-source RAG engine that combines deep document understanding with agent capabilities to create grounded, production-ready AI systems from complex unstructured data.
PaddleOCR
A global leading OCR toolkit and document AI engine that converts PDFs and images into structured JSON or Markdown data for LLM-ready applications.
superglue
An AI-powered tool builder that allows users to create production-grade integrations and tools using natural language, featuring self-healing capabilities for API changes.
agents-best-practices
A provider-neutral Agent Skill that provides blueprints and best practices for designing rigorous, production-safe agent harnesses to manage tool execution, permissions, and observability.