901

evidently: an open-source framework to evaluate, test, and monitor ML and LLM-powered systems

An open-source Python framework to evaluate, test, and monitor ML and LLM-powered systems, helping developers detect data drift and ensure output quality.

902

BentoML: a unified model serving framework for building and deploying production-ready AI inference APIs

A Python framework for building and deploying high-performance model inference APIs and multi-model serving systems for any AI/ML model.

903

agentset

Agentset is an open-source platform for building, evaluating, and shipping production-ready RAG and agentic applications, providing end-to-end tooling from ingestion to hosting.

904

llm-for-zotero

A research agent system for Zotero that allows users to chat with PDFs, summarize papers, and manage their library using LLMs with grounded citations.

905

nestia

A set of helper libraries for NestJS that provides high-performance validation, automatic SDK generation, and AI-powered development tools for typed API servers.

906

kernel-memory

Kernel Memory is a multi-modal AI service for efficient dataset indexing and Retrieval Augmented Generation (RAG), providing tools for data ingestion pipelines and natural language querying with citations.

907

vearch

Vearch is a cloud-native distributed vector database that enables efficient similarity search of embedding vectors for AI applications.

908

hamilton

A lightweight Python library for creating portable and expressive data transformation DAGs, used to structure ML workflows, ETL pipelines, and RAG systems.

909

seekdb

A MySQL-compatible state store for AI agents that provides high-performance streaming writes, hybrid vector/full-text search, and copy-on-write sandboxes for safe exploration.

910

autoflow

AutoFlow is an open-source Graph RAG knowledge base tool that enables users to create conversational search experiences using a built-in website crawler and TiDB Vector.

911

swirl-search

An open-source federated metasearch and RAG platform that enables unified search and AI summaries across enterprise data sources without moving or indexing the data.

912

rag-web-ui

An intelligent dialogue system that allows users to build custom Q&A services by combining their own document knowledge bases with LLMs using Retrieval-Augmented Generation.

913

fastembed

A lightweight, fast Python library for generating text, image, and multimodal embeddings using ONNX Runtime to avoid heavy PyTorch dependencies.

914

EvoAgentX

EvoAgentX is an open-source framework for building and automatically evolving LLM-based agents and workflows, featuring built-in evaluation and a rich library of tools.

915

AnyCrawl

AnyCrawl is a high-performance web crawling and scraping toolkit that enables the extraction of structured JSON data from websites using LLMs, making web content LLM-ready.

916

SimpleMem

A unified memory stack for LLM agents that uses semantically lossless compression to store and retrieve text and multimodal memories efficiently.

917

GenerativeAIExamples

A collection of reference implementations and tutorials for building generative AI systems, RAG pipelines, and agentic workflows using the NVIDIA software ecosystem and NIM microservices.

918

Lealone

Lealone is a high-performance, self-evolving general agent that enables full-stack application and enterprise AI service development through natural language and SQL-like commands.

919

holmesgpt

An open-source AI agent for SREs that automates production incident investigation and root cause analysis across any infrastructure stack.

920

ai

A type-safe, provider-agnostic TypeScript SDK for building streaming chat, tool-calling agents, and multimodal AI applications across multiple JS frameworks.

921

AdalFlow

AdalFlow is a PyTorch-like library for building and auto-optimizing LLM workflows, including chatbots, RAG, and agents, by replacing manual prompting with automated textual gradient descent.

922

chonkie

A lightweight, high-performance text chunking library for RAG pipelines that provides diverse splitting strategies and seamless integrations with vector databases and embedding providers.

923

OpenMemory

A cognitive memory engine for AI agents that provides long-term, multi-sector memory and temporal reasoning, moving beyond simple vector-based RAG.

924

llm-graph-builder

A tool that uses LLMs and LangChain to transform unstructured data from various sources into structured Knowledge Graphs stored in Neo4j.

925

LLM-Engineers-Handbook

A production-ready framework and codebase for building end-to-end LLM systems, covering training, RAG, and deployment on AWS.

926

openagent

OpenAgent is an open-source personal AI assistant platform that combines LLMs, RAG, and autonomous agent loops into a single self-hostable binary for browser, shell, and office automation.

927

ComfyUI-Copilot

An intelligent assistant for ComfyUI that automates workflow generation, debugging, and parameter tuning to streamline AI image generation development.

928

MineContext

MineContext is a proactive, context-aware AI partner that captures screen activity and digital content to automatically generate summaries, to-do lists, and insights.

929

PixelRAG

PixelRAG is a visual RAG system that renders documents as screenshots instead of parsing them to text, allowing AI to retrieve and reason over visual elements like tables and charts.

930

helix-db

HelixDB is a graph-vector database built in Rust that consolidates graph, vector, relational, and document data into a single platform for AI memory and knowledge graphs.

931

honcho

Honcho is memory infrastructure for stateful AI agents that allows them to maintain a persistent, evolving understanding of people and projects through background reasoning and peer-centric representations.

932

UltraRAG

A lightweight RAG development framework based on the Model Context Protocol (MCP) that enables low-code orchestration of complex workflows and rapid prototyping via a visual IDE.

933

trafilatura

A Python package and command-line tool for discovering and extracting clean, structured text and metadata from the web, removing HTML noise to create high-quality datasets.

934

airweave

An open-source context retrieval layer for AI agents and RAG systems that syncs data from 50+ integrations into a unified, LLM-friendly search interface.

935

vespa

Vespa is a high-performance platform for search, recommendation, and personalization that allows for real-time inferences and data organization using vectors, tensors, and text at scale.

936

Upsonic

A Python framework for building autonomous and traditional AI agents, featuring secure workspace execution, custom tool integration, and a unified OCR pipeline.

937

paper-qa

PaperQA2 is an agentic RAG system for scientific literature that provides high-accuracy, grounded answers with in-text citations from PDFs and other document formats.

938

garden-skills

A curated collection of production-ready skills for AI coding agents (Claude Code, Cursor, Codex) to perform specialized tasks in web design, video production, and image generation.

939

claude-context

An MCP plugin that adds semantic code search to AI coding agents, allowing them to efficiently retrieve relevant snippets from large codebases using a vector database.

940

turbovec

A Rust-based vector index with Python bindings that implements Google's TurboQuant algorithm to provide extreme memory compression and fast SIMD-accelerated vector search for RAG applications.

941

zvec

Zvec is an open-source, in-process vector database that provides low-latency similarity search and hybrid retrieval directly embedded within applications.

942

coze-studio

Coze Studio is an open-source, low-code visual development platform for creating, debugging, and deploying AI agents, workflows, and AI apps.

943

DeepTutor

DeepTutor is an agent-native personalized tutoring workspace that integrates tutoring, research, and mastery practice into a single system with shared memory and multi-engine RAG.

944

kotaemon

An open-source, customizable RAG UI for chatting with documents, featuring hybrid retrieval, multi-modal parsing, and advanced citations.

945

claude-mem

A persistent memory compression system for Claude Code and other AI CLIs that preserves project context and tool observations across sessions.

946

opendataloader-pdf

An open-source PDF parser and accessibility tool that extracts structured data (Markdown, JSON) for AI pipelines and automates the creation of Tagged PDFs for accessibility compliance.

947

ragflow

RAGFlow is an open-source RAG engine that combines deep document understanding with agent capabilities to create grounded, production-ready AI systems from complex unstructured data.

948

PaddleOCR

A global leading OCR toolkit and document AI engine that converts PDFs and images into structured JSON or Markdown data for LLM-ready applications.

949

superglue

An AI-powered tool builder that allows users to create production-grade integrations and tools using natural language, featuring self-healing capabilities for API changes.

950

agents-best-practices

A provider-neutral Agent Skill that provides blueprints and best practices for designing rigorous, production-safe agent harnesses to manage tool execution, permissions, and observability.