401

herdr: a terminal-based agent multiplexer for monitoring and managing persistent AI agent sessions

A terminal-based agent multiplexer that allows users to monitor, manage, and persist sessions for multiple AI agents.

402

ChatLab: a local-first chat history analyzer that uses AI agents to extract insights from multiple messaging platforms

An open-source desktop app that uses AI agents and a SQL engine to privately analyze and extract insights from social chat histories across multiple platforms.

403

page-agent: a client-side GUI agent that enables natural language control of web pages via text-based DOM manipulation

A JavaScript library that embeds an AI agent directly into web pages for text-based DOM manipulation and automation of web workflows via natural language.

404

hermes-webui: a lightweight web interface for the Hermes autonomous agent with full CLI parity and integrated workspace management

A lightweight, dark-themed web interface for the Hermes Agent, providing full CLI parity for managing an autonomous agent with persistent memory and self-improving skills.

405

quant-mind: an agent-native information processor that transforms raw financial data into structured and cited knowledge

An information processor for quantitative finance that transforms raw papers and news into structured, cited, and timestamped financial knowledge for RAG and agentic reasoning.

406

oh-my-claudecode: a multi-agent orchestration layer for Claude Code with autonomous workflows and multi-model coordination

A multi-agent orchestration layer for Claude Code that provides autonomous workflows, multi-model coordination via tmux workers, and Socratic requirement gathering.

407

burr: a state-machine framework for developing and monitoring stateful AI agents and workflows

A framework for building stateful AI applications and agents by modeling them as state machines with a built-in telemetry UI for real-time tracing and debugging.

408

ax: a language-agnostic programming model for typed LLM generation, agents, and workflows

A multi-language framework for building LLM applications using typed signatures, agents, and workflows, eliminating manual prompt engineering across TypeScript, Python, Java, C++, Go, and Rust.

409

memsearch: cross-platform semantic memory for AI coding agents

A semantic memory engine that provides persistent, cross-platform context and workflow distillation for AI coding agents.

410

evalscope: a comprehensive evaluation framework for benchmarking model capabilities and inference performance

A one-stop LLM evaluation framework that provides model capability benchmarking, inference performance stress testing, and result visualization.

411

morphik-core: a multimodal retrieval engine for searching and extracting data from visually rich documents

A multimodal retrieval engine that enables AI applications to search and understand visually rich documents, such as PDFs and charts, without losing spatial or visual context.

412

llm-d: a distributed inference serving stack that optimizes LLM production deployments on Kubernetes across multiple accelerators

A high-performance distributed inference serving stack for Kubernetes that optimizes LLM production deployments across various hardware accelerators to maximize throughput and reduce latency.

413

AutoRAG: a self-evolving librarian agent that curates document search results into actionable knowledge units

A self-evolving librarian agent that searches document collections and curates results into direct answers instead of raw file paths.

414

claude-mem: a persistent memory system for Claude Code that preserves project context across sessions

A persistent memory compression system for Claude Code that preserves project context across sessions using semantic summaries and hybrid search.

415

hyperresearch: a deep research harness for Claude Code featuring a 16-step adversarial pipeline and a persistent source vault

A deep research harness for Claude Code that uses a 16-step adversarial pipeline and a persistent, searchable vault to produce high-fidelity research reports.

416

ruler: a centralized instruction manager that synchronizes AI coding rules across multiple AI agents

A tool that centralizes AI coding assistant instructions into a single source of truth and automatically distributes them to the configuration files of various supported AI agents.

417

agent-service-toolkit: a full-stack template for deploying LangGraph agents with a FastAPI backend and Streamlit UI

A full-stack toolkit for building AI agent services using LangGraph, FastAPI, and Streamlit, providing a complete template from agent logic to user interface.

418

potpie: a context graph that transforms codebases and development lifecycles into actionable knowledge for AI agents

Potpie creates a living context graph from your codebase and development tools, providing AI agents with the deep project knowledge required to plan, debug, and write code.

419

text-to-cad: what it is, what problem it solves & why it's gaining traction

A library of agent skills that enables AI to generate, inspect, and prepare CAD models and robot description files for fabrication and simulation.

420

agents: a framework for building real-time programmable multi-modal voice agents with integrated job scheduling

A framework for building real-time, programmable multi-modal AI agents that can see, hear, and speak, featuring integrated job scheduling and telephony support.

421

cube: an open-source semantic layer for defining governed metrics and dimensions across BI tools and AI agents

An open-source semantic layer that allows you to define metrics and dimensions once in code and expose them via APIs to BI tools and AI agents.

422

agents: a multi-harness marketplace of agentic plugins and domain-expert agents for AI coding assistants

A marketplace of production-ready agentic building blocks that provides a single source of truth for plugins, agents, and skills across multiple AI coding assistants.

423

novu: Open‑source communication infrastructure that lets products and AI agents talk to users on any channel

Novu is an open‑source platform that offers a single API and unified conversation model for sending and receiving messages across inbox, email, SMS, push, and chat channels. It serves both product notification needs and AI‑agent communication (ACI), handling provider integrations, workflows, digests, and embeddable UI components.

424

ciso-assistant-community: a centralized cybersecurity management hub for governance, risk, and compliance

CISO Assistant is an open-source cybersecurity management platform designed to simplify Governance, Risk, and Compliance (GRC) through centralized data and framework decoupling.

425

tracecat: an agentic security automation platform combining AI agents, low-code workflows, and case management

An agentic security automation platform that combines AI agents, low-code workflows, and case management to automate security operations.

426

osaurus: a native macOS AI harness with local agents, secure sandboxing, and on-device privacy filtering

A native macOS AI harness that provides local agents, memory, and secure code execution sandboxes, allowing users to own their AI context and identity offline.

427

mistral.rs: a high-performance, zero-config LLM inference engine with native agentic runtime and multimodal support

A fast, flexible LLM inference engine that provides zero-config execution for Hugging Face models with native support for multimodality and agentic tools.

428

opencodex: a universal provider proxy that enables any LLM to work with OpenAI Codex and Claude Code

A local proxy that translates Codex's API to allow any LLM provider (Anthropic, Gemini, Ollama, etc.) to be used within OpenAI Codex and Claude Code.

429

inference: a versatile model serving library for deploying language, speech, and multimodal models across heterogeneous hardware

A versatile model serving library that simplifies the deployment of language, speech, and multimodal models across heterogeneous hardware.

430

NarratoAI: a one-stop automated tool for AI-powered movie and TV show commentary and video editing

An automated one-stop tool for creating movie and TV show commentary videos, automating script writing, video clipping, and voiceover generation.

431

superset: an orchestration platform for managing multiple CLI-based coding agents in parallel

Superset is a code editor and orchestration platform that allows developers to run multiple CLI-based AI coding agents in parallel using isolated git worktrees.

432

Toonflow-app: what it is, what problem it solves & why it's gaining traction

Toonflow is an AI-powered workbench for short drama production that automates the workflow from screenwriting and storyboarding to final video generation.

433

bisheng: an enterprise LLM devops platform with visual workflow orchestration and high-precision document parsing

An open LLM application devops platform for enterprises that provides visual workflow orchestration, expert-level agent guidance, and high-precision document parsing.

434

astron-agent: an enterprise-grade agentic workflow platform integrating AI orchestration with RPA for cross-system automation

An enterprise-grade Agentic Workflow development platform that combines AI orchestration and RPA to build scalable, production-ready intelligent agents.

435

rowboat: a desktop AI coworker with a local knowledge graph for long-term work memory

A desktop AI coworker that indexes your work into a local knowledge graph to provide long-term memory and integrated surfaces for email, coding, and meeting notes.

436

MNN: a lightweight and efficient deep learning framework for on-device inference and training

MNN is a lightweight, high-performance deep learning framework designed by Alibaba for efficient model inference and training on mobile, embedded, and IoT devices.

437

agency-agents-zh: a comprehensive library of 268 specialized AI agent roles with professional workflows for AI coding tools

A library of 268 specialized AI agent roles with defined personas and workflows, compatible with 18 AI programming tools to provide expert-level outputs across 20 business departments.

438

omlx: an optimized LLM inference server for Apple Silicon with tiered KV caching and a macOS menu bar interface

oMLX is an optimized inference server for Apple Silicon that enables efficient local LLM execution through tiered KV caching and a native macOS management interface.

439

airllm: a memory-efficient inference engine that runs massive LLMs on consumer GPUs by loading one layer at a time

AirLLM is a library that enables running massive LLMs (up to 671B parameters) on low-end GPUs by loading only one model layer at a time into VRAM.

440

pydantic-ai: a type-safe GenAI agent framework with structured outputs and durable execution

A type-safe Python agent framework by the Pydantic team that enables developers to build production-grade GenAI applications with structured outputs and seamless observability.

441

llmfit: a hardware-aware model recommender that matches LLMs to your system's RAM and GPU

A terminal tool that detects your hardware and recommends the best-fitting LLM models based on memory, speed, quality, and context.

442

langextract: a structured information extraction library that maps data back to source text for precise verification

A Python library that uses LLMs to extract structured, grounded information from unstructured text documents with built-in support for long-document processing and visualization.

443

NotFair: goal-driven autonomous marketing agents that optimize business metrics via disciplined local loops

A local-first framework for autonomous marketing agents that turn business goals into measured metrics and execute disciplined improvement loops to optimize SEO and ad spend.

444

xtuner: a training engine for ultra-large MoE models supporting up to 1T parameters and multimodal training

A next-generation LLM training engine optimized for ultra-large-scale MoE models up to 1T parameters, supporting multimodal training and advanced reinforcement learning.

445

vllm: a high-throughput LLM serving library featuring PagedAttention for efficient memory management

A high-throughput library for LLM inference and serving that uses PagedAttention to optimize memory management and increase efficiency.

446

dify: an open-source LLM app development platform with visual workflows and integrated RAG pipelines

Dify is an open-source LLM app development platform that combines AI workflows, RAG pipelines, and agent capabilities to help developers move from prototype to production.

447

astron-rpa: an enterprise RPA platform with low-code visual design and native AI agent integration

An enterprise-grade Robotic Process Automation (RPA) desktop application that enables low-code automation of Windows desktop software and web pages with native AI agent integration.

448

astrid: a capability-secure OS for composable AI agent software using WASM sandboxing

A capability-secure operating system for AI agents that uses WASM sandboxing and cryptographic grants to enforce security boundaries instead of relying on prompts.

449

pro-workflow: a persistent memory and knowledge plane that prevents AI coding agents from repeating mistakes

A persistent memory and knowledge plane for AI coding agents that eliminates repetitive corrections through self-correcting SQLite-backed memory and auto-growing research wikis.

450

ktransformers: a CPU-GPU hybrid inference and fine-tuning framework optimized for ultra-large MoE models

A flexible framework for high-performance LLM inference and fine-tuning using CPU-GPU heterogeneous computing, specifically optimized for ultra-large MoE models on consumer hardware.