router-for-me/CLIProxyAPI
A proxy server that provides unified OpenAI, Gemini, and Claude compatible API interfaces for various AI providers, allowing CLI tools to access models using existing subscriptions via OAuth.
rtk-ai/rtk
A high-performance CLI proxy that compresses bash output for AI agents, reducing input tokens by up to 90% through smart filtering and grouping.
Kuddev/pebrel
A GPU-accelerated terminal and SSH workspace designed to integrate and manage AI CLI sessions, such as Claude Code and Codex, in a single native environment.
garrytan/gstack
A software factory framework that turns AI coding agents into a virtual engineering team with specialized roles for product planning, architectural review, QA, and security audits.
Git-Agni/prod-FARM-IOS-Core
An open-source application for managing physical iOS devices to run scheduled automation workflows, featuring a built-in TikTok plugin and remote control capabilities.
miqdadbadjuber/anti-slop
A set of rules and filters for AI coding agents to prevent the generation of generic "AI slop" in UI, copy, and code comments.
max-sixty/worktrunk
A CLI for git worktree management designed to let developers run multiple AI agents in parallel by simplifying the creation and switch between separate working directories.
jingyaogong/minimind
MiniMind is a complete, open-source codebase for training a very small (~64M parameter) large language model entirely from scratch on a tiny budget (about 3 RMB and 2 hours). It covers the full modern LLM lifecycle—pretraining, SFT, LoRA, DPO, RLAIF (PPO/GRPO/CISPO), tool use, and agentic RL—and includes experimental extensions for MoE, vision, and diffusion models.
multica-ai/multica
An open-source workspace that integrates multiple AI coding agents into a single board, allowing users to assign issues to agents as if they were human teammates.
t8y2/dbx
A lightweight, Rust-based database manager supporting 90+ databases with a built-in AI SQL assistant and MCP server for AI agent connectivity.
zhaoxuya520/reverse-skill
A cybersecurity skills router that provides AI agents with structured methodologies and tool routing for reverse engineering, penetration testing, and CTF challenges.
ggml-org/llama.cpp
A plain C/C++ implementation for high-performance LLM and VLM inference across a wide range of hardware with minimal setup.
agentverse-os/AgentVerse-OS
A personal cloud operating system for developers that provides isolated workspaces for AI agents and a catalog of 944 self-hosted apps, all accessible via a secure browser-based desktop.
decolua/9router
A smart AI router and proxy that saves 20-40% of tokens and provides automatic fallback between subscription, cheap, and free AI models for coding tools.
aipoch/open-science
An open-source, local-first AI research workbench that enables reproducible science by combining AI agents, Python/R execution, and traceable provenance for all research artifacts.
Continuum-AI-Corp/OrcaRouter-Lite
A self-hosted, OpenAI-compatible LLM router that optimizes costs and reliability using automatic model selection and a managed fallback safety net.
DeusData/codebase-memory-mcp
A high-performance code intelligence engine that builds a structural knowledge graph of codebases to provide AI agents with efficient, low-token access to architectural insights.
open-webui/open-webui
A self-hosted, extensible AI platform that provides a unified interface for local and cloud-based models, featuring built-in RAG, agent creation, and enterprise-grade user management.
Wei-Shaw/sub2api
An AI API gateway platform that enables the distribution and management of subscription quotas for various AI model providers.
Neroued/ninfer
A from-scratch C++/CUDA inference engine optimized for the NVIDIA RTX 5090 to deliver maximum single-GPU performance for Qwen models.
yyjeqhc/webcodex
A bridge that lets AI agents like ChatGPT and Claude work directly with code and developer tools on your local machine without moving repositories to the cloud.
Sliverkiss/workbuddy2api
WorkBuddy2API is a self‑hosted Go service that turns one or many Tencent CodeBuddy accounts into an OpenAI‑compatible `/v1/chat/completions` API. It handles OAuth login, token refresh, multi‑account pooling, rate‑limit/cool‑down logic, session stickiness, cost‑aware routing and scheduled tasks (sign‑in, activity reporting, “cat travel”). Deploy with Docker‑Compose, add accounts via `login.sh`, and call the gateway just like any OpenAI endpoint.
linshenkx/prompt-optimizer
An AI prompt optimization tool that iteratively improves prompts for text and image generation models to enhance output quality and accuracy.
MakazhanAlpamys/Soup
Soup is a Python CLI (with optional web UI) that lets you fine‑tune LLMs—including 8B models—on low‑VRAM GPUs using a single YAML config and one command. It offers layer‑streaming, QLoRA, many training objectives, export to GGUF/ONNX/TensorRT, and an OpenAI‑compatible server, all with zero‑SSH, auto‑detected settings, and extensive documentation.
microsoft/AI-Engineering-Coach
A VS Code extension and GitHub Copilot app canvas that analyzes local AI coding session logs to provide insights into prompting patterns, context health, and developer productivity.
BerriAI/litellm
An open-source AI Gateway and Python SDK that provides a unified OpenAI-compatible interface for calling 100+ different LLM providers.
genspark-ai/genoffice
GenOffice is an open‑source, cross‑platform desktop Office suite (Docs, Sheets, Slides, PDF, HTML, Markdown) that edits real `.docx`, `.xlsx`, `.pptx` and other formats. It embeds an AI agent that can modify documents directly—producing tracked changes, live formulas, new slides, etc.—and all edits are reversible. The suite works locally; only AI calls go over the network, and you can plug in any major LLM provider or your own endpoint. Installers are provided for macOS, Windows and Linux.
Crosstalk-Solutions/project-nomad
Project NOMAD is an offline-first knowledge and education server that bundles local AI chat (with RAG), offline Wikipedia, courses, maps, and data tools into a self-contained Docker-orchestrated system, keeping critical information available without internet.
vllm-project/vllm
A high-throughput library for LLM inference and serving that uses PagedAttention to optimize memory management and increase efficiency.
unstablebuild/rune
A fast, GPU-rendered, keyboard-driven IDE that integrates a modular AI coding agent and Unix-style composable tools for power users.
experientiallabs/experiential
An open-source gateway and router for agent workflows that unifies multiple LLM providers under one API and optimizes model routing based on production traffic.
FlashML-org/FreeToken
An edge-native MoE serving engine that enables running frontier-scale open-weight models on consumer hardware by optimizing CPU-GPU co-execution and memory management.
crwdla/tokentab
A local-first tool that aggregates session logs from AI coding assistants like Claude Code and Gemini CLI to track token usage and costs across models and projects.
chaseai-yt/claudex-loop
A cross-provider workflow for Claude Code and Codex that ensures AI-generated code is independently reviewed and inspected by a second AI model to prevent errors.
tradesdontlie/tradingview-mcp
An MCP bridge that connects AI assistants to the TradingView Desktop app via Chrome DevTools Protocol for AI-assisted chart analysis and Pine Script development.
vinzdg/codenotch
A macOS app that displays a screen-edge notch showing real-time usage limits and session status for various AI coding assistants.
LukasNiessen/terrashark
A Terraform and OpenTofu skill for AI agents that eliminates hallucinations and reduces token usage through a failure-mode-first diagnostic workflow.
paperclipai/paperclip
An open-source orchestration platform for managing teams of AI agents through org charts, budgets, and goal-aligned task management.
guillaumemeyer/watermarks-remover
A service and agent skill for stripping AI provenance marks and watermarks from text and files across multiple vendors and formats.
duolahypercho/codex-router
Codex Router is a community‑maintained bridge that lets the Codex editor/CLI send prompts to many external LLM providers (Anthropic, Kimi, DeepSeek, Grok, Gemini, Claude, etc.). It stores credentials locally, offers a guided installer, a macOS menu‑bar/desktop‑widget UI, and a Homebrew‑installable CLI. Experimental bridges let it reuse existing Claude, Cursor, and Gemini agents without exposing their OAuth tokens.
h4ckf0r0day/obscura
A lightweight, stealthy headless browser engine written in Rust for AI agents and web scraping, serving as a high-performance, low-memory replacement for headless Chrome.
optiscaler/OptiScaler
A middleware tool that lets users replace the temporal upscalers and frame generation technologies in games that already support DLSS, FSR, or XeSS.
ollama/ollama
Ollama is a cross‑platform runtime that lets you download and run open‑source large language models locally. It offers a CLI, a REST API (localhost:11434), and official Python/JS SDKs, plus many community integrations (web UIs, LangChain, AutoGPT, etc.). Install with a one‑line script or Docker, then launch models like `ollama run gemma4` or use `ollama launch` to connect to coding assistants or personal bots.
zhihui-hu/one-ip
A network diagnostics and browser detection toolbox that provides IP analysis, global connectivity tests, and AI service status monitoring.
kruzovic7/ai-data-extractor
A toolkit to extract and normalize local chat history from multiple AI coding assistants into JSONL format for backup, analytics, or fine-tuning.
Devin-AXIS/iPolloWork
An enterprise-grade, local-first agent workbench that unifies multiple AI agent engines into a single workspace for coordinating tasks and creating editable content.
opendatalab/MinerU
MinerU is a professional document parsing tool that converts PDFs, images, and Office files into structured Markdown or LaTeX, featuring tiered parsing quality and stable locators for AI agents.
QuantumNous/new-api
A self-hosted AI gateway that centralizes the management of multiple AI model providers, providing a consistent API, routing, and usage tracking for teams and applications.