kai-scheduler/KAI-Scheduler

A scalable Kubernetes scheduler that optimizes GPU resource allocation and fairness for large-scale AI and machine learning workloads.

chen0416ccc-cpu/codex-windows-fast-patch-skill

A collection of repair scripts and agent skills for Windows Codex Desktop that fixes broken features like Fast Mode, Computer Use, and plugin markets after software updates.

remorses/playwriter

A browser automation tool that lets AI agents control your existing Chrome browser session, preserving logins and extensions to bypass bot detection.

google-ai-edge/LiteRT

Google's high-performance on-device runtime for deploying ML and Generative AI models across mobile, web, and IoT platforms with advanced hardware acceleration.

mozilla-ai/otari

An OpenAI-compatible LLM gateway that provides a unified endpoint for multiple providers, enabling centralized credential management, budget enforcement, and usage tracking.

shap/shap

A game theoretic approach to explain the outputs of any machine learning model by attributing credit to input features using Shapley values.

lutzroeder/netron

A visualizer for neural network and deep learning models that supports a wide variety of model formats including ONNX, TensorFlow, and PyTorch.

x-cmd/x-cmd

A modern POSIX shell toolkit that provides a standard library of 300+ modules and 600+ on-demand CLI packages, optimized for AI agents and system portability.

waybarrios/vllm-mlx

A vLLM-style inference server for Apple Silicon that brings continuous batching, paged KV cache, and dual OpenAI/Anthropic API compatibility to macOS.

kangarooking/kangarooking-skills

A collection of structured, open-source Agent Skills compatible with the Agent Skills standard, providing reusable workflows for content creation, image generation, and AI project engineering.

qixing-jk/all-api-hub

A browser extension that provides centralized management for AI relay accounts, API credentials, and self-hosted gateways, featuring automated check-ins and price comparisons.

doccker/cc-use-exp

A maintainable configuration system that synchronizes coding rules, skills, and workflows across multiple AI coding assistants like Claude Code, Cursor, and Gemini CLI.

vortex-data/vortex

A next-generation columnar file format and toolkit that provides significantly faster reads, scans, and writes for data systems backed by object storage.

CaviraOSS/LongMemory

A cognitive memory engine for AI agents that provides durable, temporal, and governed local-first storage to move beyond simple RAG.

sktime/sktime

A Python library providing a unified interface for time series analysis, including forecasting, classification, and anomaly detection, compatible with the scikit-learn ecosystem.

openhackai/OpenHack

OpenHack is an open‑source, LLM‑driven security scanner that analyses a codebase through a recon‑→‑hunter‑→‑validator‑→‑verifier pipeline, presenting findings in an interactive terminal UI and optionally verifying exploits in Docker.

lmnr-ai/lmnr

An open-source observability platform for AI agents that provides tracing, monitoring, and evaluation tools to debug and optimize agentic workflows.

kortix-ai/suna

An open-source AI Management System that lets companies manage agents, skills, and memory as a git repository and execute tasks in isolated cloud sandboxes.

PortSwigger/mcp-server

A Burp Suite extension that integrates the security tool with AI clients via the Model Context Protocol (MCP), allowing AI assistants to interact with Burp Suite.

mshumer/Claude-of-Duty

A browser-based first-person shooter built with Three.js and WebGL2, featuring entirely procedurally generated assets and a codebase written by AI agents.

NVIDIA/cuda-samples

A collection of reference samples for CUDA developers that demonstrate the features of the CUDA Toolkit across C++ and Python.

superlinked/sie

A self-hosted inference engine that serves multiple open-source models for agent tasks—such as retrieval, OCR, and generation—through a single, unified API.

Scottcjn/Rustchain

A DePIN blockchain that rewards the preservation of vintage hardware and provides AI-augmented hardware fingerprinting for Sybil-resistant AI agent identity.

cristicretu/diri

A native workspace for coding agents that allows developers to run multiple terminal-based AI agents side-by-side with isolated Git worktrees and integrated code review.

Shpigford/chops

A macOS application for discovering, organizing, and editing AI coding agent skills and agents across multiple tools like Claude Code, Cursor, and Windsurf.

fivetaku/insane-search

A resilient public-page reader for Claude Code that bypasses bot detection and WAFs to retrieve public web content without API keys.

KatrielMoses/voidaccess

A self-hostable OSINT tool that automates dark-web research by collecting, enriching, and mapping threat intelligence into structured formats.

WoJiSama/skill-based-architecture

A framework for organizing project rules and instructions into a verifiable "Skill" to ensure coding agents follow business logic and a strict definition of done.

microsoft/skills-for-fabric

A collection of reusable AI assistant instructions and MCP servers that help AI coding tools understand and operate Microsoft Fabric workloads, APIs, and best practices.

langgptai/LangGPT

LangGPT is a structured, reusable prompt‑design framework that treats prompts like a tiny programming language. It defines a clear template (role, profile, goal, skills, rules, workflow, etc.), supports variables, commands, and conditional logic, and works across major LLMs. The project provides example libraries, ready‑made ChatGPT/Claude assistants, a Claude Code skill, and an ecosystem of related tools (PromptVer, PromptShow, Minstrel) plus model‑specific prompt collections.

andrewyng/aisuite

A lightweight Python library providing a unified Chat Completions API and an Agents API to build AI agents across multiple LLM providers with minimal code.

lycorp-jp/sim-use

A cross-platform CLI that enables AI agents to observe and act on iOS and Android screens using token-efficient UI outlines and alias-based interactions.

RunanywhereAI/runanywhere-sdks

A cross-platform SDK that enables developers to run LLMs, vision, speech, and image generation models locally on phones, browsers, and desktops using a unified API.

Flux159/mcp-server-kubernetes

An MCP server that enables AI assistants to manage Kubernetes clusters, perform diagnostics, and handle Helm operations using a standardized interface.

AGenUI/AGenUI

A high-performance A2UI SDK for building native generative UI experiences on iOS, Android, and HarmonyOS, enabling LLMs to render interactive components in real-time.

ROCm/TheRock

TheRock is the open-source build and release system for HIP and ROCm, providing a modular workflow to integrate and distribute ROCm components.

oracle-devrel/oracle-ai-developer-hub

A technical hub of reference apps, notebooks, and guides for building AI agents and RAG systems using Oracle AI Database and OCI services.

traceloop/openllmetry

An open-source observability framework built on OpenTelemetry that provides complete tracing and monitoring for LLM applications, providers, and vector databases.

tenstorrent/tt-metal

tt‑metal (TT‑Metalium) is Tenstorrent’s open‑source stack for writing low‑level kernels (C++/Python) and high‑level neural‑network ops (TT‑NN) that run on Tenstorrent AI chips. It includes model demos, performance reports, and a suite of debugging/profiling tools, aimed at developers who want to run or optimise LLMs, vision models, and other AI workloads on Tenstorrent hardware.

alpacahq/alpaca-mcp-server

An MCP server that enables AI assistants to perform trading operations and retrieve market data for stocks, options, and crypto via the Alpaca Trading API.

Storybloq/storybloq

A project memory and workflow system for AI coding agents that persists stories, plans, and handovers in Git-tracked files to maintain context across sessions.

ROCm/FastFlowLM

A lightweight, NPU-first runtime for running LLMs, VLMs, and other AI models on AMD Ryzen™ AI NPUs with high power efficiency and no GPU required.

InternLM/lmdeploy

A toolkit for compressing, deploying, and serving LLMs and VLMs, offering high-performance inference engines and quantization to maximize throughput.

Sophomoresty/turnstile-bypass

A browser automation tool that bypasses Cloudflare Turnstile widgets and interstitial waiting rooms by patching browser click coordinates to avoid bot detection.

borawong/AiMaMi

A native desktop companion for OpenAI Codex that provides a GUI for managing accounts, routing, sessions, and local configurations to avoid manual file editing.

blacktop/ida-mcp-rs

A headless IDA Pro MCP server that enables AI agents to perform reverse engineering, decompilation, and binary analysis.

xorbitsai/inference

Xorbits Inference (Xinference) is a versatile library for easily deploying and serving large language, speech, and multimodal models across heterogeneous hardware.

jbarbier/CLAUDE.md

A set of opinionated instructions and a workflow contract for AI coding agents to ensure high-quality, predictable code and professional communication.