ollama: what it is, what problem it solves & why it's gaining traction
Ollama is a tool that allows users to run and manage open-source large language models locally, providing a CLI, REST API, and libraries for easy integration into applications.
transformers: a centralized model-definition framework for state-of-the-art pretrained models across multiple modalities
A model-definition framework providing a unified API to access, train, and deploy state-of-the-art pretrained models across text, vision, audio, and multimodal domains.
opentag: a local-first bridge that turns collaboration threads into governed AI coding agent work loops
OpenTag connects collaboration platforms like Slack and GitHub to local AI coding agents, allowing teams to trigger and govern agent work loops directly from their existing work threads.
herdr: a terminal-based agent multiplexer for managing and coordinating multiple AI agents
A terminal-based agent multiplexer that allows users to monitor, manage, and coordinate multiple AI agents through persistent sessions and a socket API.
shannon: an autonomous white-box AI pentester that validates vulnerabilities through real exploitation
An autonomous AI pentester for web applications and APIs that combines source-code analysis with live exploitation to provide proof-of-concept vulnerabilities.
zep: a suite of integrations and tools for implementing long-term agent memory via Zep Cloud
A collection of examples, integrations, and tools for building AI agent memory using the Zep Cloud managed platform.
adk-go: a code-first Go toolkit for building and deploying modular AI agents
A code-first Go toolkit for building, evaluating, and deploying modular AI agents and multi-agent systems.
VCPToolBox: VCPToolBox – a full‑stack platform that gives LLM‑based agents a persistent, world‑aware existence
VCPToolBox is a production‑grade framework that gives LLM‑based agents a continuous, world‑aware existence. It provides persistent memory (RiverMemo), automatic context “gravity”, autonomous scheduling, and a unified plugin system, all packaged with Node/Python front‑ends and Docker deployment.
ODS: an all-in-one automated deployment system for private local AI servers
ODS (Osmantic Deployment System) is an automated installer that turns any PC, Mac, or Linux machine into a private, all-in-one AI server with local inference, chat, agents, and RAG.
OnlySwitch: a macOS menu bar utility with customizable system toggles and AI-powered natural language control
A macOS menu bar utility that provides a centralized set of toggle switches and AI-powered natural language control to simplify system settings and routine tasks.
jcode: a high-performance AI coding harness with semantic memory and collaborative agent swarms
A high-performance, RAM-efficient AI coding harness featuring semantic memory, collaborative agent swarms, and a specialized TUI for developers.
spiceai: an accelerated SQL and LLM-inference engine for data-grounded AI apps with real-time CDC replication
A portable, accelerated SQL query and LLM-inference engine that provides real-time, sandboxed replicas of operational data for AI agents and analytics without requiring ETL pipelines.
sherpa-onnx: a multi-platform toolkit for local speech processing using ONNX
A versatile toolkit for running speech and audio AI models locally on various platforms and hardware using the ONNX runtime.
moonshine: a low-latency on-device toolkit for building real-time voice agents across multiple platforms
An open-source AI toolkit for building low-latency, on-device voice agents and applications with high-accuracy speech-to-text and text-to-speech capabilities.
diffusion-pipe: a pipeline parallel training framework for large-scale image and video diffusion models
A pipeline parallel training script that allows users to train large-scale image and video diffusion models across multiple GPUs to overcome VRAM limitations.
runanywhere-sdks: a unified cross-platform SDK for local AI inference across NPUs, GPUs, and CPUs
A cross-platform SDK suite that allows developers to run LLMs, vision, speech, and image models locally on any device using a unified API and automatic hardware acceleration.
stable-diffusion.cpp: a lightweight C++ inference engine for local image and video diffusion model generation
A lightweight, pure C/C++ inference engine for diffusion models that enables efficient local image and video generation across multiple platforms and hardware backends.
ai-toolkit: an all-in-one training suite for fine-tuning image, video, and audio diffusion models on consumer hardware
An all-in-one training suite for diffusion models that enables easy fine-tuning of image, video, and audio models on consumer-grade hardware.
InvokeAI: a professional creative engine for visual media with a unified canvas and node-based workflows
A professional creative AI engine for generating and refining visual media, featuring a unified canvas and node-based workflows for high-precision image creation.
diffusers: a modular toolbox for running and training state-of-the-art diffusion models across multiple modalities
A modular library for state-of-the-art pretrained diffusion models used to generate images, audio, and 3D molecular structures.
mockserver-monorepo: a multi-protocol mock server and proxy for simulating APIs and performing chaos engineering
An HTTP(S) mock server and proxy for testing that allows developers to simulate APIs, record network traffic, and perform chaos engineering to test system resilience.
rf-detr: a real-time transformer architecture for high-accuracy object detection, instance segmentation, and keypoint detection
RF-DETR is a real-time transformer architecture for object detection, instance segmentation, and keypoint detection that balances state-of-the-art accuracy with low latency.
codex-slides: an agent-native AI slide studio that transforms code and documents into production-ready presentations
An open-source AI slide studio for Codex coding agents that transforms prompts, repos, or files into professional presentations through a steerable, live-editing workflow.
bitsandbytes: a k-bit quantization library for reducing memory consumption during LLM inference and training
A PyTorch library that enables accessible large language models through 4-bit and 8-bit quantization for memory-efficient inference and training.
PINTO_model_zoo: a collection of pre-converted and quantized models across multiple AI frameworks for edge deployment
A comprehensive repository of neural network models inter-converted and quantized across multiple frameworks like TensorFlow, PyTorch, ONNX, and CoreML for easier edge deployment.
opencvsharp: a cross-platform .NET wrapper that brings native OpenCV computer vision functionality to C#
A cross-platform .NET wrapper for OpenCV that brings comprehensive image processing and computer vision functionality to C# developers.
opencv: a comprehensive open-source library for computer vision and AI development
OpenCV is an open-source computer vision library that provides tools for developers to process and interpret visual data for AI applications.
OmniVoice-Studio: a local-first voice AI studio for zero-shot cloning, cinematic dubbing, and multi-language synthesis
An open-source, local-first alternative to ElevenLabs that offers zero-shot voice cloning, video dubbing, and dictation across 646 languages without requiring cloud accounts or API keys.
WindsurfAPI: a multi-standard API gateway that exposes Windsurf's 100+ AI models as OpenAI and Anthropic compatible endpoints
A Node.js proxy that converts Windsurf/Devin's internal AI model access into standard OpenAI, Anthropic, and Gemini compatible APIs for use in external tools.
fsrs4anki: a machine-learning based spaced-repetition scheduler that optimizes Anki review intervals
A modern spaced-repetition scheduler for Anki that uses machine learning to optimize card review intervals based on a user's own memory patterns.
m_flow: a graph-routed RAG system that uses path-cost scoring for cognitive-style memory retrieval
M-flow is a graph-based RAG framework that uses hierarchical memory structures and path-cost scoring to provide more accurate, reasoning-based retrieval for AI agents.
Skywork-R1V: a multimodal reasoning model with advanced visual chain-of-thought capabilities and SOTA performance across complex benchmarks
Skywork-R1V is an open-source multimodal reasoning model that uses reinforcement learning and visual chain-of-thought to achieve state-of-the-art performance in complex visual logic, math, and physics tasks.
llm-space: a desktop application for building, debugging, and evaluating AI agents
A desktop application for agent builders to prototype, debug, trace, and evaluate AI agent workflows locally.
Helios: a real-time 14B video generation model for high-quality minute-scale synthesis
Helios is a 14B real-time long video generation model capable of producing high-quality, minute-scale videos at up to 19.5 FPS on a single H100 GPU.
LightX2V: a high-performance inference framework for efficient image and video synthesis across diverse hardware
A lightweight inference framework for image and video generation that optimizes speed and memory usage through step distillation, quantization, and efficient offloading.
ort: a high-performance Rust wrapper for ONNX Runtime and other pure-Rust runtimes
A Rust interface for hardware-accelerated inference and training of ONNX machine learning models, enabling efficient deployment across various hardware and platforms.
huggingface.js: a comprehensive JS/TS toolkit for managing Hugging Face repositories and executing AI model inference
A collection of JavaScript and TypeScript libraries for interacting with the Hugging Face Hub and running inference on thousands of ML models.
spikingjelly: a PyTorch-native framework for large-scale spiking neural network training and inference
A PyTorch-native framework for spiking neural networks (SNNs) that supports large-scale training, inference, and deployment to neuromorphic hardware.
physicsnemo: a scalable deep-learning framework for building and training physics-informed AI models for science and engineering
An open-source deep-learning framework for building and training physics-informed AI models for science and engineering, optimized for NVIDIA GPUs.
lightly: a modular computer vision framework for self-supervised learning
A computer vision framework designed for self-supervised learning, providing modular building blocks for implementing advanced SSL algorithms like SimCLR and DINO.
boxmot: a pluggable multi-object tracking framework with unified Python and C++ APIs for any detector
A pluggable multi-object tracking library providing a unified Python and C++ API for running and deploying various tracking algorithms with any detector.
yolov3: a real-time object detection framework featuring multi-scale predictions and edge-optimized model variants
A PyTorch implementation of the YOLOv3 real-time object detection model, providing tools for training, inference, and deployment across various hardware.
onnxruntime: a cross-platform inference and training accelerator for machine learning models
A cross-platform machine learning accelerator that optimizes inference and training for models from various frameworks like PyTorch and TensorFlow.
supervision: a model-agnostic computer vision toolkit for dataset management and result visualization
A computer vision toolkit that provides model-agnostic building blocks for data loading, visualization, and dataset management to accelerate the development of vision applications.
streamlit: a Python framework for rapidly building and sharing interactive data and AI applications
A Python framework that transforms scripts into interactive web apps, enabling data scientists to build and share dashboards and AI chat apps without front-end experience.
yolov5: a fast and easy-to-use computer vision model for object detection, segmentation, and classification
A fast and accurate computer vision model based on PyTorch for real-time object detection, image segmentation, and image classification.
mirage: a unified virtual file system that lets AI agents interact with diverse data sources using standard bash commands
Mirage is a unified virtual file system that allows AI agents to interact with diverse data sources like S3, Slack, and Google Drive using standard bash commands.
liteflow: a rules engine framework that orchestrates complex business logic and AI agents using a lightweight DSL
A modern rules engine framework for Java that decouples complex business orchestration and allows AI Agents to be integrated as standard components within business rules.
SenseNova-Skills: a suite of end-to-end office capabilities for AI agents to automate research, data analysis, and presentation generation
A collection of end-to-end office skills for AI agents to handle image generation, PPT creation, Excel data analysis, and deep research.
holaOS: a local-first workspace for running multiple AI agents with shared memory and interactive app surfaces
A local-first workspace that allows multiple AI agents to share a single memory, set of tools, and interactive application surfaces.