OpenMined/TenSEAL
TenSEAL is a Python library for homomorphic encryption on tensors, built on Microsoft SEAL. It lets you run arithmetic operations on encrypted data, enabling privacy-preserving machine learning without exposing raw information.
wangziqi06/724-office
A framework-free, self-evolving AI agent system built in pure Python that provides 24/7 personal assistance with multi-tenant routing and a three-layer memory system.
langwatch/scenario
An agent testing framework that uses simulations, user simulators, and judge agents to evaluate the behavior and robustness of AI agents in multi-turn conversations.
beizhu-1209/AIHelms
An enterprise AI resource management platform that helps organizations control AI costs, manage model access, and quantify AI productivity through centralized identity and budget tracking.
agent-sh/agentsys
A modular runtime and orchestration system for AI agents that automates software development lifecycles through structured pipelines, quality gates, and specialized agent roles.
google-deepmind/xmanager
A framework for packaging, running, and tracking machine learning experiments across local environments, Kubernetes, and Google Cloud Platform.
NVIDIA/nvbench
A C++17 library for simplifying CUDA kernel benchmarking, offering parameter sweeps and detailed throughput and bandwidth measurements.
go-skynet/go-llama.cpp
Go bindings for llama.cpp that enable high-performance local LLM inference within Go applications using the GGUF format.
nengo/nengo
Nengo is a Python library for building and simulating large-scale neural models, supporting both spiking and non-spiking simulations.
microsoft/typeagent-py
A Python implementation of TypeAgent KnowPro for building structured RAG systems, serving as an experimental prototype for structured retrieval.
lightonai/pylate
A library for simplifying the fine-tuning, inference, and retrieval of ColBERT late interaction models, providing high-speed indexing and flexible training options.
SeldonIO/MLServer
An open-source inference server that provides a standardized REST and gRPC interface for serving machine learning models across various frameworks.
ConardLi/easy-agent
A terminal-native coding agent built with TypeScript and Node.js that integrates LLMs with local file and shell tools for automated software development.
nianticlabs/spz
A C++ library and file format for compressed 3D Gaussian splats that reduces file sizes by up to 10x compared to .ply files with minimal visual difference.
aws-samples/sample-mobile-ai-assistant
A cross-platform AI assistant app for Android, iOS, and macOS that enables real-time chat, multimodal analysis, and the instant creation of web applications using various LLM providers.
instadeepai/jumanji
A diverse suite of scalable, hardware-accelerated reinforcement learning environments written in JAX for high-speed RL research and experimentation.
ThinkWatchProject/ThinkWatch
An enterprise-grade secure gateway that provides auditing, governance, and access control for AI API calls and MCP tool invocations across an organization.
scaleapi/llm-engine
An open-source engine for fine-tuning and serving large language models, providing tools to deploy models on hosted or self-managed Kubernetes infrastructure.
WeianMao/triattention
A KV-cache compression method for long-reasoning LLMs that uses trigonometric frequency-domain compression to reduce memory usage by 10.7x and boost throughput by 2.5x without accuracy loss.
AQBot-Desktop/AQBot
A local-first desktop AI workbench that unifies multiple LLM providers, AI agents, knowledge bases, and MCP tools into a single, privacy-focused interface.
MiniAiLive/Android-FaceRecognition
An Android SDK for facial recognition and 3D passive liveness detection to prevent spoofing in security and fintech applications.
ckreiling/mcp-server-docker
An MCP server that enables LLMs to manage Docker containers, images, and volumes using natural language commands.
raiyanyahya/recall
A fully-local session memory tool for Claude Code and OpenCode that automatically logs activity and creates compact summaries to eliminate the cold-start problem.
inworld-ai/tts
A training and modeling framework for SpeechLM-based text-to-speech systems, supporting pre-training, SFT, and RLHF alignment across single or multi-GPU clusters.
vybestack/llxprt-code
A provider-agnostic CLI AI coding assistant that lets developers use any LLM provider, including local models and existing subscriptions, to edit codebases and automate workflows.
meta-pytorch/torchforge
A PyTorch-native agentic RL library that separates infrastructure concerns from model concerns to simplify scalable reinforcement learning experimentation.
openyak/openyak
OpenYak is an open‑source desktop GUI (Electron + React) that lets you chat with Codex or Claude Code agents while keeping all conversation history, files and project data locally (Rust + SQLite). It integrates the official runtimes via stdio, shows referenced files (Markdown, HTML, PDF, DOCX, code), and offers a shared Playwright‑driven browser where both user and agent can control the same page. Run it with Node 26, Rust 1.90, and your own Codex/Claude CLI binaries; the app is in v2 alpha and licensed Apache‑2.0.
lyon-industries/graphrag-workbench
A local workbench that turns documents into an inspectable 3D knowledge graph by wrapping Microsoft GraphRAG with a visual interface for indexing and exploration.
innocommerce/innoshop
An open-source e-commerce system based on Laravel 13 that integrates AI models and Model Context Protocol (MCP) for intelligent automation.
probabl-ai/skore
Skore is a Python library that structures machine learning experiments by reducing evaluation boilerplate and providing built-in methodological guidance for data scientists.
groq/groq-appgen
An interactive web application that uses Groq's LLM API to generate and modify web applications based on natural language queries.
google/yggdrasil-decision-forests
A fast and extensible library for training, evaluating, interpreting, and serving decision forest models like Random Forest and Gradient Boosted Trees.
layumi/University1652-Baseline
A multi-view, multi-source benchmark and baseline for drone-based geo-localization, enabling the matching of drone, satellite, and street-view images to locate buildings.
OpenRouterTeam/ai-sdk-provider
An OpenRouter provider for the Vercel AI SDK that enables access to 300+ LLMs and embedding models through a unified interface.
Osilly/Vision-DeepResearch
A framework and benchmark for developing Multimodal Large Language Models capable of deep research through iterative visual and textual search across images and videos.
Ryan-yang125/ChatLLM-Web
A private, browser-native AI studio and agent workspace that runs LLMs locally on your device using WebGPU, requiring no API keys or backend servers.
machinewrapped/llm-subtrans
LLM‑Subtrans is a Python‑based desktop and CLI tool that translates subtitle files (SRT/ASS/VTT) using LLM APIs (Gemini, OpenAI, Anthropic, etc.). It supports pluggable subtitle formats, optional audio transcription, project‑file resume, and many advanced CLI options. Install via pre‑built Windows/macOS packages or from source with a virtual‑env installer; run through a simple GUI or command line, providing the provider’s API key.
a-bonus/google-docs-mcp
An MCP server that connects AI assistants to Google Docs, Sheets, Drive, Gmail, and Calendar, enabling LLMs to read and manage Workspace data.
dyndynjyxa/aio-coding-hub
A local AI CLI unified gateway that provides a single entry point for tools like Claude Code and Gemini CLI, featuring intelligent failover, usage tracking, and workspace isolation.
Azure/co-op-translator
An AI-powered tool that keeps multilingual GitHub documentation current by automating the translation of Markdown, Jupyter notebooks, and images while preserving repository structure and links.
PozzettiAndrea/ComfyUI-TRELLIS2
Custom ComfyUI nodes that implement Microsoft's TRELLIS.2 model to generate high-quality 3D meshes with PBR materials from a single image.
google/paxml
A framework for configuring and running large-scale machine learning experiments on top of Jax, optimized for Cloud TPU and GPU infrastructure.
ZubeidHendricks/youtube-mcp-server
An MCP server that enables AI language models to interact with YouTube by providing tools to search videos, retrieve transcripts, and analyze channel data.
neiltron/apple-health-mcp
A local Node.js server that lets AI‑assistant clients (via MCP) run SQL queries on a personal Apple Health CSV export, using an in‑memory DuckDB database. It provides schema discovery, ad‑hoc queries, and health‑summary reports while keeping data on the user’s machine.
ROCm/composable_kernel
A library providing a programming model for writing performance-critical machine learning kernels that are portable across different GPU and CPU architectures.
Taskosaur/Taskosaur
An open-source project management platform that uses conversational AI and browser automation to execute tasks and manage workflows via natural language.
ksenxx/kiss_ai
KISS Sorcar is an open‑source, local‑first AI‑agent framework that runs a daemon (`kiss‑web`) and provides VS Code, web/mobile, and Python interfaces for issuing natural‑language tasks. It supports any LLM you configure (OpenAI, Anthropic, Gemini, local endpoints, etc.), mixes multiple models in one workflow, and includes 43 built‑in third‑party agents for messaging, email, smart‑home, and productivity services. Custom Python “extension agents” let you define prompts, tools, budgets, and safety hooks per task. All prompts stay on your machine; the system handles git worktree isolation, voice wake‑word, scheduled automations, and tool calling, making it a privacy‑focused, extensible personal AI assistant.
FutureMLS-Lab/OSCAR
OSCAR is a 2-bit KV cache quantization method that uses offline spectral covariance-aware rotations to reduce memory usage by 8x while maintaining near-BF16 accuracy.