GAA-UAM/scikit-fda

A Python package for Functional Data Analysis (FDA) that provides tools for the representation, preprocessing, and statistical analysis of data depending on a continuous parameter.

castorini/anserini

Anserini is a toolkit built on Apache Lucene designed to make information retrieval research reproducible and bridge the gap between academic baselines and real-world search applications.

rapidsai/jupyterlab-nvdashboard

A JupyterLab extension that provides real-time GPU usage dashboards and one-click GPU acceleration for pandas and scikit-learn.

codalab/codalab-competitions

CodaLab is an open-source web platform for organizing and operate scientific challenges and competitions in machine learning and advanced computation research.

usc-isi-i2/kgtk

A comprehensive framework for creating and analyzing large-scale hyper-relational knowledge graphs using a scalable, disk-based TSV format.

combust/mleap

A performant, portable execution engine and serialization format for deploying machine learning pipelines from Spark and Scikit-learn without their heavy dependencies.

kuafuai/aipexbase

A Backend-as-a-Service infrastructure that provides AI-native backend capabilities, allowing developers to build full-stack AI applications without writing backend code.

meta-pytorch/monarch

A distributed programming framework for PyTorch that uses scalable actor messaging, supervision trees for fault tolerance, and RDMA for high-performance memory transfers.

Tiramisu-Compiler/tiramisu

A polyhedral-based compiler for data-parallel computations that optimizes algorithms for CPUs, GPUs, FPGAs, and distributed systems, specifically for deep learning and tensor algebra.

Microck/opencode-studio

A local GUI for managing OpenCode configurations, allowing users to toggle MCP servers, edit skills, and manage agents without editing JSON files.

robertcprice/nCPU

A complete computer where every layer, from the ALU and OS to the compiler, is implemented as a trained neural network or GPU kernel, enabling differentiable program synthesis.

lablup/backend.ai

A container-based computing cluster platform that orchestrates ML frameworks and provides multi-tenant access to heterogeneous hardware accelerators like GPUs and NPUs.

Kamalnrf/claude-plugins

A plugin manager and skills installer that provides a centralized registry and CLI tools to discover and install capabilities for AI coding agents like Claude Code and Cursor.

ekondis/mixbench

A benchmark tool for evaluating the performance bounds of GPUs and CPUs by testing mixed operational intensity kernels across various precision levels.

computerlovetech/agr

A package manager for AI agent skills that allows teams to install, version, and sync skills from Git repositories across multiple AI tools.

flowagi-eu/nyno

Nyno is a backend-as-code framework that allows developers to build and deploy AI workflows using a simple configuration language instead of writing complex backend code.

aws-samples/genai-quickstart-pocs

A collection of proof-of-concept samples demonstrating various Generative AI use cases using Amazon Bedrock, featuring Streamlit frontends for rapid prototyping.

arun1729/cog

A persistent, embedded graph database for Python that uses a fluent API for queries and supports SIMD-optimized vector similarity search.

dfm/tinygp

A lightweight Gaussian Process library built on JAX that supports GPU acceleration and automatic differentiation.

lab-v2/pyreason

PyReason is an explainable graphical inference tool that uses logical rules and facts to reason over graph structures using temporal and real-valued logic.

shenjingnan/xiaozhi-client

A client that connects AI servers to Model Context Protocol (MCP) servers, enabling AI agents to access external tools and data sources.

kubernetes-sigs/jobset

A Kubernetes-native API for managing groups of Jobs as a single unit, specifically designed to simplify the deployment of distributed AI/ML training and HPC workloads.

python-adaptive/adaptive

A Python library for parallel active learning of mathematical functions that intelligently selects the best points to sample in a parameter space to reduce computational cost.

retentioneering/retentioneering-tools

An open-source Python toolkit and MCP server for reproducible product analytics on clickstream and event log data, enabling deep behavioral analysis of user journeys.

jazzenchen/VibeAround

A centralized hub for AI coding agents that simplifies launching, API profile management, and session continuity across desktop, CLI, and messaging apps.

ServBay/ServBay

A comprehensive local web development environment manager for macOS and Windows that bundles web servers, databases, and local AI capabilities via Ollama.

cazala/synaptic

A JavaScript neural network library for Node.js and the browser that allows the creation and training of flexible, architecture-free neural networks.

JohnSnowLabs/spark-nlp

A state-of-the-art NLP library built on Apache Spark that provides scalable, production-ready NLP annotations and pre-trained models for text, image, and speech across 200+ languages.

aws/sagemaker-python-sdk

An open-source library for training and deploying machine learning models on Amazon SageMaker, supporting popular frameworks and foundation model fine-tuning.

facebookresearch/schedule_free

A collection of PyTorch optimizers that eliminate the need for learning rate schedules and predefined stopping times, enabling faster and more flexible model training.

mckinsey/causalnex

A toolkit for causal reasoning and "what-if" analysis using Bayesian Networks to identify causal relationships and estimate the impact of interventions.

Accenture/AmpliGraph

A TensorFlow-based library for relational learning that predicts missing links and generates embeddings for concepts within knowledge graphs.

facebookresearch/fvcore

A lightweight core library providing essential shared functionality and utilities for computer vision frameworks developed by FAIR.

flexflow/flexflow-train

A deep learning framework that accelerates distributed DNN training by automatically searching for the most efficient parallelization strategies.

NVIDIA/MatX

A C++20 library for numerical computing that brings NumPy-like tensor expressions to NVIDIA GPUs and CPUs, featuring JIT kernel fusion for maximum performance.

milvus-io/pymilvus

A Python SDK for Milvus, providing a programmatic interface to manage and search high-dimensional vector embeddings in the Milvus vector database.

stared/livelossplot

A Python package that provides live training loss and metric plots directly in Jupyter Notebooks for Keras, PyTorch, and other deep learning frameworks.

owlbarn/owl

A comprehensive scientific computing system for OCaml that provides n-dimensional arrays, linear algebra, and deep learning capabilities for high-performance analytical code.

aws/deep-learning-containers

A collection of pre-built, security-patched Docker images for running AI/ML workloads on AWS, providing optimized environments for frameworks like PyTorch, TensorFlow, and vLLM.

manaflow-ai/manaflow

Manaflow is an open-source orchestrator that lets you run multiple AI coding agents in parallel across isolated VS Code workspaces with integrated live previews and PR tools.

mlr-org/mlr3

An object-oriented machine learning framework for R that provides efficient, standardized building blocks for tasks, learners, and resampling.

apache/systemds

Apache SystemDS is an open-source machine learning system that manages the end-to-end data science lifecycle from data preparation to model serving using a high-level language.

BabitMF/bmf

A cross-platform multimedia processing framework by ByteDance that enables high-performance video transcoding, editing, and AI inference with strong GPU acceleration.

huggingface/dataset-viewer

The backend API for the Hugging Face dataset viewer, allowing users to browse, filter, and inspect datasets on the Hub through pre-computed data.

kendryte/nncase

A neural network compiler for AI accelerators that optimizes and deploys models from TFLite, Caffe, and ONNX formats onto specialized hardware like the K210, K510, and K230.

SynaLinks/synalinks-skills

A collection of standardized agent skills that teach AI coding agents how to use the Synalinks framework idiomatically, preventing syntax errors and hallucinations.

astroautomata/SymbolicRegression.jl

A Julia library for symbolic regression that searches for human-readable mathematical expressions to fit datasets, optimizing the trade-off between accuracy and complexity.

omnimind-ai/OmniInfer

A high-performance, cross-platform inference engine for running LLMs and VLMs locally on desktop, mobile, and edge devices.