Waikato/moa

MOA is an open-source Java framework for real-time data stream mining and large-scale machine learning, providing a suite of algorithms and evaluation tools.

inclusionAI/Ming

Ming-flash-omni 2.0 is an open-source omni-MLLM that unifies multimodal perception and generation across text, image, audio, and video in a single MoE-based architecture.

aws-neuron/aws-neuron-sdk

AWS Neuron is an SDK for high-performance deep learning and generative AI, enabling the deployment and profiling of workloads on AWS Inferentia and Trainium accelerators.

Deep-MI/FastSurfer

A fast, deep-learning based neuroimaging pipeline for volumetric and surface-based thickness analysis of brain MRI images, serving as a high-speed alternative to FreeSurfer.

yzfzzz/depth-detect

A high-performance C++ and TensorRT framework that fuses YOLO object detection with monocular/binocular depth estimation to track objects and estimate their motion states.

prov-gigapath/prov-gigapath

A whole-slide foundation model for digital pathology that uses a tile-and-slide encoder architecture to analyze massive pathology images.

Stevenic/vectra

A local, file-backed, in-memory vector database that supports metadata filtering, cosine similarity, and local or API-based embeddings.

Compresr-ai/Context-Gateway

A proxy that sits between AI agents and LLM APIs to compress conversation history in the background, preventing delays when context limits are are reached.

princepainter/ComfyUI-PainterI2V

A ComfyUI node for Wan2.2 that fixes slow-motion issues in image-to-video generation by increasing motion amplitude and improving camera movement responsiveness.

david8862/keras-YOLOv3-model-set

A comprehensive TensorFlow Keras pipeline for YOLOv2, v3, and v4 object detection, supporting diverse backbones, advanced training techniques, and on-device deployment.

ucinlp/autoprompt

An automated prompt generation tool that uses gradient-guided search to create optimized prompts for masked language models to perform NLP tasks.

mlcommons/ck

A community-driven automation framework for AI/ML research that enables reproducible and portable workflows across diverse hardware, software, and datasets.

uxlfoundation/oneDAL

A C++ and DPC++ library that provides hardware-accelerated machine learning routines for tabular data across CPUs, GPUs, and distributed systems.

bytedance/vidi

A family of Large Multimodal Models (LMMs) for video understanding and creation, enabling tasks like temporal retrieval, spatio-temporal grounding, and automated video editing.

aryn-ai/sycamore

An AI-powered document processing engine for ETL and RAG that transforms unstructured documents into high-quality data chunks for vector databases.

talmolab/sleap

A deep-learning framework for multi-animal pose tracking that uses a human-in-the-loop GUI to rapidly label and quantify animal behavior.

linkedin/venice

A derived data storage platform designed for planet-scale workloads, often used as the stateful backend for ML feature stores to enable low-latency online inference.

jingkaihe/matchlock

A CLI tool and SDK for running AI agents in isolated, ephemeral microVMs with network allowlisting and secure secret injection via MITM proxy.

tidyverse/ellmer

A package that makes it easy to use large language models from R, supporting a wide range of providers and features like streaming and tool calling.

orbital-materials/orb-models

A library of pretrained neural network potentials for fast, high-accuracy atomic simulations of molecules and crystals, replacing expensive DFT calculations.

chennuo0125-HIT/LIO-SAM_based_relocalization

A relocalization system based on LIO-SAM that allows a robot to determine its position within a previously saved map.

hackingmaterials/matminer

A data mining library for materials science that provides tools for retrieving datasets and featurizing materials data for analysis.

cgnomads/GSOPs

A SideFX Houdini plug-in that provides a comprehensive toolset for importing, editing, and animating Gaussian splatting scenes for VFX production.

f-dangel/backpack

A PyTorch extension that efficiently computes quantities beyond the standard gradient, such as per-sample gradients and gradient variance, by reusing information during the backward pass.

francescofugazzi/3dgsconverter

A high-performance tool for converting 3D Gaussian Splatting files between multiple formats with GPU-accelerated filtering and compression.

krishn404/Git-Friend

An AI-powered GitHub assistant that provides AI chat for Git help, automated README generation, and Gitmoji support to streamline repository management.

scikit-tda/scikit-tda

A collection of Topological Data Analysis (TDA) Python libraries designed to make topological data tools accessible to non-topologists.

PKU-Alignment/safety-gymnasium

A scalable and customizable Safe Reinforcement Learning benchmark library providing a standardized set of environments for evaluating safety-constrained AI agents.

Samsung/ONE

ONE is a high-performance on-device neural network inference framework that enables AI models from TensorFlow and PyTorch to run efficiently across various hardware accelerators.

sambanova/bloomchat

BLOOMChat is a 176‑billion‑parameter multilingual chat LLM fine‑tuned from the open‑source BLOOM model. The repo provides data‑prep, tokenisation, and training scripts (the latter for SambaNova’s RDU hardware) plus detailed GPU inference instructions using Hugging Face’s Bloom inference code. It’s a genuine, large‑scale LLM project aimed at researchers and developers who want to reproduce or run the model.

deepgenteam/deepgen

DeepGen 1.0 is a lightweight 5B parameter unified multimodal model that integrates image generation, editing, and reasoning into a single framework.

remyxai/VQASynth

A pipeline for transforming image datasets into spatial VQA datasets to improve the 3D spatial reasoning and distance estimation capabilities of Vision-Language Models.

matiasdelellis/facerecognition

A facial recognition app for Nextcloud that detects, analyzes, and groups faces in images locally to provide private, person-based photo organization.

ZimoLiao/scholaraio

An academic harness for AI agents that provides the structured evidence, library management, and repeatable workflows needed for professional research and academic writing.

FaceAISDK/FaceRecognition_ReactNative

A React Native demo and integration guide for an offline face recognition SDK that supports 1:1 verification and liveness detection on iOS and Android.

mlverse/torch

An R package that brings PyTorch functionality to the R ecosystem, enabling tensor operations and automatic differentiation for machine learning.

maquina-app/rails-mcp-server

A Ruby implementation of the Model Context Protocol (MCP) server that allows LLMs to analyze and interact with Rails projects through structured tools and documentation.

TheAuditorTool/Auditor

A local, database-first code intelligence and SAST platform that turns codebases into queryable facts to reduce token costs for AI agents and security teams.

AutoLab-SAI-SJTU/Paper2Rebuttal

An AI-powered multi-agent system that helps researchers analyze reviewer comments, search for supporting literature, and draft formal academic paper rebuttals.

whitzard-ai/jade-db

JADE is a safety evaluation platform that uses linguistic mutation to generate high-risk datasets for testing the safety guardrails and alignment of LLMs and AI agents.

SciML/Catalyst.jl

A symbolic modeling package for the high-performance simulation and analysis of chemical reaction networks and dynamical systems.

DataDog/toto

A foundation model for multivariate time series forecasting optimized for observability metrics, providing zero-shot and probabilistic predictions.

OpenGVLab/VideoChat-Flash

VideoChat-Flash is a multimodal large language model designed for efficient long-context video understanding, capable of processing videos up to three hours long using hierarchical compression.

polyaxon/traceml

An engine for ML and data tracking, visualization, and explainability that integrates with major deep learning frameworks to log metrics, inputs, and artifacts.

brontoguana/krasis

An LLM runtime that enables running massive Mixture-of-Experts (MoE) models on consumer NVIDIA GPUs by managing expert residency between VRAM and CPU RAM.

WenjieDu/SAITS

A self-attention-based framework for multivariate time series imputation that solves the speed and memory issues of RNN-based models.

echoVic/orca-agent

A DeepSeek-native coding agent for the terminal that can read code, edit files, and run commands to autonomously solve development tasks.

milvus-io/milvus-sdk-java

A Java client library for Milvus, enabling Java developers to integrate vector database capabilities into their AI applications.