jd-opensource/JoyAI-Video-Edit
A real-time, instruction-guided video editing system that uses autoregressive diffusion to edit live video streams or uploads as frames arrive.
TheOrcDev/videorc
Videorc is an open‑source, AI‑enhanced desktop studio that lets creators capture screen/camera, stream to multiple platforms, and automatically generate transcripts, titles, chapters, and highlights—all from a single, scene‑based UI.
WeChatCV/Wan-Alpha
A video generation framework that enables the creation of high-quality videos with stable transparency (alpha channels) using a Shiftable RGB-A Distribution Learner.
AlayaLab/AlayaWorld
An interactive autoregressive world model that enables long-horizon, playable video generation with real-time camera control and consistent spatial memory.
JayWebtech/autoshorts
A local-first desktop app that uses AI to identify viral moments in long-form video or audio and automatically crops them into vertical short-form clips.
NVlabs/LongLive
An NVFP4-powered parallel infrastructure for long video generation that optimizes training and inference for real-time, high-throughput performance.
VelornLabs/velorn
An open-source AI video workstation that combines a full timeline editor with AI generation workflows via ComfyUI and MCP-based agent automation.
NVIDIA/flashdreams
A high-performance inference and serving library for interactive autoregressive video and world models, optimized for real-time applications in robotics and autonomous vehicles.
shengshu-ai/minWM
A full-stack open-source framework for turning text-to-video foundation models into action-conditioned, real-time video world models.
ronak-create/FableCut
A browser‑only, non‑linear video editor whose timeline is a JSON file that can be edited live by humans or AI agents via MCP, REST or direct file writes. Zero npm dependencies, optional ffmpeg for fast export, and built‑in AI background removal.
amap-cvlab/ABot-World
ABot-World is a real-time interactive world simulator that enables infinite, action-conditioned video generation on a single desktop GPU.
cclank/lanshu-create-ai-presenter-video
A generalized AI digital human video production workflow that transforms scripts and images into synchronized presenter videos with subtitles and effects.
Netflix/vmaf
An Emmy-winning perceptual video quality assessment tool by Netflix that objectively measures video degradation using a fusion of multiple metrics.
Anil-matcha/Open-AI-Micro-Drama-Generator
An agentic AI pipeline that transforms a simple idea or script into a complete cinematic micro-drama video by automating story development, character design, and video production.
ModelTC/Minimax-H3-Turbo
A collection of distilled LoRA checkpoints for MiniMax-H3 that enables fast, few-step video and audio generation via Diffusers and ComfyUI.
vrgamegirl19/comfyui-vrgamedevgirl
A collection of ComfyUI custom nodes featuring an AI Video Builder for scene-by-scene production, video enhancement, and LoRA training for AI video and music video creation.
Agentchengfeng/chengfeng-videocut-skills
A Codex Plugin installation entry point that provides AI agents with the tools and methods needed to perform video editing via the chengfeng-videocut runtime.
Soul-AILab/SoulX-FlashTalk
SoulX-FlashTalk is a system for real-time, infinite streaming of audio-driven avatars using self-correcting bidirectional distillation to create continuous talking head animations.
RafaelGodoyEbert/ViralCutter
An open-source, local AI tool that transforms long YouTube videos into viral short-form clips with automatic cropping, dynamic captions, and face tracking.
thu-ml/Causal-Forcing
A framework for high-quality real-time interactive video generation using Causal ODE or Causal Consistency Distillation to enable efficient few-step autoregressive diffusion.
Robbyant/lingbot-video
An open-source MoE video generation model dedicated to embodied intelligence, designed to produce physically rational videos for robotics and physical world understanding.
SamurAIGPT/Seedance-2.5-API
A Python wrapper for ByteDance's Seedance 2.5 API that enables the generation of high-fidelity AI videos with realistic human faces and consistent character identities.
james-see/ltx-video-mac
A native SwiftUI macOS app for local AI video generation on Apple Silicon, supporting LTX and MiniMax H3 models with integrated audio and voiceover tools.
huytranvan2010/AI-auto-generate-video
An automated pipeline that converts Vietnamese text articles into 9:16 short-form videos using AI for scripting and deterministic HTML-to-video rendering.
shinkuan/Akagi
Akagi is a Rust‑based desktop assistant for online riichi mahjong (Mahjong Soul, Tenhou, etc.). It captures game traffic via a MITM proxy or Chromium DevTools, runs a built‑in neural‑net bot (and optional cloud‑inference model) to compute shanten, waits, win probabilities, opponent risk, and suggested discards, and shows these numbers in a draggable HUD. Completed matches are stored locally and visualised in a History tab with rank pie charts, PT trend lines, and detailed stats. The app is distributed as a single portable binary for Windows, macOS (Apple Silicon) and Linux, uses Tauri 2 + React for the UI, and is licensed under Apache 2.0.
Lightricks/ComfyUI-LTXVideo
A collection of custom ComfyUI nodes and workflows that extend the LTX-2 video generation model with features like HDR conversion, multilingual dubbing, and generative upscaling.
BrokenSource/DepthFlow
DepthFlow is an open-source image-to-video converter that transforms static pictures into 3D parallax animations using depth estimation and optimized shaders.
LeonQ8/ComfyUI-ALLinONE-MinimaxH3
A unified ComfyUI node that simplifies the MiniMax H3 video pipeline by consolidating text-to-video, image-to-video, and advanced editing tools into a single interface.
WhatDreamsCost/WhatDreamsCost-ComfyUI
WhatDreamsCost‑ComfyUI is a collection of custom ComfyUI nodes that add video‑ and audio‑editing capabilities (timeline editor, load/trim nodes, IC‑LoRA support, audio in‑painting, etc.) for AI‑generated media workflows.
ItusiAI/Open-Magiviz
An AI-powered video creation platform that automates the workflow from script generation and character design to storyboard creation and final video rendering.
NO6KIKO/gorest-2d-animation-spritesheet-generator
Gorest is a Node‑based, browser‑run editor that lets you build 2‑D game scenes, generate or import spritesheets, attach rich metadata, and preview animations—all with optional Codex‑driven prompts. It stores assets locally and includes a fake‑3D rotator for turntable‑style sprites.
naqashafzal/AI-Content-Studio
An automated AI video and podcast generator that converts long-form YouTube videos into viral shorts and creates AI-powered podcasts from text prompts.
agan-j/xiaoniu
An AI-powered video translation tool that automates speech translation, voice cloning, and lip-syncing to help creators localize short dramas and promotional videos for global audiences.
rediumvex/ai-video-generator-claude
A collection of prompt engineering skills for Claude that generate professional, cinematic video prompts for Seedance 2.0 on Higgsfield.
google-deepmind/tapnet
Google DeepMind’s TAP‑Net repository offers datasets, pretrained models (TAPIR, BootsTAPIR, TAPNext/Next++), training code, and Colab/real‑time demos for the “Tracking Any Point” task—precise point‑level video tracking useful for vision research and robot imitation.
NVlabs/rcm
A score-regularized continuous-time consistency model for distilling large video diffusion models, enabling high-quality, diverse video generation in 2-4 steps.
toki-plus/video-mover
An automated video distribution pipeline that handles everything from file monitoring and AI-generated captions to multi-platform uploading using browser automation.
mira-wm/mira
MIRA is a real-time world model of Rocket League that uses a 5B parameter latent diffusion model to simulate 2v2 matches based on player actions.
SamurAIGPT/Text-To-Video-AI
An AI-powered video creation tool that transforms text prompts into fully edited videos with scripts, voiceovers, B-roll, and synchronized captions.
bytedance/Bernini
Bernini is a unified video generation and editing framework that combines an MLLM-based semantic planner with a DiT-based renderer to improve instruction following and complex video edits.
hzwer/Practical-RIFE
A practical implementation of RIFE and SAFA for video frame interpolation and enhancement, allowing users to increase video frame rates for smoother motion.
aigc-apps/VideoX-Fun
A video generation pipeline supporting CogVideoX-Fun and Wan 2.1 models for text, image, and video-to-video generation with advanced structural and camera controls.
aqm857886159/Nomi
An open-source, local-first desktop workbench for AI video that lets users combine multiple generation sources and local ComfyUI to lower production costs.
showlab/Code2Video
An agentic framework that generates high-quality educational videos by synthesizing executable Manim code rather than pixels, ensuring mathematical precision and coherence.
zhouwei713/seedance-prompt
A prompting framework and skill for AI video generators that replaces generic descriptors with detailed simulations of real-world recording equipment and human imperfections to eliminate the "AI look."
geekjourneyx/hyperframes-motion-director
An Agent Skill that transforms text content into structured motion-video productions, featuring a professional two-phase workflow for planning and asset generation.
cclank/lanshu-awesome-ai-video-kit
A comprehensive AI video prompt engineering kit featuring 543 tested prompts, 15 supported models, and automated monitoring of official endpoints to ensure prompting techniques stay current.
nv-tlabs/omni-dreams
A world model that generates photorealistic, multi-camera video in real time for autonomous driving simulation using text, HD maps, and trajectory poses.