Stonewuu/ai-fusion-video

An Agent-driven video creation platform that integrates scriptwriting, storyboarding, and AI image/video generation into a unified workspace.

thedivergentai/GD-Agentic-Skills

GD Agentic Skills is a curated library of 99 Godot 4.7+ “Domain Skills” plus a master orchestrator skill, built as an external long‑term memory for AI coding agents. It provides best‑practice GDScript snippets, migration guides, performance budgets, and a vision skill that lets agents capture and score editor screenshots. Install via the `skills` CLI—either the all‑in‑one `godot‑master` for full‑project guidance or individual domain skills for focused features—while avoiding the “install all” trap that would flood the model’s context.

lcy362/agnes-video-generator

A free, open-source AI video generator that uses cloud APIs to create narrated, auto-subtitled multi-scene videos from text or images without requiring a high-end GPU.

HITsz-TMG/VideoClaw

An AI director system that automates the entire creative video production pipeline from scriptwriting and character design to storyboarding and final editing.

Carasibana/ComfyUI-H3-FaceRefine

A ComfyUI custom node set that improves the quality of small faces in MiniMax H3 videos by detecting, cropping, regenerating, and stitching them back into the original clip.

Rimagination/h3lite

A local deployment skill for MiniMax H3 that enables low-VRAM NVIDIA GPUs on Windows to generate videos with native audio via AI agents.

aqm857886159/Nomi

An open-source, local-first desktop workbench for AI video that lets users combine multiple generation sources and local ComfyUI to lower production costs.

KlingAIResearch/LivePortrait

An efficient portrait animation tool that animates static images of humans and animals using driving videos or images with precise stitching and retargeting control.

pyang5166/gbro-collage-broll

A tool that turns short voiceover lines into editorial halftone paper-collage B-roll animations using a three-stage approval workflow and Gemini Omni Flash.

ModelTC/LightX2V

A lightweight inference framework for efficient image and video generation that optimizes performance through quantization, distillation, and advanced parallelism.

amap-cvlab/ABot-World

ABot-World is a real-time interactive world simulator that enables infinite, action-conditioned video generation on a single desktop GPU.

LingGuoAI/LingGuo-Drama

An open-source workbench for AI short drama and video production that manages the pipeline from scriptwriting and asset generation to final video synthesis.

myccarl/ai-shortVideo-pipeline

myAiVideos is an open‑source, Docker‑compose‑based system that automates the whole short‑video creation workflow (topic → script → visuals → audio → post‑production → distribution) for Chinese content. It uses a FastAPI orchestrator plus a Java Spring‑Boot gateway for auth, routing, circuit‑breaking and metering, and integrates multiple LLM and multimodal models (DeepSeek, GLM, Kling, TTS services). The pipeline is layered, fail‑over capable, and fully observable via Langfuse and Prometheus.

HKUDS/VideoAgent

An all-in-one agentic framework for understanding, editing, and generating creative videos through natural language prompts.

huytranvan2010/AI-auto-generate-video

An automated pipeline that converts Vietnamese text articles into 9:16 short-form videos using AI for scripting and deterministic HTML-to-video rendering.

TNTwise/REAL-Video-Enhancer

A cross-platform GUI application for AI-powered video frame interpolation and upscaling, supporting multiple GPU backends and a variety of specialized models.

FireRedTeam/FireRed-OpenStoryline

A conversational AI video creation tool that automates scriptwriting, media sourcing, and editing through natural language prompts.

jianjieyiban/JJYB_AI_VideoAutoCut

A local-first AI video creation workstation that automates the workflow from material analysis and scriptwriting to voiceover generation and audio-visual synchronization.

cosmo-wander-ai/cosmo-edge

CosmoEdge 1.1 is an Apache‑2.0 C++ edge‑AI engine for video analytics and vision‑language models. It runs on Sophon (BM1688/CV186X), Rockchip (RK3576/RV1126B) NPUs and on x86/macOS via Docker, offering a browser‑based pipeline builder, REST/MQTT/WebSocket integration, and a “Model Guard” for protected commercial models. The repo includes Docker scripts for building device‑specific runtimes, benchmark reports, and extensive docs.

MeiGen-AI/OPSD-V

An on-policy self-distillation framework for post-training few-step autoregressive video generators to reduce error accumulation and improve motion dynamics in long videos.

FujiwaraChoki/supoclip

An open-source, AI-powered video clipping tool that converts long-form videos into viral vertical clips with automatic face-cropping, subtitles, and virality scoring.

Agentchengfeng/chengfeng-videocut-skills

A Codex plugin suite that automates talking-head video editing by coordinating AI-driven cutting, subtitle management, and exporting via a dedicated cross-platform Runtime.

vrgamegirl19/comfyui-vrgamedevgirl

A collection of ComfyUI custom nodes featuring an AI Video Builder for scene-by-scene production, video enhancement, and LoRA training for AI video and music video creation.

NVlabs/LongLive

An NVFP4-powered parallel infrastructure for long video generation that optimizes training and inference for real-time, high-throughput performance.

Lightricks/LTX-Desktop

LTX Desktop is an open‑source Electron app for generating and editing videos (and images) with LTX diffusion models. It runs locally on NVIDIA GPUs (≥16 GB VRAM) or Apple‑silicon Macs (≥15 GB free RAM) using the LTX 2.5 Fast or 2.3 Fast models, and automatically switches to a paid cloud API (LTX 2.5/2.3 Pro) on unsupported hardware. Features include text‑/image‑/audio‑to‑video, video retake/extend, LoRA style adapters, prompt enhancement via Gemini, and a timeline‑based video editor. Installation is via a GitHub release installer; the app stores data in standard OS‑specific folders and requires a free LTX API key for cloud text encoding. The codebase (Apache‑2.0) is split into React UI, Electron shell, and a Python FastAPI backend, with full dev scripts for local building. Optional anonymous telemetry can be disabled.

naqashafzal/AI-Content-Studio

An automated AI video and podcast generator that converts long-form YouTube videos into viral shorts and creates AI-powered podcasts from text prompts.

zai-org/SCAIL-2

SCAIL-2 is an open-source model for end-to-end character animation and replacement, removing the need for intermediate pose representations to support complex motions and diverse identities.

vllm-project/llm-compressor

A library for quantizing and compressing large language models (weights, activations, KV‑cache, etc.) so they run efficiently with the vLLM inference engine. Supports many algorithms (GPTQ, AWQ, AutoRound, REAP pruning, etc.) and formats (FP8, NVFP4, INT4, mixed‑precision). Provides one‑line API, DDP + disk‑offloading for multi‑terabyte models, and pre‑quantized checkpoints for popular LLMs.

rediumvex/ai-video-generator-claude

A collection of prompt engineering skills for Claude that generate professional, cinematic video prompts for Seedance 2.0 on Higgsfield.

Robbyant/lingbot-world

An open-source world simulator that generates high-fidelity, interactive video environments with real-time interactivity and long-term temporal consistency.

BrokenSource/DepthFlow

DepthFlow is an open-source image-to-video converter that transforms static pictures into 3D parallax animations using depth estimation and optimized shaders.

NO6KIKO/gorest-2d-animation-spritesheet-generator

Gorest is a Node‑based, browser‑run editor that lets you build 2‑D game scenes, generate or import spritesheets, attach rich metadata, and preview animations—all with optional Codex‑driven prompts. It stores assets locally and includes a fake‑3D rotator for turntable‑style sprites.

ModelTC/Minimax-H3-Turbo

A collection of distilled LoRA checkpoints for MiniMax-H3 that enables fast, few-step video and audio generation via Diffusers and ComfyUI.

OpenImagingLab/FlashVSR

FlashVSR is a one-step streaming framework for real-time diffusion-based video super-resolution that achieves high-resolution upscaling with significantly reduced latency.

LeonQ8/ComfyUI-ALLinONE-MinimaxH3

A unified ComfyUI node that simplifies the MiniMax H3 video pipeline by consolidating text-to-video, image-to-video, and advanced editing tools into a single interface.

s1dashu/director

director is an AI‑agent skill that guides a language model through the full video‑production pipeline—script, visual style, characters, voice, shot‑by‑shot prompts—and uses LibTV or 即梦 CLI tools to generate ready‑to‑edit video clips. It supports three modes (Animated Explainer, Storytime Animation, Cinematic Drama) and is installed as a skill directory for agents like Codex.

cclank/lanshu-create-ai-presenter-video

A generalized AI digital human video production workflow that transforms scripts and images into synchronized presenter videos with subtitles and effects.

JayWebtech/autoshorts

A local-first desktop app that uses AI to identify viral moments in long-form video or audio and automatically crops them into vertical short-form clips.

aigc-apps/VideoX-Fun

A video generation pipeline supporting CogVideoX-Fun and Wan 2.1 models for text, image, and video-to-video generation with advanced structural and camera controls.

Agents365-ai/video-podcast-maker

An automated pipeline that transforms a topic into a professional 4K video podcast by integrating AI research, TTS audio, and Remotion-based video rendering.

mira-wm/mira

MIRA is a real-time world model of Rocket League that uses a 5B parameter latent diffusion model to simulate 2v2 matches based on player actions.

AlayaLab/AlayaWorld

An interactive autoregressive world model that enables long-horizon, playable video generation with real-time camera control and consistent spatial memory.

Tencent-Hunyuan/HunyuanVideo-1.5

A lightweight 8.3B parameter video generation model that enables high-quality text-to-video and image-to-video synthesis on consumer-grade GPUs.

Soul-AILab/SoulX-FlashTalk

SoulX-FlashTalk is a system for real-time, infinite streaming of audio-driven avatars using self-correcting bidirectional distillation to create continuous talking head animations.

thu-ml/Causal-Forcing

A framework for high-quality real-time interactive video generation using Causal ODE or Causal Consistency Distillation to enable efficient few-step autoregressive diffusion.

ItusiAI/Open-Magiviz

An AI-powered video creation platform that automates the workflow from script generation and character design to storyboard creation and final video rendering.

bytedance/Bernini

Bernini is a unified video generation and editing framework that combines an MLLM-based semantic planner with a DiT-based renderer to improve instruction following and complex video edits.

jnMetaCode/ai-shortfilm-prompts

A cinematic video prompting system and Claude Code plugin that uses a 5-stage cinematography-based framework to create professional-grade AI short films.