xixihhhh/clipforge
ClipForge is an open‑source AI tool that turns a single product image (or a short text prompt) into a ready‑to‑publish short video for Douyin, Kuaishou, Xiaohongshu, TikTok, etc. It automates script writing, storyboard creation, image‑to‑video synthesis, TTS, subtitles, BGM, and final FFmpeg stitching. Two tracks exist – a completely free path using public assets and Edge TTS, and a paid AI‑generation path that charges per second of generated content. Features include director vs. beginner modes, cross‑shot consistency, selective re‑generation, a 391‑template library, hot‑topic auto‑generation, batch production, built‑in compliance (AIGC labeling, ad‑law word scanning), CLI/MCP/agent integration, and an AGPL‑3.0 license. Ideal for e‑commerce sellers, marketers, creators, and developers who need fast, low‑cost product videos.
xhongc/ai_story
AI Story is a Docker‑based web app that automates the whole pipeline from a story prompt to a finished short video, using multiple AI models for script writing, storyboard creation, image generation, camera‑move planning, and video rendering.
nkxx188/ComfyUI-MiniMaxH3-Easy
A practical node suite for ComfyUI that simplifies MiniMax H3 video generation, offering unified media management and tools for creating long, coherent videos via context segments.
taco-group/SparkVSR
SparkVSR is an interactive video super-resolution framework that allows users to guide the restoration process using sparse high-quality keyframes to ensure better temporal consistency and controllable results.
video-production-buddy/video-production-buddy
An agent-first AI video production pipeline that replaces one-shot prompting with a governed workflow of planning, approval gates, and verified rendering.
yizhi-chengzi/video-ai-talking
A local web app that turns a silent face‑cam clip into a full talking‑head video using Alibaba Bailian VideoRetalk for lip‑sync, TTS from Bailian or Volcengine, optional DeepSeek script generation, and FFmpeg for final rendering.
LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler
A custom ComfyUI node set for Minimax H3 that provides neural latent upscaling and tiled re-sampling to accelerate high-resolution video generation while avoiding VAE round-trip overhead.
Carasibana/ComfyUI-H3-FaceRefine
ComfyUI‑H3‑FaceRefine is a custom‑node pack for ComfyUI that fixes MiniMax H3’s poor rendering of small faces in video. It detects faces, crops them to a larger canvas, runs H3 on the crops, and stitches the refined faces back. Includes auto‑ and manual‑face‑selection modes, hard‑cut detection, per‑frame denoise, and ready‑to‑use example workflows.
pireel/pireel
Pireel Studio is an open-source AI video editor that lets users create videos using a combination of manual timeline editing, chat-based descriptions, and AI agents.
NikoDemon80/ComfyUI-H3-Motion-Context
A ComfyUI extension for MiniMax H3 that enables seamless chaining of video and audio clips by using latent-based continuity to prevent quality loss and motion and sound jumps.
veedstudio/open-edit
An agent-driven video editing pipeline that allows coding agents to transcribe, design, and render MP4 videos based on natural language descriptions.
KlingAIResearch/LivePortrait
An efficient portrait animation tool that animates static images of humans and animals using driving videos or images with precise stitching and retargeting control.
Stonewuu/ai-fusion-video
An Agent-driven video creation platform that integrates scriptwriting, storyboarding, and AI image/video generation into a unified workspace.
MartinDelophy/ai-video-editor
Timeline Studio is a browser‑only, local‑first AI video editor that runs on WebGPU/ONNX. It offers multilingual TTS, AI‑generated music, automatic captions, smart framing, watermark removal, vocal separation, and digital‑human avatars, all without uploading media. Projects are saved as portable `.timeline` files and can be edited via a rich UI or a head‑less CLI skill that lets LLM agents inspect, plan, and apply edits programmatically. The code is MIT‑licensed; model weights are mirrored on Hugging Face and ModelScope under their original licenses.
lcy362/agnes-video-generator
A free, open-source AI video generator that uses cloud APIs to create multi-scene videos with narration and subtitles, eliminating the need for expensive GPUs or subscriptions.
Lightricks/LTX-Desktop
LTX Desktop is an open‑source Electron app for generating and editing videos (and images) with LTX diffusion models. It runs locally on NVIDIA GPUs (≥16 GB VRAM) or Apple‑silicon Macs (≥15 GB free RAM) using the LTX 2.5 Fast or 2.3 Fast models, and automatically switches to a paid cloud API (LTX 2.5/2.3 Pro) on unsupported hardware. Features include text‑/image‑/audio‑to‑video, video retake/extend, LoRA style adapters, prompt enhancement via Gemini, and a timeline‑based video editor. Installation is via a GitHub release installer; the app stores data in standard OS‑specific folders and requires a free LTX API key for cloud text encoding. The codebase (Apache‑2.0) is split into React UI, Electron shell, and a Python FastAPI backend, with full dev scripts for local building. Optional anonymous telemetry can be disabled.
zai-org/SCAIL-2
SCAIL-2 is an open-source model for end-to-end character animation and replacement, removing the need for intermediate pose representations to support complex motions and diverse identities.
FireRedTeam/FireRed-OpenStoryline
A conversational AI video creation tool that automates scriptwriting, media sourcing, and editing through natural language prompts.
crisng95/flowkit
Flow Kit is a Python‑FastAPI + Chrome‑extension system that automates end‑to‑end AI video creation via Google Flow, handling reference images for visual consistency, scene generation, TTS, up‑scaling, thumbnailing and YouTube upload.
MeiGen-AI/OPSD-V
An on-policy self-distillation framework for post-training few-step autoregressive video generators to reduce error accumulation and improve motion dynamics in long videos.
Robbyant/lingbot-world
An open-source world simulator that generates high-fidelity, interactive video environments with real-time interactivity and long-term temporal consistency.
liyue-aigc/seedance-2-5-video-director
A video direction skill for Dreamina/Jimeng Seedance 2.5 that transforms creative ideas and reference images into structured video scripts and optimized prompts.
nexu-io/html-video
A meta-layer for turning text, articles, or GitHub repos into animated MP4 videos using local coding agents and pluggable HTML-based rendering engines.
Tencent-Hunyuan/HunyuanVideo
HunyuanVideo is a large-scale open-source video foundation model with over 13 billion parameters that generates high-quality videos with superior motion and text alignment.
jnMetaCode/ai-shortfilm-prompts
A curated library of AI video prompts and a Claude Code skill that turn any idea into a ready‑to‑run, multi‑shot prompt for 2026 text‑to‑video models. Includes genre templates, camera‑move libraries, negative‑prompt prefabs, a web‑based no‑install builder, and detailed docs on the 5‑stage prompt structure.
myccarl/ai-shortVideo-pipeline
myAiVideos is an open‑source, Docker‑compose‑based system that automates the whole short‑video creation workflow (topic → script → visuals → audio → post‑production → distribution) for Chinese content. It uses a FastAPI orchestrator plus a Java Spring‑Boot gateway for auth, routing, circuit‑breaking and metering, and integrates multiple LLM and multimodal models (DeepSeek, GLM, Kling, TTS services). The pipeline is layered, fail‑over capable, and fully observable via Langfuse and Prometheus.
ModelTC/LightX2V
LightX2V is an open‑source inference framework for fast image and video generation (T2V, I2V, T2I, I2I, T2AV, etc.). It supports many state‑of‑the‑art models (MiniMax‑H3, Wan, HunyuanVideo, Qwen‑Image, etc.) and provides speed‑up techniques such as 4‑step distilled LoRAs, FP8/NVFP4 quantization, tensor/sequence parallelism, and disaggregated deployment on a wide range of hardware. Benchmarks show up to 3.9× faster inference than competing frameworks on H100 GPUs. The project offers Docker images, pip install, Gradio/ComfyUI front‑ends, extensive docs, and an online demo.
ChatCut-Inc/agent-plugin
A collection of plug‑ins that let AI coding assistants (Codex, Claude Code, Grok Bot, Cursor) edit ChatCut video projects via natural‑language prompts—import media, add graphics, generate audio, transcribe, export, etc. Requires a ChatCut account and one of the supported hosts; authentication is handled through ChatCut’s MCP endpoint.
LingGuoAI/LingGuo-Drama
An open-source workbench for AI short drama and video production that manages the pipeline from scriptwriting and asset generation to final video synthesis.
Rimagination/h3lite
A local deployment skill for MiniMax H3 that enables low-VRAM NVIDIA GPUs on Windows to generate videos with native audio via AI agents.
jianjieyiban/JJYB_AI_VideoAutoCut
A local-first AI video creation workstation that automates the workflow from material analysis and scriptwriting to voiceover generation and audio-visual synchronization.
ChenShuo2004/cs-board
cs-board is a locally‑run AI video‑creation workstation that turns Chinese text, a reference audio clip, and optional style/character images into a fully‑rendered whiteboard‑style MP4. It handles voice cloning (via IndexTTS), script segmentation, AI‑generated illustrations, hand‑drawn animation, subtitles, and audio‑video sync, offering 12 visual templates, custom character/style support, dynamic infographics, and LAN‑based collaborative queues.
zenstory-ai/video-recap-skills
An AI-powered video recap system that transforms raw videos into narrated summaries by automating video understanding, scriptwriting, voiceover generation, and editing.
FujiwaraChoki/supoclip
An open-source, AI-powered video clipping tool that converts long-form videos into viral vertical clips with automatic face-cropping, subtitles, and virality scoring.
HITsz-TMG/VideoClaw
An AI director system that automates the entire creative video production pipeline from scriptwriting and character design to storyboarding and final editing.
SandAI-org/MAGI-2-preview
A 114B-parameter unified audio-video generation model that produces 1080p video clips with synchronized sound from text or image prompts.
TNTwise/REAL-Video-Enhancer
A cross-platform GUI application for AI-powered video frame interpolation and upscaling, supporting multiple GPU backends and a variety of specialized models.
HKUDS/VideoAgent
An all-in-one agentic framework for understanding, editing, and generating creative videos through natural language prompts.
OpenImagingLab/FlashVSR
FlashVSR is a one-step streaming diffusion framework for real-time video super-resolution that achieves high-speed upscaling and scalability to ultra-high resolutions.
pyang5166/gbro-collage-broll
A tool that turns short voiceover lines into editorial halftone paper-collage B-roll animations using a three-stage approval workflow and Gemini Omni Flash.
pexoai/pexo-skills
A set of open-source agent skills that automate the full video production pipeline, transforming text, images, or URLs into finished multi-shot videos with music and subtitles.
xiejunjie524/handdraw-story-video
A tool for creating short vertical story videos by animating the transition from hand-drawn line art to colored images.
huangserva/ComfyUI_MiniMaxH3_Director
A collection of ComfyUI workflows for MiniMax H3 Director, enabling text-to-video, image-to-video, and advanced video-to-video editing with reference materials.
Vchitect/VBench
VBench is an open‑source benchmark suite that automatically evaluates video‑generation models across dozens of quality dimensions (technical, aesthetic, and trustworthiness). It includes prompt collections, metric pipelines, human‑aligned scores, and public leaderboards for T2V, I2V, long‑video, and intrinsic‑faithfulness evaluation.
Agents365-ai/video-podcast-maker
An automated pipeline that transforms a topic into a professional 4K video podcast by integrating AI research, TTS audio, and Remotion-based video rendering.
twinnydotdev/twinny
Twinny is a free VS Code extension that adds AI‑driven code completion, an interactive chat sidebar, workspace‑wide context via embeddings, and optional P2P inference sharing. It works with many LLM providers (OpenAI, Anthropic, Mistral, etc.) and can run locally, offering customizable prompts and features like diff view, commit‑message generation, and offline operation.
Anil-matcha/Open-AI-UGC
An open-source AI video ad studio for creating realistic UGC vertical ads using multiple SOTA video models, designed as a self-hostable alternative to Arcads and MakeUGC.
s1dashu/director
director is an AI‑agent skill that guides a language model through the full video‑production pipeline—script, visual style, characters, voice, shot‑by‑shot prompts—and uses LibTV or 即梦 CLI tools to generate ready‑to‑edit video clips. It supports three modes (Animated Explainer, Storytime Animation, Cinematic Drama) and is installed as a skill directory for agents like Codex.