hypit-ai/hypit
Hypit is an open‑source framework that lets LLM agents clone or generate videos via a declarative SVML workflow. Drop a reference video (or a description), the AI builds a complete production pipeline—footage, captions, B‑roll, effects—and renders it using headless Chromium. The system supports cheap mass‑variant generation, pluggable components, and optional zero‑cost rendering (no external AI calls). Install with `npx skills add hypit-ai/hypit -g`; then use the `/hypit` skill in any coding‑agent session to create or remix videos. Licensed under a permissive Apache‑2.0‑derived license.
mcncarl/jianying-headless
A local automation tool for Jianying Pro on macOS that enables programmatic generation of editable video drafts and native MP4 exports for AI video workflows.
calesthio/OpenMontage
An open-source, agentic video production system that automates research, scripting, asset generation, and editing to create professional videos from plain-text prompts.
heygen-com/hyperframes
HyperFrames is an open‑source Node.js framework that lets you describe a video as ordinary HTML + CSS (with special `data-*` timing attributes) and render it to a deterministic MP4. It ships a set of “skills” that AI coding agents can load, so an LLM can be prompted to create product‑launch videos, PR walkthroughs, explainer clips, motion graphics, etc., automatically generating the HTML, pulling assets, lint‑checking, previewing, and finally rendering via a CLI (`hyperframes render`). The system supports many animation runtimes (GSAP, Lottie, Three.js, CSS, etc.), a unified media manager, optional cloud/Lambda rendering, and a catalog of reusable design blocks. Ideal for AI‑driven video generation pipelines and designers who want a web‑first, code‑centric video authoring workflow.
harry0703/MoneyPrinterTurbo
An automated AI short-video generation tool that transforms themes or keywords into full videos with scripts, voiceovers, subtitles, and matched footage.
jub0t/Concat
An open-source, cross-platform video editor that provides a local-first alternative to CapCut with offline AI captions, speech, and cutouts.
ArcReel/ArcReel
An open-source, self-hosted AI video production workbench that transforms novels, scripts, and product materials into consistent, controllable short videos.
Vincentwei1021/video-shotcraft
An AI agent skill that uses Remotion to turn product screenshots and copy into cinematic promo videos using a library of 157 shot recipe cards.
Vincentwei1021/anything2explainer
An AI-driven pipeline that converts topics or documents into narrated motion-graphics explainer videos by using AI agents to write Remotion code.
dramaclaw/dramaclaw
DramaClaw is a self‑hosted, Docker‑based platform that turns scripts into AI‑generated short films. It offers a node‑based infinite canvas (XiaHua) for image, video, audio, and 3D set creation, plus a structured pipeline (Series) that handles ingest, episode planning, storyboard generation, voice‑over, video composition, and export. An integrated AI assistant (Xia Director) can run tasks and suggest next steps. All heavy model inference is routed through an OpenAI‑compatible gateway, letting you use any provider you like. Released under Elastic 2.0, it can be used commercially with a simple “Powered by DramaClaw” attribution.
zhouxiaoka/autoclip
An AI-powered video clipping tool that automatically downloads videos from YouTube and Bilibili, analyzes content via LLMs to extract highlights, and generates short-form collections.
eternityspring/reelbench-skills
A collection of skills for Claude Code and Codex that automates video shot analysis and creates synchronized video-information overlays.
browser-use/video-use
An open-source AI video editing tool that allows users to edit raw footage via chat by converting video into structured text and on-demand visuals for an LLM to process.
cosmo-wander-ai/cosmo-edge
CosmoEdge 1.1 is an Apache‑2.0 C++ edge‑AI engine for video analytics and vision‑language models. It runs on Sophon (BM1688/CV186X), Rockchip (RK3576/RV1126B) NPUs and on x86/macOS via Docker, offering a browser‑based pipeline builder, REST/MQTT/WebSocket integration, and a “Model Guard” for protected commercial models. The repo includes Docker scripts for building device‑specific runtimes, benchmark reports, and extensive docs.
diffusionstudio/editor
A professional video editor designed for AI agents that treats video compositions as code, allowing agents to and humans to collaboratively edit videos via a bidirectional JSX-based workflow.
AIMixer/ComfyUI_MiniMaxH3_Director
A ComfyUI plugin that provides a UI‑driven “director” node for multi‑segment video generation with MiniMax‑H3 models, supporting text, image, reference, and video‑to‑video tasks, audio generation, motion continuity, second‑pass refinement, and import/export of project packs.
luoluoluo22/jianying-editor-skill
An AI-powered automation skill for Jianying (CapCut China) that allows users to create full video drafts—including scripts, voiceovers, and effects—using natural language.
samuelgursky/davinci-resolve-mcp
DaVinci Resolve MCP is an open‑source server that exposes the full DaVinci Resolve Studio scripting API as a JSON‑over‑HTTP service. AI agents can call high‑level tools (project creation, media import, timeline editing, grading, Fusion, Fairlight, render setup, etc.) and receive a uniform result envelope for safe, auditable automation. It works with the paid Studio edition via the normal external‑scripting interface, and includes a Lua‑based bridge to reach the free edition. An optional Node‑based “advanced” server can edit Resolve project files offline. The package ships a local browser control panel and integrates with many IDEs, making it a practical foundation for on‑premise AI post‑production assistants.
Vincentwei1021/video-talkcraft
An AI agent skill that transforms scripts and voiceovers into professional explainer videos with automated character-level audio syncing and a library of 108 motion graphics templates.
ATH-MaaS/Pixelle-Video
An AI-powered automated short-video engine that generates scripts, visuals, and voiceovers from a single topic to create complete videos without manual editing.
kaomei/stickman-video-director
A Codex Skill that transforms text into a structured director's proposal and optimized prompts for creating consistent, one-minute stickman animations using Gemini Omni Flash.
SkyNotSilent/awesome-MiniMax-H3-cases
A community‑run catalog of MiniMax H3 video cases, full prompts, and hardware‑specific tutorials, with searchable web UI, JSON data, and AI‑agent Skills; MIT‑licensed code, source‑attributed content.
AlayaLab/Evoke
A high-speed, three-step world model that uses an external world state bank to generate consistent, endless video rollouts with mid-flight re-prompting capabilities.
shengshu-ai/Vidu-S
A suite of real-time interactive video generation models for high-resolution digital avatars, live video editing, and immersive spatial video.
thu-ml/TurboDiffusion
TurboDiffusion is a video generation acceleration framework that speeds up end-to-end diffusion generation by 100-200x using attention acceleration and timestep distillation.
cartesiancs/cartcut
An AI-powered video editor featuring layer-based editing and motion graphics tools to help creators produce professional effects without heavy software.
0xsline/OpenChatCut
An open-source, local-first AI video editor that combines a professional multitrack timeline with conversational agents for editable, agent-driven video production.
letorig/video-generator-client
An asynchronous Python client and web UI that provides a unified interface for generating videos across multiple providers including Seedance, Kling, MiniMax, and Wan.
shuyu-labs/BigBanana-AI-Director
BigBanana AI Director is a Docker‑deployed web platform that turns a script (or even a single sentence) into a short film / motion comic. It uses AntSK’s text, image, and video models, drives each shot with start‑ and end‑keyframes, interpolates motion with the Veo model, and provides a project‑wide asset library, shot grid UI, and a simple timeline editor for final export.
hao-ai-lab/FastVideo
A unified post-training and real-time inference framework designed to accelerate video generation through sparse distillation and attention optimizations.
Colafornia/short-video-generator-AI
An open‑source Python tool that automatically creates vertical short clips from YouTube or local videos. It transcribes with local Whisper, uses an LLM (OpenAI, Gemini, or MuAPI) to rank viral moments, optionally adds an AI‑generated hook, and renders the clips via a simple CLI or a local web UI.
Junchao-cs/SolarWM
SolarWM is an open‑source framework that provides a massive, camera‑annotated video dataset, a three‑stage training recipe, and ready‑to‑use 5 B–33 B video world‑model checkpoints (Wan2.2, LTX‑2.5, MiniMax‑H3). It lets you train and run long‑horizon interactive video models with a unified CLI and fully documented pipelines.
YILS-LIN/short-video-factory
An AI-powered desktop tool that automates the creation of product marketing and general content short videos from text prompts and video clips.
Alisa0808/vox-director
An agent skill that automates the production of Vox-style paper-collage explainer videos, handling everything from script and keyframes to motion and voice-over.
gamedev-skills/awesome-gamedev-agent-skills
A library of 73 AI‑agent “skills” (engine‑specific code guides, discipline, genre and workflow templates) plus a router that auto‑detects your game engine and loads only the needed pieces, enabling agents like Claude, Cursor, Gemini‑CLI, etc., to generate game code on demand.
liangdabiao/Seedance2-Storyboard-Generator
An AI video production workflow that uses Claude Code and Seedance 2.0 to transform stories into consistent, multi-episode video series.
GVCLab/PersonaLive
A real-time, streamable diffusion framework for generating infinite-length expressive portrait animations from a single reference image and a driving video.
nateherkai/hyperframes-student-kit
An open‑source, AI‑assisted video‑editing kit that lets you edit footage via natural‑language prompts to Codex or Claude Code. It ships with 14 agent‑driven “skills”, 400+ motion‑graphics cards, scene templates, transcription helpers (ElevenLabs, Whisper), and demo projects. Install Node 22+, Git, FFmpeg, then run `npm run demo` and use commands like `$edit-video` or `/short-form-edit` to cut silences, fix mistakes, and produce reels or Shorts. Optional API keys unlock cloud transcription and asset generation; otherwise it works locally with synthetic data.
fqscfqj/Y2A-Auto
Y2A‑Auto is an open‑source, Docker‑ready pipeline that downloads YouTube videos, optionally runs Whisper/Voxtral ASR, translates and quality‑checks subtitles, generates titles/tags via OpenAI, encodes the video (CPU or GPU), and uploads it to AcFun and/or bilibili. It includes a Flask web UI, YouTube monitoring, CookieCloud sync, content‑moderation, and multi‑channel notification support, all configurable through a JSON file and protected by password/lockout features.
Lightricks/LTX-2
LTX-2 is a DiT-based audio-video foundation model that generates high-fidelity, synchronized audio and video from text or image prompts.
Emily2040/seedance-2.0
An agent skill for AI video generation that transforms rough ideas into directed prompts, managing shot planning, reference binding, and clip continuity.
Anil-matcha/AI-Youtube-Shorts-Generator
An open-source AI tool that converts long-form YouTube videos into viral vertical shorts by using LLMs for highlight detection and Whisper for transcription.
Robbyant/lingbot-world-v2
LingBot‑World‑V2 is a research‑grade video‑generation model that creates infinite, interactive worlds from text prompts. It offers fast (real‑time) and high‑quality variants (14 B and 1.3 B), supports diverse actions, and uses a pilot‑director agent architecture. The repo provides inference code, model download links, and a quick‑start guide; deployment code is not released.
palmier-io/palmier-pro
A macOS video editor with built-in generative AI and MCP support, enabling AI agents to create and edit content directly on the timeline.
gnipbao/story-to-handdrawn-video
A tool that converts story text or images into vertical hand-drawn animations with 20 built-in styles, automatable via AI agent skills.
OpenVDN/vdn-minimax-h3
VDN‑Minimax‑H3 (VDN‑H3) is a hybrid‑attention video diffusion model built on MiniMax H3. It adds a lightweight linear‑attention branch and LoRA adapters, enabling near‑real‑time 768p video generation (e.g., 14‑s clip in ~11 s on a single B200 GPU). The repository provides full training scripts (four staged recipe) and optimized inference code using FP8 and FlashAttention‑4, with detailed setup instructions and performance tables.
MemeCalculate/moyin-creator
A production-grade AI filmmaking tool that automates the entire pipeline from script parsing and character consistency to batch video generation.
HKUDS/ViMax
ViMax is an open‑source, agent‑driven framework that turns a text idea, screenplay, or novel excerpt into a complete AI‑generated video. It handles scriptwriting, storyboarding, character consistency, and final rendering via configurable LLM, image, and video APIs. Use a terminal TUI or a browser‑based Web UI; install with `uv sync`, configure API keys in `configs/agent.local.yaml`, and run `vimax tui` or `npm run dev` for the UI.