calesthio / OpenMontage
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files.
View calesthio/OpenMontageVideo AI on GitHub splits into two jobs: generating footage, and cutting, subtitling, encoding or streaming it. This board tracks both. It is built from topic pages for text-to-video, video generation and video editing, ranked by stars, so a diffusion wrapper and a battle-tested encoder can appear side by side. Each card carries language, stars and forks and links to a page with the full description, license, activity dates, README and related video projects. Whether you need something to render clips on a GPU box or a library to decode a stubborn container, this is the fastest survey of what people are actually starring in the open-source video stack.
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files.
View calesthio/OpenMontageUnrestricted Open-source alternative to AI video platforms — Free AI image & video generation studio with 600+ models (Flux, Midjourney, Kling, Sora, Veo). No content filters.
View Anil-matcha/Open-Generative-AI🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
View ATH-MaaS/Pixelle-VideoToonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animated short dramas.
View HBAI-Ltd/Toonflow-app🚀 Truly open-source AI avatar(digital human) toolkit for offline video generation and digital human cloning.
View duixcom/Duix-Avatar首家工业级全流程 AI 影视生产平台。Industry-first professional AI Agent platform for controllable film & video production. From shorts to live-action with Hollywood-standard workflows.
View waooAI/waoowaootext and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
View zai-org/CogVideoHunyuanVideo: A Systematic Framework For Large Video Generation Model
View Tencent-Hunyuan/HunyuanVideo"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
View HKUDS/ViMaxSANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
View NVlabs/SanaAI video skill for Claude Code & Codex — cinematic product videos with Remotion: 152 shot recipe cards, 209 motion previews, a production-ready template
View Vincentwei1021/video-shotcraftImplementation of Imagen, Google's Text-to-Image Neural Network, in Pytorch
View lucidrains/imagen-pytorchBackground Remover lets you Remove Background from images and video using AI with a simple command line interface that is free and open source.
View nadermx/backgroundremover🚀🎬 ShortGPT - Experimental AI framework for youtube shorts / tiktok channel automation
View RayVentura/ShortGPTAutoClip : AI-powered video clipping and highlight generation · 一款智能高光提取与剪辑的二创工具
View zhouxiaoka/autoclip[SIGGRAPH Asia 2022] VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild
View OpenTalker/video-retalkingA framework for efficient model inference with omni-modality models
View vllm-project/vllm-omniOpenShot Video Editor is an award-winning free and open-source video editor for Linux, Mac, and Windows, and is dedicated to delivering high quality video editing and animation solutions to the world.
View OpenShot/openshot-qtThis repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
View promptslab/Awesome-Prompt-EngineeringFunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
View modelscope/FunClipA curated list of recent diffusion models for video generation, editing, and various other applications.
View showlab/Awesome-Video-DiffusionA general-purpose AIGC video engine: script to finished film in one pipeline — dramas, ads, product videos, otome games, and more. | 通用 AIGC 视频引擎 —— 从剧本到成片一条流水线,漫剧、广告、电商、乙游皆可
View dramaclaw/dramaclawVideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
View AILab-CVC/VideoCrafterOpen-source alternative to Opus Clip, Vidyo.ai, Klap & SubMagic. Turn long-form YouTube videos into viral 9:16 shorts using LLM highlight detection, Whisper transcription
View Anil-matcha/AI-Youtube-Shorts-Generator[CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation
View antgroup/echomimic_v2Programmatic video for coding agents — HTML to video on your laptop. Turn HTML, CSS & data into real MP4s with pluggable render engines, 21 templates, AI soundtrack. Apache-2.0, no per-render fees.
View nexu-io/html-videoHunyuanVideo-1.5: A leading lightweight video generation model
View Tencent-Hunyuan/HunyuanVideo-1.5Open source AI clip generator: turns long videos into viral 9:16 shorts with AI moment detection, face tracking, subtitles and dubbing.
View mutonby/openshortsAI Agent 驱动的开源可自部署视频工作台:将小说与剧本转为角色、场景、道具资产、分镜、视频和剪映草稿,支持跨镜头一致性、多供应商与费用追踪 | Self-hosted AI video workspace for stories, storyboards and short-form video production
View ArcReel/ArcReelA unified inference and post-training framework for accelerated video generation.
View hao-ai-lab/FastVideoThe Ruby-native AI framework. Chats, agents, tools, images, audio, and video through one consistent API, in plain Ruby or Rails.
View crmne/ruby_llm轻量、灵活、易上手的Python剪映草稿生成及导出工具,构建全自动化视频剪辑/混剪流水线。本项目的CapCut版本正于 https://github.com/GuanYixuan/pyCapCut 内开发
View GuanYixuan/pyJianYingDraftClone any viral video with AI agents. Not just a script, the whole workflow: swap the face, the words, the B-roll, ship 100 variants in one command, and get your 100M views.
View hypit-ai/hypitMulti-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.
View SamurAIGPT/Generative-Media-Skills[ECCV 2024] Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
View fudan-generative-vision/champ[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
View Picsart-AI-Research/Text2Video-ZeroOn-device subtitle generation that connects directly to DaVinci Resolve, Premiere, and After Effects.
View tmoroney/auto-subs[ICCV 2025] Official implementations for paper: VACE: All-in-One Video Creation and Editing
View ali-vilab/VACE[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, image, and video models.
View thu-ml/SageAttention[CVPR 2026] PersonaLive! : Expressive Portrait Image Animation for Live Streaming
View GVCLab/PersonaLiveTurboDiffusion: 100–200× Acceleration for Video Diffusion Models
View thu-ml/TurboDiffusion🎬 Fully automated YouTube channel management with AI agents. Creates, optimizes & publishes videos 24/7. Works with FREE Gemini API or OpenAI. No coding required!
View darkzOGx/youtube-automation-agentFireRed-OpenStoryline is an AI video editing agent that transforms manual editing into intention-driven directing through natural language interaction, LLM-powered planning
View FireRedTeam/FireRed-OpenStorylineDiffusion model papers, survey, and taxonomy
View YangLing0818/Diffusion-Models-Papers-Survey-Taxonomy50+ open-source generative AI apps you can clone, deploy, and monetize — image generators, video tools, virtual try-ons, AI SaaS templates, and platform integrations.
View Anil-matcha/awesome-generative-ai-apps[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
View jy0205/Pyramid-FlowInternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, SAM, interactive image editing
View OpenGVLab/InternGPT[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
View Doubiiu/DynamiCrafterRecord your screen, ship a demo. Free and open-source, GPU-accelerated, no watermarks, no subscriptions. Windows, macOS, Linux. Actively maintained.
View getopenscreen/openscreenMuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising
View TMElyralab/MuseVResearchStudio: Our AI co-author, from research problem to final publication.
View microsoft/ResearchStudioLightweight Image Video Action Generation Inference Framework
View ModelTC/LightX2VOpen-source, ad-free Android multimedia recorder with background video recording, screen recording, live streaming, and remote camera control
View anonfaded/FadCamFLUX, Stable Diffusion, SDXL, SD3, LoRA, Fine Tuning, DreamBooth, Training, Automatic1111, Forge WebUI, SwarmUI, DeepFake, TTS, Animation, Text To Video, Tutorials, Guides, Lectures, Courses
View FurkanGozukara/Stable-DiffusionHigh-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
View Tencent/MimicMotionMatrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
View SkyworkAI/Matrix-Game[CSUR] A Survey on Video Diffusion Models
View ChenHsing/Awesome-Video-Diffusion-ModelsFreeCut is a professional-grade video editor that runs entirely in your browser. Professional video editing, zero installation.
View walterlow/freecutA web-based Video Editing SDK built on WebCodecs. 基于 WebCodecs 构建的网页视频编辑 SDK。
View WebAV-Tech/WebAVAI-native video production toolkit for Claude Code
View digitalsamba/claude-code-video-toolkit🎬 2000+ curated Seedance 2.0 video generation prompts — cinematic, anime, UGC, ads, meme styles. Includes Seedance API guides, character consistency tips, and advanced video workflows.
View YouMind-OpenLab/awesome-seedance-2-promptsImplementation of Make-A-Video, new SOTA text to video generator from Meta AI, in Pytorch
View lucidrains/make-a-video-pytorchOpen Source API and interchange format for editorial timeline information.
View AcademySoftwareFoundation/OpenTimelineIOTurn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas Cloud + ffmpeg. An agent skill.
View Alisa0808/vox-directoran editor for spoken-word audio with automatic transcription
View bugbakery/audapolis[EMNLP2026] "VideoAgent: All-in-One Agentic Framework for Video Understanding and Editing, and Remaking"
View HKUDS/VideoAgentAI-powered animated comic generator — transform scripts into fully animated videos with AI-driven character design, storyboarding, and video synthesis.
View LingyiChen-AI/AIComicBuilder[ICCV'25 Best Paper Finalist] ReCamMaster: Camera-Controlled Generative Rendering from A Single Video
View KlingAIResearch/ReCamMasterOpen-source, local-first conversational AI video editor with a professional multi-track timeline, Agent Skills, MCP integration, and Remotion rendering.
View 0xsline/OpenChatCutLocal-first Video Knowledge Base. Index your video library with multi-modal analysis (YOLO, DeepFace, Whisper), search semantically via natural language, Docker-ready.
View IliasHad/edit-mind🚀 AI 全自动化视频生成员工 | Your First AIGC Coworker. Chat an Idea. Get a Film. 🦞
View HITsz-TMG/VideoClawOfficial implementations for paper: DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic Models
View ali-vilab/dreamtalk🎬 seedance2接入 开源本地 AI 短剧 & 漫剧生成工具 —— 从故事到成片一站式完成,数据不出本机,短剧工作流管理平台,高灵活度,AI真人剧,AI漫剧本地搞定。 Open-source local AI short drama maker: story → storyboard → video, fully offline, your data stays yours. 纳米流水线
View xuanyustudio/LocalMiniDramaOfficial Pytorch Implementation for "TokenFlow: Consistent Diffusion Features for Consistent Video Editing" presenting "TokenFlow" (ICLR 2024)
View omerbt/TokenFlowThe all-in-one local AI studio for your desktop: chat, image and video generation and a coding agent in one free, open source app. Windows and Linux. No Docker, no terminal, no cloud required.
View PurpleDoubleD/locally-uncensoredTopic → 4K narrated video for coding agents. v5.3.0: local TTS (edge free + azure, no external engine), manifest-based Asset Engine, Remotion composition, cost-gated AI generation
View Agents365-ai/video-podcast-makerOfficial MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.
View MiniMax-AI/MiniMax-MCPOfficial implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
View menyifang/MIMOOpenShot Video Library (libopenshot) is a free, open-source project dedicated to delivering high quality video editing, animation, and playback solutions to the world.
View OpenShot/libopenshotAI video agents framework for next-gen video interactions and workflows.
View video-db/DirectorPhantom: Subject-Consistent Video Generation via Cross-Modal Alignment
View Phantom-video/PhantomText To Video Synthesis Colab
View camenduru/text-to-video-synthesis-colab[ICCV 2023] Tracking Anything with Decoupled Video Segmentation
View hkchengrex/Tracking-Anything-with-DEVA自然语言驱动的无限画布工作流 Agent,让 AI 视频创作第一次真正变成可编辑的工作流。 AICON 面向创作者,提供从剧本拆解、分镜生成、素材生成、视频合成到内容分发的一整套能力。 不是只给你一个输入框,而是让你用自然语言和无限画布一起驱动创作,把文本、图片、视频节点组织成完整链路,真正把“从灵感到 成片”放进一个系统里完成。
View 869413421/ai-moive-studio【融光】 - 基于 Agent 的全流程AI短剧/漫剧/视频创作平台 - Java & agentscope2.0 | Agent-based end-to-end AI creation platform for short dramas, motion comics, and videos – built on Java & & agentscope 2.0.
View Stonewuu/ai-fusion-video😎 The open-source, Haskell-built video editor for GIF makers.
View lettier/gifcurry[ICCV 2023] StableVideo: Text-driven Consistency-aware Diffusion Video Editing
View wenhaochai/StableVideo