AI Agent 驱动的开源可自部署视频工作台:将小说与剧本转为角色、场景、道具资产、分镜、视频和剪映草稿,支持跨镜头一致性、多供应商与费用追踪 | Self-hosted AI video workspace for stories, storyboards and short-form video production
-
Updated
Aug 9, 2026 - Python
AI Agent 驱动的开源可自部署视频工作台:将小说与剧本转为角色、场景、道具资产、分镜、视频和剪映草稿,支持跨镜头一致性、多供应商与费用追踪 | Self-hosted AI video workspace for stories, storyboards and short-form video production
[ECCV 2026] Official PyTorch implementation of RefAlign: Representation Alignment for Reference-to-Video Generation
ComfyUI custom nodes for Veo 3.1 video generation — text-to-video, image-to-video, reference-to-video, extend, and 4K upscale via MuAPI
Original MiniMax H3 prompt library, MuAPI Python client, API examples, and current model notes.
MiniMax H3 (Hailuo 3.0) front-end for ComfyUI — text-to-video, image-to-video, first-and-last-frame and reference-to-video with native audio. A local LLM writes H3's structured prompts and self-corrects them. Character consistency from a subject photo, live preview, VRAM handoff.
AI-K SK MiniMax H3 local prompt engineering toolkit for ComfyUI|支持 T2VA、I2VA、L2VA、FL2VA、Ref2VA,多参考映射、Audio 编号、说话者绑定、冲突校验与盲评交叉验证。
ComfyUI custom nodes for MiniMax H3 text-to-video, image-to-video, and multimodal reference-to-video generation through Muapi.
C# SDK for Vidu (Shengshu Tech) - AI-powered video generation with text-to-video, reference-to-video (multi-subject consistency), image-to-video, and more
Vidu is a generative video AI platform from Shengshu Technology (ShengShu / 生数科技), built on the company's U-ViT diffusion-transformer architecture. The Vidu API turns text prompts, still images, and reference subjects into short video clips with features like text-to-video, image-to-video, reference-to-video (multi-entity / character consistency)…
Add a description, image, and links to the reference-to-video topic page so that developers can more easily learn about it.
To associate your repository with the reference-to-video topic, visit your repo's landing page and select "manage topics."