视频生成

待看 Learn how to automate faceless short-form + long-form video content and dominate YouTube, TikTok, Facebook & Instagram on autopilot — from idea → script → video → scheduled posts.

待看 火宝短剧 - 基于AI的一站式短剧生成平台 《一句话生成完整短剧,从剧本到成片全自动化》

强依赖 Sora 或豆包视频服务接口地址

目前最好的 story-flicks 成功。需要LLM和图片生成的api,可以蹭免费的阿里百炼llm-qwen3-max和硅基流动的Kwai-Kolors/Kolors 已经二次开发

成功,完全免费,short-video-maker,完全免费,另一个版本

成功,可以免费运行,但是效果不好,视频和文案不匹配,用到:LLM可选/ngrok/,MoneyPrinterTurbo

short-video-factory 效果很差且需要手动给视频素材

门槛:需要本地comfyi或者购买runninghub

MoneyPrinterPlus 门槛:需要开通一些云服务

Stable Video Diffusion (by Stability AI)

Long Video Generation

ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)

字字动画 - 完全免费的AIGC视频生成软件,主要用于AI短剧,AI电影,小说推文

code to video

remotion 3人以下公司免费

Prompt to Video,有门槛,需要OPENAI_API_KEY和ELEVENLABS_API_KEY

文字=》视频

Step‑Video‑TI2V

https://github.com/univa-agent/univa

https://github.com/hpcaitech/Open-Sora

https://github.com/SCUTlihaoyu/open-chat-video-editor

https://github.com/FoundationVision/Waver

https://github.com/YBYBZhang/ControlVideo

pika labs https://runwayml.com/

https://github.com/lllyasviel/FramePack

利用AI大模型,一键生成高清短视频 Generate short videos with one click using AI LLM.

自动化视频制作工具ShortGPT https://mp.weixin.qq.com/s/tdgld7kH4GFhgtjOK1LQ7w

Image to Video

https://github.com/camenduru/stable-video-diffusion-colab

Unlimited-length talking video generation​​ that supports image-to-video and video-to-video generation

https://mp.weixin.qq.com/s/jD1hoQNjUv9eCCH__dzOQQ

https://mp.weixin.qq.com/s/aLXCrH4sUK8HY-h2D5zPSw

钉钉+SD+UE5 https://mp.weixin.qq.com/s/fNxb2B5PiTTzTSXMuXHvOg

数字人

wan2.2

教学视频

教学动画

雾象是一款由大型语言模型(LLM)驱动的动画引擎 agent 。用户输入抽象概念或词语,雾象会将其转化为高水平的生动动画。

Video generation via code

videotutor

talking stickman 火柴人

Papagayo-NG → Synfig

Papagayo-NG (lip-sync) Audio → Load Audio File Paste Transcript Lip-sync → Auto-sync File → Export → Synfig (.dat) Synfig Studio New File 1920×1080, 30fps Head Circle tool → draw head Eyes Two small circles Mouth Shapes (IMPORTANT) Create 5 mouths as separate layers: Mouth_A Mouth_E Mouth_O Mouth_U Mouth_M Put all mouth layers into a Group called: mouth Apply Papagayo Lip-Sync in Synfig Select the mouth group Canvas → Properties → Lip Sync Load Synfig (.dat) Map mouths

3D场景

https://github.com/princeton-vl/infinigen

漫画

一个利用 AI 制作漫画的工具,支持脚本创作、分镜设计和角色风格控制。

https://animatediff.github.io/

极虎漫剪 https://mp.weixin.qq.com/s/eKkcFNx77DJM4Usoc72fjw

https://wonderdynamics.com/

视频变换

对口型 lipsync

https://lipsync.video/

Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

LatentSync

MuseTalk

Wav2Lip

SadTalker

d-id-tts-talkinghead (OpenTalker community)

VideoReTalking (by OpenTalker)

EmoTalker

换脸

Real-time face swap and video deepfake with a single click and only a single image.

剪辑

CapCutAPI is a powerful editing API that empowers you to take full control of your AI-generated assets, including images, audio, video, and text. It provides the precision needed to refine and customize raw AI output, such as adjusting video speed or mirroring an image.

AutoClip : AI-powered video clipping and highlight generation · 一款智能高光提取与剪辑的二创工具

虎牙,斗鱼,抖音,BiliBili,TikTok,Twitch🔥热门🔥智能直播视频剪辑发布AI机器人,自动化🤖,全智能化⚙(智能生成切片,标题,封面,简介),可视化👓,平台热门监控🌡,丰富插件随意扩展🕹,快速部署⚡,视频账号打造自动发布🌟,支持DIY

opusclip

Deepfake 视频换脸 https://mp.weixin.qq.com/s/9RJGpxvKieMY4Mu__mHfFQ https://colab.research.google.com/drive/1NG9AoH3QDtC7h97z1Yodmn_CiiGh8Y1T?usp=sharing#scrollTo=0aHr4Fo-7IRy

视频解析 VLM

字幕

An AI-powered video transcription and summarization tool that supports multiple video platforms including YouTube, Tiktok, Bilibili, and 30+ platforms.

VideoLingo-全自动视频搬运工具

配音(音色克隆)

智能视频多语言AI配音/翻译工具 - Linly-Dubbing — “AI赋能,语言无界”

AI视频翻译配音工具,100种语言双向翻译,一键部署全流程,可以生抖音,小红书,哔哩哔哩,视频号,TikTok,Youtube等形态的内容成适配

PyVideoTrans一键视频翻译+配音+字幕 https://mp.weixin.qq.com/s/6J35rQO8v69mpJfypFl8Tw

图像解析

A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone

Real-time webcam demo with SmolVLM and llama.cpp server

Unlocking Video Understanding with SmolVLM-2: A Comprehensive Guide

detection 标记

视频转图文

AI 视频图文创作助手是一款 Web 工具, 基于 AI 大模型, 一键将视频和音频转化为各种风格的文档

直播/AR增强现实

将 CAM++、SenseVoice​ 和 Silero VAD​ 结合起来,完全可以打造一个非常酷的、具备多模态交互能力的实时视频直播应用。这个想法很棒,它能让虚拟形象不仅“听得见”,还能“看得见”是谁在说话,并进行个性化的交互