图像与视频生成
文生图、图生视频、图像编辑与设计辅助
数据截至 8/1 18:48(构建快照,正在获取最新)
按仍在维护排序:先筛掉近 30 天没有提交的项目,再按热度排。 高 star 但早已停更的项目不会出现在这里。
- 1
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
- 2
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
- 3
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.
- 4
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
- 5
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.
- 6
Unrestricted Open-source alternative to AI video platforms — Free AI image & video generation studio with 500+ models (Flux, Midjourney, Kling, Sora, Veo). No content filters. Self-hosted, MIT licensed.
- 7
Toonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animated short dramas. Features AI scriptwriting, storyboarding, character and video generation. A cross-platform desktop app for efficient content creation.
- 8
🍌 World's largest Nano Banana Pro prompt library — 10,000+ curated prompts with preview images, 16 languages. Google Gemini AI image generation. Free & open source.
- 9
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
- 10
Production ready toolkit to run AI locally
- 11
🚀 World's largest GPT Image 2 prompt library, updated daily — 2000+ curated prompts with preview images, 16 languages. OpenAI's next-gen image model with pixel-perfect text rendering, cross-image consistency, and commercial-grade illustration. Free & open source.
- 12
Multi-Platform Package Manager for Stable Diffusion
- 13
Run Stable Diffusion on Mac natively
- 14
A repository of models, textual inversions, and more
- 15
SD.Next: All-in-one WebUI for AI generative image and video creation, captioning and processing
- 16
截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS
- 17
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
- 18
This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc
- 19
A framework for efficient model inference with omni-modality models
- 20
A general fine-tuning kit geared toward image/video/audio diffusion models.
另有 31 个项目因近期无提交或缺少数据未列入。
图像与视频生成要落到业务里,还差什么?
开源项目给的是能力,不是方案。数据怎么接、权限怎么管、上线后谁维护,这些才是落地的真正成本。