图像与视频生成

文生图、图生视频、图像编辑与设计辅助

数据截至 8/1 18:48(构建快照,正在获取最新)

仍在维护排序:先筛掉近 30 天没有提交的项目,再按热度排。 高 star 但早已停更的项目不会出现在这里。

  1. 1
    ComfyUIGitHub图像与视频生成GPL-3.0

    The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

    123.1kstar今天有提交Python
  2. 2
    LocalAIGitHub图像与视频生成MIT

    LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

    48.1kstar今天有提交Go
  3. 3
    OpenMontageGitHub图像与视频生成AGPL-3.0

    World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant into a full video production studio.

    44.5kstar7 天前提交Python
  4. 4
    diffusersGitHub图像与视频生成Apache-2.0

    🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

    34.2kstar今天有提交Python
  5. 5
    InvokeAIGitHub图像与视频生成Apache-2.0

    Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.

    27.7kstar今天有提交Python
  6. 6
    Open-Generative-AIGitHub图像与视频生成MIT

    Unrestricted Open-source alternative to AI video platforms — Free AI image & video generation studio with 500+ models (Flux, Midjourney, Kling, Sora, Veo). No content filters. Self-hosted, MIT licensed.

    25.3kstar今天有提交JavaScript
  7. 7
    Toonflow-appGitHub图像与视频生成Apache-2.0

    Toonflow 是开源一站式 AI 短剧创作工具,将小说、剧本快速转化为动画短剧。集成 AI 编剧、智能分镜、角色与视频生成,跨平台桌面端轻量部署,助力创作者低成本批量产出视觉内容。Toonflow is an open-source AI tool that turns stories and scripts into animated short dramas. Features AI scriptwriting, storyboarding, character and video generation. A cross-platform desktop app for efficient content creation.

    13.2kstar4 天前提交TypeScript
  8. 8
    awesome-nano-banana-pro-promptsGitHub图像与视频生成

    🍌 World's largest Nano Banana Pro prompt library — 10,000+ curated prompts with preview images, 16 languages. Google Gemini AI image generation. Free & open source.

    13.0kstar今天有提交TypeScript
  9. 9
    openvinoGitHub图像与视频生成Apache-2.0

    OpenVINO™ is an open source toolkit for optimizing and deploying AI inference

    10.6kstar今天有提交C++
  10. 10
    runanywhere-sdksGitHub图像与视频生成

    Production ready toolkit to run AI locally

    10.3kstar今天有提交C++
  11. 11
    awesome-gpt-image-2GitHub图像与视频生成

    🚀 World's largest GPT Image 2 prompt library, updated daily — 2000+ curated prompts with preview images, 16 languages. OpenAI's next-gen image model with pixel-perfect text rendering, cross-image consistency, and commercial-grade illustration. Free & open source.

    9.0kstar今天有提交TypeScript
  12. 12
    StabilityMatrixGitHub图像与视频生成AGPL-3.0

    Multi-Platform Package Manager for Stable Diffusion

    8.6kstar6 天前提交C#
  13. 13
    MochiDiffusionGitHub图像与视频生成GPL-3.0

    Run Stable Diffusion on Mac natively

    7.9kstar6 天前提交Swift
  14. 14
    civitaiGitHub图像与视频生成Apache-2.0

    A repository of models, textual inversions, and more

    7.2kstar今天有提交TypeScript
  15. 15
    sdnextGitHub图像与视频生成Apache-2.0

    SD.Next: All-in-one WebUI for AI generative image and video creation, captioning and processing

    7.2kstar今天有提交Python
  16. 16
    eSearchGitHub图像与视频生成GPL-3.0

    截屏 离线OCR 搜索翻译 以图搜图 贴图 录屏 万向滚动截屏 屏幕翻译 Screenshot Offline OCR Search Translate Search for picture Paste the picture on the screen Screen recorder Omnidirectional scrolling screenshot Screen translator 支持Windows Linux macOS

    6.8kstar今天有提交TypeScript
  17. 17
    stable-diffusion.cppGitHub图像与视频生成MIT

    Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++

    6.6kstar1 天前提交C++
  18. 18
    Awesome-Prompt-EngineeringGitHub图像与视频生成Apache-2.0

    This repository contains a hand-curated resources for Prompt Engineering with a focus on Generative Pre-trained Transformer (GPT), ChatGPT, PaLM etc

    6.2kstar今天有提交TypeScript
  19. 19
    vllm-omniGitHub图像与视频生成Apache-2.0

    A framework for efficient model inference with omni-modality models

    5.8kstar今天有提交Python
  20. 20
    SimpleTunerGitHub图像与视频生成AGPL-3.0

    A general fine-tuning kit geared toward image/video/audio diffusion models.

    2.9kstar3 天前提交Python

另有 31 个项目因近期无提交或缺少数据未列入。

图像与视频生成要落到业务里,还差什么?

开源项目给的是能力,不是方案。数据怎么接、权限怎么管、上线后谁维护,这些才是落地的真正成本。