Followed5d ago
illustrious-lora-trainer
github.com/nec-o

针对 RTX 3060 6GB 优化的 SDXL / Illustrious 风格 LoRA 训练脚本,纯 diffusers + peft,一键输出 ComfyUI 通用格式

Followed5d ago
Stable Video Diffusion
github.com/shadowburn0

[NVIDIA ONLY] Stable Video Diffusion Streamlit App. Currently supports Nvidia GPU machines only.

Followed5d ago
DiffBIR
github.com/bycloud-AI

DiffBIR: Towards Blind Image Restoration with Generative Diffusion Prior

Followed5d ago
openshorts
github.com/mutonby

Free & open source AI video platform — Clip Generator, AI Shorts (UGC with AI actors) & YouTube Studio. Self-hosted, no watermarks.

Followed5d ago
Dolce Colore
github.com/ntropey

Locally colorize black-and-white public domain movies and historical photos with DDColor. Runs on CPU, Intel Graphics, Apple Silicon, or NVIDIA CUDA.

Followed5d ago
Omnigen 2
github.com/6Morpheus6

Unified Image Understanding and Generation. Text-to-Image Generation, In-context Generation, Instruction-guided Image Editing, Visual Understanding (Minimum Requirements 12GBV RAM / 48GB RAM, Recommended Requirements 24GB VRAM / 32GB RAM)

Followed5d ago
SenseNova-U1 Image
github.com/shinshekai

High-quality image generation and editing powered by SenseNova-U1-8B-MoT (NEO-Unify architecture). Supports text-to-image, image-to-image editing, and reasoning mode.

Followed5d ago
YouTube Video Factory MVP
github.com/LuMa670

Local AI YouTube video generator: script, scenes, voiceover, thumbnail and MP4 export.

Followed5d ago
Audio Flamingo 3
github.com/PierrunoYT

NVIDIA's Audio Flamingo 3 - Large Audio-Language Model for speech, sound, and music understanding with Gradio web interface

Discovered6d ago
Opal
github.com/debpalash

Opal — The open-source everything media player & browser. Jellyfin/Stremio/Kodi alternative for macOS, Linux & Windows.

Followed6d ago
crispz-klein
github.com/mikecastrodemaria

FLUX.2 Klein 4B studio (Apache 2.0, 4 steps), Fooocus-style, 100% local: txt2img, multi-reference editing in the SAME pipeline (up to 4 refs, no second model), inpaint/outpaint, ESRGAN+refine upscale, LoRA, styles, Describe/Improve & Vision Mix (Ollama), Remove BG, Face Swap. ~15 GB VRAM for everything. Note: klein is distilled, so negative prompts and CFG have no effect. Fork of crispz-qwen-edit. https://github.com/mikecastrodemaria/crispz-klein

Followed6d ago
seedance-2-generator
github.com/SamurAIGPT

Open-source Seedance 2.0 video generator — production-ready Next.js SaaS for text-to-video and multi-image reference editing. Stripe billing, credits, NextAuth, and Prisma out of the box.

Discovered6d ago
dalle-mini
github.com/borisdayma

DALL·E Mini - Generate images from a text prompt

Discovered6d ago
dalle-mini/dalle-mini · Hugging Face
huggingface.co/dalle-mini

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Discovered6d ago
m-a-p/SheetSage2 · Hugging Face
huggingface.co/m-a-p

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Followed6d ago
AI-Youtube-Shorts-Generator
github.com/anil-matcha

Open-source alternative to Opus Clip, Vidyo.ai, Klap & SubMagic. Turn long-form YouTube videos into viral 9:16 shorts using LLM highlight detection, Whisper transcription, and auto vertical cropping — free, no watermarks, no per-clip credits.

Discovered6d ago
multiset-unity-sdk
github.com/multiset-ai

Contribute to MultiSet-AI/multiset-unity-sdk development by creating an account on GitHub.

Discovered6d ago
yue2-studio
github.com/krakenunbound

Native Windows music studio for YuE2-3B, with local model downloads, song editing, lyrics, covers and video tools.

Discovered6d ago
OOTDiffusion
github.com/levihsu

[AAAI 2025] Official implementation of "OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on"