Global radar

New projects people are discovering or following across Pinokio.
Followed1 min
Maestro
github.com/Blizaine

An all-in-one, 100% local AI video, image & music studio. Its Director mode turns a single prompt into a full music video or short film — LLM-planned, shot by shot. Built on the WanGP pipeline (Wan 2.1/2.2, LTX-2.3, Qwen, Hunyuan Video, Flux). Requires an NVIDIA GPU (6GB+ VRAM).

Followed2 min
REAL-Video-Enhancer
github.com/manat0912

Interpolate, Upscale, Decompress, and Denoise videos Locally on Linux/Windows/MacOS.

Followed2 min
Comfyui
github.com/pinokiofactory

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface. https://github.com/comfyanonymous/ComfyUI

Followed5 min
MiniMax H3 ComfyUI
github.com/ThomasEricB

MiniMax H3 omni-modal video generation in ComfyUI. Text/image/video/audio in, video with native 32kHz stereo audio out (768p default, 1080p+ supported). Disk-optimized: pruned INT8 + NVFP4 weights (~63GB instead of ~290GB). NVIDIA only.

Followed6 min
VoxCPM2 Portable
github.com/timoncool

ElevenLabs at home. Multilingual TTS with Voice Design, Voice Cloning, and end-to-end LoRA fine-tuning straight from a video or podcast. Built on VoxCPM2 by OpenBMB. 30 languages incl. Russian.

Followed9 min
CogStudio
github.com/pinokiofactory

[NVIDIA ONLY] Advanced Web UI for CogVideo (text to video, image to video, video to video, extend video, etc) -- Generate videos with less than 10GB VRAM

Followed19 min
LingBot-Map
github.com/cocktailpeanut

NVIDIA launcher for LingBot-Map, a streaming 3D reconstruction viewer from Robbyant.

Followed19 min
Wan2GP - AMD
github.com/6Morpheus6

[AMD ONLY] Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video, Flux and more. (On Windows supported by all dedicated AMD GPUs from RDNA 2 - RDNA 4)

Followed22 min
aura-sr-upscaler
github.com/pinokiofactory

AuraSR-v2 - An open reproduction of the GigaGAN Upscaler from fal.ai https://huggingface.co/spaces/gokaygokay/AuraSR-v2

Followed24 min
e2-f5-tts
github.com/pinokiofactory

F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching https://huggingface.co/spaces/mrfakename/E2-F5-TTS

Followed26 min
FaceFusion
github.com/shadowburn0

Next generation face swapper and enhancer

Followed27 min
Video Studio KH
github.com/theng12

Apple Silicon video studio — text-to-video and video-to-video with LTX-Video, Wan 2.2, HunyuanVideo, and CogVideoX. Powered by PyTorch (MPS) + Diffusers.

Followed31 min
HeartMuLa Studio
github.com/cocktailpeanut

A professional, Suno-like music generation studio for HeartLib. https://github.com/fspecii/HeartMuLa-Studio

Followed37 min
Onetrainer
github.com/6Morpheus6

[NVIDIA, ROCM] One app to train them all. LORA training and Model finetuning for Z-Image, Qwen Image, FLUX.1, Flux.2 Dev and Klein, Chroma, SD 1.5 - 3.5, SDXL, Würstchen-v2, Stable Cascade, PixArt-Alpha, PixArt-Sigma, Sana, Hunyuan Video and inpainting models.

Followed38 min
Faceswap
github.com/manat0912

Deepfakes Software For All

Followed39 min
TRELLIS.2
github.com/Deathdadev

One-click installer for Microsoft TRELLIS.2: High-quality 3D asset generation from images with PBR textures.

Followed41 min
Applio
github.com/pinokiofactory

A simple, high-quality voice conversion tool focused on ease of use and performance.

Followed41 min
MuScriptor
github.com/cocktailpeanut

Local multi-instrument audio-to-MIDI transcription from Kyutai and Mirelo.

Followed42 min
Wan2GP
github.com/pinokiofactory

1-click WanGP Launcher. Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP

Followed43 min
Chatbot-Ollama
github.com/cocktailpeanutlabs

open source chat UI for Ollama https://github.com/ivanfioravanti/chatbot-ollama