Global radar

New projects people are discovering or following across Pinokio.
Followed7h ago
ACE-Step Studio
github.com/timoncool

Suno at home. Local AI music generation studio — full songs with vocals, lyrics, covers, and music videos. Built on ACE-Step 1.5 XL.

Followed7h ago
OpenAudio
github.com/pinokiofactory

Multilingual Text-to-Speech with Voice Cloning (Supports: English, Japanese, Korean, Chinese, French, German, Arabic, and Spanish) https://github.com/fishaudio/fish-speech

Followed7h ago
Openvoice2
github.com/cocktailpeanutlabs

Openvoice 2 Web UI - A local web UI for Openvoice2, a multilingual voice cloning TTS https://x.com/myshell_ai/status/1783161876052066793

Followed7h ago
Qwen3-TTS
github.com/Xeronal81

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team

Followed7h ago
VoxCPM
github.com/IAnMove

Tokenizer-free multilingual TTS and voice cloning with low-VRAM and VoxCPM2 Web UI/API launch modes.

Followed7h ago
DramaBox
github.com/PierrunoYT

Expressive TTS with voice cloning, prompt-driven speech synthesis built on LTX-2.3 by Resemble AI

Followed7h ago
IP-Adapter-FaceID
github.com/cocktailpeanutlabs

Enter a face image and transform it to any other image. Demo for the h94/IP-Adapter-FaceID model https://huggingface.co/spaces/multimodalart/Ip-Adapter-FaceID

Followed7h ago
OpenClaw (aka ClawdBot)
github.com/cocktailpeanut

The AI that actually does things https://openclaw.ai

Followed7h ago
RVC-realtime
github.com/Feedjer

[WINDOWS/LINUX ONLY] Easily train a good VC model with voice data <= 10 mins!: https://github.com/RVC-Project/Retrieval-based-Voice-Conversion-WebUI

Followed7h ago
Flowise
github.com/Feedjer

Drag & drop UI to build your customized LLM flow: https://github.com/FlowiseAI/Flowise

Followed7h ago
Z-Fusion
github.com/ai-anchorite

Z-Image, Flux2 Klein, & SeedVR2 with a Gradio UI. Uses a built-in ComfyUI backend for speed and efficiency! [8GB+VRAM, 16GB+ RAM]

Followed8h ago
stable-diffusion-webui-ux
github.com/linhclone1399

Stable Diffusion web UI UX: https://github.com/anapnoe/stable-diffusion-webui-ux

Followed8h ago
Stable Audio 3
github.com/cocktailpeanut

Launcher for Stable Audio 3 Small Music, Small SFX, and NVIDIA Medium using public cocktailpeanut Hugging Face mirrors. https://github.com/Stability-AI/stable-audio-3

Followed8h ago
Transcribr
github.com/PierrunoYT

Bulk transcribe many YouTube videos, whole playlists, or your own uploaded audio/video files at once with faster-whisper. Outputs txt, srt, vtt, or json.

Followed8h ago
TripoSR
github.com/cocktailpeanutlabs

a state-of-the-art open-source model for fast feedforward 3D reconstruction from a single image, developed in collaboration between Tripo AI and Stability AI. https://huggingface.co/spaces/stabilityai/TripoSR

Followed8h ago
Wan 2.1
github.com/remphanstar

[NVIDIA ONLY] Super Optimized Gradio UI for Wan2.1 video for GPU poor machines (5GB+ VRAM). Generate up to 12 sec videos https://github.com/deepbeepmeep/Wan2GP

Followed8h ago
Unlimited OCR – Searchable PDF
github.com/Krill99

Local Baidu Unlimited-OCR with live rendered Markdown and optional Searchable Scan or Reconstructed PDF output.

Followed8h ago
Whisper-WebUI
github.com/pinokiofactory

A Web UI for easy subtitle using whisper model.