A Gradio-based web UI for voice cloning and voice design, powered by Qwen3-TTS & VibeVoice. Can use Whisper or VibeVoice-ASR for automatic transcription. Improved from the FrankyB origial version
Global radar
New projects people are discovering or following across Pinokio.
The ultimate space for work and life — to find, build, and collaborate with agent teammates that grow with you. We are taking agent harness to the next level — enabling multi-agent collaboration, effortless agent team design, and introducing agents as the unit of work interaction.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.

We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Japanese GUI + Whisper auto-transcription for Qwen3-TTS. RTX 5090 tested.
Music Generation Foundation Model v1.5
A webui for propainter. Easily pick up objects from the video and eliminate them.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Scalable and Versatile 3D Generation from images
MusePose: a Pose-Driven Image-to-Video Framework for Virtual Human Generation
plug whisper audio transcription to a local ollama server and ouput tts audio responses
Contribute to predatorONE/MuseTalk-Pinokio development by creating an account on GitHub.
Partial MPS support for ComfyUI nodes for LivePortrait to use them on a MacBook
SECourses Musubi Tuner - 1-Click to Install App for LoRA Training and Full Fine Tuning Qwen Image, Qwen Image Edit, Wan 2.1 and Wan 2.2 Models with Musubi Tuner with Ready Presets
Paste a YouTube link, and our AI will detect the most engaging moments, add captions, and edit them into perfect viral shorts ready to post.