Global radar
New projects people are discovering or following across Pinokio.
LFM2-Audio-1.5B is Liquid AI's first end-to-end audio foundation model. Designed with low latency and real time conversation in mind
Turn any image into a video! (Web UI created by fffiloni: https://huggingface.co/spaces/fffiloni/MS-Image2Video)
A multi-voice AI audiobook generator built on Qwen3-TTS — annotate scripts with an LLM, assign unique voices to each character, per-line style instructions for delivery, clone voices from reference audio, design new voices from text descriptions, train custom voices with LoRA fine-tuning, and export to MP3 or Audacity multi-track projects
[ICLR 26 Oral] Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
Create and display beautiful presentations with AI integration in Pinokio
Gradio-based web interface for the LuxTTS voice cloning and text-to-speech model, enabling users to generate customized speech from text using uploaded or recorded audio references with adjustable parameters like speed, guidance scale, and inference steps.
Multi-voice text-to-speech for stories and audiobooks, with local and cloud TTS engines plus optional LLM text processing.
Standalone local hi-res fix: enlarge images with Real-ESRGAN and reinject clean detail with Z-Image Turbo img2img. 100% local — no ComfyUI, no SwarmUI, no cloud.
UVR5 - The ultimate vocal remover application. Ported from https://huggingface.co/spaces/Eddycrack864/UVR5 with added multiplatform GPU support. Orginal project: https://github.com/Anjok07/ultimatevocalremovergui (12GB install. GPU Conversion setting requires 6GB+ VRAM)
Run PrismML Bonsai and Ternary-Bonsai language models locally on macOS, Linux, and Windows with llama.cpp.
Contribute to Eliovp-BV/paiton-vllm-plugin development by creating an account on GitHub.
Open-source alternative to Higgsfield AI, Freepik, Krea, Openart AI — Free AI image generation & cinema studio with 20+ models (Flux, SDXL, Midjourney, Ideogram). Self-hosted, customizable, MIT licensed.
