Global radar

New projects people are discovering or following across Pinokio.
Followed11h ago
VideoCrafter 2
github.com/cocktailpeanutlabs

[Runs fast on NVIDIA GPUs. Works on M1/M2/M3 Macs but slow] VideoCrafter is an open-source video generation and editing toolbox for crafting video content. It currently includes the Text2Video and Image2Video models https://github.com/AILab-CVC/VideoCrafter

Followed11h ago
Drivebay
github.com/lordstryfe

Password-locked file browser for every drive on this machine.

Followed11h ago
VoxCPM2 Portable
github.com/timoncool

ElevenLabs at home. Multilingual TTS with Voice Design, Voice Cloning, and end-to-end LoRA fine-tuning straight from a video or podcast. Built on VoxCPM2 by OpenBMB. 30 languages incl. Russian.

Followed11h ago
HunyuanVideo
github.com/pinokiofactory

[NVIDIA ONLY] Super Optimized Gradio UI for Hunyuan Video Generator that works on GPU poor machines. Generate up to 10~14 sec videos https://github.com/deepbeepmeep/HunyuanVideoGP

Followed11h ago
AI Prompt Studio
github.com/arnold2006

A local prompt-writing assistant for Ideogram 4, MiniMax H3 video, and plain-text image models, powered by a local LLM (llama-server).

Followed11h ago
Qwen3-TTS (AMD GPU)
github.com/trevortai

Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, with AMD GPU support via ROCm. Windows and Linux.

Followed11h ago
ACE-Step 1.5
github.com/cocktailpeanut

The most powerful local music generation model that outperforms most commercial alternatives.

Followed11h ago
browser-use
github.com/pinokiofactory

Run AI Agent in your browser. https://github.com/browser-use/web-ui

Followed11h ago
MuseTalk
github.com/manat0912

MuseTalk is a cutting-edge video-to-video (V2V) lip-sync solution engineered to deliver highly accurate and natural mouth movements synchronized to audio input. Precision LipSync: Realistic and seamless synchronization of speech audio to facial movements. Efficiently designed to run on 8–12 GB VRAM,

Followed11h ago
Qwen3-TTS MLX WebUI Enhanced
github.com/Blizaine

High-quality text-to-speech with Beautiful Web UI & API, optimized for Apple Silicon using MLX. Features include Custom Voice (preset speakers), Voice Design (natural language), and Voice Cloning. With enhanced features for saving custom voices and long-form / endless TTS streaming.

Followed11h ago
Cosyvoice3
huggingface.co/spaces
Followed11h ago
Wan2GP
github.com/matrokweb

Fast AI Video Generation per GPU poor (Wan2.1, Hunyuan, LTV). Gradio UI su http://127.0.0.1:7860

Followed12h ago
Wan2GP
github.com/shadowburn0

Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP

Followed12h ago
Wan2GP
github.com/jillesmc

Super Optimized Gradio UI for AI video creation for GPU poor machines (6GB+ VRAM). Supports Wan 2.1/2.2, Qwen, Hunyuan Video, LTX Video and Flux. https://github.com/deepbeepmeep/Wan2GP

Followed12h ago
Open WebUI
github.com/daddoon

User-friendly WebUI for LLMs, supported LLM runners include Ollama and OpenAI-compatible APIs https://github.com/open-webui/open-webui

Followed12h ago
SwarmUI
github.com/SUP3RMASS1VE

A Modular AI Image Generation Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility. Supports Stable Diffusion, Flux, etc. AI image models, with plans to support AI video, audio, and more in the future.

Followed12h ago
MagicAnimate Mini
github.com/cocktailpeanut

[NVIDIA GPU Only] An optimized version of MagicAnimate https://github.com/sdbds/magic-animate-for-windows

Followed12h ago
TRELLIS.2
github.com/Deathdadev

One-click installer for Microsoft TRELLIS.2: High-quality 3D asset generation from images with PBR textures.

Followed12h ago
Euraika Avatar Studio
github.com/Euraika-Labs

Local-first AI avatar video studio powered by duixcom/Duix-Avatar, Docker, and a consent-aware browser studio.

Followed12h ago
OmniVoice
github.com/PierrunoYT

Zero-shot multilingual TTS (600+ languages) with voice cloning and voice design — Gradio UI (app/app.py)