Global radar

New projects people are discovering or following across Pinokio.
Followed24d ago
Qwen-Image-Edit-2511-LoRAs-Fast-Lazy-Load
github.com/PRITHIVSAKTHIUR

Demonstration for the Qwen-Image-Edit-2511 model with lazy-loaded LoRA adapters for advanced single- and multi-image editing. Supports 7+ specialized LoRAs including photo-to-anime, multi-angle camera control, pose transfer (Any-Pose), upscaling, style transfer, light migration, and manga tone. Features fast inference (4 steps default).

Followed25d ago
SongGeneration Studio
github.com/Ivan-Stanisci

AI Song Generation with Full Style Control - Generate complete songs with lyrics, vocals, and instrumental tracks using Tencent AI Lab's SongGeneration (LeVo) model. [NVIDIA ONLY]

Followed25d ago
portable-hermes-agent
github.com/rookiemann

Hermes Agent made portable desktop for Windows — 100 tools, GUI, local models via LM Studio, TTS, Music, ComfyUI, workflows, tool maker. No install. No Docker. No admin rights.

Followed25d ago
Voice Studio KH
github.com/theng12

Apple Silicon TTS — curated MLX-first models for cloning, multilingual speech, narration, and expressive voices.

Followed25d ago
Auris BadBaDaki
github.com/tbl4dr

Auris-BadBaDaki is Offline audiobook reader for EPUB, PDF, and TXT with local OmniVoice TTS, character-aware voices, per-book narrator control, and synced text highlighting. Everything runs locally after setup. No API keys. No hosted TTS dependency.

Discovered25d ago
OpenVoiceV2 - a Hugging Face Space by myshell-ai
huggingface.co/spaces

Enter the text you want spoken, choose a language/style, and upload a brief recording of the speaker you wish to imitate. After agreeing to the MIT license, the app creates an audio file that reads...

Followed25d ago
DocToSpeech (OpenVoice)
github.com/C0m3b4ck

Fully offline document-to-speech converter using OpenVoice V2. Tone color conversion and voice cloning via MeloTTS. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.

Followed25d ago
ZipVoice
github.com/SUP3RMASS1VE

Fast and High-Quality Zero-Shot voice clone Text-to-Speech with Flow Matching

Followed25d ago
Omnigen 2
github.com/6Morpheus6

Unified Image Understanding and Generation. Text-to-Image Generation, In-context Generation, Instruction-guided Image Editing, Visual Understanding (Minimum Requirements 12GBV RAM / 48GB RAM, Recommended Requirements 24GB VRAM / 32GB RAM)

Followed25d ago
Faster-Whisper-XXL-GUI
github.com/cbro33

GUI for Faster‑Whisper‑XXL transcription tool: download YouTube audio, transcribe local files, manage models, and export multiple formats with themes and auto yt‑dlp updates.

Followed25d ago
dgrauet/ltx-2.3-mlx-q4 · Hugging Face
huggingface.co/dgrauet

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Followed25d ago
video-object-remover.git
github.com/6Morpheus6

[NVIDIA ONLY] Remove Objects in videos with inpainting. Recommended requirements 16 - 24 GB VRAM / 48 GB RAM, Minimal requirements 12GB VRAM / 32 GBRAM

Followed25d ago
YuE
github.com/multimodal-art-projection

YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open

Discovered26d ago
beads
github.com/gastownhall

Beads - A memory upgrade for your coding agent

Discovered26d ago
likec4
github.com/likec4

Visualize, collaborate, and evolve the software architecture with always actual and live diagrams from your code