Global radar

New projects people are discovering or following across Pinokio.
Discovered6mo ago
alltalk_tts
github.com/erew123

AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of advanced features, such as a settings page, low VRAM support, DeepSpeed, narrator, model finetuning, custom models, wav file maintenance. It can also be used with 3rd Party software via JSON calls.

Discovered6mo ago
JayLL13/VoxCPM-1.5-VN · Hugging Face
huggingface.co/JayLL13

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Discovered6mo ago
MiniCPM-o
github.com/OpenBMB

A Gemini 2.5 Flash Level MLLM for Vision, Speech, and Full-Duplex Multimodal Live Streaming on Your Phone

Followed6mo ago
FastMovieAI
github.com/xhadmincn

FastMovieAI 是一个功能完整的开源短剧/短视频创作平台,采用前后端分离架构,提供 AI 驱动的视频内容创作能力。平台集成了用户管理、支付系统、内容管理、视频生成等完整功能模块,适合内容创作者和视频制作团队使用。

Discovered6mo ago
2cent-tts
github.com/taylorchu

Contribute to taylorchu/2cent-tts development by creating an account on GitHub.

Discovered6mo ago
TRELLIS.2-Text-to-3D-RERUN
github.com/PRITHIVSAKTHIUR

A Gradio app with Rerun visualization for Microsoft's TRELLIS.2-4B model that generates textured 3D assets (GLB) from text or images using a two-stage pipeline: text-to-image (Z-Image-Turbo) then image-to-3D (TRELLIS.2).

Discovered6mo ago
ClearerVoice-Studio
github.com/modelscope

An AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction, etc.

Discovered6mo ago
Fabric
github.com/danielmiessler

Fabric is an open-source framework for augmenting humans using AI. It provides a modular system for solving specific problems using a crowdsourced set of AI prompts that can be used anywhere.

Discovered6mo ago
videosos
github.com/timoncool

Enable AI models for video production in the browser

Discovered6mo ago
resemble-enhance
github.com/resemble-ai

AI powered speech denoising and enhancement

Discovered6mo ago
Real-Time-Voice-Cloning
github.com/CorentinJ

Clone a voice in 5 seconds to generate arbitrary speech in real-time

Discovered6mo ago
Open-Interface
github.com/AmberSahdev

Control Any Computer Using LLMs.

Discovered6mo ago
X-AnyLabeling
github.com/CVHub520

Effortless data labeling with AI support from Segment Anything and other awesome models.

Discovered6mo ago
ComfyUI-Chatterbox
github.com/wildminder

ComfyUI Chatterbox TTS & Voice Conversion Node

Discovered6mo ago
dia
github.com/nari-labs

A TTS model capable of generating ultra-realistic dialogue in one pass.

Discovered6mo ago
sesame-csm
github.com/akashjss

A Conversational Speech Generation Model with Gradio UI and OpenAI compatible API. UI and API support CUDA, MLX and CPU devices.

Discovered6mo ago
agentkits-marketing
github.com/aitytech

Enterprise-grade AI marketing automation for Claude Code, Cursor, GitHub Copilot, and any AI assistant supporting agents & skills