Global radar

New projects people are discovering or following across Pinokio.
Discovered6mo ago
ACE-Step/Ace-Step1.5 · Hugging Face
huggingface.co/ACE-Step

We’re on a journey to advance and democratize artificial intelligence through open source and open science.

Discovered6mo ago
seamless_communication
github.com/facebookresearch

Foundational Models for State-of-the-Art Speech and Text Translation

Discovered6mo ago
FacePoke
github.com/jbilcke-hf

Select a portrait, click to move the head around (please use your own space / GPU!)

Followed6mo ago
subtitle-translator
github.com/rockbenben

⚡️ Blazing-fast batch subtitle translation for SRT/ASS/VTT/LRC — 70+ languages, AI-powered 批量字幕翻译

Discovered6mo ago
auto-subtitle-translate
github.com/YJ-20

Automatically generate, translate, and overlay subtitles for any video.

Discovered6mo ago
Coqui-TTS-XTTS-v2-
github.com/Jaden-J

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

Discovered6mo ago
Qwen3
github.com/QwenLM

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.

Discovered6mo ago
ZipVoice
github.com/k2-fsa

Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching

Followed6mo ago
FramePack-ZLUDA
github.com/githubcto

Lets make video diffusion practical!

Discovered6mo ago
langchain
github.com/langchain-ai

🦜🔗 The platform for reliable agents.

Followed6mo ago
qwen3-tts-apple-silicon
github.com/kapi2800

Run Qwen3-TTS text-to-speech locally on Mac (M1/M2/M3/M4). Voice cloning, voice design, custom voices. 100% offline using MLX.

Discovered6mo ago
Style-Bert-VITS2
github.com/litagin02

Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles.

Discovered6mo ago
OpenVoice - a Hugging Face Space by myshell-ai
huggingface.co/spaces

Enter text and upload a reference audio to create a synthesized speech that matches the speaker's voice and chosen style. Supports English and Chinese.

Discovered6mo ago
SadTalker - a Hugging Face Space by vinthony
huggingface.co/spaces

Upload a source image and audio file to create a video of the image's face moving and speaking as if it were saying the audio. You can also use reference videos to enhance the animation.

Followed6mo ago
MiniMax-M2.1
github.com/MiniMax-AI

MiniMax M2.1, a SOTA model for real-world dev & agents.

Followed6mo ago
ACE-Step
github.com/ace-step

ACE-Step: A Step Towards Music Generation Foundation Model

Discovered6mo ago
open-omr
github.com/gregoriomomm

Code for a proof of concept for OMR (Optical Mark Recognition) using opensource tools