Wanted
1,939 projectsNon-launcher projects without a Pinokio launcher yet.
An AI focused photo manipulation tool based on Gradio
Python script that scrapes the currently trending YouTube videos in a variety of countries
SBX Bank Bridge System
Contribute to HumanAIGC-Engineering/OpenAvatarChat development by creating an account on GitHub.
Open-source TTS for European languages with full voice cloning - fork of KugelAudio and based on VibeVoice
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
Local Audio Transcription Tool
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持Windows, macOS, Linux)
An MCP App for playing and inspecting local audio files in an MCP host.
A simple GUI to use Whisper.
Zonos-v0.1 is a leading open-weight text-to-speech model trained on more than 200k hours of varied multilingual speech, delivering expressiveness and quality on par with—or even surpassing—top TTS providers.
Auto detecting, masking and inpainting with detection model.
FireRed-Image-Edit is a powerful image editing foundation model achieving open-source state-of-the-art performance with precise instruction following, high-fidelity generation, superior identity consistency, and seamless multi-element fusion.
Inference server for MioTTS, a lightweight and fast LLM-based TTS model.
Use claude-code for free in the terminal, VSCode extension or discord like OpenClaw (voice supported)
Helios: Real Real-Time Long Video Generation Model
AI short drama & micro-drama video generator — turns any idea into a complete short-form drama using multi-agent AI pipeline (screenwriter → storyboard → frames → video). Seedance 2 VIP, Kling 3.0 Pro, Veo 3.1, Sora 2.
A comprehensive Model Context Protocol (MCP) server that enables AI agents to create fully mixed and mastered tracks in REAPER with both MIDI and audio capabilities.
A minimal PDF reader with built-in text-to-speech, powered by Kokoro TTS
[CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
