Wanted
1,941 projectsNon-launcher projects without a Pinokio launcher yet.
Gradio based WebUI with a SAM (segment-anything)
[CVPR 2026] Towards Real-Time Diffusion-Based Streaming Video Super-Resolution — An efficient one-step diffusion framework for streaming VSR with locality-constrained sparse attention and a tiny conditional decoder.
Swift client for the fal.ai model APIs
ComfyUI node for background removal, implementing InSPyreNet the best method up to date
PyTorch implementation of Real-ESRGAN model
[ECCV 2024, Oral] DynamiCrafter: Animating Open-domain Images with Video Diffusion Priors
Perfect Green Screen Keys made EZ!
Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.
Wan2.1 for Mac.
The official JavaScript (Node) library for the ElevenLabs API.
🚀🪐🌕🌑☄️🛸 Opensource equivalent of Google's Antigravity
深度学习辅助漫画翻译工具, 支持一键机翻和简单的图像/文本编辑 | Yet another computer-aided comic/manga translation tool powered by deeplearning
Get up and running with Kimi-K2.5, GLM-5, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Use Gemini to decide which Pexels videos to download for your project
Lets make loop video diffusion practical!
A cross-platform desktop application for running AI models from [WaveSpeedAI](https://wavespeed.ai), as well as many free local AI models including Z-Image.
Contribute to WaveSpeedAI/wavespeed-comfyui development by creating an account on GitHub.
An AI Hedge Fund Team
Local-first AI video intelligence platform. Index your video library with multi-modal analysis (YOLO, DeepFace, Whisper), search semantically via natural language, Docker-ready.
VietTTS: An Open-Source Vietnamese Text to Speech
