Turn any image into a video! (Web UI created by fffiloni: https://huggingface.co/spaces/fffiloni/MS-Image2Video)
Global radar
Text-to-Speech for 16 Indian languages: Assamese, Bengali, Bodo, English (Indian accent), Hinglish, Gujarati, Hindi, Kannada, Malayalam, Manipuri, Marathi, Odia, Punjabi, Rajasthani, Tamil, and Telugu. SOTA models based on FastPitch and HiFi-GAN V1.
Multi-voice text-to-speech for stories and audiobooks, with local and cloud TTS engines plus optional LLM text processing.
Contribute to lllyasviel/stable-diffusion-webui-forge development by creating an account on GitHub.
Launcher principal y simple para ComfyUI. Instala ComfyUI + Manager, descarga modelos listos para usar y te guia hacia las plantillas oficiales.
Community Pinokio package for OpenClaw preconfigured to use localhost-only Ollama or LM Studio endpoints.
Remove objects from an image https://huggingface.co/spaces/OzzyGT/diffusers-image-fill
Pre-mastering & audio enhancement for AI-generated music. 12-stage processing chain with platform presets (Suno, Udio), before/after spectrogram, and broadcast-ready LUFS normalization.
All-in-one AI music studio by GANTASMO. Stable Audio 3 and Magenta RealTime 2 generation, Chimera multi-track fusion, Demucs stems, MIDI and notation, DJ and VJ rigs, DAW project import, VST3 and .gan plugins, and a RAG-backed in-app assistant.
An enhanced version of Fooocus giving you access to all of the latest AI image generation models
🎬 Professional Video Dubbing Pipeline with Parakeet-TDT-0.6b-v2, Gemini AI, and Edge TTS. Complete solution for automated video dubbing with step-by-step processing and batch video creation from multiple audio files.
FlashVSR - Video and Image Upscaler: [Runs on 12GB vram, 32GB ram] Diffusion-Based Streaming Video Super-Resolution
[NVIDIA Only] Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation https://github.com/fudan-generative-vision/hallo
