ComfyUI HiTem3D Integration - Generate 3D models from images using HiTem3D API
Global radar
Enter text and select a voice to convert it to speech. Adjust the speech rate and pitch to your preference and hear the result as an audio file.
Real-time Vision Language Model interaction via webcam - WebRTC-based web interface
Create 🔥 videos with Stable Diffusion by exploring the latent space and morphing between text prompts
An autonomous agent that takes work, does work, gets paid, and gets better at it.
A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot text-to-speech
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Chatterbox TTS supporting 23 languages
The most powerful local music generation model that outperforms most commercial alternatives
Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.
🚀 Unleash AMD GPU Performance: Fix PyTorch ROCm detection for 4x AI/ML speedup on RX 6000/7000 series for Pinokio and developers / custom setups
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Platform to build admin panels, internal tools, and dashboards. Integrates with 25+ databases and any API.
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
A Gradio-based web application for performing image editing tasks using the FireRed-Image-Edit-1.0 model with accelerated 4-step inference. Supports single and multi-image editing through natural language prompts.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
