An AI powered mirror
cocktailpeanut
@cocktailpeanutCreations by @cocktailpeanut
154 totalConvert your videos to densepose and use it on MagicAnimate https://github.com/Flode-Labs/vid2densepose
Bring portraits to life!
Build and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
MagicAnimate: Temporally Consistent Human Image Animation using Diffusion Model
Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
Upload a clean 20 seconds WAV file of the vocal persona you want to mimic, type your text-to-speech prompt and hit submit! A local version of https://huggingface.co/spaces/fffiloni/instant-TTS-Bark-cloning
enhance the resolution and spatiotemporal continuity of text-generated videos and image-generated videos
An all-in-one LLMs Chat UI for Apple Silicon Mac using MLX Framework.
User-friendly WebUI for LLMs (Formerly Ollama WebUI)
LLM Web UI and API
A Real-Time Text-to-Image Generation Model
Marching cubes implementation for PyTorch environment.
Contribute to cocktailpeanut/dust3r development by creating an account on GitHub.
llama.cpp with BakLLaVA model describes what does it see (https://github.com/Fuzzy-Search/realtime-bakllava)
Contribute to cocktailpeanut/stable-diffusion-webui-forge development by creating an account on GitHub.
An intuitive GUI for GLIGEN that uses ComfyUI in the backend
A Web UI for easy subtitle using whisper model.
moondream1 is a tiny (1.6B parameter) vision language model trained by @vikhyatk that performs on par with models twice its size. It is trained on the LLaVa training dataset, and initialized with SigLIP as the vision tower and Phi-1.5 as the text encoder. https://huggingface.co/spaces/vikhyatk/moondream1
Focus on prompting and generating

