Store
Minimal Flux Web UI powered by Gradio & Diffusers
Phased Consistency Model - generate high quality images with 2 steps https://huggingface.co/spaces/radames/Phased-Consistency-Model-PCM
Contribute to compphoto/BoostingMonocularDepth development by creating an account on GitHub.
An arbitrary face-swapping framework on images and videos with one single trained model!
[NVIDIA GPU ONLY] One click installer for Intel's ldm3d
Dense Text-to-Image Generation with Attention Modulation
An open source implementation of Microsoft's VALL-E X zero-shot TTS model
Demo showcasing ~real-time Latent Consistency Model pipeline with Diffusers and a MJPEG stream server (https://github.com/radames/Real-Time-Latent-Consistency-Model)
Demo showcasing ~real-time Latent Consistency Model pipeline with Diffusers and a MJPEG stream server (https://github.com/radames/Real-Time-Latent-Consistency-Model)
Text-to-Video (T2V) generation framework from Vchitect https://github.com/Vchitect/LaVie
An AI powered mirror
A Realtime Creation Engine
Vid2DensePoseFeatured
Convert your videos to densepose and use it on MagicAnimate https://github.com/Flode-Labs/vid2densepose
Estimating the Focal Length of a Monocular Image
Integrates Florence2 and SAM2 models for detailed image captioning and object detection. Florence2 generates detailed captions that are then used to perform phrase grounding. The Segment Anything Model 2 (SAM2) converts these phrase-grounded boxes into masks. https://huggingface.co/spaces/SkalskiP/florence-sam
some archived legacy forge extensions
[IJCV] FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds. AI拟音大师,给你的无声视频添加生动而且同步的音效 😝
Aplikasi ini digunakan untuk menghasilkan suara berbasis teks dengan berbagai pilihan pembicara. Teknologi yang digunakan meliputi model text-to-speech (TTS) yang canggih dengan konversi teks ke fonem. Model yang dipakai dilatih khusus untuk bahasa Indonesia, Jawa dan Sunda.
Bring portraits to life!
