Global radar
Local multimodal app powered by Liquid AI LFM2.5-Audio-1.5B and LFM2.5-VL-1.6B models, delivering real-time voice chat, text-to-speech synthesis, long-form audio transcription, and multi-image vision reasoning.
[NVIDIA GPU ONLY] Unofficial Implementation of Animate Anyone https://github.com/MooreThreads/Moore-AnimateAnyone
[NVIDIA ONLY] Efficient Implementation of Animate Anyone (13G VRAM + 2G model size) https://github.com/sdbds/Moore-AnimateAnyone-for-windows
⚡️ Efficient 6B parameter image generation model with sub-second inference. Generate high-quality, photorealistic images with only 8 inference steps. Features bilingual text rendering (Chinese & English) and Single-Stream Diffusion Transformer architecture.
Turn your eBooks into audiobooks using the OmniVoice text-to-speech model
Multi-Voice Text-to-Speech for Stories and Audiobooks. Supports Kokoro and Chatterbox TTS engines with GPU acceleration.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Transform flat 2D sprite PNGs into layered 2.5D asset packs (albedo, normal map, emission, shadow) for Unity, UE5, or custom renderers.
Artist is a training-free text-driven image stylization method. You give an image and input a prompt describing the desired style, Artist give you the stylized image in that style. The detail of the original image and the style you provide is harmonically integrated https://huggingface.co/spaces/fffiloni/Artist
Automatically create music videos. Synchronize the cuts to the music's beat.
Run PrismML Bonsai and Ternary-Bonsai language models locally on macOS, Linux, and Windows with llama.cpp.
Video to 3D: 4D Face Reconstruction from any Video or Image Sequence. Normal Map, Depth Map and 3D Mesh Generation.
