Launcher updates

More
cocktailpeanutlabs/deusupdated 2y ago
A Realtime Creation Engine
0 check-insNVIDIAAMDApple
pinokiofactory/florence-samv2.0updated 2y ago
Integrates Florence2 and SAM2 models for detailed image captioning and object detection. Florence2 generates detailed captions that are then used to perform phrase grounding. The Segment Anything Model 2 (SAM2) converts these phrase-grounded boxes into masks. https://huggingface.co/spaces/SkalskiP/florence-sam
1 check-inNVIDIAAMDApple
pinokiofactory/accdiffusionv2.0updated 2y ago
0 check-insNVIDIAAMDApple
Feedjer/stable-diffusion-webui-ux.pinokiov1.5updated 2y ago
Stable Diffusion web UI UX: https://github.com/anapnoe/stable-diffusion-webui-ux
4 check-insNVIDIAAMDApple
Feedjer/AniPortrait.pinokiov1.5updated 2y ago
AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation:https://github.com/Zejun-Yang/AniPortrait
0 check-insNVIDIAAMDApple
Feedjer/Langflow.pinokiov1.5updated 2y ago
Langflow is a dynamic graph where each node is an executable unit. Its modular and interactive design fosters rapid experimentation and prototyping, pushing hard on the limits of creativity: https://github.com/langflow-ai/langflow
0 check-insNVIDIAAMDApple
Feedjer/HunyuanDiT.pinokiov1.5updated 2y ago
Hunyuan-DiT : A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding/ https://github.com/Tencent/HunyuanDiT
0 check-insNVIDIAAMDApple
Feedjer/Omost.pinokiov1.5updated 2y ago
Your image is almost there!:https://github.com/lllyasviel/Omost
0 check-insNVIDIAAMDApple
Feedjer/Flowise.pinokiov1.5updated 2y ago
Drag & drop UI to build your customized LLM flow: https://github.com/FlowiseAI/Flowise
0 check-insNVIDIAAMDApple
Feedjer/cambrian.pinokiov1.5updated 2y ago
[Need 24GB VRAM] Cambrian-1 is a family of multimodal LLMs with a vision-centric design: https://github.com/cambrian-mllm/cambrian
0 check-insNVIDIAAMDApple
pinokiofactory/doughv1updated 2y ago
Dough is a open source tool for steering AI animations with precision
1 check-inNVIDIAAMDApple
cocktailpeanutlabs/moondream1v1.1updated 2y ago
moondream1 is a tiny (1.6B parameter) vision language model trained by @vikhyatk that performs on par with models twice its size. It is trained on the LLaVa training dataset, and initialized with SigLIP as the vision tower and Phi-1.5 as the text encoder. https://huggingface.co/spaces/vikhyatk/moondream1
0 check-insNVIDIAAMDApple
rimsila/fooocus-API-pinokiov1.5updated 2y ago
1 check-inNVIDIAAMDApple
GivEN29/autogen-studio-pinokioupdated 2y ago
Declaratively define and modify agents and multi-agent workflows through a point and click, drag and drop interface (e.g., you can select the parameters of two agents that will communicate to solve your task).
0 check-insNVIDIAAMDApple
cocktailpeanut/bark.pinokioupdated 2y ago
Upload a clean 20 seconds WAV file of the vocal persona you want to mimic, type your text-to-speech prompt and hit submit! A local version of https://huggingface.co/spaces/fffiloni/instant-TTS-Bark-cloning
@cocktailpeanut0 check-insNVIDIAAMDApple
cocktailpeanut/ms-video2video.pinokioupdated 2y ago
enhance the resolution and spatiotemporal continuity of text-generated videos and image-generated videos
@cocktailpeanut0 check-insNVIDIAAMDApple
cocktailpeanut/xinference.pinokioupdated 2y ago
LLM Web UI and API
@cocktailpeanut0 check-insNVIDIAAMDApple
cocktailpeanut/sdxl-turboupdated 2y ago
A Real-Time Text-to-Image Generation Model
@cocktailpeanut2 check-insNVIDIAAMDApple
cocktailpeanutlabs/paligemmav1.5updated 2y ago
an open vision-language model by Google. PaliGemma is designed as a versatile model for transfer to a wide range of vision-language tasks such as image and short video caption, visual question answering, text reading, object detection and object segmentation https://huggingface.co/spaces/google/paligemma
0 check-insNVIDIAAMDApple