Global radar
All-in-one AI music studio on Stable Audio 3 and a CUDA port of Magenta RealTime. It generates audio from text, separates stems with Demucs, transcribes to MIDI and notation, edits a multitrack timeline with a real-time Web Audio FX rack and automation, masters, DJs with stem decks, runs a live VJ engine, and plays from a Quest by hand over ADB.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Paste long text, clean it into readable sections, summarize each section, and ask questions in-browser with WebGPU. Choose between LFM2.5 230M, 350M, 1.2B-Instruct, and 1.2B-Thinking.
Contribute to HumanAIGC/lite-avatar development by creating an account on GitHub.
In order to make it easier to use the ComfyUI, I have made some optimizations and integrations to some commonly used nodes.
LEMAS‑TTS is a multilingual zero‑shot text‑to‑speech system, supporting 10 languages: Chinese English Spanish Russian French German Italian Portuguese Indonesian Vietnamese
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
generate a video from an image with a text prompt
Type a description into the box and click “Run” to have the app create several pictures that match your prompt. The generated images appear in a gallery below, and you can download a screenshot of ...
Official Implementation of SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning
🗂 The essential checklist for modern web development, for humans and AI agents
TripoSplat converts a single 2D image into high-quality and variable number of 3D Gaussians, developed by TripoAI.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
