We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Global radar
Use multiple FLUX.2-Klein LoRAs
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
PDFCraft is a free, privacy-focused PDF toolkit that runs entirely in your browser. With 90+ professional tools, you can edit, convert, merge, split, and secure your PDF files without ever uploading them to a server.
Temporally consistent video editing. A local version of https://huggingface.co/spaces/weizmannscience/tokenflow
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Open-source social media scheduling tool with AI. Schedule posts to X, LinkedIn, Reddit, Discord, Threads, TikTok, YouTube, Pinterest, Dribbble, Slack, Mastodon, Facebook, GitHub, and more.
User-friendly WebUI for LLMs, supported LLM runners include Ollama and OpenAI-compatible APIs https://github.com/open-webui/open-webui
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Turn any image into a DLSS 5 meme (using FLUX.2-klein-9b-kv)
Unified Image Understanding and Image Generation with Data and Model Scaling https://github.com/peanutcocktail/Janus
[SIGGRAPH Asia 2026 Journal Track] InfiniSplat: Implicit Gaussian Decoding for Large-Baseline Monocular View Synthesis
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Create and display beautiful presentations with AI integration in Pinokio
Transform lyric transcriptions into karaoke-style MP4 videos. Built on Python-Lyric-Transcriber, this Gradio UI uses Whisper for transcription, an LLM for lyric edits, and Demucs for vocal separation. A fun tool for karaoke fans, though outputs may vary.
Fully offline document-to-speech converter using Resemble Chatterbox TTS. 350M parameter model with paralinguistic tags and voice cloning. Converts EPUB, PDF, DOCX, HTML, and TXT files to audio. No API keys or cloud services required.
[NVIDIA Only] Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation https://github.com/fudan-generative-vision/hallo
