Open Vocabulary Image Segmentation using Segment Anything Model and MetaCLIP combo
Global radar
State-of-the-art open-source speech recognition model supporting 14 languages. 2B parameter ASR model from Cohere Labs.
A female-presenting Telegram bot that listens and replies as a caring, affectionate girlfriend.
Turn hand-drawn story illustrations into 35–45 second line-reveal and gradual-coloring videos with HyperFrames.
Think with AI beyond the chat box. A shared canvas for handwriting, equations, diagrams, and spatial reasoning.
Ultimate Vocal Remover 5 with Gradio UI. Separate an audio file into various stems, using multiple models
Contribute to fromlifetolines/ai-data-studio-dashboard development by creating an account on GitHub.
A self-improving swarm of local-LLM agents that mine, smelt, build, farm, and fight their way through Minecraft as a coordinated team. Built on mineflayer + Ollama.
Digital Avatar Conversational System - Linly-Talker. 😄✨ Linly-Talker is an intelligent AI system that combines large language models (LLMs) with visual models to create a novel human-AI interaction method. 🤝🤖 It integrates various technologies like Whisper, Linly, Microsoft Speech Services, and SadTalker talking head generation system. 🌟🔬
Instantly Replace Faces and Backgrounds with AI
Premium AI-Powered Audiobook Generator with 47 voices, PDF processing, and voice cloning
🔍 An LLM-based Multi-agent Framework of Web Search Engine (like Perplexity.ai Pro and SearchGPT)
real time face swap and one-click video face swap with only a single image. You can use one face or ten faces to replace in realtime using insightface, mouth mask, face tracking
Contribute to alphaXiv/openresearch-cli development by creating an account on GitHub.
Free Local-first Full-Stack AI App Builder & Automation — Build, Test & Deploy with LLMs - Antigravity, Lovable, Bolt opensource Alternative ✨ 🌟 Star if you like it!
NocoBase is an open-source AI + no-code platform for building business systems fast. Instead of generating everything from scratch, AI works on top of production-proven infrastructure and a WYSIWYG no-code interface, so you get both speed and reliability.
[NVIDIA ONLY] Autocomplete any voice(s), powered by Hertz AI (Standard Intelligence)
