1,288 skills found · Page 1 of 43
AUTOMATIC1111 / Stable Diffusion WebuiStable Diffusion web UI
gradio-app / GradioBuild and share delightful machine learning apps, all in Python. 🌟 Star to support our work!
Zeyi-Lin / HivisionIDPhotos⚡️HivisionIDPhotos: a lightweight and efficient AI ID photos tools. 一个轻量级的AI证件照制作算法。
DrewThomasson / Ebook2audiobookGenerate audiobooks from e-books, voice cloning & 1158+ languages!
camenduru / Stable Diffusion Webui Colabstable diffusion webui colab
abus-aikorea / Voice ProGradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTube download, Demucs vocal isolation, and multilingual translation.
AbdBarho / Stable Diffusion Webui DockerEasy Docker setup for Stable Diffusion with user-friendly UI
modelscope / FunClipFunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
GiovanniPasq / Agentic Rag For DummiesA modular Agentic RAG built with LangGraph — learn Retrieval-Augmented Generation Agents in minutes.
ant-research / MagicQuill[CVPR'25] Official Implementations for Paper - MagicQuill: An Intelligent Interactive Image Editing System
OpenGVLab / Ask Anything[CVPR2024 Highlight][VideoChatGPT] ChatGPT with video understanding! And many more supported LMs such as miniGPT4, StableLM, and MOSS.
rsxdalv / TTS WebUIA single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Audio, MMS, StyleTTS2, MAGNet, AudioGen, MusicGen, Tortoise, RVC, Vocos, Demucs, SeamlessM4T, and Bark!
OpenGVLab / InternGPTInternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, SAM, interactive image editing, etc. Try it at igpt.opengvlab.com (支持DragGAN、ChatGPT、ImageBind、SAM的在线Demo系统)
jhj0517 / Whisper WebUIA Web UI for easy subtitle using whisper model.
om-ai-lab / OmAgent[EMNLP-2024] Build multimodal language agents for fast prototype and production
rupeshs / FastsdcpuFast stable diffusion on CPU and AI PC
Mukosame / Anime2SketchA sketch extractor for anime/illustration.
camenduru / Text Generation Webui ColabA colab gradio web UI for running Large Language Models
liltom-eth / Llama2 WebuiRun any Llama 2 locally with gradio UI on GPU or CPU from anywhere (Linux/Windows/Mac). Use `llama2-wrapper` as your local llama2 backend for Generative Agents/Apps.
flowtyone / Flowty Realtime Lcm CanvasA realtime sketch to image demo using LCM and the gradio library.