Ai-Model
-
Image-Text-to-Text β’ 25B β’ Updated β’ 125k β’ 638 -
openai/whisper-large-v3-turbo
Automatic Speech Recognition β’ 0.8B β’ Updated β’ 7.62M β’ β’ 3.23k -
SWivid/F5-TTS
Text-to-Speech β’ Updated β’ 740k β’ 1.19k -
D-Edit
π84Edit images by selecting objects and applying text prompts
-
FacePoke
π2.21kImport a portrait, click to move the head!
-
Expression Editor
π¨1.66kQuickly edit the expression of a face
-
F5-TTS
π£2.9kF5-TTS & E2-TTS: Zero-Shot Voice Cloning (Unofficial Demo)
-
FLUX.1 [dev]
π₯9.51kGenerate images from text prompts
-
Open NotebookLM
π1.09kPersonalised Podcasts For All - Available in 13 Languages
-
PMRF
πΌ313A gradio demo for Posterior-Mean Rectified Flow (PMRF)
-
stabilityai/stable-diffusion-3.5-large
Text-to-Image β’ 8B β’ Updated β’ 63.4k β’ β’ 3.69k -
genmo/mochi-1-preview
Text-to-Video β’ 10B β’ Updated β’ 3.19k β’ β’ 1.34k -
Freepik/flux.1-lite-8B-alpha
Text-to-Image β’ 8B β’ Updated β’ 322 β’ 427 -
rhymes-ai/Allegro
Text-to-Video β’ 3B β’ Updated β’ 85 β’ 263 -
CohereLabs/aya-expanse-8b
Text Generation β’ 8B β’ Updated β’ 11.5k β’ 440 -
deepseek-ai/Janus-1.3B
Any-to-Any β’ 2B β’ Updated β’ 2.35k β’ 598 -
Pangea
π50A Fully Open Multilingual Multimodal LLM for 39 Languages
-
Etched/oasis-500m
Updated β’ 346 β’ 499 -
microsoft/OmniParser
Image-Text-to-Text β’ Updated β’ 458 β’ 1.71k -
OuteAI/OuteTTS-0.1-350M
Text-to-Speech β’ 0.4B β’ Updated β’ 121 β’ 303 -
tencent/Tencent-Hunyuan-Large
Text Generation β’ Updated β’ 223 β’ 617 -
nvidia/Llama-3.1-Nemotron-70B-Instruct-HF
Text Generation β’ 71B β’ Updated β’ 7.9k β’ 2.07k -
tencent/HunyuanVideo
Text-to-Video β’ Updated β’ 785 β’ β’ 2.24k -
zai-org/CogVideoX-5b
Text-to-Video β’ 6B β’ Updated β’ 18.2k β’ β’ 682 -
LanguageBind/Open-Sora-Plan-v1.2.0
Updated β’ 12 β’ 47 -
microsoft/phi-4
Text Generation β’ 15B β’ Updated β’ 633k β’ β’ 2.29k -
TRELLIS
π’4.78kScalable and Versatile 3D Generation from images
-
Reverse Face Search β Find Anyone by Photo
π879Free reverse face search β find anyone by photo
-
Kolors Virtual Try-On
π10.2kGenerate virtual tryβon images of a person wearing a chosen garment
-
DeepSeek-R1 WebGPU
π§565Next-generation reasoning model that runs locally in-browser
-
AnyCoder
π3.3kGenerate code snippets with AI for web and app frameworks
-
tencent/Hunyuan3D-2
Image-to-3D β’ Updated β’ 124k β’ 1.8k -
openbmb/MiniCPM-o-2_6
Any-to-Any β’ 9B β’ Updated β’ 273k β’ 1.3k -
deepseek-ai/DeepSeek-R1-Distill-Llama-70B
Text Generation β’ 71B β’ Updated β’ 309k β’ β’ 797 -
Magic Face
π€ͺ333Transform Your Face Into Legendary Characters!
-
Llasa 3b Tts
π₯315Zero Shot voice cloning with llasa 3b (Unofficial Demo)
-
mistralai/Mistral-Small-24B-Instruct-2501
24B β’ Updated β’ 56.9k β’ 964 -
Pyramid Flow
β±670Generate videos from text prompts (optional image guidance)
-
microsoft/OmniParser-v2.0
Updated β’ 2.85k β’ 1.35k -
Zyphra/Zonos-v0.1-hybrid
Text-to-Speech β’ 2B β’ Updated β’ 1.43k β’ 1.11k -
agentica-org/DeepScaleR-1.5B-Preview
Text Generation β’ 2B β’ Updated β’ 8.09k β’ 584 -
stepfun-ai/Step-Audio-Chat
Audio-Text-to-Text β’ 132B β’ Updated β’ 76 β’ 461 -
hexgrad/Kokoro-82M
Text-to-Speech β’ Updated β’ 11.5M β’ β’ 6.67k -
black-forest-labs/FLUX.1-dev
Text-to-Image β’ 12B β’ Updated β’ 481k β’ β’ 14.1k -
NousResearch/DeepHermes-3-Llama-3-8B-Preview
Text Generation β’ 8B β’ Updated β’ 1.24k β’ β’ 370