HuggingFaceTB/SmolVLM-256M-Instruct Image-Text-to-Text ⢠0.3B ⢠Updated Apr 8, 2025 ⢠317k ⢠345
Running on Zero Featured 1.76k Dia 1.6B šÆ 1.76k Generate realistic dialogue from a script, using Dia!
HuggingFaceTB/SmolVLM2-500M-Video-Instruct Image-Text-to-Text ⢠Updated Apr 8, 2025 ⢠271k ⢠125
microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition ⢠6B ⢠Updated Dec 10, 2025 ⢠305k ⢠1.58k
Running Featured 354 Kokoro Text-to-Speech (WebGPU) š£ 354 High-quality speech synthesis powered by Kokoro TTS
mlx-community/SmolVLM2-500M-Video-Instruct-mlx Video-Text-to-Text ⢠Updated Feb 20, 2025 ⢠2.01k ⢠18
Running on Zero Featured 3.58k InstantID š» 3.58k Generate personalized images preserving your face identity
Runtime error Featured 33 CLIPnCROP š 33 Extract and crop image sections based on text description