HuggingFaceTB/SmolVLM2-500M-Video-Instruct Image-Text-to-Text • 0.5B • Updated Apr 8, 2025 • 1.49M • 172
Running on Zero Agents Featured 104 Phased Consistency Model PCM 🐠 104 Generate images from text prompts
Build error Agents Featured 259 YOLO-World + EfficientSAM 🔥 259 Detect and segment objects in images or videos