SmolVLM2-2.2B-Instruct
Multimodal 2.2B vision‑language model capable of understanding text and images.
SmolVLM2 is a lightweight vision‑language model from Hugging Face capable of understanding text and images for tasks such as visual question answering, image captioning, and optical character recognition.
Not supported
This model is currently not supported on any IoT chipset.
To see performance metrics for this model on other chipsets, click the button below.
View for other chipsetsTechnical Details
Applicable Scenarios
- Dialogue
- Content Generation
License
Model:APACHE-2.0
Terms of Use:Qualcomm® Generative AI usage and limitations
Tags
- llm
- vlm
- generative-ai
Supported IoT Devices
- Arduino VENTUNO Q
- Dragonwing IQ-9075 EVK
- Dragonwing IQ-X5121
- Dragonwing IQ-X7181
- Dragonwing Q-8750
Supported IoT Chipsets
- Qualcomm® QCS5121
- Qualcomm® Dragonwing™ IQ-X7181
- Qualcomm® Dragonwing™ IQ-8275
- Qualcomm® Dragonwing™ Q-8750
- Qualcomm® Dragonwing™ IQ-9075
Looking for more? See models created by industry leaders.
Discover Model Makers








