Qualcomm® AI HubAI Hub

SmolVLM2-2.2B-Instruct

Multimodal 2.2B vision‑language model capable of understanding text and images.

SmolVLM2 is a lightweight vision‑language model from Hugging Face capable of understanding text and images for tasks such as visual question answering, image captioning, and optical character recognition.

Not supported

This model is currently not supported on any All Models chipset.

To see performance metrics for this model on other chipsets, click the button below.

View for other chipsets

Technical Details

Applicable Scenarios

  • Dialogue
  • Content Generation

Supported Form Factors

  • Compute
  • Phone
  • IoT

License

Tags

  • llm
  • vlm
  • generative-ai

Supported Devices

  • Arduino VENTUNO Q
  • Dragonwing IQ-9075 EVK
  • Dragonwing IQ-X5121
  • Dragonwing IQ-X7181
  • Dragonwing Q-8750
  • Samsung Galaxy S25
  • Samsung Galaxy S26
  • Snapdragon X Elite CRD
  • Snapdragon X Plus 8-Core CRD
  • Snapdragon X2 Elite CRD

Supported Chipsets

  • Qualcomm® QCS5121
  • Qualcomm® Dragonwing™ IQ-X7181
  • Qualcomm® Dragonwing™ IQ-8275
  • Qualcomm® Dragonwing™ Q-8750
  • Qualcomm® Dragonwing™ IQ-9075
  • Snapdragon® 8 Elite For Galaxy Mobile
  • Snapdragon® 8 Elite Gen 5 For Galaxy Mobile
  • Snapdragon® X Elite
  • Snapdragon® X Plus 8-Core
  • Snapdragon® X2 Elite

Looking for more? See models created by industry leaders.

Discover Model Makers