Qualcomm® AI HubAI Hub

SmolLM2-1.7B-Instruct

Compact language model capable of solving a wide range of tasks while being lightweight.

A 1.7B parameter instruction‑tuned variant of SmolLM2, fine‑tuned for conversational and instruction‑following tasks, optimized for efficient on‑device inference on Qualcomm Snapdragon platforms.

Not supported

This model is currently not supported on any Mobile chipset.

To see performance metrics for this model on other chipsets, click the button below.

View for other chipsets

Quick Start

1

Install Windows CLI App

2

Run the Model

Paste into CLI and run the following code.

For application and server integration, see Docs.

Technical Details

Architecture:Llama-based
Response Rate:Rate of response generation after the first response token.
TTFT:Time To First Token is the time it takes to generate the first response token. This is expressed as a range because it varies based on the length of the prompt. The lower bound is for a short prompt (up to 128 tokens) and the upper bound is for a prompt using the full context length (8192 tokens).

Applicable Scenarios

  • Dialogue
  • Content Generation
  • Text Processing

Supported Mobile Form Factors

  • Phone
  • Tablet

License

Tags

  • llm
  • generative-ai

Supported Mobile Devices

  • Samsung Galaxy S25
  • Samsung Galaxy S26

Supported Mobile Chipsets

  • Snapdragon® 8 Elite For Galaxy Mobile
  • Snapdragon® 8 Elite Gen 5 For Galaxy Mobile

Related Models

See all models

Looking for more? See models created by industry leaders.

Discover Model Makers