SmolLM2-1.7B-Instruct
Compact language model capable of solving a wide range of tasks while being lightweight.
A 1.7B parameter instruction‑tuned variant of SmolLM2, fine‑tuned for conversational and instruction‑following tasks, optimized for efficient on‑device inference on Qualcomm Snapdragon platforms.
Not supported
This model is currently not supported on any Compute chipset.
To see performance metrics for this model on other chipsets, click the button below.
View for other chipsetsQuick Start
1
Install Windows CLI App
2
Run the Model
Paste into CLI and run the following code.
For application and server integration, see Docs.
Technical Details
Architecture:Llama-based
Response Rate:Rate of response generation after the first response token.
TTFT:Time To First Token is the time it takes to generate the first response token. This is expressed as a range because it varies based on the length of the prompt. The lower bound is for a short prompt (up to 128 tokens) and the upper bound is for a prompt using the full context length (8192 tokens).
Applicable Scenarios
- Dialogue
- Content Generation
- Text Processing
License
Model:APACHE-2.0
Terms of Use:Qualcomm® Generative AI usage and limitations
Tags
- llm
- generative-ai
Supported Compute Devices
- Snapdragon X2 Elite CRD
Supported Compute Chipsets
- Snapdragon® X2 Elite
Related Models
See all modelsLooking for more? See models created by industry leaders.
Discover Model Makers










