Qwen 2.5 1.5B
Better than Llama 1B on non-English (~1.6 GB VRAM). Low-resource.
- Runs on
- Your device, in the browser
- Size
- 1.5B
- Best for
- Coding
Qwen 2.5 1.5B runs in your browser, on your own device. The model downloads once and stays there; what you type is never sent to Alibaba or to us , and no API key is involved.
Context window
32,768 tokens
≈ 25,000 words · about 98 pages
the model's own limit — how much your device can hold depends on its memory
Max output
8,192 tokens
≈ 6,140 words in a single answer
Modalities
- TextInput and output
- ImageNot supported
- AudioNot supported
- VideoNot supported
Features
- StreamingSupported
- Function callingNot supported
- Structured outputsSupported
- Fine-tuningNot supported
Where it comes from
| Model card | Alibaba's card for Qwen 2.5 1.5B |
|---|---|
| Weights | The build your browser downloads |
| Runtime | WebLLM |
Ready to try it?
Polymeti is currently invite-only.
Log inRuns on your device · No API key · Your chats stay in your browser