OpenHubAI
Back to browse
Text generationapache-2.0
Qwen/Qwen2.5-0.5B-Instruct
In plain English

Good for a fast, on-device assistant for short replies and simple instructions. Too small for long documents or complex reasoning, and it will make things up under pressure.

A 0.5 billion parameter instruction tuned model from Qwen's smallest tier, built for devices where every megabyte and millisecond counts. It trades reasoning depth for speed and a tiny footprint.

Required documentation

Complete
Intended use

Intended for short, low-latency assistant tasks on constrained hardware such as phones or edge devices. Not intended for long-context work or tasks requiring strong reasoning.

Training data provenance

Base weights and instruction tuning data are Qwen's own general web, code, and multilingual corpora, redistributed here with quantised GGUF builds added for local inference.

Licence & terms

Apache-2.0. Free for commercial and research use with attribution retained in redistributions.

MRMaya Rehanirep 4.8Prototype7ae5576
7h ago
Accuracy3.0
Fine-tune3.5
Documentation3.5
Speed4.8
TNTomasz Nowakrep 4.5Research7ae5576
7h ago
Accuracy3.2
Fine-tune3.0
Documentation3.5
Speed4.9
PSPriya Shahrep 4.6Prototype7ae5576
7h ago
Accuracy2.8
Fine-tune3.5
Documentation3.0
Speed5.0

Discussion0

Posting as a signed-in userSign in to post

No questions yet. Be the first to ask one.

Structured rating
3.6

from 3 reviews of 7ae5576

Only 3 reviews on 7ae5576 so far. Earlier versions carry more, but they are not this build.
Accuracy3.0
Fine-tune3.3
Documentation3.3
Speed4.9
Run it now
ollama run Qwen/Qwen2.5-0.5B-Instruct
Will it fit?
BuildFileVRAM
Q4_K_M397 MB0.9 GB
Q8_0531 MB1.1 GB

Set a hardware profile in the browse sidebar to see which builds fit.

Inference providers

No hosted providers listed yet.

No cloud required. Run it with
vLLMSGLang
Model tree
Base modelQwen/Qwen2.5-0.5B
Adapters0
Finetunes0
Quantizations0
Merges0
At a glance
Parameters
Licenceapache-2.0
Downloads7.0M
Reviews3
Docs completenessComplete