Qwen2
Grouped-query attention with qkv bias. The architecture behind the Qwen2 and Qwen2.5 families.
Language
3 supported combinations
Apple Silicon
Model families on this architecture
A family names one architecture. Support is declared per model and per target, so the same architecture can be open on one backend and absent from another.
Every supported combination
| Model | Command | Checkpoint on disk | Veizik device memory | Context | Hardware |
|---|---|---|---|---|---|
| Qwen2.5-0.5B | 0.5b | 942 MiB | 327 MiB | 32,768 | Apple Silicon |
| Qwen2.5-1.5B | 1.5b | 2,944 MiB | 883 MiB | 32,768 | Apple Silicon |
| Qwen2.5-7B | 7b | 14,525 MiB | 3,923 MiB | 32,768 | Apple Silicon |
Veizik prepares supported weights in memory; no converted checkpoint is written to disk.