Qwen3.6-27B

BaseRT .base builds of Qwen/Qwen3.6-27B for fast local inference on Apple Silicon (Metal).

A dense 27B reasoning model with hybrid attention (Gated DeltaNet + periodic full attention). Q4 needs ~24 GB Apple Silicon (tight); Q8 needs 48 GB+.

Files

File Precision Size
Qwen3.6-27B-Q4.base 4-bit 15 GB
Qwen3.6-27B-Q8.base 8-bit 27 GB

Usage

curl -LsSf https://basecompute.co/install.sh | sh
basert pull basecompute/Qwen3.6-27B
basert chat basecompute/Qwen3.6-27B

Released under the apache-2.0 license, inherited from the base model.

Downloads last month
183
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for basecompute/Qwen3.6-27B

Base model

Qwen/Qwen3.6-27B
Finetuned
(303)
this model