Rewritten 2026-08-30: router/mHC/KDA aux kept unquantized. 4bit=QUASAR-init (KL -8.5% vs RTN). ~29-33 tok/s on M3 Ultra. Receipts in each repo.
ALISVOLATPROPRIIS PRO
avlp12
AI & ML interests
MLX quantization, Apple Silicon inference, MoE models, local LLM deployment, medical AI, geopolitical early warning systems, humanoid robotics, semiconductor/memory markets, multi-agent systems
Recent Activity
updated a model about 19 hours ago
avlp12/GLM-5.3-Flash-Alis-MLX-4bit updated a model 3 days ago
avlp12/GLM-5.3-Flash-Alis-MLX-8bit updated a model 3 days ago
avlp12/GLM-5.3-Flash-Alis-MLX-6bitOrganizations
None yet