Compresses a 27B-class multimodal model into end-to-end ternary weights to run 27B reasoning on-device: 5.9–8.6 GB deployed footprint, 262K-token context, ~98.2% of FP16 benchmark performance; ships MLX and GGUF packs and runs on Apple MLX and CUDA.