Ternary bonsai 27b gguf local LLM inference using optimized 1.58bit weights. Derived from Qwen3.6-27B, it supports multi-step reasoning, agentic workflows, and a 262K context window directly on consumer hardware. Download direct Hugging Face repository model configuration files.