3-bit VLM base + 4-bit MTP drafter. Pair both for accelerated instruct decode. reasoning_effort baked to low.
Jorge Leon
leonsarmiento
AI & ML interests
AI for Environment
Recent Activity
new activity about 7 hours ago
scottlowry/Qwen3.8-27B-oQ4e-mtp:dude, this is one is finally working for me. thanks! liked a model about 8 hours ago
scottlowry/Qwen3.8-27B-oQ4e-mtp updated a model about 8 hours ago
leonsarmiento/Qwen3.8-27B-3bit-mlx