Issue 2026-09-14 · Industry · 模型发布 · 开源 · 产品

UkisAI releases Swift‑Qwen3.8‑27B

UkisAI post‑trained Qwen 3.8 27B to penalize tokens tied to “overthinking,” using On‑Policy Distillation to cut thinking tokens by 58%, speed up 1.95×, and keep accuracy loss under 1%. The model is open‑sourced on Hugging Face and a free NVIDIA‑backed OpenAI‑compatible API is available (5 RPM limit).

r/LocalLLaMA3 d ago
Read original ↗