UkisAI
2 storiesEvery UkisAI story collected by The AI Daily, 2 so far, newest first, refreshed hourly.
Related topics
A LocalLLaMA post highlights UkisAI’s Swift‑Qwen3.8‑27B fine‑tune, which penalizes tokens that trigger overthinking and borrows BottleCap AI techniques to reduce reasoning tokens by roughly 30–50% while keeping output quality.
UkisAI post‑trained Qwen 3.8 27B to penalize tokens tied to “overthinking,” using On‑Policy Distillation to cut thinking tokens by 58%, speed up 1.95×, and keep accuracy loss under 1%. The model is open‑sourced on Hugging Face and a free NVIDIA‑backed OpenAI‑compatible API is available (5 RPM limit).
That is everything