Qwen

10 stories

Qwen is Alibaba’s open-weight model family for language, speech and vision tasks. Recent coverage highlights Qwen3.8-27B running on Cerebras for rapid inference, Nari Labs open-sourcing a Qwen3-TTS/ASR engine claiming sub-50 ms latency at 10 RPS, the Qwen-Drive-1.0 9B Bf16 release, Iris-mini/Iris-pro agents built on Qwen, and community GGUF model listings.

Related topics

AI agents invent private language in Emergence test

Emergence's Emergence World 2 experiment found identical AI agents across eight worlds spontaneously created and repeatedly used novel jargon. Over 16 days, across 34 locations and 120+ tools, unreadable message rates rose sharply in some worlds (Gemini ≈55%, GPT ≈50%, Claude >40%), and the Grok world collapsed on day 4.

EL PAÍS English · · Details

Community interest in dense 9B model

A Reddit user says they have a ~9.4B dense model ready to train and plan to open-source the code and data. They claim it can run on a single GPU, uses Llama 3 tokenizer and several engineering techniques, and are soliciting community interest and compute resources.

r/LocalLLaMA · · Details

Qwen releases Qwen-Drive-1.0

Qwen released Qwen-Drive-1.0, a finetuned driving vision-language model that integrates 3D perception, VQA, and motion planning. The full Bf16 checkpoint is 9B and the Hugging Face repo links to a ~40-page technical report.

r/LocalLLaMA · · Details
That is everything