LocalLLaMA
6 storiesLocalLLaMA is a community subreddit focused on running and sharing models locally. Recent coverage highlights a new model-download site called “The Hugging Bay,” users reporting models that work on 12GB VRAM (notably Qwen 3.6 35B 3A), the release of a music model YuE2-3B, and community posts advising against FOMO-driven spending and suggesting [Finetune] tagging.
Related topics
A LocalLLaMA Reddit post reports that HuggingFace has begun censoring certain models, citing audnai/penclaw-GLM-5.3-abliterated-for-offensive-cy as an example. The community says disabled messages should be more specific.
A Reddit post announces “The Hugging Bay,” a new website offering model downloads as a backup if Hugging Face begins censoring or restricting access. The post appeared on r/LocalLLaMA by /u/Thrumpwart.
A Reddit LocalLLaMA thread asks 12GB VRAM 3080 users what models work well locally; Qwen 3.6 35B 3A was cited as common, and the OP seeks models that leverage the extra 4GB for agent-style personal assistant workflows.
A r/LocalLLaMA user shared a new music model, YuE2-3B, with an online demo link and described it as solid. The post links to the demo but gives no further details on training or weights.
Advice to hobbyists: don’t let FOMO drive costly local LLM setups—learn via APIs or small models that fit existing hardware; you can gain more by doing less rather than buying expensive GPUs.
A community member suggests distinguishing major 'new model' releases from smaller finetunes by tagging or prefixing finetune posts (e.g., '[Finetune]') so 'new model' is reserved for substantial pretrained or major updates.
That is everything