Issue 2026-09-12 · Industry · 模型 · 开源 · 研究

Agnes-3.0-Flash multimodal model

Agnes-3.0-Flash, posted on r/LocalLLaMA, is a 33B multimodal model with a 262,144-token context window. It uses a hybrid-attention decoder (72 layers: 54 delta-rule recurrent + 18 global-attention), includes vision/video understanding, and targets long-context and tool-calling use cases.

r/LocalLLaMA5 d ago
Read original ↗