DeepSeek
8 storiesDeepSeek is an organization that develops and provides large AI models and APIs aimed at low‑cost inference and local deployment. Recent coverage notes its launch of an ultra‑low‑cost model charging as little as a fraction of a cent per million tokens, an internal beta for deepseek‑v4.1‑flash (native multimodal, faster, same pricing as v4‑flash, 20 concurrent requests per account), the soft retirement of V4 Pro, and discussion on whether tech reports like DeepSeek’s carry weight in PhD applications.
Related topics
DeepSeek V4.1 Flash achieved code execution on all 11 vulnerable targets while 4 fixed targets remained secure; accepted runs cost $4.65 (complete cost $5.14). The benchmark used 2,349 Bash commands, ~2h38m active model time, 268.3M input tokens (266.2M cached) and ~2M output tokens; audit found six planned exploits and five unexpected routes.
Foreign Policy reports Anthropic accused Moonshot AI, DeepSeek and others of distilling Claude at scale, including one case routing at least 300,000 user requests to Claude over 10 days. The data reportedly included surveillance material, internal code, credentials and other sensitive information, raising security concerns.
After DeepSeek v4.1, an engineer who wrote its main Attention operator reflects that small models are rapidly improving: AI can already read low-level code and optimize operators, and within six months to a year AI-written operators may match or exceed human work, due to AI's speed and parallelism advantages.
The BBC outlines how AI and generative models work, their uses in recommendations, medicine, and assistants, and the concerns around energy use, ethics, and regulation. The piece calls for balancing development with safeguards.
A Reddit post asks whether tech reports for large models (not arXiv papers) — e.g. Kimi K3, DeepSeek, Gemini, Mistral Leanstral — carry similar weight to a first‑author A* paper in PhD applications, seeking community input.
Bloomberg reports DeepSeek released an AI model charging as little as a fraction of a cent per million tokens, increasing pressure on rivals like Anthropic, OpenAI, and Z.ai.
Deepseek announced the soft retirement of V4 Pro, signaling the product line is being phased out and will no longer receive active promotion or updates.
DeepSeek opened internal beta testing for deepseek-v4.1-flash (expires-on-0910), a new-architecture intermediate V4.1 Flash with native multimodal support, faster speed and lower cost. Pricing matches deepseek-v4-flash and the rate limit is 20 concurrent requests per account. The announcement was posted by Chubby on X.
That is everything