DeepSeek

8 stories

DeepSeek is an organization that develops and provides large AI models and APIs aimed at low‑cost inference and local deployment. Recent coverage notes its launch of an ultra‑low‑cost model charging as little as a fraction of a cent per million tokens, an internal beta for deepseek‑v4.1‑flash (native multimodal, faster, same pricing as v4‑flash, 20 concurrent requests per account), the soft retirement of V4 Pro, and discussion on whether tech reports like DeepSeek’s carry weight in PhD applications.

Related topics

DeepSeek v4.1 Flash excels at hacking

DeepSeek V4.1 Flash achieved code execution on all 11 vulnerable targets while 4 fixed targets remained secure; accepted runs cost $4.65 (complete cost $5.14). The benchmark used 2,349 Bash commands, ~2h38m active model time, 268.3M input tokens (266.2M cached) and ~2M output tokens; audit found six planned exploits and five unexpected routes.

Hacker News · · Details

Are Chinese firms handing US data to US firms?

Foreign Policy reports Anthropic accused Moonshot AI, DeepSeek and others of distilling Claude at scale, including one case routing at least 300,000 user requests to Claude over 10 days. The data reportedly included surveillance material, internal code, credentials and other sensitive information, raising security concerns.

Foreign Policy · · Details

DeepSeek engineer on AI replacing operators

After DeepSeek v4.1, an engineer who wrote its main Attention operator reflects that small models are rapidly improving: AI can already read low-level code and optimize operators, and within six months to a year AI-written operators may match or exceed human work, due to AI's speed and parallelism advantages.

r/LocalLLaMA · · Details

Do tech reports matter for PhD apps

A Reddit post asks whether tech reports for large models (not arXiv papers) — e.g. Kimi K3, DeepSeek, Gemini, Mistral Leanstral — carry similar weight to a first‑author A* paper in PhD applications, seeking community input.

r/MachineLearning · · Details

DeepSeek Flash 4.1 internal beta via API

DeepSeek opened internal beta testing for deepseek-v4.1-flash (expires-on-0910), a new-architecture intermediate V4.1 Flash with native multimodal support, faster speed and lower cost. Pricing matches deepseek-v4-flash and the rate limit is 20 concurrent requests per account. The announcement was posted by Chubby on X.

r/LocalLLaMA · · Details
That is everything