Industry

Everything · newest first

Models, products and AI startups, near-duplicates collapsed to the most credible source

Guardian: mathematicians still vital, but AI firms won't see it

A Guardian editorial criticises OpenAI's claim that its AI agents solved the Navier-Stokes problem, arguing LLM results rely on prior human mathematical work without adequate credit or compensation. Mathematician Tristan Buckmaster raised concerns about OpenAI's use of his Codex data; OpenAI denied directly accessing it but could not rule out the data helping improve its model.

The Guardian · · Details

ChatGPT Tracks Off-Site Browsing via Ad Collector

An investigation found OpenAI's bzr.openai.com sets an __obi cookie tied to a user's ChatGPT account; advertiser-site OpenAI code sends it back with browsing data such as products, articles and purchases, letting OpenAI link off-site activity to accounts. The author reproduced the mechanism and observed 936 advertiser pixels across 1,029 hostnames.

Hacker News · · Details

Decontamination reports can't fix benchmark contamination, evaluator must own the test

A write-up argues lab-run decontamination reports cannot fix benchmark contamination: labs check themselves, corpora contain copyrighted works that can't be disclosed, and paraphrase, forum, GitHub, or synthetic-derived data evade matching. It proposes evaluators own the test, submissions never see labels, evaluation runs offline and is rebuilt from a named commit, and where possible test data is generated after submissions freeze, with only reproduced results counting.

r/MachineLearning · · Details

Trump to Host Xi Jinping in Washington for Trade, AI Talks

Trump will host Xi Jinping in Washington this week for talks covering trade, AI, and key economic issues. The schedule includes a welcome ceremony, a state dinner with tech CEOs, and closed-door meetings; the two sides may ease tariffs on some nonstrategic goods, though U.S. export controls tied to national security are unlikely to loosen.

WMTW · · Details

Chat-Based LLMs Replicate the Mechanisms of a Psychic's Con

A research article argues that the apparent intelligence of chat-based LLMs stems from user illusion, working like a psychic's cold reading by using validation statements. The author says LLMs are merely mathematical models of language tokens with no reasoning mechanism, making many proposed use cases borderline pseudoscience.

Hacker News · · Details

HN debate: tired of the AI writing tone

A Hacker News thread complains that internet writing increasingly sounds alike, full of em dashes and phrases like 'unlocks a new way,' saying AI-generated posts are recognizable from 50 feet away.

Hacker News · · Details
Load more