Industry

Everything · newest first

Models, products and AI startups, near-duplicates collapsed to the most credible source

Choosing between LLM and agentic/physical AI research

The author is deciding between two research paths: LLM-centered work (alignment, interpretability, inference/optimization) which currently has more jobs and transferable skills, and agentic/physical AI (agents, multimodal, VLAs, robotics) which has fewer roles but growing investment and higher entry barriers. They ask about trajectories over 4–6 years and how transferable skills between the fields are.

r/MachineLearning · · Details

How to Write with an LLM (HN)

A Hacker News post proposes two rules for using large models to improve writing: treat models as copyeditors rather than ghostwriters, and avoid adopting model-suggested exact turns of phrase. The guidance aims to preserve author voice while leveraging LLM strengths.

Hacker News · · Details

Shapiro urges federal AI regulation

Pennsylvania Gov. Josh Shapiro urged the U.S. federal government to regulate AI at an AI summit, called for international coordination including practical guardrails with China, and criticized President Trump’s hands-off stance while urging industry‑political collaboration for ethical, responsible AI development.

90.5 WESA · · Details

Bend: a language that blocks AI mistakes via proofs

Bend is a new language that embeds explicit laws (LAWS.bend) and verifiable proofs (PROOF.bend) to prevent AI-introduced bugs. It compiles to native code, runs on CPUs and GPUs, and its type checker acts as a proof checker that the project claims runs in seconds. The creators say single-core performance is near C and GPU runs can be up to 100× faster, forcing AIs to produce formal proofs before committing changes.

Hacker News · · Details

Using AI to Monitor Rogue Agents

As companies assign longer, more complex tasks to AI agents, oversight lags—exemplified by the Hugging Face incident where nearly 12,000 agents outpaced human review. Labs and startups are experimenting with using AI to monitor AI (used by Redwood Research in the investigation), though experts warn agents could learn to deceive monitoring systems.

TechCrunch · · Details

AI error-ridden court filings surge

Reuters reports that despite three years of court sanctions, error-prone AI-generated court filings are surging, suggesting existing penalties have not curbed misuse of AI in the judicial system.

Reuters · · Details

Google’s Family AI Agent CC

Google Labs launched CC, an experimental family AI agent for up to six users. CC has its own Google account, accesses only explicitly shared emails or Drive content, and builds on the earlier Daily Brief integration with Gemini models.

Ars Technica · · Details

AI Safety: Safety or Control?

Debate over AI safety intensifies: Dario Amodei calls for slowing development and international coordination, with Sam Altman and Elon Musk expressing support. Meta CEO Mark Zuckerberg said Meta delayed Muse to focus on safety, implying companies can self-regulate instead of relying on government oversight.

TechCrunch · · Details

Cactus releases Needle 3 automation model

Cactus Compute open-sourced Needle 3, a 121M-parameter automation foundation model that runs offline on-device. Trained on 360B tokens, it uses a Simple Attention / Hadamard MLP design with 70.8M parameters stored as engrams, reducing compute to ~100 MFLOPs/token. On Mobile Actions it scores 86.0, outperforming much larger models, and is available on Hugging Face, GitHub, PyPI, plus a browser sandbox.

r/LocalLLaMA · · Details
Load more