We’ve had an AI breakthrough at Figure and will be showcasing this tomorrow
What leading AI figures posted on X
Nathan Lambert praised the public MiMo-V2.6 large-scale RL resource as very impressive.
Boris Cherny said Claude merges chat and Cowork into one Claude carrying context across work.
Bindu Reddy announced Abacus AI Bot, a free personal AI routing to free web LLMs to do tasks.
Alexandr Wang said Muse saved users $9,649.71 across 100 real stories.
Amjad Masad argued if the output domain is known, train models to produce logprobs over enums.
We’ve had an AI breakthrough at Figure and will be showcasing this tomorrow
excited to help Shopify merchants advertise their products in ChatGPT: [Quoting @harleyf]: ChatGPT Ads for @Shopify is live. We are @OpenAI's first commerce partner. OpenAI pulls straight from Shopify Catalog, so merchants' products are already there. Mer…
wall-to-wall deployment of astra for engineers at databricks: [Quoting @pwendell]: Today we rolled out Astra to every engineer at Databricks (N=~3500). Some notes that may be helpful to others: 1. Astra unambiguously out performs our previous highest-end mode…
We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. Mo…
One of the coolest at-scale RL resources made public yet! You love to see it. [Quoting @_LuoFuli]: Nearly half a year of silence. We spent it studying one problem: how far RL can scale. MiMo-V2.6 is in the middle of its RL run right now. Three things we scal…
Use Grok Voice in fal to build intelligent, low-latency agents that resolve real customer issues [Quoting @fal]: Grok Voice is live on fal. The latest speech model from @SpaceXAI that answers in 0.70 seconds and finishes its tool calls before the sentence end…
We're seeing extraordinary results from @typesafeai. Default mode in 𝚏𝚡 is auto, with a safety reviewer analyzing every command. That reviewer runs on GPT Luna today. Jev is up to 18x faster (p95) *and* more accurate. It's coming to @vercel AI Gateway and l…
We rolled out GPT-6 Astra to every Databricks engineer today. It beats Claude Opus 5 on the hardest, long-horizon tasks and increased our coding spend by 60%. I think it’s the strongest model. Expensive. But hopefully worth it. [Quoting @pwendell]: Today we r…
Grok Build now gets better the more you use it. It remembers conventions, decisions, and project facts across sessions. /memory to browse /dream to organize recent notes into topics https://t.co/A2MMZT5eJX https://t.co/UD8lJGnfOB
Super interesting read! [Quoting @JacquelineSYC19]: A few weeks ago I wrote that our whole team at Artie uses Hermes (an AI agent harness from @NousResearch) to do their work. Here's how we got there: what we tried first, what it cost us, how we built a Herm…
"muse is the best app - AI or otherwise - I've used since ChatGPT in terms of bringing back that original 'wow' factor." you're making me emotional 🥹🥹🥹 [Quoting @0xShayan]: Muse is the best app - AI or otherwise - I've used since ChatGPT in terms of bring…
You can just build (physical) things. [Quoting @natalieyeo]: "Hello World" from this cute physical Codex pet! To all the @OpenAIDevs Day 5 of building with Astra https://t.co/aTeijhJQT1
Gave my mom my (very simple) agent setup. She named him Morgan. He just closed a $600 ad deal for her site. People are massively overthinking their setups. Bitter Lesson applies to agents too: just ask the model, let it spawn what it needs. https://t.co/ldJzE…
Codex voice for road trips: [Quoting @jonathanroomer]: I’ve got Codex (voice) in CarPlay. And it’s fantastic. I can now build things during road trips, while all 5 kids make a noise and my wife asks why I talk to the AI more than I talk to her. It runs throug…
Introducing the Hermes Agent plugins catalog! You can access it in your Hermes Desktop app's capabilities section to easily search, browse, and explore community and official plugins! Submit your own as well! [Quoting @NousResearch]: Hermes Agent now has a Pl…
Any guess what this Union Alpha beast on OpenRouter is? Claude Fable 5.1-level, over 10x cheaper. Waiting for it to be open-sourced. https://t.co/okoQnHeMw6
Also today: Claude Docs, Claude Slides, and Claude Design are in every conversation. Ask Claude for a presentation and you get one you can open, edit, and export as PowerPoint or PDF. Same for a document or a design. There's no separate tool to navigate to. T…
Claude Code showed that AI could do real work, not just answer questions. Developers hand Claude a feature, come back to shipped code. That's where much of the industry's serious engineering runs now. Cowork proved knowledge workers could do the same: hand Cl…
Three key open-weight labs have seen monthly dollars spent on their models grow by 10x or more so far in 2026. The closed models have seen comparatively slower growth in spending, though from much higher starting points. https://t.co/iOKARyF2as
When I talk to senior managers, increasingly hearing stories of the blurring of jobs inside organizations (we found this at our P&G study as well): coding, design, product management, all collapsing & overlapping as everyone uses AI We need new models…
muse voice transcribe is really good!! [Quoting @IsaacKing314]: I regret to inform you all that Meta AI is finally good, at least in the fields where their competitors have stopped trying. Their new voice transcription model is nearly an order of magnitude fa…
Haha... Literally no one is pacing the frontier - Opus 5.2 in testing - Grok 4.8 ships in a couple of weeks - Jev is a new ultra fast classifier - OpenAI already has Astra+ in testing We continue to accelerate
saying weird stuff about AI and the singularity The Anthropic Institute 🤝 the DeepMind Institute [Quoting @ShaneLegg]: My journey to develop AGI spans 25 yrs, including 10+ yrs thinking about technical & societal perspectives at Google DeepMind. AGI is…
Thrilled to complete the definitive and join forces with AlephAlpha! 🇨🇦🇩🇪 [Quoting @cohere]: Cohere and Aleph Alpha announce the signing of a definitive agreement, becoming the first foundational AI model developer anchored on both sides of the Atlantic …
🚨 Announcing Abacus AI Bot - 100% FREE PERSONAL AI AI can manage your calendar, book tickets, or make reservations. These simple tasks should be TOTALLY FREE Today, we announce our new product - Abacus AI bot - Finds and runs on FREE LLMs on the web - Autom…
Today, we are announcing a partnership with @mozilla to bring privacy, control and choice to people using AI to browse online. 🦊🐈 https://t.co/Bsd6N4jNdV https://t.co/OczDk9utfL
Study from Google about AI use in science, complex impacts: acceleration (7 hours saved per week) along with shifts in the kind of work (more verification) and what research gets done (possibly safer topics). Also a good diagram of the jagged frontier https:/…
muse saved this person over $5000 on WiFi! [Quoting @raunaqbn]: @Muse just continues to blow my mind! Today I had it call Xfinity to haggle down my internet bill. it got through the phone tree to a human, hit the verification text it couldn't read, and patche…
muse saved people $9,649.71 across 100 different stories imagine how much money this’ll save people when a billion people have muse!! we are working hard to expand access and get it everywhere! [Quoting @SingularityRes]: Meta's Muse AI has saved real people $…
i swear muse code is actually good!! the same model behind muse app powers muse code! [Quoting @ryanmcadams]: Never thought I'd say this, but @alexandr_wang has been telling people how good Muse code is, and I wasn't buying it. Spent real time in it today and…