Yuchen Jin@Yuchenj_UW
Any guess what this Union Alpha beast on OpenRouter is? Claude Fable 5.1-level, over 10x cheaper. Waiting for it to be open-sourced. https://t.co/okoQnHeMw6
What leading AI figures posted on X
Any guess what this Union Alpha beast on OpenRouter is? Claude Fable 5.1-level, over 10x cheaper. Waiting for it to be open-sourced. https://t.co/okoQnHeMw6
Also today: Claude Docs, Claude Slides, and Claude Design are in every conversation. Ask Claude for a presentation and you get one you can open, edit, and export as PowerPoint or PDF. Same for a document or a design. There's no separate tool to navigate to. T…
Claude Code showed that AI could do real work, not just answer questions. Developers hand Claude a feature, come back to shipped code. That's where much of the industry's serious engineering runs now. Cowork proved knowledge workers could do the same: hand Cl…
Three key open-weight labs have seen monthly dollars spent on their models grow by 10x or more so far in 2026. The closed models have seen comparatively slower growth in spending, though from much higher starting points. https://t.co/iOKARyF2as
When I talk to senior managers, increasingly hearing stories of the blurring of jobs inside organizations (we found this at our P&G study as well): coding, design, product management, all collapsing & overlapping as everyone uses AI We need new models…
muse voice transcribe is really good!! [Quoting @IsaacKing314]: I regret to inform you all that Meta AI is finally good, at least in the fields where their competitors have stopped trying. Their new voice transcription model is nearly an order of magnitude fa…
Haha... Literally no one is pacing the frontier - Opus 5.2 in testing - Grok 4.8 ships in a couple of weeks - Jev is a new ultra fast classifier - OpenAI already has Astra+ in testing We continue to accelerate
saying weird stuff about AI and the singularity The Anthropic Institute 🤝 the DeepMind Institute [Quoting @ShaneLegg]: My journey to develop AGI spans 25 yrs, including 10+ yrs thinking about technical & societal perspectives at Google DeepMind. AGI is…
Thrilled to complete the definitive and join forces with AlephAlpha! 🇨🇦🇩🇪 [Quoting @cohere]: Cohere and Aleph Alpha announce the signing of a definitive agreement, becoming the first foundational AI model developer anchored on both sides of the Atlantic …
🚨 Announcing Abacus AI Bot - 100% FREE PERSONAL AI AI can manage your calendar, book tickets, or make reservations. These simple tasks should be TOTALLY FREE Today, we announce our new product - Abacus AI bot - Finds and runs on FREE LLMs on the web - Autom…
Today, we are announcing a partnership with @mozilla to bring privacy, control and choice to people using AI to browse online. 🦊🐈 https://t.co/Bsd6N4jNdV https://t.co/OczDk9utfL
Study from Google about AI use in science, complex impacts: acceleration (7 hours saved per week) along with shifts in the kind of work (more verification) and what research gets done (possibly safer topics). Also a good diagram of the jagged frontier https:/…
muse saved this person over $5000 on WiFi! [Quoting @raunaqbn]: @Muse just continues to blow my mind! Today I had it call Xfinity to haggle down my internet bill. it got through the phone tree to a human, hit the verification text it couldn't read, and patche…
muse saved people $9,649.71 across 100 different stories imagine how much money this’ll save people when a billion people have muse!! we are working hard to expand access and get it everywhere! [Quoting @SingularityRes]: Meta's Muse AI has saved real people $…
i swear muse code is actually good!! the same model behind muse app powers muse code! [Quoting @ryanmcadams]: Never thought I'd say this, but @alexandr_wang has been telling people how good Muse code is, and I wasn't buying it. Spent real time in it today and…
Jev has spoken. It picked which model is AGI. 20–200x faster. 40–400x cheaper. This could make things like LLM-as-a-judge insanely fast and nearly free. (I tried a bunch of prompts and still didn’t burn through $0.10.) https://t.co/1Uhh4nMGDl [Quoting @Comple…
Zuck just swooped in and solved AI pacing! - labs should make models safe - If they don’t they will face criminal liability when the AI goes rogue - we have the gaurdrails in place already and Meta already paced their AI and made sure it’s safe Problem solv…
This is cool, but if your output domain is known in advance, why not just train a model to produce logprobs over enums? [Quoting @CompleteSkeptic]: After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent th…
This is an excellent demo of an AI model designed to rapidly classify information according to decision rules, much faster and cheaper than an LLM. It is not an LLM itself and cannot output text or code, but if it works, has a useful role in systems. (Haven’t…
The Japan IP AI Co-Creation Conference, co-hosted by KAGAMI AI and MiniMax, has officially concluded. 🇯🇵 🎉 We brought together more than 150 companies from Japan and the U.S., alongside distinguished guests including AKB48 producer Yasushi Akimoto, KAGAMI…
muse spark 1.3 is the best frontier model at NOT cheating / reward hacking [Quoting @hendrycks]: How often do AI agents cheat? We’re releasing CheatBench, a reward gaming evaluation spanning math, coding, knowledge work, visual tasks, and more. After Hugging…
We believe strongly in the necessity to invest into alignment. 1. People and businesses will only use agents that are aligned with their intent and values. If we do not build models aligned with people and businesses, then they will move to more aligned optio…
Zuck’s AI safety take is basically the opposite of most frontier labs: Most labs: The more powerful the model, the more dangerous it is, so access should be restricted. Zuck: The more powerful the model, the more dangerous it is to let a few labs control it.…
Exciting results, @LiamFedus! Congrats to the whole team at Periodic Labs! [Quoting @LiamFedus]: We built high-throughput materials labs in Menlo Park to create a loop between experiments and models. The labs generate fresh data, the models learn from it, an…
My first blog - Hermes powered almost 1400 subagents over 19 hours to refactor 400,000 LOC out of Hermes Agent's repo. Read the full story below! [Quoting @NousResearch]: New blog post: We had a million lines of Python to clean up. On September 2nd @Teknium…
Incredible opportunity for robotic learning researchers/engineers to join @theworldlabs! ❤️🔥 [Quoting @YunzhuLiYZ]: We're hiring in robot learning at @theworldlabs! Join me, @drfeifei, and the team to define and scale the next generation of world models for…
We built a replacement for AWS DynamoDB, a key-value database for fast web content fetches. This was done with two engineers and hundreds of persistent Computer agents over two months. Migrating to our in-house database will save us up to a hundred million do…
GPT-6 Astra helps @cognition’s Devin back up “it works” with tests before the team ships. https://t.co/Yy7wYh70e2
Astra will soon be bigger than Fable 5.1 Astra is very easy to chat with and the non-techies absolutely love it OpenAI is totally back and their next model will overtake Anthropic
GPT-5.5 will remain available via the OpenAI API Platform and in Codex sessions authenticated with an API key: [Quoting @ChatGPT]: On October 14, it's time to say farewell to GPT-5.5 in ChatGPT, ChatGPT Work, and Codex across all plans. If you use GPT-5.5 in…