


@emollick · X



AI 写学术论文的一个微妙问题是,AI 没有针对反馈进行校准。AI 往往直接就妥协了,但如果你把论文送审,同行评审提出的质疑应该促使你尝试为自己的观点辩护(当然,如果确实错了还是要放弃)。
Hi! I work on Cowork. Probably no surprise, but we want to make our users as powerful as possible with Claude - and you're right that explaining and keeping up with the capabilities is hard. Affected users got both an email and an in-app notice, but I'm grateful for the feedback - I think this change was more surprising to people than I'd like it to be. This one will be a longer explainer, bear with me! The "old" version of Cowork runs model inference in the cloud, executing tool calls in an Anthropic-provided VM we shipped to your computer. We added the VM for capability, safety, and security reasons - mapping in just the data you explicitly added to your session. People loved what they were able to do with Claude but didn't love the disk, battery, and performance cost of running the VM locally. Also, people didn't love that closing your laptop means the work stops. The "new" version of Cowork runs model inference and the VM in the cloud. Each session gets its own sandbox, not sharing state with other sessions. When the VM needs something on the users' device (like a file), the desktop app is responsible for that file access tool call. To be clear, Claude will only access files in folders that the user explicitly added to the session, the same way Claude only saw files explicitly added to the VM in the "old" version. Sandboxes are destroyed when the session ends! We think this solves a lot of problems we've heard about (like using Cowork from a phone, keeping work running, or getting all the same power without losing battery to the VM) - but we'll keep iterating on new feedback we hear!
关于 Claude 或 OpenAI 使用情况的报道,太多都集中在同一小批拥有竞品和庞大自有云基础设施的科技公司上。 外面还有非常非常多的公司,它们并非直接竞争对手,它们的使用趋势会更有参考价值。
🚨BREAKING: Microsoft and META are aggressively cutting employee use of Claude ahead of Anthropic's IPO Microsoft has cut internal claude spend by more than 33%, nuked per-employee token budget from $100k/month to $10k/month, and forced Copilot to auto-route to cheaper models META used Claude code to build Muse, then cut active users from 60,000 to 30,000 (50% decline) after launch, and replaced it with Muse Code Palantir and Nvidia are also scaling back claude over soaring prices and data privacy fears it’s OVER…
眼下 AI 政策,尤其是各大实验室的做法,让我想起这首诗。 如果你相信超级智能近在眼前,你就不必为如何构建 AI 让世界变得更好而做艰难抉择。只需等待 ASI,它会替我们决定。但如果这没有发生……
另一个急需的产品:一个为团队打造的邮件系统,让人类和自主智能体都能访问。 我们需要一个带有适当权限的系统,让你的AI(以及受邀的其他人)可以在其中工作、互相评论,并在需要时把事项提请你的注意。
Do you ever think of all the ChatGPT pro users who aren’t on Twitter and are just completely confused at why their usage randomly resets and why they receive banked resets
我其实想要的是为与AI协作而优化的个人操作系统,包括细粒度的安全理念、更便捷地动态查看和理解AI正在做什么的方式,以及独立的次级“审计”智能体,来检查你系统上AI的输出结果。
whenever I see - "claude/chatgpt has started debugging your browser" - have to give an agent full disk and/or accessibility access I'm convinced we need fundamentally better security primitives for agents to do work as us - and we are living in an awkward intermediate era.
AI won’t take your job. Someone who knows how to use AI better than you, will take your job.
I had access to GPT-5. I think it is a very big deal as it is very smart & just does stuff for you Full write up in comments, but this is “make a procedural brutalist building creator where i can drag and edit buildings in cool ways" & "make it better" a bunch. I touched no code


Ethan Mollick 转述一篇综述称,现有证据仅支持一个较窄的结论:AI 可能已影响受暴露程度最高的初级白领岗位的招聘边际,但这一归因存在争议,总量层面的劳动力市场扰动尚未在数据中出现。
有人让我把这篇文中的 Bitter Lesson 视频传到 YouTube,所以来了:https://youtu.be/OAwat51S_Sk
I wrote about the thing I underestimated most about progress in AI: its ability to self-organize to accomplish tasks. Also, what that means for agents like Muse and Dots, along with a music video explaining why we keep relearning The Bitter Lesson. https://open.substack.com/pub/oneusefulthing/p/the-dot-and-the-swarm?r=i5f7&utm_medium=ios
是的!我觉得这里有太多"我们还处在早期"的自我陶醉了。公司里很多高管聪明、有动力,对自己的业务了解很深。很多人技术也很过硬。他们绝对在拥抱AI。但改变一家公司是一个更漫长的过程。
Everyone uses AI by now. Senior leaders are vibecoding apps. They've adopted claude. They're finding workflows to automate. Folks have their own claws. AI capex is keeping the S&P afloat. Corporate America's geared up. It's all very much mainstream. We're not so early anymore.
每家公司的客服智能体即将被 Dots & Muses 等用语音/聊天来谈判争取更优惠交易的需求淹没。人们把这类工作委托给自己的智能体、并因此省钱的案例正在不断涌现,而且只会越来越火。
这是当今时代最重要的问题之一:谁来决定AI的发展方向? Daron主张在AI决策过程中引入更多民主参与,尽管这会带来诸多挑战。
Second question on AI. We are told repeatedly that AI is going to transform every aspect of our lives – jobs, productivity, inequality, science, communication, daily activities, social order, and politics, among others. But this promise (or threat) is coupled with the rhetoric that such an important technology, with all of the risks and competitive pressures that it entails, should be left to experts or to “technocracy” (perhaps construed broadly to include some regulators). These two statements are hard to reconcile in a democratic society. If anything is half as important as AI is said to be (and I agree, AI is potentially very important and transformative), then involving democratic voice is essential. If something will shape our future in a democratic society, then its direction is for democratic institutions to decide. My instinct is that democratic voice is essential, and relying too much on technocracy could be both dangerous and counterproductive. The counterargument that AI’s direction can and should be entrusted to technocracy would go something along the following lines. First, democratic decision-making has become imperiled in our age of polarization. Second, AI is sufficiently complex that most citizens won’t have a deep enough understanding to meaningfully contribute to the debate (and even to the question of what we want from AI). Third, competition between different labs, and perhaps competition between the US and China, creates enough discipline for a socially beneficial direction of AI to be adopted. Fourth, today’s AI leaders are enlightened and ethical enough that within the framework created by competition, they can be broadly trusted. There are many aspects of this counterargument that I do not find convincing. Taking them in order: polarization can be overcome, and big decisions and challenges sometimes bring societies together; in fact, delegating key decisions to technocracy without democratic input may diminish trust in institutions and experts, and may worsen polarization. Second, democratic voice does not require citizens to write code or design new models; the debate should be informative enough that citizens can weigh in about what type of future they want and how they trade off the costs and benefits of different options. Third, competition doesn’t seem to be a good disciplining framework; on the contrary, competition sometimes brings the worst out of both organizations and people. Fourth, if three decades of work on political economy and institutions has taught me anything, it is that we should not bank on the ethical grounding of unconstrained leaders. But, still, I do not mean to immediately dismiss the technocracy option if there are more compelling arguments for it. The question is, then, whether there are any circumstances under which such important decisions can be delegated to AI experts and technocracy. One final secondary question: even if we managed to get democratic input in the United States or even in Europe, AI will shape the lives of everyone on this planet. How do we ensure that the voice of nearly 6 billion people who don’t live in the US, Europe and China also contributes to the debates on AI?
关于 AI 生意,有一件事需要知道:拥有前沿模型的实验室可以发布半成品,而这些产品效果出奇地好,因为 AI 自己能搞明白并随机应变。这就像在产品里内置了一名前线部署工程师和客服人员。
METR 正迅速成为 AI 领域事实上的行业标准制定机构 它开始看起来像是金融领域的 FINRA 的 AI 版本:不是政府监管机构,而是审查公司并定义可接受实践的机构。不知道立法是否也会将其写入法典
剧情(大概算剧透)?https://t.co/aVwRw8P5OJ
值得注意的是,RSI 可能并非全部,原因有几点:AI 能力可能受限(受算力、架构等限制),AI 驱动的研究可能受限(受发现有趣问题的能力等限制),公司可能无法利用这些成果,等等。






无论你现在听到多少关于 AI 的消息,今天都是你听到 AI 消息最少的一天。https://t.co/UJp6h80yCb
Every FrontierMath Tier 4 problem has now been solved by AI, with GPT-6 Astra solving the last problem standing. Mathematicians often commented that AI found unintended shortcuts when solving their Tier 4 problems. Not so for this last one, which was created by Jay Pantone. https://t.co/ieEZaomhaE