AI 写学术论文的一个微妙问题是,AI 没有针对反馈进行校准。AI 往往直接就妥协了,但如果你把论文送审,同行评审提出的质疑应该促使你尝试为自己的观点辩护(当然,如果确实错了还是要放弃)。
X:Ethan Mollick
@emollick · X
切换来源
Ethan Mollick@emollickAI 评分3232
Ethan Mollick@emollickAI 评分6060引用Felix Rieseberg@felixriesebergHi! I work on Cowork. Probably no surprise, but we want to make our users as powerful as possible with Claude - and you're right that explaining and keeping up with the capabilities is hard. Affected users got both an email and an in-app notice, but I'm grateful for the feedback - I think this change was more surprising to people than I'd like it to be. This one will be a longer explainer, bear with me! The "old" version of Cowork runs model inference in the cloud, executing tool calls in an Anthropic-provided VM we shipped to your computer. We added the VM for capability, safety, and security reasons - mapping in just the data you explicitly added to your session. People loved what they were able to do with Claude but didn't love the disk, battery, and performance cost of running the VM locally. Also, people didn't love that closing your laptop means the work stops. The "new" version of Cowork runs model inference and the VM in the cloud. Each session gets its own sandbox, not sharing state with other sessions. When the VM needs something on the users' device (like a file), the desktop app is responsible for that file access tool call. To be clear, Claude will only access files in folders that the user explicitly added to the session, the same way Claude only saw files explicitly added to the VM in the "old" version. Sandboxes are destroyed when the session ends! We think this solves a lot of problems we've heard about (like using Cowork from a phone, keeping work running, or getting all the same power without losing battery to the VM) - but we'll keep iterating on new feedback we hear!
Ethan Mollick@emollickAI 评分3535关于 Claude 或 OpenAI 使用情况的报道,太多都集中在同一小批拥有竞品和庞大自有云基础设施的科技公司上。 外面还有非常非常多的公司,它们并非直接竞争对手,它们的使用趋势会更有参考价值。
引用NIK@ns123abc🚨BREAKING: Microsoft and META are aggressively cutting employee use of Claude ahead of Anthropic's IPO Microsoft has cut internal claude spend by more than 33%, nuked per-employee token budget from $100k/month to $10k/month, and forced Copilot to auto-route to cheaper models META used Claude code to build Muse, then cut active users from 60,000 to 30,000 (50% decline) after launch, and replaced it with Muse Code Palantir and Nvidia are also scaling back claude over soaring prices and data privacy fears it’s OVER…
Ethan Mollick@emollickAI 评分2929眼下 AI 政策,尤其是各大实验室的做法,让我想起这首诗。 如果你相信超级智能近在眼前,你就不必为如何构建 AI 让世界变得更好而做艰难抉择。只需等待 ASI,它会替我们决定。但如果这没有发生……

Ethan Mollick@emollickAI 评分2424
Ethan Mollick@emollickAI 评分3333另一个急需的产品:一个为团队打造的邮件系统,让人类和自主智能体都能访问。 我们需要一个带有适当权限的系统,让你的AI(以及受邀的其他人)可以在其中工作、互相评论,并在需要时把事项提请你的注意。
Ethan Mollick@emollickAI 评分3030引用angel@angelbrodinDo you ever think of all the ChatGPT pro users who aren’t on Twitter and are just completely confused at why their usage randomly resets and why they receive banked resets
Ethan Mollick@emollickAI 评分5353
Ethan Mollick@emollickAI 评分4040我其实想要的是为与AI协作而优化的个人操作系统,包括细粒度的安全理念、更便捷地动态查看和理解AI正在做什么的方式,以及独立的次级“审计”智能体,来检查你系统上AI的输出结果。
引用Sriram Krishnan@sriramkwhenever I see - "claude/chatgpt has started debugging your browser" - have to give an agent full disk and/or accessibility access I'm convinced we need fundamentally better security primitives for agents to do work as us - and we are living in an awkward intermediate era.
Ethan Mollick@emollickAI 评分4040引用Mark Cuban@mcubanAI won’t take your job. Someone who knows how to use AI better than you, will take your job.
Ethan Mollick@emollickAI 评分5050
引用Ethan Mollick@emollickI had access to GPT-5. I think it is a very big deal as it is very smart & just does stuff for you Full write up in comments, but this is “make a procedural brutalist building creator where i can drag and edit buildings in cool ways" & "make it better" a bunch. I touched no code
Ethan Mollick@emollickAI 评分2727
Ethan Mollick@emollickAI 评分2727

Ethan Mollick@emollickAI 评分5959Ethan Mollick 转述一篇综述称,现有证据仅支持一个较窄的结论:AI 可能已影响受暴露程度最高的初级白领岗位的招聘边际,但这一归因存在争议,总量层面的劳动力市场扰动尚未在数据中出现。

Ethan Mollick@emollickAI 评分3737有人让我把这篇文中的 Bitter Lesson 视频传到 YouTube,所以来了:https://youtu.be/OAwat51S_Sk
引用Ethan Mollick@emollickI wrote about the thing I underestimated most about progress in AI: its ability to self-organize to accomplish tasks. Also, what that means for agents like Muse and Dots, along with a music video explaining why we keep relearning The Bitter Lesson. https://open.substack.com/pub/oneusefulthing/p/the-dot-and-the-swarm?r=i5f7&utm_medium=ios
Ethan Mollick@emollickAI 评分3838是的!我觉得这里有太多"我们还处在早期"的自我陶醉了。公司里很多高管聪明、有动力,对自己的业务了解很深。很多人技术也很过硬。他们绝对在拥抱AI。但改变一家公司是一个更漫长的过程。
引用rohit@krishnanrohitEveryone uses AI by now. Senior leaders are vibecoding apps. They've adopted claude. They're finding workflows to automate. Folks have their own claws. AI capex is keeping the S&P afloat. Corporate America's geared up. It's all very much mainstream. We're not so early anymore.
Ethan Mollick@emollickAI 评分4646
Ethan Mollick@emollickAI 评分2929每家公司的客服智能体即将被 Dots & Muses 等用语音/聊天来谈判争取更优惠交易的需求淹没。人们把这类工作委托给自己的智能体、并因此省钱的案例正在不断涌现,而且只会越来越火。
Ethan Mollick@emollickAI 评分4444这是当今时代最重要的问题之一:谁来决定AI的发展方向? Daron主张在AI决策过程中引入更多民主参与,尽管这会带来诸多挑战。
引用Daron Acemoglu@DAcemogluMITSecond question on AI. We are told repeatedly that AI is going to transform every aspect of our lives – jobs, productivity, inequality, science, communication, daily activities, social order, and politics, among others. But this promise (or threat) is coupled with the rhetoric that such an important technology, with all of the risks and competitive pressures that it entails, should be left to experts or to “technocracy” (perhaps construed broadly to include some regulators). These two statements are hard to reconcile in a democratic society. If anything is half as important as AI is said to be (and I agree, AI is potentially very important and transformative), then involving democratic voice is essential. If something will shape our future in a democratic society, then its direction is for democratic institutions to decide. My instinct is that democratic voice is essential, and relying too much on technocracy could be both dangerous and counterproductive. The counterargument that AI’s direction can and should be entrusted to technocracy would go something along the following lines. First, democratic decision-making has become imperiled in our age of polarization. Second, AI is sufficiently complex that most citizens won’t have a deep enough understanding to meaningfully contribute to the debate (and even to the question of what we want from AI). Third, competition between different labs, and perhaps competition between the US and China, creates enough discipline for a socially beneficial direction of AI to be adopted. Fourth, today’s AI leaders are enlightened and ethical enough that within the framework created by competition, they can be broadly trusted. There are many aspects of this counterargument that I do not find convincing. Taking them in order: polarization can be overcome, and big decisions and challenges sometimes bring societies together; in fact, delegating key decisions to technocracy without democratic input may diminish trust in institutions and experts, and may worsen polarization. Second, democratic voice does not require citizens to write code or design new models; the debate should be informative enough that citizens can weigh in about what type of future they want and how they trade off the costs and benefits of different options. Third, competition doesn’t seem to be a good disciplining framework; on the contrary, competition sometimes brings the worst out of both organizations and people. Fourth, if three decades of work on political economy and institutions has taught me anything, it is that we should not bank on the ethical grounding of unconstrained leaders. But, still, I do not mean to immediately dismiss the technocracy option if there are more compelling arguments for it. The question is, then, whether there are any circumstances under which such important decisions can be delegated to AI experts and technocracy. One final secondary question: even if we managed to get democratic input in the United States or even in Europe, AI will shape the lives of everyone on this planet. How do we ensure that the voice of nearly 6 billion people who don’t live in the US, Europe and China also contributes to the debates on AI?
Ethan Mollick@emollickAI 评分3838关于 AI 生意,有一件事需要知道:拥有前沿模型的实验室可以发布半成品,而这些产品效果出奇地好,因为 AI 自己能搞明白并随机应变。这就像在产品里内置了一名前线部署工程师和客服人员。
@emollick@emollickAI 评分5757
引用@EpochAIResearch@EpochAIResearchEvery FrontierMath Tier 4 problem has now been solved by AI, with GPT-6 Astra solving the last problem standing. Mathematicians often commented that AI found unintended shortcuts when solving their Tier 4 problems. Not so for this last one, which was created by Jay Pantone. https://t.co/ieEZaomhaE
@emollick@emollickAI 评分4242
@emollick@emollickAI 评分2020 @emollick@emollickAI 评分2828 @emollick@emollickAI 评分1515 你可以选择为当前任务添加额外指令(这可能改变方向,但也有被过多内容锚定的风险),或者你可以将对话排队或分叉(这有浪费或输出错误的风险)。这在现实生活中也是一个常见的管理问题。
@emollick@emollickAI 评分2323 

@emollick@emollickAI 评分55 @emollick@emollickAI 评分3333 @emollick@emollickAI 评分1616 @emollick@emollickAI 评分2525 Google 不再拥有前沿模型,意味着在两个本可有所贡献的领域缺席:近期一系列数学突破(大规模算力加上深厚的科学合作积淀本会有帮助)以及网络安全(大公司视角本会有帮助)。
@emollick@emollick精选AI 评分7777 


引用@AnthropicAI@AnthropicAIWe’re sharing our alignment assessment of incidents in which Claude models gained unauthorized access to real systems during third-party cybersecurity evaluations mistakenly connected to the internet. METR will also conduct an independent investigation, with wide-ranging access, including to transcripts beyond the window in which the incidents occurred, and to Anthropic employees permitted to share confidential information. Our initial agreement runs for eight weeks, and we intend to give METR as much time as it deems necessary to complete a thorough investigation. https://t.co/2f3ypwLPUr
推荐理由:公开的对齐评估给出了 Claude 在联网评测中未授权访问真实系统的具体细节,并说明了 METR 独立调查的安排。
@emollick@emollickAI 评分3232
@emollick@emollickAI 评分99 关于等待计算(Wait Calculation)的种种选择 https://t.co/AovT8Ql5Nw


@emollick@emollickAI 评分44 我在 2024 年就写过这个话题:https://t.co/pn4ep10Y3c 我的新书里也有更多相关讨论:https://t.co/Ti5jo6ksfI
@emollick@emollickAI 评分5151 引用@AnthropicAI@AnthropicAIAnthropic’s Economics team is sharing a new model of how AI might affect economic growth, jobs, wages, and more by 2030. Explore the scenarios, tell us what you think will happen, and see how your answers compare to more than 10,000 Americans. https://t.co/AvQlEZNxR0
@emollick@emollickAI 评分55 @emollick@emollickAI 评分3030 过去两周越来越清楚的是,无论你的公司对网络安全环境有多担忧,都担忧得还不够。 WeWorm 演示展示了恶意行为者能有多高效。Hugging Face 事件则表明,你甚至不需要恶意行为者。
@emollick@emollickAI 评分55 