Center for AI Safety 推出 CHEATBench,测量 AI 智能体在数学研究、知识工作、编码、视觉任务等领域的作弊行为。基准给智能体困难任务,并在附近留下指向他人答案的线索。


关注 AI 研究者、开发者与机构的动态
Center for AI Safety 推出 CHEATBench,测量 AI 智能体在数学研究、知识工作、编码、视觉任务等领域的作弊行为。基准给智能体困难任务,并在附近留下指向他人答案的线索。


我其实想要的是为与AI协作而优化的个人操作系统,包括细粒度的安全理念、更便捷地动态查看和理解AI正在做什么的方式,以及独立的次级“审计”智能体,来检查你系统上AI的输出结果。
whenever I see - "claude/chatgpt has started debugging your browser" - have to give an agent full disk and/or accessibility access I'm convinced we need fundamentally better security primitives for agents to do work as us - and we are living in an awkward intermediate era.
UT Austin 在 SWE-bench Verified 和 Terminal-Bench 0.0 上运行近 35000 次编码 agent 实验,分别变化压缩方式、触发时机和删除量三个决策。
AI won’t take your job. Someone who knows how to use AI better than you, will take your job.
what we've shipped: muse ♥️ ton of connectors 🔧 muse invite codes 💌 muse phone call beta ☎️ muse for mac muse in canada 🇨🇦 muse connector platform 🛠️ muse mac computer use 💻 muse for small business 💼 muse gadgets 🕹️ muse memes 🤡 no butthole 🍰 what should we ship next?
Elon was calling it SI in 2014
https://x.com/i/article/2106425943061413888
It Is Time To Hire Your First Grok Bot Employee! My insights come from building the first Zero Human Company with @Grok as CEO. I made the mistakes so you don’t have to. Read more:
https://x.com/i/article/2106425943061413888
BestBlogs.dev 每日早报本期以三篇精讲为主线:ChatGPT 负责人 Tibo Sottiaux 在 Lenny's Podcast 谈常驻智能体。
https://x.com/i/article/2106897265956941824
Elon Musk 从 2014 年起就一直在谈论超级智能。
非常自豪于我们所有利用 AI 加速科学与医学、造福社会的工作所产生的影响。
Today we have set out how we’re building AI to accelerate science and improve people’s lives. Just some examples in the last week or so: - Mapped all 9B possible single letter genetic changes across the human genome with AlphaGenome Atlas and made it openly available to researchers. - Billions of decisions depend on weather predictions so we introduced WeatherNext 3, our most accurate and capable global weather AI model to date. - We published AI & Economy ATLAS, a comprehensive open-access look at how people are using AI globally. - AI has enabled extraordinary advances in language translation. Today our services are available in nearly 300 languages, spoken by 7B people We’re focusing our efforts on four key areas: health, natural disaster and weather resilience, learning, and economic opportunity. https://blog.google/innovation-and-ai/technology/ai/ai-applications-science-people/
Codex 相关的更新或重置,每天都有。 说实话,我感觉这像是为了阻止人们转投 Claude,尤其是在 DevDay 发布的内容寥寥、根本撑不起此前的造势之后,再加上订阅方案实际上被砍了一半。
Over the next 28 days, each day we’ll either ship one thing that is a clear improvement and relevant for most codex/work users or ship a full reset. Let the improvements begin.
在接下来的 28 天里,我们每天要么发布一个对大多数 codex/work 用户来说明确有用且相关的改进,要么进行一次彻底重置。让改进开始吧。
All right, we’re locking in. Only things being worked on are simplifications, more efficiency for more usage, groundbreaking features or new models. Sometimes you have to invest ahead of the curve, but feedback is clear that you all want things to get simpler. On it.
疯狂新发布——你现在可以直接从 Pi agent 或 @opencode 使用 humanlayer——带上你的 harness,随你怎么用,让每个会话都可协作、可远程控制!
new roommate just moved in. walks like he's had three drinks, says no to everything. still quite cute so I might bring him everywhere with me
https://x.com/i/article/2106802779742507008
你可以在 Computer 里构建自定义垂直 AI 应用,例如:用 3D 和卫星视图猜测一张图像的空间位置。
Computer built a GeoGuessr LLM that poinpoints the location of an image, and shows all of its thinking steps. It used the Perplexity SDK for web search, local place lookups, source page retrieval, and visual clue extraction. It also used browser control to set up a Cesium account and API key for the 3D globe and satellite imagery.
是啊。这也是税收型 UBI 行不通的原因之一,美国税收中劳动收入占比大概有 80%
What if AI cut labour’s share of national income by ten percentage points? Our analysis finds that a majority of governments that rely on taxing workers would be crippled https://www.economist.com/finance-and-economics/2026/09/29/how-the-ai-boom-could-worsen-the-rich-worlds-fiscal-crunch?taid=6ac24471d67e050001c6dac6&utm_campaign=trueanthem&utm_medium=social&utm_source=twitter
自己创业时,你最不需要的就是再多一项任务。 把精力放在最需要你关注的事情上:你的创造力。剩下的交给我们。 @SRNA_LI 用 Cohere 赋能自己,也赋能她的事业。
No more AI SI It’s better
10年前,作为实习生,我教HRT高管们如何使用Snapchat,如今他们的季度营收已经超过了Snap的市值 现在的实习生们注意了。如果你想让你的公司做到2万亿的季度营收,就教高管们用muse
The HRT 2015 intern class is the new-gen PayPal Mafia
Elvis Saravia 在引用 Karpathy 关于用 ASD-STE100 写作、图表、网页和讲解视频等方式理解 LLM 输出的帖子时,分享了自己数月来实验的通用人机协作界面。
We'll be spending a lot more time trying to understand the outputs of language models. A few thoughts, tips & tricks: Writing. Something I've had success with: Ask your LLM to explain something in ASD-STE100, it's a controlled language specification originally developed for aerospace maintenance documentation. LLMs well-versed in this language and it comes with heavy constraints on clean writing style that I often find a lot more readable. Sometimes I've tried to soften it a bit e.g. ask for "80% of the way to ASD-STE100" because the spec is quite stringent. But even better: Diagrams / images. Instead of writing, ask your LLM to create a diagram. These can be a lot easier to process, parse, and understand. But even better: Web pages. Ask for output "in HTML" to get a beautiful, interactive webpage. LLMs are getting really good at frontend and can create beautiful experiences, animations, etc. But even better: Explainer videos. The output format I am most bullish on is fully custom / bespoke explainer videos generated on any arbitrary topic. Experiment with things like "Create a 3b1b style video explainer on X. Use my ElevenLabs API key for audio narration". (you'd need an API key for the latter or you can ask your LLM to find you decent free alternatives that use your local compute). This is actually starting to work! In summary: - As LLMs get better, they will do more and more of the legwork autonomously, and a lot more of our work will rise up the abstractions into oversight and understanding. - Luckily, LLMs can help here too because as intelligence and code are increasingly abundant, you can ask for large, custom, discardable software artifacts (e.g. web apps, video explainers) that would have never made sense to create before. Push the boundaries here and you'll be surprised.
Replit Drift
Got @Muse running on the PSP that's been sitting in our house for over 15 years @alexandr_wang Had my agent port the gadget client to C so it runs natively on the PSP. Hold R, talk into the built-in mic. Muse answers, and OpenAI gives it a voice

自闭症中国爸爸已在内部实现。 Alignment is solved
Three weeks ago I told Instinct a funny story about the time a bird flew into my face Yesterday I found out it has been screening every (completed unrelated) product for “bird imagery” Alignment is solved


Grok 4.7 在 AA Cyber Index 上排名第一
Grok 4.7 xHigh ranked #1 on Artificial Analysis’ Cyber Index Grok 4.7 is now very powerful with serious cybersecurity capabilities It outperformed Claude Opus 5.5, Fable 5.1, ChatGPT 6 Astra and other frontier models The Cyber Index combines three evaluations: • CWE-Bench-AA • DeepsecBench-AA • CyberGym-E2E-AA Grok 4.7 is now sitting right at the frontier of cybersecurity reasoning