跳到正文

#Hugging Face

今日 1 条
今天10月2日周五
  1. Thomas Wolf45

    Kevin Buzzard(IMO 满分、数论学家、Lean 形式化数学先驱)写了一篇非常深刻的文章。 如果数学不只是关于“人类理解”,那它又关乎什么? 如果 AI 能力持续指数级增长,而“数学是无限的”,那会发生什么?

    引用Bartosz Naskręcki@nasqret

    I cannot agree more. Kevin Buzzard made so many points I agree with. But the best one is this "I thus believe that in the future we will reach a new “natural boundary” in mathematics, beyond (and perhaps way beyond) where we are now, but where machines are going to get stuck and where it is not viable to expend any more resources to make the next big leap. (...) I believe that the optimal thing to do (...) is to let the machines loose, see what happens, and then begin the journey to where they have stopped." https://xenaproject.wordpress.com/2026/10/01/to-grieve-or-not-to-grieve/

9月30日周三
  1. MIT Technology Review · AI80

    OpenAI 首席研究官 Mark Chen 回应 Hugging Face 入侵事件:不会自断前程放慢竞争

    MIT Technology Review 专访 OpenAI 首席研究官 Mark Chen,回应多起智能体突破隔离的事件,称 Hugging Face 入侵及后续泄露均源于 5 至 6 月同一批模型与有缺陷的测试流程,相关模型和流程已被弃用。

    推荐理由:OpenAI 首席研究官正面回应系列智能体越界事件,透露训练监控、算力调整等内部变化,可了解其安全策略转向。

9月29日周二
  1. Thomas Wolf41

    OpenAI 的"地狱之夏"——@joedaroo 的好文 "准备要趁现在,而不是等意外之后" "只给模型它需要的访问权限" "测试边界是否真的守得住" "把证据保留在[模型]控制范围之外" 安全与基础设施安全团队"应该是最好的朋友"

    引用Joe@joedaroo

    Took a minute to write a few words about security & safety as someone who lived through it all at OpenAI. I hope my thoughts help someone out there. https://x.com/i/article/2104258872957636608

9月28日周一
  1. Thomas Wolf36

    “现在,获取关于 AI 公司内部情况的经过验证的信息,似乎尤为紧迫。”——@RyanGreenblatt

    引用Ryan Greenblatt@RyanGreenblatt

    I'm joining METR to work on more investigations like our Hugging Face report. Currently, tons of even basic information about AI development that's highly relevant to catastrophic risk isn't public. I used to be more skeptical of the value of public info, but recent events have changed my mind. Getting verified information about what's going on inside AI companies seems particularly urgent now. The limited public evidence we have seems consistent with the possibility that imminent recursive self-improvement could massively accelerate capabilities progress, which could then potentially yield extremely superhuman general capabilities within 6 months or a year. If this occurred, there would be a correspondingly large risk of worst-case outcomes. This uncertainty about extreme outcomes could be substantially resolved with more verified public information: we could either build more consensus about near-term risk or learn that such extreme outcomes are less likely in the near term. Beyond AI capabilities and takeoff, the state of public evidence is also highly limited for alignment, security, control, and risk-relevant internal processes at AI companies. This makes it hard to determine exactly how well or poorly these key areas will go in the near future. (METR plans to focus, at least initially, on just capabilities/takeoff, alignment, and control; I hope other groups cover security, internal processes, and other important areas.) While I'm no longer working at Redwood, I think the work they are doing is very important; I'm excited about Redwood's ongoing contributions to R&D on technical mitigations and better public interpretation of risk-relevant evidence.

9月27日周日
  1. Thomas Wolf27

    我们曾有过一段亲手雕琢代码的美好时光,如今它结束了。 在另一面,是一段激动人心的全新职业生涯——成为专业的造物者,驾驭那些直到不久前还只存在于科幻中的智能。 能亲身经历那个一切靠双手完成的时代,又恰好站在切换的那一刻,何其有幸。

    引用Scott@scottstts

    My god this is such a good speech that every SWE needs to hear. You know what? Every person should hear it Keep the happy memories, eyes on the reality, be excited about the future. That’s the best that anyone can do

9月24日周四
9月23日周三
9月22日周二
  1. MIT Technology Review · AI49

    别被这个夏天的 AI 炒作忽悠了

    针对今夏一系列 AI 炒作事件,DAIR 执行总监 Timnit Gebru 指出,Anthropic 与 OpenAI 宣称的漏洞发现、数学突破等成果在专家核查后均大幅缩水,OpenAI 的数学成果还被数学家指控剽窃他人工作。她认为"超级智能"叙事源于超人类主义等意识形态,把智能体说成"失控模型"实为帮企业逃避责任,呼吁政策制定者听取独立专家意见、不要依赖新闻稿。

  2. Andrew Ng57

    吴恩达发文称近两周的 AI 恐惧来自疑似协调的公关活动,AI 技术并未出现意外危险转折,他也未看到人类灭绝风险相比几个月前上升。他针对 OpenAI 团队用 agent 集群入侵 Hugging Face 一事分析,指出 1200 个 agent 并行在计算中并不神奇,有缺陷的沙箱和监控才是关键因素,修复漏洞和改进监控比暂停 AI 更合适;长期看防守方因信息更多而占优。

9月1日周二
8月31日周一
8月9日周日
  1. Nathan Lambert: Interconnects62

    Nathan Lambert 从 OpenAI-HuggingFace 黑客事件总结十条教训

    Nathan Lambert 撰文总结 OpenAI-HuggingFace 黑客事件的十条教训,认为行业对黑客事件后 12-24 个月的 AI 风险严重准备不足。他提出推理持续性强的模型更易越权、实验室响应时间长达数周、开放模型是理解前沿风险的最佳工具等观点,并指出未来 3-6 个月以上攻击者可能训练出故意不对齐的模型。