跳到正文

#Microsoft

今日 7 条
今天10月2日周五
  1. 🚨 AI News | TestingCatalog39

    10月2日AI简报:Grok 4.7 已在网页和移动端全面上线,成为所有模式的基座模型,并登陆 Google Gemini Enterprise Agent 平台。

    引用🚨 AI News | TestingCatalog@testingcatalog

    DAILY AI BRIEF 🗞 — Oct 1 GOOGLE 🔥: - Gemini 4 Argon is with Fairwind trusted testers and the US government. 1M output tokens. Broader rollout ASAP. - Built for coding, enterprise knowledge work, and cyber defense. Google's chart: 77.9% on DeepSWE v1.1, a new SOTA. - Artificial Analysis: 53 on the Intelligence Index, tying GPT-6 Astra. List $4/$20 per 1M, 50% promo to $2/$10, cache reads $0.10. - Skills are rolling out globally in Gemini. Gems migrate into Skills in November. Opal shuts down Nov 17. - Security review mode spotted in Google AI Studio, next to a Plan mode still in development. ANTHROPIC 🔥: - Claude[.]dev is live: engineering deep dives, Claude Code and API guides, plus easter eggs. - Founder House is set for SF Tech Week Oct 6–8 and Stockholm Oct 14. - Skills attachment menu spotted on Claude mobile. SPACEX AI 🔥: - Grok Bot got new developer upgrades. Elon: try the latest. Team engineer bots in Slack can open Projects and hand coding to cloud agents. OPENAI 🔥: - Shareable profiles are live in ChatGPT, bundling Sites and plugins so others can find and reuse what you built. PERPLEXITY 🔥: - pplx-embed-v2-context-9b-preview is on Hugging Face. Leads ConTEB answer and evidence retrieval. 1 KB vectors vs Voyage's 8 KB. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing. ** This daily brief also arrives in the daily email format; subscribe on the blog.

  2. Microsoft AI News61

    Microsoft AI 发布 MAI-Transcribe-2-Streaming 及 MAI-Voice-2.1 语音模型

    Microsoft AI 发布实时流式转写模型 MAI-Transcribe-2-Streaming,在 Artificial Analysis 的最终与部分转写准确率均排名第一,支持 60 种语言,接收到音频后约 100ms 内产出首个 partials,内部评测显示实时听写场景出词速度比最接近的竞品快 2x,年底前 introductory 价 $0.54 每小时音频。

    推荐理由:官方公布三款语音模型的具体定价、延迟数字和 Artificial Analysis 排名,可据此评估搭建语音智能体的成本与速度。

10月1日周四
  1. The Decoder78

    OpenAI 称已阻止窃取其模型推理内容的行动,但同样手法在 Azure 上仍然有效

    OpenAI 称 7 月拦截了一起窃取其模型思维链用于蒸馏的行动,7 月 24 至 25 日流量激增至来自超 4000 名用户的 16000 次请求,关联账号网络超 15000 个,7 月 28 日已全部封禁,OpenAI 将核心团伙与 Moonshot AI(Kimi 模型开发商)相关人员联系起来,并注明这些是未遂提取。

    推荐理由:原文串起 OpenAI 的蒸馏攻击拦截和研究者的复测结果,能帮读者看清同一模型在不同云平台防护不一致的问题。

9月30日周三
  1. Mark Zuckerberg66

    Mark Zuckerberg 表示,美国各大前沿实验室负责人已承诺实施严格的内部控制和多层审计与审查,认为这能让公众更有信心实验室技术会按预期运行。引用的 @DavidSacks 推文称,各前沿实验室在白宫签署 White House Accord on Super Intelligence,承诺产品安全开发责任并接受新的内部控制和外部审计。

    引用David Sacks@DavidSacks

    Only President Trump could convene all the leaders of the top companies developing chips, data centers and frontier models for Super Intelligence. This new Industrial Revolution has already created a million new jobs and is spurring a bigger infrastructure build-out than the railroads, canals and grid combined. I was honored to witness history as the leaders of the frontier lab companies signed the White House Accord on Super Intelligence, accepting responsibility for the safe development of their products and imposing new internal controls and external audits. This is far better than waiting years for some international agreement that would probably never happen. President Trump continues to ensure that U.S. remains the technology leader while putting Americans first.

    推荐理由:协议文本列明四层控制与审计的具体安排,读者可以据此了解各实验室安全承诺的实际内容。

9月29日周二
  1. Microsoft Research61

    Microsoft Research 发布生物研究领域 AI 系统 Quine

    Microsoft Research 推出 Quine,一个面向生物学的多模态世界模型与交互式 harness,连接科学工具、文献和研究人员。在与 Broad Institute 合作中,Quine 用于预测可驱动胰腺癌肿瘤细胞状态转变的化合物,排名第一的候选化合物在湿实验中产生了最大的预期细胞状态转变,从缩小化合物范围到确定候选名单仅用了一个周末。

    推荐理由:官方披露了系统构成和胰腺癌湿实验验证结果,读者可以据此评估AI世界模型在生物研究中的实际作用。

  2. Microsoft Research24

    微软研究院亚洲新加坡分院成立一周年:推进前沿 AI 研究、合作与人才培养

    微软研究院亚洲新加坡分院(MSRA – Singapore)作为微软在东南亚的首个研究实验室,成立一年来围绕下一代 AI 模型与智能体系统、领域专用 AI、AI 原生研究实践、生态与人才培养四大方向展开工作。该实验室与新加坡医疗生态伙伴合作推进多模态医疗 AI 和自进化诊断智能体,并与 EDB、IMDA 等机构在物流运输、工业 AI 等领域开展合作。

9月28日周一
9月26日周六
  1. GitHub Blog · AI & ML60

    GitHub Copilot app 新手教程:如何用 canvases 构建自定义工作流

    GitHub Copilot app 提供 canvas 功能,运行 /create-canvas 技能并用自然语言描述工作流、界面操作和 agent 职责三要素,即可生成看板、清单等自定义界面并保存为可共享的 extension。

    推荐理由:官方教程给出从描述到生成可复用 canvas 界面的具体步骤和三个提问框架,读者可以直接照做迁移到自己的工作流。

9月25日周五
  1. GitHub Blog · AI & ML61

    GitHub Copilot 博客:为什么聊天界面往往不是正确的 UI,用 canvas 试试

    GitHub Copilot 博客作者提出聊天(chat)很多时候是错误的 AI 交互界面,介绍 GitHub Copilot app 中的 canvas,它是在应用内运行、无浏览器外壳的全栈小应用,可与 Copilot agent 双向通信。

    推荐理由:作者作为 Copilot 团队成员提出用 canvas 自定义界面替代聊天框,并结合工作流自动化等实例说明何时值得让智能体先造工具。

9月24日周四
9月21日周一
  1. Mustafa Suleyman36

    这份跨党派的人类主义 AI 宣言中有很多非常好的提议。仍有一些值得我们讨论,但总体上是正确方向。我鼓励大家都去看一看。

    引用Max Tegmark@tegmark

    I'm delighted to share that @mustafasuleyman, CEO of Microsoft AI, co-founder of Google DeepMind and Inflection AI, has signed the Pro-Human AI Declaration. If you too support it, please join him and over a million others by signing it here – the momentum is building! Let's build tools not beings & keep humans in charge. https://humanstatement.org

9月19日周六
9月17日周四
  1. GitHub Blog · AI & ML78

    GitHub 用 Copilot 将 Copilot agent runtime 迁移到 Rust

    GitHub 用 Copilot 智能体将 Copilot agent runtime 从 TypeScript 完整重写为 832,378 行生产 Rust,共 128 个 PR,于 8 月 21 日完成。

    推荐理由:作者以第一手移植经历拆解了 AI 智能体团队完成大规模重写的具体策略、会话数据和教训,方法细节对类似工程迁移有直接参考价值。

9月16日周三
9月15日周二
9月14日周一
  1. Mustafa Suleyman40

    这是一个非常直白且符合常识的观点:技术的目的是服务人类,加速人类繁荣。 任何无法实现这一目标的技术都是失败的,应当被拒绝。 我们还没有到那一步。但开始为这种可能性做准备是正确的。

    引用Satya Nadella@satyanadella

    Any pursuit of superintelligence has to be grounded in the core principle that if the AI we build is not helping humanity and under human control, it's not worth pursuing. We also need to accelerate and spread the benefits of AI, such that they are diffused broadly across countries, communities, and companies. This requires a frontier ecosystem in which both closed and open-source models can thrive. And for firms, it’s imperative that they retain full control over their unique and tacit knowledge. Every organization should be able to build its own continuous learning loop/hill climbing machine, without becoming dependent on any one model provider, and have the ability to embed its own knowledge into models and weights they control. So, in this context, we welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal. We also welcome ideas like "embedded evaluators" and the broader efforts to develop the mechanisms to make this more than just talk. The key is that this cannot be controlled by a handful of entities, but must have broad representation across the ecosystem, countries, and fields, including academia. This is the approach we are taking: broad access and choice at every layer of the AI stack; enterprise control of learning loops and models; and the “Code of Conduct” that underlies our own first party MAI models that we’ll publish tomorrow for public consultation.

9月12日周六
9月11日周五
9月10日周四
9月5日周六
9月4日周五
  1. GitHub Blog · AI & ML60

    GitHub Copilot app 新手教程:如何同时运行多个 agent 会话

    GitHub 官方博客发布面向新手的教程,介绍如何在 GitHub Copilot app 中同时运行多个 agent 会话。每个会话可运行在独立的 Git worktree 上互不干扰,各自保留上下文,用户通过 sessions 视图追踪进度,教程以 tailspin-toys 仓库演示了并行添加功能、无障碍审查和运行测试的做法。

    推荐理由:官方教程讲清了并行 agent 会话如何借助 Git worktree 隔离运行,读者可以照着示例直接上手尝试。

9月3日周四
9月1日周二
8月31日周一
8月21日周五