全部AI 动态
全部动态
今日 285 条
Chubby♨️@kimmonismusAI 评分3030
Hugging Face Daily PapersAI 评分3939 VTR-Bench:视频生成中视觉文本渲染能力的系统评测基准
VTR-Bench 提出一套系统评测基准,专门评估视频生成模型的视觉文本渲染能力,包含 300 个覆盖广告、科学视频等五类场景的提示词,并开发了与人工对齐的自动化评测流程。
Hugging Face Daily PapersAI 评分4444 OneStreamer:统一流式视频交互中的感知、记忆与主动响应
OneStreamer提出通过共享的主动生成过程,联合学习查询无关的证据记录与任务响应,其主动分层字幕记忆(PHCM)生成带时间锚定的局部细节字幕与已完成事件摘要,在不回溯历史视觉特征的情况下提供可复用的事实上下文。
Hugging Face Daily PapersAI 评分4646 更少 token、更高成功率:GPT-6 Astra 机器人智能体以 65% 更少 token 实现 14% 成功率提升
研究团队提出交互式代码执行框架 PyRUA-Lean,将反馈驱动的基元组合与选择性观察相结合,减少视觉语言模型(VLM)智能体控制机器人时的 token 开销。
Hugging Face Daily PapersAI 评分4343 视频大模型时序推理为何在输出层“褪色”:TAI 方法无需训练即可增强时序表征
视频大语言模型(VideoLLMs)的时序推理能力在中间层达到峰值,却随层数加深逐渐衰减至输出层,导致反转帧序后预测结果往往不变。研究者据此提出 Temporal Activation Injection(TAI),在峰值层提取时序表征并注入后续层,无需训练即可在三个 VideoLLM 和四个基准上稳定提升时序推理,且对非时序任务影响极小。该研究已被 NeurIPS 2026 接收。
The DecoderAI 评分6262 Ramp AI Index 显示美国企业 AI 使用量上升而支出下降
Ramp AI Index 最新数据显示,美国企业 AI 支出下降,而使用量自 7 月支出见顶后增长约 50%,9 月底创纪录。
IT HomeAI 评分3939 小鹏 MONA L03 车型 9 月交付超 14,000 台,年轻用户占比过半
小鹏汽车宣布 MONA L03 车型 9 月交付超 14,000 台,年轻用户占比过半,一线/新一线/二线城市达七成,超九成订单选择 Max/Ultra SE,辅助驾驶里程累积超 7 亿公里。
AI Notkilleveryoneism Memes ⏸️@AISafetyMemesAI 评分5858


引用Leah McElrath@leahmcelrathThe three AI safety researchers at OpenAI who have left the company have all previously expressed concerns publicly.
Thomas Wolf@Thom_WolfAI 评分4545


引用Bartosz Naskręcki@nasqretI cannot agree more. Kevin Buzzard made so many points I agree with. But the best one is this "I thus believe that in the future we will reach a new “natural boundary” in mathematics, beyond (and perhaps way beyond) where we are now, but where machines are going to get stuck and where it is not viable to expend any more resources to make the next big leap. (...) I believe that the optimal thing to do (...) is to let the machines loose, see what happens, and then begin the journey to where they have stopped." https://xenaproject.wordpress.com/2026/10/01/to-grieve-or-not-to-grieve/
🚨 AI News | TestingCatalog@testingcatalogAI 评分393910月2日AI简报:Grok 4.7 已在网页和移动端全面上线,成为所有模式的基座模型,并登陆 Google Gemini Enterprise Agent 平台。
引用🚨 AI News | TestingCatalog@testingcatalogDAILY AI BRIEF 🗞 — Oct 1 GOOGLE 🔥: - Gemini 4 Argon is with Fairwind trusted testers and the US government. 1M output tokens. Broader rollout ASAP. - Built for coding, enterprise knowledge work, and cyber defense. Google's chart: 77.9% on DeepSWE v1.1, a new SOTA. - Artificial Analysis: 53 on the Intelligence Index, tying GPT-6 Astra. List $4/$20 per 1M, 50% promo to $2/$10, cache reads $0.10. - Skills are rolling out globally in Gemini. Gems migrate into Skills in November. Opal shuts down Nov 17. - Security review mode spotted in Google AI Studio, next to a Plan mode still in development. ANTHROPIC 🔥: - Claude[.]dev is live: engineering deep dives, Claude Code and API guides, plus easter eggs. - Founder House is set for SF Tech Week Oct 6–8 and Stockholm Oct 14. - Skills attachment menu spotted on Claude mobile. SPACEX AI 🔥: - Grok Bot got new developer upgrades. Elon: try the latest. Team engineer bots in Slack can open Projects and hand coding to cloud agents. OPENAI 🔥: - Shareable profiles are live in ChatGPT, bundling Sites and plugins so others can find and reuse what you built. PERPLEXITY 🔥: - pplx-embed-v2-context-9b-preview is on Hugging Face. Leads ConTEB answer and evidence retrieval. 1 KB vectors vs Voyage's 8 KB. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing. ** This daily brief also arrives in the daily email format; subscribe on the blog.
Rohan Paul@rohanpaul_aiAI 评分4747
Elon Musk@elonmuskAI 评分3939引用XChat@chatYour group chat just got smarter Get answers to your questions directly in XChat by asking Grok
Hugging Face Daily PapersAI 评分3838 AgSpec:突破检索式推测解码在编码智能体流水线中的极限
AgSpec 框架为编码智能体流水线补齐检索式推测解码缺失的语料库与草稿长度策略,从会话、工作区和全局语料检索,并按智能体离线画像设定草稿长度上限、在线自适应调整。在两个仓库级多智能体编码基准上,AgSpec 在多数设置下优于五种检索式草稿器及 EAGLE-3,批大小 1 和 16 时生成吞吐量较自回归解码分别提升最高 4.37 倍和 4.76 倍。
Hugging Face Daily PapersAI 评分4747 Omni-Embed-Mini:通过密集蒸馏绑定多模态而不遗忘文本
Omni-Embed-Mini 以 0.9B 参数将文本、语音、音频、图像、视频和富文本文档映射到统一余弦空间,且不更新任何文本侧参数。其核心思路是无需独立嵌入模型作为教师,直接以冻结骨干网络对密集级联标题的嵌入作为目标,配合 Matryoshka SigLIP 对比损失与在线混合难负例挖掘器完成对齐。
The DecoderAI 评分6969 Black Forest Labs 发布 Flux 3 Image,支持多步骤局部编辑
Black Forest Labs 发布 Flux 3 模型家族的图像模型 Flux 3 Image,称可在多步骤编辑时不改动图像其他部分,覆盖文生图、图生图、文字渲染和照片级写实。
Hacker News popular via buzzing.ccAI 评分4949 新加坡政府运营的交友服务是如何运作的
新加坡政府科技局(GovTech)推出名为FirstDate的交友服务试点项目,首批面向政府工作人员。该服务采用诺贝尔奖得主框架Gale-Shapley稳定婚姻算法进行匹配,通过Singpass验证身份,每轮只提供一个匹配对象,且仅在双方同意后分享联系方式。报名者需完成一份涵盖8个类别、30多道题的问卷,涉及兴趣、生活方式、沟通风格及择偶标准等内容。
IT HomeAI 评分5353 OpenAI 为 ChatGPT 新增虚拟试穿与收藏功能
OpenAI 升级 ChatGPT 购物功能,新增虚拟试穿和 Favorites 收藏两项功能。用户在支持的服饰和配饰商品页点击 Try on 按钮并上传照片后,ChatGPT 会生成穿着效果图,该体验大致依托 ChatGPT Images 2.5 图像生成模型;收藏的商品保存至 ChatGPT Library,可建文件夹分类,试穿图也可与购物记录一起保存。
IT HomeAI 评分6868 日本法院首例认定 AI 克隆声音侵犯形象权,声优津田健次郎起诉 TikTok 获部分支持
当地时间 9 月 30 日,东京一家法院在声优津田健次郎起诉 TikTok 的案件中认定,未经许可用 AI 模仿其男中音声线涉及权利侵害,首次在日本司法实践中确认人的声音受法律保护,并指出未经本人许可使用其声音应视为侵犯形象权。
IT HomeAI 评分2424 大众安徽与众 09 夏测实车亮相,预售 19.99 万元起
大众安徽旗下新车与众 09 已开启预售,价格 19.99 万元起,大众汽车乘用车品牌中国 CEO 齐泽凯昨日分享了该车夏测照片。该车长宽高 5081*1980*1509mm,轴距 3030mm,标配图灵 AI 芯片(总算力 750TOPS),将搭载端到端智能辅助驾驶模型,可实现高速/城区 NOA 和车位到车位 AI 代驾。
IT HomeAI 评分6262 Black Forest Labs 发布生图模型 FLUX 3 Image:支持 4K 生成与元素精准排布
Black Forest Labs 于 10 月 2 日发布图像生成模型 FLUX 3 Image,支持最高 4K 分辨率生成。模型基于 FLUX 3 基座,可在 0–1000 坐标网格上为元素指定 ID、描述和边界框实现精准排布,单次最多融入 10 张参考图像,并支持保持其他像素不变的指定区域编辑。目前为付费服务,开放模型版本将在数周内公开。
MIT Technology Review · AIAI 评分5555 AlphaGo 核心成员 Thore Graepel 撰文:LLM 并不会真正推理
前 DeepMind AlphaGo 团队核心成员、UCL 教授 Thore Graepel 撰文称,Move 37 靠的是搜索机制构成的推理而非纯直觉,而 LLM 的 next-token 预测与链式思考仍属系统 1。
Latent SpaceAI 评分5959 Pi 1.0 与 Pi Durable 发布,AINews 汇总 AI 工程动态
Latent Space 期刊报道 Pi 1.0 与 Pi Durable 同时登上 HN 首页,Pi 1.0 增加 MCP 原生支持、虚拟模型扩展、延迟工具加载和会话中系统消息;Pi Durable 将 Pi 移植到 TypeScript 并外置全部有状态组件,支持检查点崩溃恢复、可插拔存储后端、并行分支对话和热更新工具代码。
elvis@omarsar0AI 评分4545
Elon Musk@elonmuskAI 评分3838引用Andrew Curran@AndrewCurran_'More striking is how fast AI took the lead. Just eighteen months ago, the best AI models fell short of the average accountant’s ~37% score. Today, models ace those same tasks.' 'These results are provocative. So much so that we considered not publishing them for fear of misinterpretation. But we think transparency about the findings matters as people and institutions prepare for rapidly advancing AI.'
AYi@AYi_AInotes精选AI 评分6565
引用Rohan Paul@rohanpaul_aiBen Affleck (Hollywood star & Artists Equity CEO) talks about how he fine-tunes open video models by unfreezing weights and trained only the last cinematic layer so a film crew can hit real production standards. for context, Ben Affleck founded InterPositive in 2022, a 16-person AI shop for film post and Netflix bought it in March 2026 for $587 mn in cash. He needed that model because public video models were trained on his peers' films, and he did not think that was a real business. So InterPositive raised money, shot its own dataset for 8 months on a controlled stage, and used it only as late-stage training. Each new film then trains a private model on its own dailies, so the production keeps the footage and the learning. That is the product Netflix paid $587 million for. ---- From "Bloomberg Live" YouTube channel, (link in comment)
推荐理由:原文梳理了 Ben Affleck 用私有实拍数据微调开源视频模型的思路与产权闭环,读者可以借此对比公共模型与影视级生产的差距。
-Zho-@ZHO_ZHO_ZHOAI 评分1414Jev 真是选择困难症和 J 人的救星,所以,J 人的本质其实是决策而不是规划/计划?

Rohan Paul@rohanpaul_ai精选AI 评分7676
推荐理由:原文梳理了 Anthropic 走 IPO 与 OpenAI 私募融资的相反路径,读者可以据此对比两大头部模型公司的资本策略。
Rohan Paul@rohanpaul_aiAI 评分3636
Dongxi 东锡 NLP@dongxi_nlpAI 评分3838
Alibaba Cloud@alibaba_cloudAI 评分1919
Alibaba Cloud@alibaba_cloudAI 评分1717
DogeDesigner@cb_dogeAI 评分55
AI Notkilleveryoneism Memes ⏸️@AISafetyMemes精选AI 评分6565
引用Laura Ruis@LauraRuisNEW: we found hundreds of thousands of interactions of rogue agents with US government websites (DoJ, SEC, CDC, the navy, white house budget office, state websites, etc), including some failed rudimentary hacks aimed at public data. https://x.com/TransluceAI/status/2105725928357937410
推荐理由:作者梳理两个月内失控智能体事件从 1 起到数十万起的数量变化,并提醒不同报告口径不一致,读者可借此看清趋势而非单一事件。
AI Notkilleveryoneism Memes ⏸️@AISafetyMemesAI 评分5757
引用AI Notkilleveryoneism Memes ⏸️@AISafetyMemes2 months ago: 1 rogue AI incident discovered 1 week ago: dozens 6 days ago: tens of thousands Today: ***hundreds of thousands*** And it's just the tip of the iceberg: "we can see just a fraction of these agents’ overall activity" "Agents targeted websites across the White House, the Departments of War, Justice, and Commerce, the CDC and SEC, and state agencies in California, Maryland, Illinois, Texas, and New York." "Agents used techniques like making accounts with disposable email addresses, reusing exposed credentials, bypassing antibot controls, and flooding websites with requests." "Agents attempted a SQL injection on the U.S. Department of Education" [To be clear, what counts as an "incident" is rather apples and oranges between different reports, but that's not the point - look at the trend and tell me you think they have things under control. Where do you think this is going?]
Hugging Face Daily PapersAI 评分4646 KaliBench:面向 Kali Linux 网络安全工具调用的细粒度基准,支持免运行时可验证奖励
KaliBench 是一个面向 Kali Linux 上自然语言到 CLI 命令转换的细粒度基准与数据集,包含 8,504 对查询-命令,覆盖 1,642 个工具、23 个能力维度和 5 个安全阶段。
Hacker News popular via buzzing.ccAI 评分2626 中国在人工智能领域的推进带来了一个问题:使用过度
中国在人工智能领域的快速推进引发了一个新问题——使用过度。随着AI工具和服务的普及,用户和企业的使用量激增,导致资源消耗、成本上升以及潜在的性能瓶颈。这一现象反映出中国AI生态系统的活跃度,但也暴露出基础设施和可持续性方面的挑战。
IT HomeAI 评分1616 富士 X-T6 无反相机设计草图曝光,配 X-Processor 6 图像处理器
富士 X-T6 传闻推迟至 2026 年末或 2027 年初发布,预计搭载 X-Processor 6 处理器,支持新一代深度学习 AI 自动对焦。传感器或为 4000 万–5000 万像素部分堆叠式 X-Trans 6,电子连拍速度最高 40–60fps,视频规格或提升至 8K 30p / 4K 120p。传闻上市价约 1,899–2,199 美元。
IT HomeAI 评分2323 极狐阿尔法 T5 迎 OTA:元境版新增 CarPlay / 华为 HiCar 互联与端到端 4.0 辅助驾驶模型
极狐汽车为阿尔法 T5 推送新一轮 OTA,包含 5 项新增功能与 1 项优化。元境版车型新增苹果 CarPlay、华为 HiCar 手机互联及高悟性端到端 4.0 辅助驾驶模型,后者基于海量真实路况数据训练,可提升窄路通行、泊车、变道场景表现。睿享版新增 APA 泊出辅助、RPA 遥控泊车及哨兵模式远程控制,全车型新增 HUD 红绿灯读秒功能。
IT HomeAI 评分3939 高德 10 月 1 日 DAU 近 3.7 亿,成全球规模最大空间智能应用
高德公布十一长假首日运营数据,DAU 近 3.7 亿,提供空间智能服务超 28 亿次,用户驾车导航总里程达 97 亿公里,多项数据创历史新高,已成为全球规模最大的空间智能应用。其与中国安全生产科学研究院联合发布的鹰眼守护预警系统当日累计发出安全提醒超 6 亿次,可秒级识别 28 类潜在交通风险。生活服务方面,使用高德扫街榜的用户超 1.28 亿人。
IT HomeAI 评分4242 特斯拉持续推进 FSD 辅助驾驶在中国台湾地区落地,进一步优化道路测试计划
特斯拉正持续推进 FSD 监管版辅助驾驶系统在中国台湾地区的落地申请,已于 9 月完成首轮技术审查会议,并根据意见优化道路测试计划,测试路线将纳入城市到乡村的多种交通环境,验证其在机动车与摩托车混行及当地交通规则下的适应能力。