跳到正文

全部动态

今日 59 条
9月23日周三
  1. ViggleAI46

    我们为开源社区推出了首个 Qwen-Image-2.1 turbo。 试试 Viggle-Turbo,一个经 DMD 蒸馏的 Qwen-Image-2.1,仅需 4 个采样步即可生成和编辑,无需 classifier-free guidance。 权重:https://huggingface.co/Viggle/Qwen-Image-2.1-viggle-turbo

    引用Hugging Apps@HuggingApps

    Qwen-Image-2.1 in 4 steps is here ⚡ @ViggleAI distilled Qwen-Image-2.1 into a 4-step turbo model, 6× faster, and holds up side by side with the full model ▶️ on Spaces https://hf.co/spaces/Viggle/Qwen-Image-2.1-viggle-turbo

  2. Gary Marcus28

    Gary Marcus 致信特朗普:美中可在 AI 安全与癌症、网络安全上合作而非放缓

    Gary Marcus 公开致信特朗普,建议其在与中国领导人的 AI 会谈中放弃"放缓"路线,转而推动美中合作,领域包括癌症与网络安全,类似冷战时期美苏在太空和天花上的合作。他援引《人民日报》文章称美中应"让 AI 成为中美合作新前沿"并利用政府间 AI 对话机制,认为中国实际上渴望合作。

  3. elsewhere articles68

    从 MiMo-V2.6 看大模型「斩杀线」:斩的是中间层模型

    文章以小米 9 月开源的 MiMo-V2.6 系列为切入点,分析大模型「斩杀线」概念,即在智能和成本两个维度都被超过的模型会失去被选择的理由。

    推荐理由:文章借小米 MiMo-V2.6 梳理了智能与成本双维度的行业竞争框架,读者可以据此理解 Agent 时代效率为何成为模型竞争的新坐标。

  4. Latent Space85

    Anthropic 发布 Claude Opus 5.5,OpenAI 同日跟进低价 GPT-6 Sol 与 Luna

    AINews 汇总 Claude Opus 5.5 发布:作为 Claude 5.5 家族首个模型,官方称多数任务达到 Fable 5.1 水平、比 Opus 5 快约 30% 且每任务便宜约 40%,token 单价从 $5/$25 降至 $4/$20,并成为 Claude Code 与 Claude 应用的新默认模型。

    推荐理由:汇总了 Opus 5.5 与 GPT-6 Sol/Luna 同日发布的基准、定价与第三方评测,含实际每任务成本与争议细节,便于对比两家宣传口径。

  5. Tencent Hy60

    Hy Image3.5 preview 已可在 ComfyUI 中使用,人类评测胜率较 Hy Image3.0 提升 30%。单模型同时支持文生图和图生图,最高 2K 分辨率,可正确渲染多语言文字与符号,覆盖电影感、漫画、商业摄影和插画风格,身份与产品特征在场景、服装和风格切换中保持一致。

    引用ComfyUI@ComfyUI

    Hy Image3.5 preview is now available in ComfyUI. Professional-grade image generation, +30% win rate in human eval vs Hy Image3.0 → Text to image and Image to image in one model, up to 2K → Multilingual text, symbols, and small print that render correctly → Cinematic, comic, commercial photography, and illustration styles → Identity and product features that hold through scene, outfit, and style changes

  6. elsewhere articles33

    弋途科技完成近亿元 Pre-B 轮融资,AICAR 正式上线

    弋途科技(EXTURING)近日完成近亿元 Pre-B 轮融资,由上海半导体装备材料产业投资基金与 Sands Talk Capital 联合领投,资金将用于加速 AICAR 规模化推广及自研端侧模型落地。公司已服务超 70% 国内头部车企及主流合资品牌,数十款车型量产上车,累计交付达百万级,AICAR 近期正式上线。

  7. Qwen66

    千问(Qwen)宣布 Qwen-Image-2.1 在 Arena 的 Image Edit 和 Text-to-Image 两个榜单均排名第一的开源模型。@arena 引用称其 Image Edit Arena 得分 1367,总排名第 16,距第 15 名 GPT-Image-1.5-high-fidelity 仅差 3 分。

    引用Arena.ai@arena

    Qwen-Image-2.1 by @Alibaba_Qwen just landed as the #1 open source model in the Image Edit Arena and Text-to-Image Arena! With 1367 pts in the Image Edit Arena, Qwen-Image-2.1 took the #1 spot among open. It landed #16 overall, just 3 pts from GPT-Image-1.5-high-fidelity at #15. See the leaderboard for the Text-to-Image arena below. Congrats to the @Alibaba_Qwen team on this contribution to the open source ecosystem!

    推荐理由:官方确认 Qwen-Image-2.1 登顶两个图像 Arena 的开源榜首,榜单分数可用于同类模型的横向比较。

  8. Ant Ling46

    感谢 @ValsAI 的高水准评测!"flash"这个词现在有点"误导"了。Ling-3.0-flash-fin 总参数 124B、激活 5.1B,是一款高智能密度的"flash lite"。趁免费 API 还在,赶紧用起来。我们还有 fp4 量化版本可用于本地 AI 😛

    引用Vals AI@ValsAI

    Ant Group’s Ling 3.0 Flash Fin is a finance-specialized open-weight model that delivers strong financial analysis at budget-model pricing. On Finance Agent v2, it scores 54.9% at just $0.045 per task.

  9. Apple Machine Learning Research46

    Apple 提出 probe guidance:用冻结内部状态引导流匹配模型

    Apple 研究团队提出 probe guidance,利用现有扩散模型冻结的内部状态构建引导信号,无需在推理时增加额外前向计算。该方法在连续扩散语言模型的无条件生成上刷新 SOTA,应用于 1.7B 扩散语言模型时持续提升多项选择题基准表现。研究还发现,autoguidance 中的弱模型必须来自训练的低熵区域。

  10. Google Developers Blog65

    Antigravity SDK 支持本地模型,首发接入 Gemma 4 26B A4B 与 LiteRT

    Google 宣布 Antigravity SDK 支持本地工作流,首发支持通过 Google AI Edge 的 LiteRT 运行 Gemma 4 26B A4B,可完全离线提供智能体能力,建议机器配备 24GB 以上 VRAM 或统一内存。

    推荐理由:原文给出本地运行智能体的具体配置方法和混合编排演示数据,开发者可以直接照此把智能体工作流搬到本地 GPU 上。

  11. Simon Willison83

    Anthropic 发布 Claude Opus 5.5,OpenAI 同日推出 GPT-6 Sol 与 GPT-6 Luna 并掀起价格战

    Anthropic 于 9 月 22 日发布 Claude Opus 5.5,约一小时后 OpenAI 发布 GPT-6 Sol 和 GPT-6 Luna。GPT-6 Luna 定价 $0.10/$0.50 每百万 token,比 GPT-5.6 Luna 再降一半;Opus 5.5 降价 20% 至 $4/$20,缓存读取价格下降 60%。

    推荐理由:作者用实测和价格对比表梳理了这轮降价的具体幅度,还发现 Opus 5.5 max 档过度思考撞上输出上限的问题。

  12. Gary Marcus54

    Gary Marcus 回顾 Facebook M 失败史,质疑 Meta Muse 重走老路

    Gary Marcus 撰文称 Meta 新智能体 Muse 正在重复 Facebook 2015 年 Project M 的做法,如餐厅订位、代订票务等。Project M 因运行成本高只开放给约 1 万用户,后被曝幕后有真人参与,AI 实际处理的请求不超过 30%,于 2018 年 1 月在上线不到三年后取消。作者还批评扎克伯格在隐私问题上的态度,认为他在重犯当年的错误。