

关注 AI 研究者、开发者与机构的动态


宇树发布载人变形机甲 GD01,起售价 390 万人民币。推文提到大疆新无人机可吊 600kg 物品,而 GD01 为 500kg,两者组合让人联想到《环太平洋》经典镜头。
移动端的语音输入法必须带全键盘,但是桌面端的语音输入法最好跟输入法本身解耦。 Typeless 犯了前一个错误,豆包犯了后一个。
latent.space/p/unsupervised-… 这期精彩的节目启发了这篇文章。 @swyx @jacobeffron
I hope you guys understand that this is going to keep getting worse
推荐理由:梳理了这轮 npm 供应链攻击的传播链条和被绕过的签名验证,读者可据此理解 AI 开发流面临的风险。
[AINews] TML-Interaction-Small 276B-A12B latent.space/p/ainews-thinki… Thinking Machines' Native Interaction Models - TML-Interaction-Small 276B-A12B - advances SOTA Realtime Voice and kills standard VAD all our highlights!
这就是我们打造 SenseNova U1 的原因。✨ 感谢 @feesyiam 用它来聚焦儿童福利的关键议题。可视化让艰难的对话更易被理解——这正是 AI 真正有价值的时刻。 继续创作吧。🥰
I gave it a topic. It came back with a full magazine-style infographic. Charts. Layout. Icons. Colour coding. Dense structured copy. That model is SenseNova U1. And it's open-source. 🧵
在拿了真格的 Token Grant 之后,跟他们聊了一下最近的一些思考,希望对大家有帮助。 mp.weixin.qq.com/s/KAv6l934V…
Veo 4 vs Seedance 2.0 social media is cooked Veo 4 looks like it will be much better than Seedance 2.0 and it's an omni model which means consistent voice references and image inputs just like Seedance what a time to be alive.. Prompt: A professor writes out a mathematical proof for trigonometric identities on a traditional chalkboard, explaining the step he is currently on in the equation. Video Video
GOOGLE 🔥: An upcoming Gemini Omni video model from Google is expected to be much more advanced in video editing, capable of completing tasks like removing watermarks, replacing objects in the video, and more. It is also likely that Google will release 2 versions of this model, including a Pro variant. And I assume what we see isn't Pro? Anime sample 👀 Video




官方博客中引用了 @pingToven 的话 claude.com/blog/claude-platf…
The Claude Platform on AWS is now generally available. AWS customers get the full set of Claude API features, with AWS authentication, billing, and commitment retirement.
New in Claude Code: agent view. One list of all your sessions, available today as a research preview. Video
推荐理由:原文列出了 Agent 视图的打开方式和状态标注,读者可据此判断多任务并行时如何管理多个 Agent。
People talk, listen, watch, think, and collaborate at the same time, in real time. We've designed an AI that works with people the same way. We share our approach, early results, and a quick look at our model in action. thinkingmachines.ai/blog/int… Video
nitter.net/chhillee/status/205394…
In modern ML accelerators, FLOPS have absolutely exploded. Often though, the bottleneck is not FLOPS but memory bandwidth. Similarly, model intelligence has exploded, causing the bottleneck to be human<->AI bandwidth. At Thinky, we think that it’s important to solve this. 1/4
最出色的广告大片不只是展示产品,而是让你向往它所处的世界。 设定愿景。定义美学。Luma Agents 由此构建每一张奢侈品广告视觉。 设定标准 → lumalabs.ai/app Video
After being a Claude Code devotee for a year, I finally tried Codex on a new project this weekend. Once again, in the matter of a few months, it feels like the world changed. I can see myself doing *everything* inside of Codex this week.
如果你的团队开站会汇报进展,GPT-Realtime-2 会帮你移动工单呢? 视频
lowkey the funniest videos of the batch. thinky has some comedians!! congrats to @thinkymachines on reviving the omnimodel dream that others could not
是她。这就是她。 piped.video/watch?v=iVDJ8O89…


Today we're sharing our work on interaction models. A new class of model trained from scratch to handle real-time interaction natively, instead of gluing it onto a turn-based one. piped.video/A12AVongNN4
如果这听起来有意思,请与我们联系! openai.com/daybreak/
20 位开发者。8 周中的第 3 周。不再空想,只有交付。 本周三,看谁在追逐自己的第一桶金。 《Race to Revenue》第 3 集 ⠕ 视频
情绪板一直是最棒的部分。现在它只是起点。 上传你的参考图。设定方向。Luma Agents 从那里把情绪板变成成品广告。 把它做成广告 → lumalabs.ai/app 视频