Google 在 X 上宣布推出新 AI 模型 Gemini Omni,称其可以从任意输入生成内容,并先从视频开始。帖文附带一段视频演示,并带有 #GoogleIO 标签。
X
关注 AI 研究者、开发者与机构的动态
按账号或来源筛选(519)
@Google@googleAI 评分5858 
@kimmonismus@kimmonismusAI 评分1919 
引用Chubby♨️ (@kimmonismus)@kimmonismusIo starts now!
@Google@google精选AI 评分7676 Google 表示 Gemini App 用户数在一年内翻倍,已突破 9 亿。该推文带有 #GoogleIO 话题标签,未披露其他数据。

推荐理由:原文只给出 Gemini App 用户一年内翻倍并突破9亿这一数字,可据此观察其用户增长节奏。
@kimmonismus@kimmonismusAI 评分1313 这太离谱了。Token 处理规模简直疯狂! (引用推文:Io 现在开始了!)
引用Chubby♨️ (@kimmonismus)@kimmonismusIo starts now!
@GeminiApp@geminiappAI 评分3131 #GoogleIO 2026 直播开始了!点此加入:nitter.net/i/events/2053241348807…
@MSFTResearch@msftresearchAI 评分1717 让社区通过参与AI开发流程来影响AI,可以改进AI,并帮助社区实现AI为其提供良好服务的潜力。news.microsoft.com/source/fe…

@Google@googleAI 评分2121 
@kimmonismus@kimmonismusAI 评分99 
@AYi_AInotes@ayi_ainotesAI 评分2222 补充一个大家都忽略的细节,Karpathy全程没有提AGI,只提了LLM前沿, 这说明在他眼里,下一个三年的主战场依然是LLM本身,而不是什么虚无缥缈的全新范式。
@AYi_AInotes@ayi_ainotes精选AI 评分8383
引用Andrej Karpathy (@karpathy)@karpathyPersonal update: I've joined Anthropic. I think the next few years at the frontier of LLMs will be especially formative. I am very excited to join the team here and get back to R&D. I remain deeply passionate about education and plan to resume my work on it in time.
推荐理由:作者从 Karpathy 官宣措辞中的 formative 一词切入,解读这次人事变动所指向的技术路线倾向。
@kimmonismus@kimmonismusAI 评分5656 Kim 发推写道“Karpathy 加入 Anthropic?不可能!什么情况”,该推文引用 @tbpn 的内容称 @karpathy 已加入 Anthropic。
引用TBPN (@tbpn)@tbpnBREAKING: @karpathy has joined Anthropic
@vista8@vista8AI 评分2020 Gemini Omni Flash 效果很拉胯啊! 提示词:生成墨比斯风格的科幻动画短片,银河系搭车客指南 好像根本没理解第二句话... Video

@berryxia@berryxiaAI 评分3434 还不知道在哪里看直播的兄弟们注意了! Google I/O 直播观看地址👇 #google
引用Google Gemini (@GeminiApp)@GeminiAppIt’s #GoogleIO Day One. Who’s ready to see what’s coming to Gemini? Livestream starts here at 10am PT: nitter.net/i/events/2053241348807…
@berryxia@berryxiaAI 评分5151
引用Elon Musk (@elonmusk)@elonmuskAnthropic will not be destroyed. Their AI+harness goes far beyond coding and Opus 4.7 is still better than Composer 2.5, albeit a lot more expensive. Cursor is however an important piece of the puzzle to make Grok much better.
@berryxia@berryxiaAI 评分1818 引用烟花老师 (@teach_fireworks)@teach_fireworks还有一百多就五千订阅了,不知道一觉醒来会不会有惊喜。我经常不按常理出牌,就提前写好庆祝5k订阅达成吧,哈哈🎆 我主业是一个AI架构师,也是一支烟花AI社区的联创,从23年至今大概积累了40个垂直的AI社群,大家都很纯粹 全都是免费的社群,基本上都是研发,产品和创业者和行业大佬,也欢迎大家一起进群交流,可以在这里登记信息,我会邀请大家进群 (长期有效) hqexj12b0g.feishu.cn/share/b… 虽然X算法偶尔抽风 还会误杀,不可否认X上的算法还算公平的,之前我分别在公众号,小红书和抖音尝试了蛮久自媒体,也是差不多的输出,最终还是X上正反馈更多一些,其他的平台都一言难尽。 争取今年做到一万粉。谢谢订阅我的朋友们,以后继续输出更多干货! 不过相当于X的收获,最大的惊喜是之前开源的fireworks-tech-graph 快7k star 了,靠神佬等众多大佬的喜爱转发,基本全靠X平台的传播,也合并了不少PR,我基本上没有在国内自媒体宣传过这个项目,完全靠自来水推荐,非常幸运可以感受到了流量加持后项目开源。 不过流量来得快去的也快,我内心也算比较平静,大大小小写了快20个开源项目,由于我懒得宣传,基本上是我自己在用,有几个harness 相关的skill 真的不错,大家可以去看下,总有一款你喜欢。
@op7418@op7418AI 评分2424 引用歸藏(guizang.ai) (@op7418)@op7418谷歌 Gemini Omni Flash 视频编辑测试。 你们应该能猜到我原始视频是在哪儿录的,反正效果远不如 SeeDance 2.0 Video
@berryxia@berryxiaAI 评分6363 @berryxia@berryxia精选AI 评分6565 引用Yukang Chen (@yukangchen_)@yukangchen_🚀 Excited to release LongLive 2.0! 🎬 An end-to-end infrastructure for long video generation, with FP4 and parallelism at the core of both training and inference. ⚡45.7 FPS generation speed on 5B model⚡ ✨ LongLive 2.0 supports real-video training, few-step distillation, multi-shot training/inference, sequence-parallel acceleration, NVFP4 KV cache, and async VAE decoding deployment. 🧩 To our knowledge, this is the first open-source 4-bit long video generation infra that covers both training and inference. 🙌 Welcome to check it out, try it, and share feedback! 🔗 Code: github.com/NVlabs/LongLive 📰 Paper: huggingface.co/papers/2605.1… 🎥 Demo: nvlabs.github.io/LongLive/Lo… #LongVideoGeneration #VideoGeneration #Realtime #AIInfra #EfficientAI #FP4 #Parallel #NVIDIA Video
推荐理由:开源方案把 FP4 量化与并行加速同时用在训练和推理,读者可据此了解长视频实时生成的技术路线。
@googleaidevs@googleaidevsAI 评分4343 @GeminiApp@geminiappAI 评分3131 今天是 #GoogleIO 第一天。谁准备好看看 Gemini 将迎来哪些新东西了? 直播将于太平洋时间上午 10 点开始:nitter.net/i/events/2053241348807…
@berryxia@berryxiaAI 评分2020 @LumaLabsAI@lumalabsaiAI 评分4343 
@op7418@op7418AI 评分1818 忽略口音的话,这个表现也不是很自然。 口音是因为我选了它内置的那个 Gemini TTS,所以才有这种英式口音 Video

@op7418@op7418AI 评分1818 这个分镜编排能力以及创作意识都不如 Sora 2,没有创作能力,只能编辑,做素材还行 Video

@berryxia@berryxiaAI 评分2424 引用Business (@XBusiness)@XBusinessx.com/i/article/205536179511…
@op7418@op7418AI 评分4242 谷歌 Gemini Omni Flash 视频编辑测试。 你们应该能猜到我原始视频是在哪儿录的,反正效果远不如 SeeDance 2.0 Video
引用歸藏(guizang.ai) (@op7418)@op7418哇! 谷歌新视频模型 Gemini Omni Flash 已经上线 FLow
@kimmonismus@kimmonismusAI 评分88 天哪,为什么大家觉得这是假的。我在 Google I/O 现场,Logan 显然就在开发者们身边 :D



@frxiaobei@frxiaobei精选AI 评分7373
引用Andrej Karpathy (@karpathy)@karpathyPersonal update: I've joined Anthropic. I think the next few years at the frontier of LLMs will be especially formative. I am very excited to join the team here and get back to R&D. I remain deeply passionate about education and plan to resume my work on it in time.
推荐理由:Karpathy 亲自宣布加入 Anthropic 并回归研发,读者可据此了解一线研究者在前沿实验室间的最新流向。
@op7418@op7418AI 评分44 @berryxia@berryxiaAI 评分2626 斯坦福数学家 George Pólya 用40年观察发现,聪明学生卡在难题上并非因为笨,而是没人教他们动手前该做什么——问题一出现就焦虑地立刻开算,越努力越偏。
引用Dr.Xiao.AI (@xiaoxiao_2580)@xiaoxiao_2580A Stanford mathematician spent forty years watching one brilliant student after another crash into hard problems. Not because they weren’t smart. But because no one had ever taught them what to actually do “before” they started solving. His name was George Pólya. In 1945, he published “How to Solve It”. The book sold over a million copies and has never gone out of print. Even Marvin Minsky, who built the first neural network, said that everyone should read it. Yet most people still haven’t heard of it. What Pólya kept seeing was the same failure pattern, again and again: The moment a difficult problem appears, students get anxious and immediately start calculating. Not because calculating is the right first step, but because doing somethingfeels much more comfortable than sitting with “I don’t know.” They end up working hard in completely the wrong direction. The step he found most neglected was this: Truly understand the problem first. Not just skim it. Not just think “this looks familiar.” His test was simple but ruthless: Can you restate the problem in your own words without looking at the original? If you can’t, you don’t actually understand it yet. Most people skip this step entirely. They jump straight into execution and then get stuck on a problem they never truly grasped. Pólya outlined four steps for solving problems. But in real life, the two that matter most are usually the first and the last: 1. Understand the problem deeply 2. Devise a plan (if you’re stuck, try solving a simpler version first and bring the insight back) 3. Carry out the plan 4. Look back — verify, generalize, and reflect The people who get truly good at this aren’t the ones who practice more. They’re the ones who’ve learned to slow down when every instinct is screaming at them to just start calculating — especially at the beginning, and again at the end. What struck me most after reading this: We assume hard problems are difficult because they are. Most of the time, it’s simply because we never took the time to truly understand them.
@kimmonismus@kimmonismusAI 评分1313 看看我刚遇到谁,传奇本人,@OfficialLoganK 人超好!!
引用Chubby♨️ (@kimmonismus)@kimmonismusI’ve been invited by Google to attend its annual I/O conference as part of the Builders Program, and I’m incredibly excited. It’s my first time at Google, and this time I brought a camera with me to capture the experience and create a recap video afterward. During the event, I’ll be conducting two fascinating interviews with Google employees, focusing on AI. These will be published at a later date. I’ve already met some amazing people from the community. Here’s to two unforgettable days!
@kimmonismus@kimmonismusAI 评分3737 
@frxiaobei@frxiaobeiAI 评分4040 引用Lucius (@LuciusHQ)@LuciusHQWe raised $3M to build Lucius AI - the Context Layer for Your Organization. Backed by Future Capital Discovery Fund, we’re tackling a problem we kept running into ourselves: Individuals ship 10× faster with AI. Organizations don't. Over 30% of your team's time is spent rebuilding context someone already had. It shows up everywhere a decision was already made but can't be found again - community operations, customer support, pre-sales reception, sales research, project management, internal collaboration. We're building Lucius to close that gap. Video
@p0@p0AI 评分4848 Parallel Search 现已登陆 @OpenRouter。 使用 Parallel 可获得可配置的上下文大小、域名过滤,以及干净、结构化的结果,带来最佳准确率和 token 效率。
引用OpenRouter (@OpenRouter)@OpenRouterAny tool-calling model on OpenRouter can call web search and web fetch agentically. The model decides when to search, what to search for, and how many times. We've added @p0 as a new web search provider. Learn more: openrouter.ai/announcements/…
@op7418@op7418AI 评分1717 还有 Flow 智能体模式,可以帮你构思概念和处理视频图像文件

@op7418@op7418AI 评分2020 Flow 现在功能超级强大,还内置了一系列图像和视频处理工具

@kimmonismus@kimmonismusAI 评分1010 
@OpenRouter@openrouterAI 评分5050 @OpenRouter@openrouterAI 评分5050 @OpenRouter@openrouterAI 评分4848 当你选择非原生引擎时,这两个工具都支持 allowed_domains 和 blocked_domains。 如果你在构建一个只应访问你自己的文档或可信来源的智能体,这一点至关重要。