I am absolutely stunned by how good @Muse is, it genuinely feels like living in the future. Pretty clear why Meta stock is up 10% just this week. I think Muse has surprised a lot of people. Hats off @alexandr_wang, your team has shocked the world
X
关注 AI 研究者、开发者与机构的动态
按账号或来源筛选(535)
Alexandr Wang@alexandr_wangAI 评分2525引用Comfortably Smug@ComfortablySmug
OpenAI Developers@OpenAIDevsAI 评分1919从一个想法到一款可以在 @modretro 上玩的游戏。 我们采访了 @Casey,聊了聊创意的未来:

Alexandr Wang@alexandr_wangAI 评分3434muse 能监控取消情况,帮你拿到最后一刻极难约到的预约和订位!
引用Dan Peguine@danpeguineMuse just booked for me an appointment with a very busy doctor that had no openings until June, for today. Yesterday I asked it to put a watch on the queue and book as soon as there’s a cancellation. Amazing.
fal@falAI 评分4646
OpenRouter@OpenRouter精选AI 评分6666引用Black Forest Labs@bfl_aiIntroducing FLUX 3 Image. Control every pixel. Make precise multi-turn edits without changing any other pixel. Lay out the image exactly how you want using bounding boxes. Generate in up to 4K to preserve details. Use up to 10 references to compose an image. Commercial Weights available for companies running image generation at scale. Open Weights version of FLUX 3 Image is launching in the coming weeks.
推荐理由:原文给出 FLUX 3 Image 的多轮编辑、边界框布局和 4K 输出等具体能力,并说明已上线 OpenRouter。
fal@falAI 评分5454
Higgsfield AI 🧩@higgsfieldAI 评分5353
Boris Cherny@bcherny精选AI 评分7070引用ClaudeDevs@ClaudeDevsYou can now mod Claude Code: - Change how it behaves - Customize the UI - Swap in your own features Write one with a few lines of TypeScript, or have Claude build it for you. Mods ship inside plugins, so you install them with /plugin in the CLI or desktop app. A few examples:
推荐理由:原文介绍了用提示词自定义 Claude Code 行为和界面的新机制,以及通过插件安装和分享的方式。
Artificial Analysis@ArtificialAnlysAI 评分4040Artificial Analysis 的 Coding Agent Index 新增安全拒绝报告,可查看拒绝发生在任务提示词阶段还是智能体已开始工作之后,以及拒绝后切换到了哪个回退模型。

Elon Musk@elonmuskAI 评分2020引用Nate Esparza@Nate_Esparzafeel like grok bot is fast AF now
🚨 AI News | TestingCatalog@testingcatalogAI 评分6161
Elon Musk@elonmuskAI 评分3939引用matt palmer@mattypGrok Bot is now more proactive. What does that mean? Here's one example: I switched flights, but @bot noticed I didn't update my Uber reservation Bot realized, then pinged me to update the reservation *during* my connecting flight I had @starlink, so I was able to change it before I landed and save myself a headache
ARC Prize@arcprize精选AI 评分6565
推荐理由:ARC Prize 公布了 Qwen3.8-27B 的Verified 成绩与每任务成本,并指出 chat template 缺 medium 档指引可能影响表现。
Thariq@trq212AI 评分1616我通常为开发者写作,但下一篇帖子我打算面向那些试图应对 AI 编程智能体变革的公司领导者。 你希望你的领导层能理解智能体的哪些方面?或者如果你正在经营一家公司——你遇到了什么问题?
Andrew Milich@milichabAI 评分3030引用XChat@chatYour group chat just got smarter Get answers to your questions directly in XChat by asking Grok
Emad@EMostaqueAI 评分5454引用Tavus@tavusIntroducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
Andrew Milich@milichabAI 评分3535如果你在使用 Gemini Enterprise Agent Platform,Grok 4.7 现已可用
引用SpaceXAI@SpaceXAIGrok 4.7 is now available on the Gemini Enterprise Agent Platform
SpaceXAI@SpaceXAIAI 评分2828Grok 4.7 现已在 Gemini Enterprise Agent Platform 上线

Chubby♨️@kimmonismusAI 评分4242引用Chetaslua@chetaslua🚨 Fable 5.5 is auto routing on web , this is the screenshot it edited for x without even prompted he knows tibo check https://claude.ai see if you are getting routed or not
Peter Steinberger 🦞@steipeteAI 评分3131
Sam Altman@samaAI 评分4444引用Vignesh@vigysoI am going to try and answer a bunch of questions that people have raised about Sign in with ChatGPT in the next couple of days. Before that, I want to let you all in to why we shipped Sign in with ChatGPT (SIWC) in the first place. Clarify our intentions here before getting to the specifics.
ARC Prize@arcprizeAI 评分5050
fofr@fofrAIAI 评分3636引用Max Sapo@MaxHappyverseWe've just sent @Google #TPU v6e to space (with @SpaceX falcon rocket) 🚀🤙
elvis@omarsar0AI 评分4848引用Soohyun Bae@RealSoohyunBaeYour AI voice sounds human. So why can't it say your product's name? A great AI voice reads "Porsche Taycan" as TAY-can. Porsche says TIE-kahn. It guessed from the spelling, and nobody caught it, because nobody listens to line 1,200. Today we're launching Onepin: the production step after text-to-speech. It checks every line of voiceover before it ships using the voices you already work with. Onepin can: ➤ Check people's and product names against a 4-million-word pronunciation dictionary ➤ Spell out prices and dates before the voice speaks ➤ Score every line of audio for naturalness, clarity and word accuracy ➤ Fix the one wrong word in the same voice, without re-rendering the take Works with your voice subscription on @ElevenLabs, @OpenAI, @Google and 30+ more. No phonetic spellings to type. No re-rolls. No switching providers. Free to start, no credit card required. Hear the before and after in the thread ⬇️
Karina@karinanguyenAI 评分3434
引用Thoughtful@thoughtfullabPostTrainBench v1.2 is out! A few updates: 1. Cloud GPU support. You can now run the benchmark with identical settings through Harbor + Modal using our new Harbor adapter. 2. New leaderboard leaders. Fable 5.1 takes #1 at 44.6%, followed by Opus 5.5 at 43.8% and GPT-6 (Astra) at 41.9%. 3. Evaluation fixes. Removed BFCL, fixed HumanEval and remote-code scoring, added averaging across multiple seeds, and switched contamination checks to majority vote.
Michael Truell@mntruellAI 评分2828引用Grok Bot@botGrok Bot can now suggest ways to help without you needing to ask.
Thariq@trq212AI 评分6363引用ClaudeDevs@ClaudeDevsYou can now mod Claude Code: - Change how it behaves - Customize the UI - Swap in your own features Write one with a few lines of TypeScript, or have Claude build it for you. Mods ship inside plugins, so you install them with /plugin in the CLI or desktop app. A few examples:
Arena.ai@arenaAI 评分6060
引用Xiaomi MiMo@XiaomiMiMoIntroducing Xiaomi MiMo-V2.6 — Pro & Flash. Frontier intelligence, all the modalities, built in public. 🔹 Two omnimodal models, advancing through scaled reinforcement learning 🔹 Pro performs on par with Claude Opus 5 and GPT-5.6 Sol across most agent benchmarks 🔹 Pro scores 46 on the Artificial Analysis Intelligence Index — the highest among open-source models 🔹 Stronger coding, computer use, 3D reasoning and creative capabilities 🔹 Open model weights, technical report, RL environments and training code Blog:https://mimo.xiaomi.com/mimo-v2-6
Anthropic@AnthropicAIAI 评分4242
🚨 AI News | TestingCatalog@testingcatalogAI 评分4141
elvis@omarsar0AI 评分4848
引用David Stout@DavidstoutHalf a million downloads in a month. Today, our open source family takes another step forward. Thank you for the incredible support behind our first-generation models. We’re excited to introduce TwIL-LM3-Pro. At just 3.6 billion parameters, it brings powerful reasoning to everyday computers, with quantized builds that run locally. No cloud required. In our evaluation: Formal logic: Highest recorded headline score among the small models compared—beating China’s VibeThinker-3B by 35% and Qwen3.5-4B by 24%, and Liquid AI’s LFM2.5-8B-A1B by 47%. Broader reasoning: 95% on SVAMP and 64.1% on MuSR, the highest recorded scores among the small models compared. BIG-Bench Hard’s logic subset: 95.4%, compared with VibeThinker-3B’s 61.1%. We believe AI is entering a post-training era. The advantage will increasingly belong to companies with the best pipelines and those that can produce capable, personalized intelligence faster and more efficiently, then put it on devices people already own. That’s what we’re building at webAI. And we’re only beginning to share what’s coming out of our lab. Coming soon: Meridian, our family of frontier-class models built to run on device. Our most advanced models will be available through the @thewebAI application. Join the waitlist as we expand access. Proudly built in Austin, Texas. 🇺🇸
gabriel@gabriel1AI 评分2222hey granola,你们在 sol 6.1 发布一天后才加上 sol 6,这是错的模型
引用gabriel@gabriel1hey @meetgranola can you please stop defaulting to GPT 5.6 terra non-thinking?? what is even this list of options, why would i ever use gpt 5.4 or sonnet 5
François Chollet@fcholletAI 评分5252
🚨 AI News | TestingCatalog@testingcatalogAI 评分5858Grok 4.7 正在 Grok 网页端和移动应用上线,成为所有模式下的基础模型,包括 Fast、Expert、Build 和 Heavy。作者还预计其每日 AI 简报内容将随之改善。
引用Lumina@LuminaBench🚨 Grok 4.7 is now in app This took way too long
Peter Steinberger 🦞@steipeteAI 评分2323引用yingchao@baggiiiie@GergelyOrosz at least we know it thinks it's part of the gpt family and somehow coderabbit once 😆
🚨 AI News | TestingCatalog@testingcatalogAI 评分4848
引用Mac Liu@themacliuI’m excited to announce that @arceuslegal is launching with $17M in funding, led by @greycroftvc, with participation from @craft_ventures, @spc, and others. As a founder, I always hated how helpless I felt working with law firms. I went through four or five different firms and somehow the experience was always the same. I’d be waiting on something important to our business with no idea when I’d hear back. I’d have to re-explain our business over and over again. And I dreaded jumping on calls because I knew every minute was costing me money. We started Arceus because we believe every business deserves a better law firm. One that moves faster, costs less, and puts the client first. And we’re just getting started. ↓
OpenRouter@OpenRouterAI 评分4848
🚨 AI News | TestingCatalog@testingcatalogAI 评分5353Tavus 推出 video-to-video 模型 Griffin 的研究预览版,用于实时面对面视频对话。模型可观察和聆听对话,支持被打断或在有人进入画面时作出反应。
引用Tavus@tavusIntroducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
elvis@omarsar0AI 评分5858引用Tavus@tavusIntroducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
🚨 AI News | TestingCatalog@testingcatalogAI 评分4242
引用Microsoft AI@MicrosoftAIIntroducing 3 new models: MAI-Transcribe-2-Streaming, MAI-Voice-2.1 and MAI-Voice-2.1-Flash. Accurate streaming transcription. Natural speech and less waiting between turns. Build voice agents that keep the conversation moving!