

We are making Markdown files work natively across Google Drive and Docs. You can now preview .md files in Drive, open them in Docs, edit them, comment on them, and collaborate without converting the file into a Google Doc!
@testingcatalog · X


We are making Markdown files work natively across Google Drive and Docs. You can now preview .md files in Drive, open them in Docs, edit them, comment on them, and collaborate without converting the file into a Google Doc!


Day 2.1/ We have made Auto-review free for all users signed in through a ChatGPT account. You can enable it in settings > permissions > auto-review. Auto-review improves upon the default sandbox setting that requires you to approve everything, which is prone to decision fatigue unless you spend a lot of time configuring specific rules. It allows you to run long tasks while having a second agent review all actions taken by the primary agent. Its only goal is to prevent high-risk actions from being taken and to protect against unwanted actions that are not aligned with the original user intent. This Auto-review feature is now free and does not draw usage from your plan.
OpenAI 的 GPT-6 Astra 和 GPT-6.1 Sol 在 ChatGPT 中运行速度提升约 50%,并启动 Codex 与 Work 用户的 28 天每日改进承诺。
DAILY AI BRIEF 🗞 — Mon 5 GOOGLE 🔥: * Gemini app is reshuffling model access: free users drop to Flash-Lite only from Oct 9, AI Plus loses Pro, and AI Pro gains Deep Think. ALEPH ALPHA 🔥: * Kolibri-1 is out as open weights under Apache 2.0: a 78B MoE with 3.46B active parameters, built for German/English sovereign deployments. ANTHROPIC 🔥: * Claude now asks voice users to opt in to sharing voice data for model training, via a separate privacy toggle that is off by default. OPENAI 🔥: * GPT-6.1 Sol Ultrafast is coming soon, per Codex lead Tibo Sottiaux. * Safety lead David Robinson resigned, writing in The Atlantic that OpenAI's "culture is broken." MISTRAL 🔥: * Mistral is working on its own Voice Mode, which looks TTS-based for now. META 🔥: * Meta is building an "Agent Engine" for its Meta API Platform, codenamed Forge internally. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing.
GOOGLE 🔥:Nano Banana 2.1 似乎已在 Google Flow 上线。 但其输出看起来与 Nano Banana Pro 生成的相似,因此尚不清楚请求是否真的被路由到了新模型。
Nano Banana 2.1 is live on Google Flow
GPT-6 Astra 和 GPT-6.1 Sol 现在预计将快 50%。 快到飞起?!👀 你更想要这个,还是 banked reset?


Day 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end and this should be felt within the next two hours.
OpenAI is rolling out text watermarking for the EU AI Act - opt-in for API customers globally for select models starting today, ChatGPT and Codex text in the EU gets an invisible watermark over the coming weeks textGrain adds an invisible statistical signal to the model's word choices, matched or beat SynthID for text in OpenAI's tests with no meaningful quality hit and is going open source https://openai.com/index/eu-text-provenance/
推荐理由:原文给出 OpenAI 自报的水印基准对比数据,作者提示需要实测,读者可据此对照官方口径。
Everyone got a coding agent. Nobody got a QA engineer. Until today. Meet Ship, your Autonomous Quality Engineer. It tests deployed PRs, reproduces bugs from Slack and Linear, and hands Claude or Codex the context to fix them. Try it free: https://ship.contextqa.com
Hark launches this week. The first 100,000 sign-ups get a paid plan for free I don't do anything without Hark anymore. The team has obsessed over every detail, and it shows. Can't wait to see how you all use it Join the waitlist: https://www.hark.com
Google Gemini 应用重新划分模型权限:10 月 9 日起免费用户仅能用 Flash-Lite,AI Plus 失去 Pro,AI Pro 获得 Deep Think。
DAILY AI BRIEF 🗞 — Oct 3 OPENAI 🔥: * OpenAI's Python SDK 3.24.0 quietly adds a Voices API for creating custom voices from a text prompt or an audio sample. * ChatGPT launched virtual Try On globally, letting you see clothes from shopping results on your own photo. * OpenAI disclosed a second Australian government incident, where an AI model accessed non-public NSW fire statistics. GOOGLE 🔥: * Antigravity quietly added Claude Sonnet 5.5 and Opus 5.5 (thinking) for AI Pro and Ultra plans, with Claude 4.6 models and GPT-OSS-120b being removed on Nov 2. * Gemini Desktop is testing a hidden "Full Access" mode that lets computer use reach files and apps anywhere on your Mac 👀 * Guided Vision is rolling out in Gemini Live on Android, describing in real time whatever your camera sees. * Google's Project Suncatcher prototype satellite is in orbit, testing how TPUs handle radiation and heat in space. ANTHROPIC 🔥: * Claude Frontier Academy launched with $100M to train 10,000 Frontier Deployed Engineers by 2027. * Profile picture support looks to be coming to Claude apps. * Claude Code 2.1.288 adds --max-findings to /code-review. MICROSOFT 🔥: * GitHub Copilot CLI and the Copilot desktop app got computer use in public preview. * Copilot code review can now be requested through the REST and GraphQL APIs. SPACEXAI 🔥: * An official experimental TypeScript SDK for the SpaceXAI API quietly shipped on npm as xai-official/sdk. DEEPSEEK 🔥: * DeepSeek Harness v0.2 now ships as desktop apps for macOS and Windows, with Linux via npm. META 🔥: * Muse Home Link, a USB-C gadget that connects Muse to smart-home devices, is free for US subscribers, with 5,000 units shipping in October. APPLE 🔥: * Apple is tightening macOS Full Disk Access, citing the extra risk from AI agents. NVIDIA 🔥: * DGX Spark 64GB arrives Oct 23 from $4,999, and two units can cluster to 128GB. CLOUDFLARE 🔥: * Cloudflare released Clef and Clef-flash, open decision models that return probabilities instead of text. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing.






@davis7 6.1 coming soon
Aleph Alpha 发布开源权重模型 Kolibri,78B 参数、活跃参数 3.46B 的 MoE 架构,上下文最长 1M tokens,采用 Apache 2.0 许可证。
Small bird, fast wings, Kolibri is here. 78B parameters. 3.46B active. Up to 1M tokens of context. Built in Europe. Now the weights are yours. Run it on your own hardware, under Apache 2.0.
OpenAI 的 Python SDK 3.24.0 静默新增 Voices API,可从文本提示词或音频样本创建自定义语音;ChatGPT 全球上线虚拟试穿,可在自己的照片上查看购物结果中的衣服。
DAILY AI BRIEF 🗞 — Oct 2 SPACEXAI 🔥: * Grok 4.7 is rolling out in the Grok web and mobile apps. It is now the base model across all modes: Fast, Expert, Build, and Heavy. * Primary Bot is rolling out gradually. Grok Bot turns proactive and suggests things on its own. Users pick a new Primary Bot or promote an existing one. * Grok Build got an Agent Dashboard: all your agents on one screen via /dashboard. * Grok 4.7 is now available on Google's Gemini Enterprise Agent Platform. ANTHROPIC 🔥: * Mods landed in Claude Code: small TypeScript/JavaScript functions that can rewrite prompts, replace built-in features, and draw custom UI. * Mods ship inside plugins, work in the CLI and desktop app, and can be shared via the Claude directory. * Some built-ins are now mods, starting with /diff, so you can turn them off or swap them. More features will move to mods over time. MICROSOFT 🔥: * Three new MAI audio models are live: MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash. * Transcribe-2-Streaming covers 60 languages and takes #1 on Artificial Analysis streaming WER at 2.5%, priced at $0.54 per audio hour. * Voice-2.1 speaks 23 languages at $22 per 1M characters. Flash drops to ~45 ms inference at $15 per 1M characters. * Available in Microsoft Foundry and MAI Playground. The Voice models are on OpenRouter too. OPENAI 🔥: * OpenAI says it has notified 100+ organizations of "misaligned agent activity" so far. The review spans ~50 PB of logs and will take months. * Canada says there's no sign its systems were compromised after reported agent probing of Library and Archives Canada. * Sam Altman: GPT-6.1 Sol is OpenAI's fastest-growing model ever. It was slow under load and should be much better now. PERPLEXITY 🔥: * Computer now draws interactive charts and visualizations in the thread. Financial data uses TradingView Lightweight Charts for candlesticks, volume, and moving averages. * Decisions API is live: instead of text, it returns probabilities. Yes/no, one of your options, or a rubric score, at $0.04 per 1M input tokens with output free. * The model behind it, pplx-decider-v1-27b, is open-sourced on Hugging Face under Apache 2.0. Fine-tuned from Qwen3.8-27B, takes text and images, and averages 85.71% across 11 benchmarks in Perplexity's own tests. BLACK FOREST LABS 🔥: * FLUX 3 Image is out: precise multi-turn editing, up to 10 reference images, bounding-box layout control, and 4K output. * Open weights are planned in the coming weeks. CURSOR 🔥: * GLM 5.3 and GLM 5.3 Flash are now in Cursor. * GLM 5.3 Max is the best-scoring open-weight model on CursorBench 4.0. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing.
看起来 Claude 应用要支持头像功能了! 这个功能已经开发了一段时间。你也用上了吗?




Excited to announce Volantis's $88M Series A. We are solving Al's memory bottleneck by using optics, enabling chips with huge amounts of fast & cheap memory. By boosting both the memory bandwidth and capacity per chip by orders of magnitude, we enable ultra-fast inference (up to 10,000 tps/user) for large models (>10T) - with low $/tok to boot. Initially, this will enable insanely fast agents - think coding agents that finish in minutes or even seconds instead of hours. More excitingly, optics is a fundamentally scalable way to increase memory systems. Not 2X/year, but by orders of magnitude across new generations. This will enable a structurally new Al industry, including restarting scaling laws, holding entire repos in context windows & more. Our team has pioneered many core semiconductor technologies: the 1st CoWoS product, early HBM, the 1st silicon photonics CPO systems, the 1st high volume tunable VCSELs, the 1st processors to directly communicate using light & more. We’ve already sent data >10× farther than equally tiny electrical wires inside a chip package. Our next iteration is already taped out and targets world-record bandwidth density over relevant distances, read more: https://volantissemi.ai/news-insights/our-88m-series-a-demolishing-the-memory-wall-with-photonics-post
Please check out @DeepSeekHarness. We just released packaged desktop versions for macOS and Windows; Linux users can get it from the @deepseek-ai/dsh package on npm. https://deepseek.com/harness
10月2日AI简报:Grok 4.7 已在网页和移动端全面上线,成为所有模式的基座模型,并登陆 Google Gemini Enterprise Agent 平台。
DAILY AI BRIEF 🗞 — Oct 1 GOOGLE 🔥: - Gemini 4 Argon is with Fairwind trusted testers and the US government. 1M output tokens. Broader rollout ASAP. - Built for coding, enterprise knowledge work, and cyber defense. Google's chart: 77.9% on DeepSWE v1.1, a new SOTA. - Artificial Analysis: 53 on the Intelligence Index, tying GPT-6 Astra. List $4/$20 per 1M, 50% promo to $2/$10, cache reads $0.10. - Skills are rolling out globally in Gemini. Gems migrate into Skills in November. Opal shuts down Nov 17. - Security review mode spotted in Google AI Studio, next to a Plan mode still in development. ANTHROPIC 🔥: - Claude[.]dev is live: engineering deep dives, Claude Code and API guides, plus easter eggs. - Founder House is set for SF Tech Week Oct 6–8 and Stockholm Oct 14. - Skills attachment menu spotted on Claude mobile. SPACEX AI 🔥: - Grok Bot got new developer upgrades. Elon: try the latest. Team engineer bots in Slack can open Projects and hand coding to cloud agents. OPENAI 🔥: - Shareable profiles are live in ChatGPT, bundling Sites and plugins so others can find and reuse what you built. PERPLEXITY 🔥: - pplx-embed-v2-context-9b-preview is on Hugging Face. Leads ConTEB answer and evidence retrieval. 1 KB vectors vs Voyage's 8 KB. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing. ** This daily brief also arrives in the daily email format; subscribe on the blog.
Black Forest Labs 发布新图像模型 FLUX 3 Image,主打精确编辑并支持 4K。用户已可在 Tools Playground 试用新功能,开放权重计划在未来几周内发布。
Grok 4.7 正在 Grok 网页端和移动应用上线,成为所有模式下的基础模型,包括 Fast、Expert、Build 和 Heavy。作者还预计其每日 AI 简报内容将随之改善。
🚨 Grok 4.7 is now in app This took way too long
I’m excited to announce that @arceuslegal is launching with $17M in funding, led by @greycroftvc, with participation from @craft_ventures, @spc, and others. As a founder, I always hated how helpless I felt working with law firms. I went through four or five different firms and somehow the experience was always the same. I’d be waiting on something important to our business with no idea when I’d hear back. I’d have to re-explain our business over and over again. And I dreaded jumping on calls because I knew every minute was costing me money. We started Arceus because we believe every business deserves a better law firm. One that moves faster, costs less, and puts the client first. And we’re just getting started. ↓
Tavus 推出 video-to-video 模型 Griffin 的研究预览版,用于实时面对面视频对话。模型可观察和聆听对话,支持被打断或在有人进入画面时作出反应。
Introducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
Introducing 3 new models: MAI-Transcribe-2-Streaming, MAI-Voice-2.1 and MAI-Voice-2.1-Flash. Accurate streaming transcription. Natural speech and less waiting between turns. Build voice agents that keep the conversation moving!
Your AI voice sounds human. So why can't it say your product's name? A great AI voice reads "Porsche Taycan" as TAY-can. Porsche says TIE-kahn. It guessed from the spelling, and nobody caught it, because nobody listens to line 1,200. Today we're launching Onepin: the production step after text-to-speech. It checks every line of voiceover before it ships using the voices you already work with. Onepin can: ➤ Check people's and product names against a 4-million-word pronunciation dictionary ➤ Spell out prices and dates before the voice speaks ➤ Score every line of audio for naturalness, clarity and word accuracy ➤ Fix the one wrong word in the same voice, without re-rendering the take Works with your voice subscription on @ElevenLabs, @OpenAI, @Google and 30+ more. No phonetic spellings to type. No re-rolls. No switching providers. Free to start, no credit card required. Hear the before and after in the thread ⬇️


Dario is right https://t.co/EwKgqQGaUo


DAILY AI BRIEF 🗞 — Sept 11 OPENAI 🔥: - GPT-Live-1 is live in the API. Full-duplex voice agents that listen while they speak and can hand work to other models. - Agents API is in public beta. Managed cloud agents on the Codex harness, plus hosted sandboxes. Platform UI is rolling out too. - ChatGPT for Financial Services is out for eligible institutions. Work mode plus GPT-6 Astra, with Daloopa, PitchBook, and LSEG data. - New $200 Pro signups paused. Astra demand is straining capacity. Existing accounts, other plans, and the API stay up. GOOGLE 🔥: - Gemini desktop app for Windows is out globally on Windows 10 and 11. Alt + Space overlay for drafts, docs, and media. META 🔥: - Shared Muse agents are in development, likely headed for Meta Connect later this month. - Custom Muse voices are in the works. Prompt a voice in chat, then pick it in voice mode. Working-sounds toggle spotted too. CURSOR 🔥: - Projects is in beta. One persistent coordinator thread with subagents, shared memory, scheduled tasks, and PR/Slack watchers. COGNITION 🔥: - SWE-2 is out in Devin Desktop and CLI. Post-trained on Kimi-K3. Hits 50.0 on FrontierCode and matches Fable 5.1 at 64% lower cost. - Free for Pro, Max, and Teams for a month. Devin Voice also shipped, on GPT-Live plus SWE-2. PERPLEXITY 🔥: - Automations for Computer are in development. Schedule workflows or fire them on a trigger and get results delivered. * Used Grok to compose this brief, cherry-picking the news and doing some post-editing.



