

@kimmonismus · X


Reset will land around 14pm PST tomorrow.


推荐理由:报道给出的资金与团队规模,让读者能看清英伟达自研开放权重模型的投入量级及其双线竞争位置。
No way! @synthwavedd says that the secret Model "Ox Alpha" on OpenRouter is the upcoming GLM 5.3 Flash A Flash model (!!) outperforming even the best models like Fable and GPT Sol?! If this turns out to be true, zAI's post training put western models at shame Oh, and looks like Kimi k3.1 also incoming What is happening right now. unbevlieable
差不多就是这样。我真的不知道 Claude 怎么了,但它说话的方式太绕、太长、太啰嗦,简直没意思。https://t.co/b3sQ7Ec4JE
他言出必行。为我们所有人存下了重置额度 https://t.co/wsrm4XdDzj
The banked reset will be there by 8pm PST. For all paid users of ChatGPT Work and Codex. Do with this information what you may.
As we continue to push the frontier of capabilities while improving efficiency, we're dropping API and credit pricing of GPT-5.6 Sol by over 20% for the next 3 months. https://t.co/UoTb3hcB2t
No way! @synthwavedd says that the secret Model "Ox Alpha" on OpenRouter is the upcoming GLM 5.3 Flash A Flash model (!!) outperforming even the best models like Fable and GPT Sol?! If this turns out to be true, zAI's post training put western models at shame Oh, and looks like Kimi k3.1 also incoming What is happening right now. unbevlieable
@synthwavedd 我的意思是,你明白这意味着什么吗?我刚写了关于这个的内容:https://t.co/NmOCZ48eIe
If it's true that this is indeed GLM-5.4/5.5, then it would change everything, without exaggeration. GLM-5.3 was released just 7 days ago and was an extremely significant leap compared to GLM-5.2, which was improved solely through real-time modeling (RL). Same base model. And all this in a very short time. If it's true that GLM-5.4 has become so much better just a week later thanks to RL, it would demonstrate: 1) how much faster the models are now becoming. Not only is there no end in sight, but: now more than ever, exponential growth. 2) It would force OpenAI and Anthropic to release models. Anthropic, in particular, with its upcoming IPO, now has to prove itself. And it would put pressure on slowing down in favor of security. 3) And at least as importantly: the gap between China and the US is shrinking even further, even faster. It seems to be generally accepted that this is a Chinese model. No one suspects it's a Google model. Given the same tokenizer, it's most likely GLM, and that would be the craziest thing we've seen in a long time for the reasons mentioned above. But perhaps it's MiMo. Or, quite far-fetched, Ilya Sutskever's SSI model. It remains exciting. I've rarely seen the community so impressed and confused at the same time. I'm equally confused and impressed. @davis7 used Fable to determine which model it most closely resembles. A screenshot of the test is attached. h/t Ben Davis
@synthwavedd 我们怎么走到这一步的?Google 的 Gemini Flash 现在跟中国的 Flash 比都差远了 o.O


It's me again. I come bearing great news. First of all, we have hit 20M active users for Codex some time this week. Second of all, this is cause for celebration and during the day we will credit every Codex and ChatGPT Work user with a BANKED reset that you can use at your own leisure. And we will have some other good news later too! Now, on usage limits draining faster, while we're not seeing anything abnormal, we do take it incredibly seriously and there is an ongoing investigation. I will share if we do find anything and my below post is really a clarification on a specific pattern that we did see that I wanted to call out. Go do something amazing today.
We've investigated a few messages about codex usage limits being different. That's not something we change without engaging the community and being transparent. What we did see is that when talking to affected users many were using sub2api. Converting a subscription into api traffic to then re-serve or share across many users is not something we support and this type of usage gets flagged by our fraud-prevention systems. You are completely fine if you use your subscription through Sign in With ChatGPT, either through the official clients or through one of the many OSS clients (Pi, OpenCode, ...) that support signing in with your account and using your included usage.
DeepSeek-V4-Flash-Vision-Exp is now live on the DeepSeek API Platform! 🚀 🔹 This experimental multimodal model matches DeepSeek-V4-Flash on text capabilities—including agents, reasoning, and world knowledge. 🔹 On multimodal agent benchmarks, V4-Flash-Vision-Exp makes a major leap over V4-Flash, bringing multimodal agent performance close to Opus-4.8. Try it with model='deepseek-v4-flash-vision-exp'. DeepSeek Harness 0.1.1 was released today with out-of-the-box support for the new model. 1/n
推荐理由:基准对比显示 Flash 小模型在多模态智能体评测上已接近 Opus-4.8,可供判断轻量模型的能力边界。
If it's true that this is indeed GLM-5.4/5.5, then it would change everything, without exaggeration. GLM-5.3 was released just 7 days ago and was an extremely significant leap compared to GLM-5.2, which was improved solely through real-time modeling (RL). Same base model. And all this in a very short time. If it's true that GLM-5.4 has become so much better just a week later thanks to RL, it would demonstrate: 1) how much faster the models are now becoming. Not only is there no end in sight, but: now more than ever, exponential growth. 2) It would force OpenAI and Anthropic to release models. Anthropic, in particular, with its upcoming IPO, now has to prove itself. And it would put pressure on slowing down in favor of security. 3) And at least as importantly: the gap between China and the US is shrinking even further, even faster. It seems to be generally accepted that this is a Chinese model. No one suspects it's a Google model. Given the same tokenizer, it's most likely GLM, and that would be the craziest thing we've seen in a long time for the reasons mentioned above. But perhaps it's MiMo. Or, quite far-fetched, Ilya Sutskever's SSI model. It remains exciting. I've rarely seen the community so impressed and confused at the same time. I'm equally confused and impressed. @davis7 used Fable to determine which model it most closely resembles. A screenshot of the test is attached. h/t Ben Davis
这东西太离谱了:https://t.co/lJASYQSeuL
Wtf! Ben ran this mystery model through 10 DeepSWE tasks and it scored over 80%, versus 65% for Fable and 52% for GPT-5.6-sol! this is insane. Probably a chinese company. Either a new GLM or Kimi model, I reckon.
Ox Alpha (stealth model) is free for the next week - 1M Context - Multi-modal - Zero Data Retention Generous rate limits, near unlimited usage We have capacity for 100T tokens per day, lets see what you can do




Claude Academy is now live. Whether you're figuring out what AI is or already using Claude every day, there's a path that meets you where you are. The courses and tutorials are free and open to anyone at https://t.co/WRiSRAvK6l https://t.co/llx2W0VIIy
推荐理由:报道给出 Anthropic IPO 的递交时点与募资对标的参照,读者可据此判断其上市规模量级。