聊聊安全话题:GLM5.3 无审查版已上线 HuggingFace。好戏在后头。
here u go https://huggingface.co/Infatoshi/GLM-5.3-UNCENSORED-EXL3-3.0bpw
@kimmonismus · X
聊聊安全话题:GLM5.3 无审查版已上线 HuggingFace。好戏在后头。
here u go https://huggingface.co/Infatoshi/GLM-5.3-UNCENSORED-EXL3-3.0bpw
Half a million downloads in a month. Today, our open source family takes another step forward. Thank you for the incredible support behind our first-generation models. We’re excited to introduce TwIL-LM3-Pro. At just 3.6 billion parameters, it brings powerful reasoning to everyday computers, with quantized builds that run locally. No cloud required. In our evaluation: Formal logic: Highest recorded headline score among the small models compared—beating China’s VibeThinker-3B by 35% and Qwen3.5-4B by 24%, and Liquid AI’s LFM2.5-8B-A1B by 47%. Broader reasoning: 95% on SVAMP and 64.1% on MuSR, the highest recorded scores among the small models compared. BIG-Bench Hard’s logic subset: 95.4%, compared with VibeThinker-3B’s 61.1%. We believe AI is entering a post-training era. The advantage will increasingly belong to companies with the best pipelines and those that can produce capable, personalized intelligence faster and more efficiently, then put it on devices people already own. That’s what we’re building at webAI. And we’re only beginning to share what’s coming out of our lab. Coming soon: Meridian, our family of frontier-class models built to run on device. Our most advanced models will be available through the @thewebAI application. Join the waitlist as we expand access. Proudly built in Austin, Texas. 🇺🇸


Global reset landing tomorrow 10am PST for all paid ChatGPT accounts. Apologies for the slow start with GPT-6.1 Sol, it's now back to running at expected speeds after the massive load spike in the first two days.
Because usage on your primary dot is virtually unlimited at the moment, I can’t really give a reset. I need to come up with something new fast.
As a European, I don't have access to OpenAI's dot anyway, but: somehow nobody's talking about it. My entire "for you" page is empty. It seems like nobody's really excited about it. Or is that just how it appears?
作为一个欧洲人,我反正用不了OpenAI的那个点,但是:不知为何没人在讨论它。 我的整个“为你推荐”页面都是空的。看起来没人对它真正感到兴奋。还是说只是看起来如此?
🚨 Fable 5.5 is auto routing on web , this is the screenshot it edited for x without even prompted he knows tibo check https://claude.ai see if you are getting routed or not
推荐理由:原文同时给出 WSJ 报道、OpenAI 官方回应和 GPT-6.1 Astra 取消的背景,读者可以看清安全团队人事变动的完整脉络。
This @Bloomberg story is BS. Argon has been my daily driver for a while and it's been a great experience. It's particularly awesome at agentic debugging besides day-to-day coding tasks.
Claude Opus 5.5 has taken the top spot on the Epoch Capabilities Index (ECI) with a score of 167, narrowly ahead of GPT-6 Astra. Claude Sonnet 5.5 has roughly matched Claude Fable 5.1 (165).
下一站:旧金山。@salesforce Dreamforce 见,朋友们!https://t.co/dEeHGucbMr
我理解Wang的帖子是在说,对齐和加速可以同时发生,不一定需要减速。如果我理解正确的话,这对Meta来说是一个巨大的胜利。https://t.co/lRse16HEaL
Alignment is fundamental to delivering personal superintelligence for everyone. People need agents they can trust to reliably do what they ask. MSL is rapidly scaling up the share of our efforts that goes into alignment as our models become more powerful. We do believe alignment can be the gating factor for scaling as we get closer to the frontier.
Dario Amodei:“这项技术正处于指数级发展轨道上。” 依然看不到任何瓶颈。他们担心它正在脱离掌控。但与此同时,丰裕时代已触手可及。https://t.co/PaWhFyN94X
所以事实上这不仅仅是传闻。正如 Dario Amodei 所说,递归自我改进目前正在行业中发生。 而这大概就是他们出于恐惧呼吁放缓的原因。
Rumors are spreading like a wildfire that Google DeepMind has reached RSI. Lyra is part of the reliable and huge leaker community. But Id say its more than just rumors: Google says Demis Hassabis will focus his “full attention” on shaping AGI. Reuters reports in August Sergey Brin is directing resources toward RSI, while DeepMind’s strategy chief calls it key to the AI investment thesis.
I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon. https://t.co/1YhhIybZX7
Kim(@kimmonismus)表示自己没想到马斯克会支持 Amodei 关于放缓 AI 发展的呼吁,并附上马斯克称 Dario 说得对的推文作为引用。
Dario is right https://t.co/EwKgqQGaUo
我不想搞阴谋论。但时机确实蹊跷,就在 Jacob Coxon 因对 AI 与人类灭绝的担忧走红几天后。 而偏偏就在此时,Amodei 呼吁放缓。
The urgent call for slow down officially started. Dario Amodei demands and suggest a plan for unified slow-down. Very sad to read that. At the same time, he expects abundance and cure for most disease in the next 5-10 years. „I believe that AI could cure most major diseases in the next 5-10 years, greatly accelerate economic growth rates, create a world of abundance and empowerment, and usher in a renaissance of democracy and freedom.“


The urgent call for slow down officially started. Dario Amodei demands and suggest a plan for unified slow-down. Very sad to read that. At the same time, he expects abundance and cure for most disease in the next 5-10 years. „I believe that AI could cure most major diseases in the next 5-10 years, greatly accelerate economic growth rates, create a world of abundance and empowerment, and usher in a renaissance of democracy and freedom.“


这简直蠢得离谱。40位"国会议员",即英国下议院民选议员,要求禁止超级智能。 一个在AI领域毫无建树、完全落后的大陆,其政客却对AI发出最强烈的禁令和监管要求,这怎么可能? 简直荒谬。
顺便说一句,下周我会在 newsletter 里详细讲这个 https://t.co/rhX1KI3lwV ,免费订阅 :)


推荐理由:时间线把智能体批量提交行为与平台暂停注册的应对对应起来,便于了解这类安全事件的经过。
来了:又一次 codex 重置!我们的男孩从不让人失望。https://t.co/47nByG07x5
Hi Astra users. A reset and a quick update on quality issues that have been posted around. Working with some of you, we have found and fixed the following issues: - Some skills written for previous models were triggering too often or preventing the model from checking its work. - An opt-in context management experiment that could cause early stops or replies to older messages. We've disabled it. Our rough estimate is that 4-5k users were affected by this experiment. - We've also removed some badly configured engines that resulted in a measured quality degradation for a long tail of traffic flowing through them. We’ve also made some more minor improvements and things should feel significantly better across the board. More consistent follow-through, better tracking of your latest message, and better checks on the work as it’s going through the motions. The examples posted and all the users who worked directly with us were incredibly useful in helping fix things quickly. Always grateful for this incredible community. And of course, a reset is also landing by midnight today.