| 1. Claude Fable 5 (High) |
| 2. Claude Opus 4.8 (Thinking) |
| 3. GPT 5.6 Sol (xHigh) |
| 4. Kimi K3 |
| 5. Claude Sonnet 5 (High) |
The BriefThree days after Moonshot shipped the 2.8-trillion-parameter Kimi K3 as open weights, Alibaba answered with Qwen3.8-Max, a 2.4-trillion-parameter multimodal model it says is second only to Claude Fable 5. The preview is live through Alibaba's Token Plan, Qoder, and QoderWork, and the company says open weights are coming, though there is no license, model card, or full benchmark table yet. Treat the ranking as a vendor claim until third-party evals land. The bigger story is the cadence: China's labs are now trading trillion-parameter open-weight releases on a weekly clock, and frontier-adjacent capability you can host yourself is becoming a quarterly expectation rather than a yearly surprise.
Level UpA new large-scale study (4,100 participants) measured how convincing AI-generated voice phishing has become using off-the-shelf voice models. If you run a help desk or answer customer calls, the move this week is to script one internal vishing drill: have a teammate place a scripted AI-style pretext call to your own intake line and see what gets through. One drill will teach your team more about the real exposure than any policy memo. Read the study, then run one vishing drill on your own intake line