
Claude 开始给文字打水印:原理藏在模型犹豫的地方,而「只有欧盟在管」是个误会
Anthropic 说水印不加字符、不加 token、不加钱、读者看不出来——那它到底改了什么?答案是模型挑词的那一刻。顺便把「只有欧盟强制水印」这个印象核了一遍:中国早了 11 个月,加州跟欧盟同一天。
12 posts

Anthropic 说水印不加字符、不加 token、不加钱、读者看不出来——那它到底改了什么?答案是模型挑词的那一刻。顺便把「只有欧盟强制水印」这个印象核了一遍:中国早了 11 个月,加州跟欧盟同一天。

Kimi K3, Qwen 3.8 Max, and the GA release of DeepSeek V4 Flash all shipped within a month. GLM-5.3 and Kimi K3.1 are still rumors. I also switched my daily-driver model from Claude Code Opus to Kimi K3. This post fact-checks the wave and lays out what actually changed for me.

一篇博文说 Codex 的主 session 没有纠错的自省能力,Claude 才有,还建议 context 别超过 300k。我刚把 Kimi K3 接进自己的 Web Claude Code Pilot,正想知道它能不能当 orchestrator,就去验了一遍。k3 和 Sol 都有自省能力;那条 300k 的建议也是空的,因为 Codex 在 258.4k 就自己压缩了。而「Codex 主 session 没产出」的症状,是我自己的客户端提前关流造成的——那个 bug 只长在 Codex 那条通道上。

A leaderboard scored six Claude 'PPT skills' by feature blurb and ranked a 0-star project above one with 23,000 stars. So I gave the top tools the same real client brief and built the decks. The ranking didn't survive contact with the output.

Opus 4.8 ships with dynamic workflows that can fan out to a thousand subagents. The docs aren't joking. Here's what that actually feels like from a metered Max 5x seat — and why the most-hyped feature of the release is, for small users, a beautiful gadget I'm not going to use.

Last Friday a Cursor agent running Claude Opus 4.6 deleted a startup's production database in nine seconds. Four layers of safety should have stopped it. None of them did. Here's the chain of failures, layer by layer.

My OpenClaw agent had been silently dropping half of its Telegram responses. No error, no warning. I found the bug, and nobody was fixing it — so I did.

Anthropic launched Managed Agents yesterday. I spent the day reading, thinking, and realizing the game has shifted — from who can build the best agent infrastructure to who can design the best agent strategy. Some thoughts on speed, consulting, and the uncertain road ahead.