正文 Markdown
Fable is careful. None of the other models are careful. GPT-5.6 Sol, Opus, Kimi, Grok. You can compare them all day long on capabilities, and it doesn't matter, because in order to use a model for real production work, it must first and foremost be careful. That's the only dimension that matters to me. And if you're doing real production-facing customer-facing work with AI, it's the only thing that should matter to you, too. The other models will not be truly competitive at anything but proofs until they are trained to be careful.