Well I think the reason is that Opus has a very broad range of knowledge, you can ask it about things like extinct penguins and it will give you accurate information on that without tool usage, but models like GLM have been optimised for programming and reasoning work. Anthropic are targeting their models as generalists. I am only really interested in programming myself.
You are missing the fact that every program is solving some real-world problem, and that's hard to limit in advance. For us, it is highly beneficial that the AI understands energy markets.
This "reasoning work" optimization comes combined from large enough pre-training corpus, and the expensive post-training RL which is exactly the part GLM illicitly copied. These models are taught to think, they don't start thinking from just reading the whole internet.
Sure, if you
really are just replacing the "Indian subcontractor", for whom you used to write perfect specification documents and they just "wrote code", then sure, I agree, a significantly smaller "programming only" model will do fine. But better be conscious about this limitation.
But GLM, even the Flash version, isn't a small model. It's probably quite capable, even if you want to write some extinct penguin catalogue software. But it also won't run on that 256GB or 512GB Mac Studio in a satisfactory way for agentic software development. Probably not even the flash version (not sure).
MoE itself is a pretty strong performance optimization allowing to run a large model not that slowly, but yeah, you still need the RAM to hold the model.
And personally I don't care if these models were distilled from Claude. Anthropic didn't pay for the training data in most cases. At least the distillation users are actually paying for API usage. Anthropic can get annoyed about it if they like but I don't see a contradiction here. Model outputs are not copyrightable.
That's true - it's not a copyright infringement. Just contract breach and possibly a cyber attack issue. I'm actually, this time, worried about the exact same thing Anthropic
says they are worried about (in reality they are of course worried about their
business but that's obvious): the Chinese distillates have
already proved they happily hack into actual live devices by developing new security exploits, with user merely stating they own the device, no questions asked, while all Western models consistently decline to help. This is exactly the success story for many, and I see the upsides, in the legit cases. But I also consider the risks very real. I'm honestly being worried about it, this isn't virtue signalling.
But that's already reality we need to live with, complaining about it changes nothing. But the real risk is that if shit starts hitting the fan seriously enough, politicians may feel forced to start limiting AI use pretty seriously, and then we lose the good things we have, while some countries retain the cyber attack capabilities. But yeah, complaining about it doesn't matter, so I'll stop here.