Has anyone ever produced any evidence that any Anthropic model was ever nerfed?
One would assume this would be easy since you could rerun the same evals from release day.
Has anyone ever produced any proof that Anthropic or ANY frontier AI provider has "nerfed" a model after release? Everyone says it and I'm trying to steelman it first before I attribute it to psychological phenomena.