Claude Opus 5.5 Intelligence, Performance and Price Analysis (Max)
https://artificialanalysis.ai/models/claude-opus-5-5loading story #49805100
Do these evaluations get re run a few weeks after launch? I started doing that yesterday for our internal dataset and found Sol’s performance had regressed to be equal to Luna’s. Granted this was one run, but something I’m becoming more concerned about, the model providers want to quickly prove they’re the best, people switch to them, then they pull the rug.
loading story #49806995
loading story #49806964
loading story #49804835
loading story #49807069
loading story #49805655
loading story #49806768
loading story #49808022
loading story #49807038
loading story #49804962
loading story #49805176
loading story #49805129
loading story #49806134
loading story #49805160
loading story #49806042
loading story #49805096
loading story #49805070