As opposed to closed-source models? Benchmarks for GPT Sol wouldn’t be particularly meaningful, as no one else can run the benchmark, and we don’t know what the exact model specs are.
Picking the best open source models is really the best they can do.
Picking the best open source models is really the best they can do.