@Teknium
@steipete I’m sorry that you’re this desperate that you will take such an unscientific benchmark instead of any established one. Also qwen local is one of the most random length models there is with all its looping. And we smoke you all on quality benchmarks on every open model. Here’s wildclawbench by internlm, same speed on open models, much better results.