Many thanks for your great work!
I am wondering why Qwen3 series model got worse performance than llama3 series? Since Qwen3 models got great performance on llm general benchmarks, this seems quite strange to me.
Is there any understanding behind this? Many thanks in advance!
Many thanks for your great work!
I am wondering why Qwen3 series model got worse performance than llama3 series? Since Qwen3 models got great performance on llm general benchmarks, this seems quite strange to me.
Is there any understanding behind this? Many thanks in advance!