CoWorkBench

Qwen's internal benchmark for long-horizon professional work across computer science, finance, law, medicine, and other productivity domains.

  • Not in index
  • Agentic
ModelsModels with a result.4
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.N/A

Results

4 results
#ModelResultThe score from the source.
Qwen3.8 Max74.8%
Qwen3.8-27B70.7%