18 August 2026

Alibaba's smaller Qwen model matches larger competitors on benchmark

  • Alibaba's Qwen 3.8 27B model scored 52 on the Artificial Analysis Intelligence Index, a standardized test of AI capability.
  • This smaller model matched GPT-5.6 Luna and came close to much larger models like GLM-5.2 and DeepSeek V4 Pro.
  • The result suggests a smaller model from Alibaba can perform as well as larger competitors on at least one measurement.

How it was covered

Simon WillisonDaily notes and links

Qwen 3.8 27B model scored 52 on the Artificial Analysis Intelligence Index, matching GPT-5.6 Luna and coming close to larger models like GLM-5.2 and DeepSeek V4 Pro. The newsletter emphasises that this 27B parameter model achieves competitive performance against much larger models.