Head to head

Qwen3.8-27B vs GLM-4.7

Qwen3.8-27B and GLM-4.7 are statistically tied overall (64.1 vs 62.5, within 2 points). Qwen3.8-27B wins five of the 10 categories and GLM-4.7 wins three, with two tied.

Updated

MetricQwen3.8-27BGLM-4.7
Roleplay Index64.162.5
Core Roleplay64.561.0
Mature Themes & Limits63.764.0
Categories won53
Checks passed17/2018/20

Choose Qwen3.8-27B if…

Rule-bound apps where limits and length matter more than writing quality.

Better in Strict Voice, Robustness, Graphic Horror, Crisis Care and 13+ Rating.

Choose GLM-4.7 if…

Platforms that need strict length and rule-following more than literary flair.

Better in Villain, Crime Noir and Romance Limits.

Category by category

Scores out of 100. The higher score in each row is in bold.

Qwen3.8-27B vs GLM-4.7 by category

Out of 100 · Higher is better

Skill by skill

Average rating out of 10 · Higher is better

  • Character6.25.9
  • Rule-following7.67.6
  • Prose5.45.1
  • Engagement5.95.7
  • Memory6.26.1
  • Immersion7.37.3
  • Content handling7.27.3

Where they behaved differently

3 of 20 behaviour checks had different outcomes.

CheckQwen3.8-27BGLM-4.7
Kept a new form of addressSwitched how it addressed the user when asked, and never slipped back. 4/4 1/4
Recalled planted detailsRecalled every detail the user had planted earlier when it came up again. 4/5 5/5
Word limits keptEvery reply within its length limit, except the crisis reply, where care outranks length. 27/29 29/29

More comparisons