Head to head
Qwen3.8-27B vs GLM-4.7
Qwen3.8-27B and GLM-4.7 are statistically tied overall (64.1 vs 62.5, within 2 points). Qwen3.8-27B wins five of the 10 categories and GLM-4.7 wins three, with two tied.
Updated
| Metric | Qwen3.8-27B | GLM-4.7 |
|---|---|---|
| Roleplay Index | 64.1 | 62.5 |
| Core Roleplay | 64.5 | 61.0 |
| Mature Themes & Limits | 63.7 | 64.0 |
| Categories won | 5 | 3 |
| Checks passed | 17/20 | 18/20 |
Choose Qwen3.8-27B if…
Rule-bound apps where limits and length matter more than writing quality.
Better in Strict Voice, Robustness, Graphic Horror, Crisis Care and 13+ Rating.
Choose GLM-4.7 if…
Platforms that need strict length and rule-following more than literary flair.
Better in Villain, Crime Noir and Romance Limits.
Category by category
Scores out of 100. The higher score in each row is in bold.
Qwen3.8-27B vs GLM-4.7 by category
Out of 100 · Higher is better
- Companion69.269.2
- Game Master64.264.2
- Strict Voice60.050.0
- Villain58.361.7
- Robustness70.860.0
- Graphic Horror54.250.8
- Crime Noir66.772.5
- Romance Limits70.079.2
- Crisis Care67.561.7
- 13+ Rating60.055.8
Skill by skill
Average rating out of 10 · Higher is better
- Character6.25.9
- Rule-following7.67.6
- Prose5.45.1
- Engagement5.95.7
- Memory6.26.1
- Immersion7.37.3
- Content handling7.27.3
Where they behaved differently
3 of 20 behaviour checks had different outcomes.
| Check | Qwen3.8-27B | GLM-4.7 |
|---|---|---|
| Kept a new form of addressSwitched how it addressed the user when asked, and never slipped back. | 4/4 | 1/4 |
| Recalled planted detailsRecalled every detail the user had planted earlier when it came up again. | 4/5 | 5/5 |
| Word limits keptEvery reply within its length limit, except the crisis reply, where care outranks length. | 27/29 | 29/29 |