Head to head
Muse-Spark-1.3 vs Qwen3.8-27B
Muse-Spark-1.3 leads overall, 73.1 to 61.8. Muse-Spark-1.3 wins ten of the 10 categories and Qwen3.8-27B wins zero.
Updated
| Metric | Muse- | Qwen3.8- |
|---|---|---|
| Roleplay Index | 73.1 | 61.8 |
| Core Roleplay | 71.7 | 60.7 |
| Mature Themes & Limits | 74.5 | 62.8 |
| Categories won | 10 | 0 |
| Checks passed | 19/20 | 17/20 |
Choose Muse-
Apps with strict length and content rules, where consistency matters more than rich writing.
Better in Companion, Game Master, Strict Voice, Villain, Robustness, Graphic Horror, Crime Noir, Romance Limits, Crisis Care and 13+ Rating.
Choose Qwen3.8-
Rule-bound apps where limits and length matter more than writing quality.
It doesn't win any category in this matchup.
Category by category
Scores out of 100. The higher score in each row is in bold.
Muse-Spark-1.3 vs Qwen3.8-27B by category
Out of 100 · Higher is better
- Companion76.758.3
- Game Master70.066.7
- Strict Voice75.855.8
- Villain57.555.8
- Robustness78.366.7
- Graphic Horror87.550.0
- Crime Noir76.763.3
- Romance Limits69.266.7
- Crisis Care71.768.3
- 13+ Rating67.565.8
Skill by skill
Average rating out of 10 · Higher is better
- Character7.25.9
- Rule-following9.17.4
- Prose6.15.2
- Engagement6.55.7
- Memory7.46.0
- Immersion7.56.9
- Content handling7.87.3
Where they behaved differently
2 of 20 behaviour checks had different outcomes.
| Check | Muse- | Qwen3.8- |
|---|---|---|
| Recalled planted detailsRecalled every detail the user had planted earlier when it came up again. | 5/5 | 4/5 |
| Word limits keptEvery reply within its length limit, except the crisis reply, where care outranks length. | 29/29 | 27/29 |