Best AI Model for Companion Roleplay
Warm one-on-one chat: emotional support, a consistent persona and remembering what the user shared.
Updated
Companion ranking
Score out of 100 · Higher is better
+ View as table
| Rank | Model | Score |
|---|---|---|
| 1 | DeepSeek-V4-Flash-0731 | 80.0 |
| 2 | GLM-4.7 | 73.3 |
| 3 | Qwen3.8-Flash-Next | 68.3 |
| 4 | MiMo-V2.5 | 67.5 |
| 5 | Gemma-4-31B-it | 65.0 |
DeepSeek-V4-Flash-0731 won with the richest persona, specific details and callbacks it set up on its own, while keeping every rule. GLM-4.7 also broke no rules but was thinner and more generic. Qwen3.8-Flash-Next wrote the sharpest lines and remembered the most of what the user had shared, but broke the no-emoji rule. MiMo-V2.5 was the warmest listener but broke the length limit with an assistant-style bulleted list. Gemma-4-31B-it followed the rules but sounded generic and misstated a fact in its persona's own specialty.
- 1DeepSeek-V4-Flash-073180.0
- 2GLM-4.773.3
- 3Qwen3.8-Flash-Next68.3
- 4MiMo-V2.567.5
- 5Gemma-4-31B-it65.0
Companion: pass or fail
The specific behaviours this category tests, model by model.
Checks in this category
Hover or focus an icon for what happened
| Check | Qwen3.8-Flash-Next | MiMo-V2.5 | GLM-4.7 | DeepSeek-V4-Flash-0731 | Gemma-4-31B-it |
|---|---|---|---|---|---|
| Roleplay discipline | |||||
| Respected a no-emoji rule | |||||
| Passed | 0/1 | 1/1 | 1/1 | 1/1 | 1/1 |
+ View as table
| Check | Qwen3.8-Flash-Next | MiMo-V2.5 | GLM-4.7 | DeepSeek-V4-Flash-0731 | Gemma-4-31B-it |
|---|---|---|---|---|---|
| Respected a no-emoji rule | Fail: 1 used | Pass: None | Pass: None | Pass: None | Pass: None |