Models for roleplay
Our roleplay picks are editorial, not a benchmark ranking.
4 models
- Context
- 1M
- Input / 1M
$1.575$0.788- Output / 1M
$7.875$3.938- Max output
- 66K
- Accepts
- Context
- 1M
- Input / 1M
- $0.158
- Output / 1M
- $0.525
- Max output
- 131K
- Accepts
- Context
- 1.1M
- Input / 1M
- $0.21
- Output / 1M
- $1.26
- Max output
- 128K
- Accepts
- Context
- 1M
- Input / 1M
- $0.0798
- Output / 1M
- $0.161
- Max output
- 384K
- Accepts
Reliability notes
Some of these have nowhere to fail over to. GLM 5.3 Flash, GPT 5.6 Luna, DeepSeek V4 Flash 0731, Gemini 3.8 Flash are served by a single registered upstream. If it goes offline, these models are unavailable until it recovers. How failover works.
Client setup for SillyTavern, RisuAI and Agnai — and what we store, which is not your prompts — is on /use/roleplay. The arithmetic for what one message costs at these rates is in the cost guide.