Models for roleplay

Our roleplay picks are editorial, not a benchmark ranking.

Author
Input
Providers
Limits

Context

Input price

Only

Sort

4 models

  • Gemini 3.8 Flash50% off
    Context
    1M
    Input / 1M
    $1.575$0.788
    Output / 1M
    $7.875$3.938
    Max output
    66K
    Accepts
    • Text
    • Image
    • PDF
    • Video
  • GLM 5.3 Flash
    Context
    1M
    Input / 1M
    $0.158
    Output / 1M
    $0.525
    Max output
    131K
    Accepts
    • Text
    • Image
  • GPT 5.6 Luna
    Context
    1.1M
    Input / 1M
    $0.21
    Output / 1M
    $1.26
    Max output
    128K
    Accepts
    • Text
    • Image
    • PDF
  • DeepSeek V4 Flash 0731
    Context
    1M
    Input / 1M
    $0.0798
    Output / 1M
    $0.161
    Max output
    384K
    Accepts
    • Text

Reliability notes

Some of these have nowhere to fail over to. GLM 5.3 Flash, GPT 5.6 Luna, DeepSeek V4 Flash 0731, Gemini 3.8 Flash are served by a single registered upstream. If it goes offline, these models are unavailable until it recovers. How failover works.

Client setup for SillyTavern, RisuAI and Agnai — and what we store, which is not your prompts — is on /use/roleplay. The arithmetic for what one message costs at these rates is in the cost guide.