Support

AI-powered help

Welcome!

Please introduce yourself before we start.

    Best Models for Roleplay

    Models with strong character consistency, creative prose, and long context windows — compared by price and context size

    Compare
    Use Case
    Capabilities
    Provider
    Status
    Input Price ($/M tokens)
    Output Price ($/M tokens)
    Context Size (tokens)
    22/281
    Models
    24/53
    Providers
    12
    Vision Models (filtered)
    19
    Tool-enabled (filtered)
    0
    Free Models (filtered)
    Features
    ByteDance
    glm-4.7
    $0.60$2.20$0.11
    Cerebras
    glm-4.7
    $2.25$2.75—
    EmberCloud
    glm-4.7
    $0.38$1.98$0.19
    NovitaAI
    glm-4.7
    $0.60$2.20$0.11
    Together AI
    glm-4.7
    $0.45$2.00—
    Vertex AI (OpenAI-compatible)
    glm-4.7
    $0.60$2.20—
    Z AI
    glm-4.7
    $0.60$2.20$0.11
    Alibaba Cloud
    glm-5
    $0.57$2.58—
    Alibaba Cloud(cn-beijing)
    glm-5
    $0.57$2.58—
    Baidu
    glm-5
    $1.00$3.20$0.20
    EmberCloud
    glm-5
    $0.72$2.30$0.14
    NovitaAI
    glm-5
    $1.00$3.20$0.20
    Tencent Cloud
    glm-5
    $1.00$3.20$0.20
    Vertex AI (OpenAI-compatible)
    glm-5
    $1.00$3.20$0.10
    Z AI
    glm-5
    $1.00$3.20$0.20
    Alibaba Cloud
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(cn-beijing)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(eu-frankfurt)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(singapore)
    glm-5.2
    $1.40$4.40$0.28
    Alibaba Cloud(us-virginia)
    glm-5.2
    $1.40$4.40$0.28
    Baidu
    glm-5.2
    $1.40$4.40$0.26
    ByteDance
    glm-5.2
    $1.40$4.40$0.26
    CanopyWave
    glm-5.2
    $1.40$4.40$0.26
    EmberCloud
    glm-5.2
    $1.26$3.96$0.23
    NovitaAI
    glm-5.2
    $1.40$4.40$0.26
    Tencent Cloud
    glm-5.2
    $1.40$4.40$0.26
    Z AI
    glm-5.2
    $1.40$4.40$0.26
    Cerebras
    qwen3-235b-a22b-instruct-2507
    $0.60$1.20—
    NovitaAI
    qwen3-235b-a22b-instruct-2507
    $0.09$0.58—
    Vertex AI (OpenAI-compatible)
    qwen3-235b-a22b-instruct-2507
    $0.22$0.88—
    Baidu
    kimi-k2.6
    $0.95$4.00$0.16
    CanopyWave
    kimi-k2.6
    $0.95$4.00$0.16
    Moonshot AI
    kimi-k2.6
    $0.95$4.00$0.16
    NovitaAI
    kimi-k2.6
    $0.80$3.40$0.16
    Tencent Cloud
    kimi-k2.6
    $0.86$3.57$0.14
    Alibaba Cloud
    kimi-k2.5
    $0.57$3.01—
    Alibaba Cloud(cn-beijing)
    kimi-k2.5
    $0.57$3.01—
    Alibaba Cloud(eu-frankfurt)
    kimi-k2.5
    $0.57$3.01—
    Alibaba Cloud(us-virginia)
    kimi-k2.5
    $0.57$3.01—
    EmberCloud
    kimi-k2.5
    $0.40$1.98$0.22
    Moonshot AI
    kimi-k2.5
    $0.60$3.00$0.10
    NovitaAI
    kimi-k2
    $0.57$2.30—
    MiniMax
    minimax-text-01
    $0.20$1.10—
    MiniMax
    minimax-m3
    $0.60$2.40$0.12
    Tencent Cloud
    minimax-m3
    $0.30$1.20$0.06
    Together AI
    minimax-m3
    $0.30$1.20$0.06
    Mistral AI
    mistral-small-2506
    $0.10$0.30—
    Mistral AI
    mistral-large-2512
    $0.50$1.50—
    Alibaba Cloud
    deepseek-v4.1-flash
    $0.30$1.20$0.03
    Alibaba Cloud(cn-beijing)
    deepseek-v4.1-flash
    $0.28$1.13$0.03
    Page 1 of 3

    Newsletter

    Stay ahead of the curve

    Insights on LLM routing, new model launches and cost optimization, sent to your inbox.

    • New models and providers, rounded up
    • Tips to cut LLM costs
    • Major product launches

    No spam. Unsubscribe anytime.

    System status
    AICPA SOC for Service Organizations badgeSOC 2 Type II
    compliant

    Product

  1. Features
  2. AI Gateway
  3. Observability
  4. Models
  5. Providers
  6. Rankings
  7. Add Provider
  8. Partners
  9. Lounge
  10. Changelog
  11. DevPass
  12. Compare Models
  13. Enterprise
  14. Resources

    • Legal Overview
    • Apps
    • Templates
    • Agents
    • MCP Server
    • Use Cases
    • Blog
    • Documentation
    • Developer resources
    • Integrations
    • Guides
    • Brand Assets
    • Token Cost Calculator
    • Copilot Cost Calculator
    • Referral Program
    • GitHub
    • Discord
    • Twitter
    • About
    • Contact Us

    Compliance

    • Trust Center
    • Security Portal
    • Terms
    • Privacy Policy
    • Provider Information
    • Sub-processors
    • SOC 2 Type II
    • Status

    Compare

    • All Comparisons
    • GitHub Copilot
    • OpenRouter
    • LiteLLM
    • Portkey
    • AWS Bedrock
    • Microsoft Foundry
    • Vercel AI Gateway
    • OpenRouter Alternatives
    • LiteLLM Alternatives
    • Migration Guides

    Models

    • Text Generation
    • Text to Image
    • Image to Image
    • Video Generation
    • Embeddings
    • Vision
    • Reasoning
    • Tool Calling
    • Web Search
    • Discounted
    • Best for Roleplay
    • Best for Coding
    • Best for Creative Writing
    • Best for Translation
    • Best for Math
    • Long Context
    • Cheapest
    • Open Source

    Providers

    • OpenAI
    • Anthropic
    • Google AI Studio
    • Google Vertex AI
    • Vertex AI (OpenAI-compatible)
    • Vertex AI (Anthropic)
    • Groq
    • Cerebras
    • xAI
    • DeepSeek
    • Alibaba Cloud
    • Runpod
    • NovitaAI
    • AtlasCloud
    • AWS Bedrock
    • AWS Mantle
    • Azure
    • Azure AI Foundry
    • Azure Anthropic
    • Z AI
    • Moonshot AI
    • Baidu
    • Perplexity
    • Mistral AI
    • CanopyWave
    • Inference.net
    • Together AI
    • SCX.ai (Turbo)
    • SCX.ai
    • ByteDance
    • MiniMax
    • EmberCloud
    • Meta
    • Meta Contributor
    • Sakana AI
    • Xiaomi
    • DeepInfra
    • ElevenLabs
    • Runware
    • Gonka24
    • Fireworks AI
    • RanoAI
    • Consensus Protocol
    • Tencent Cloud
    • Atria
    • TypeSafe AI

    © 2026 LLM Gateway. All rights reserved.

    A good roleplay model needs three things: prose that stays in character over hundreds of messages, a context window large enough to hold character cards and long chat histories, and per-token pricing that doesn't punish long sessions. This page lists the models the roleplay community actually uses — from budget favorites like DeepSeek and GLM to premium options like Claude — with live pricing and context sizes for every provider.

    Every model here is available through the same OpenAI-compatible endpoint, so you can plug LLM Gateway into SillyTavern, RisuAI, or your own frontend with one API key, switch models mid-conversation, and fall back automatically when a provider has an outage.

    Frequently asked questions

    What is the best AI model for roleplay?

    It depends on your budget. DeepSeek V4 and GLM-5 are the best value for money and rarely break character, Kimi K2.6 is known for expressive creative prose, and Claude Opus 4.8 and Claude Sonnet 5 write the highest-quality prose if you're willing to pay premium per-token rates. Grok's non-reasoning models are a popular fast middle ground.

    Can I use these models with SillyTavern or my own frontend?

    Yes. LLM Gateway exposes an OpenAI-compatible chat completions API, so any frontend that supports a custom base URL — SillyTavern, RisuAI, Agnai, or your own app — works by pointing it at the gateway and using your LLM Gateway API key.

    Which roleplay models have the largest context windows?

    Grok 4.1 Fast supports up to 2 million tokens, and Claude Sonnet 5, GLM-5.2, DeepSeek V4, and MiniMax Text-01 all reach 1 million tokens. That's enough to keep an entire long-running roleplay, including character cards and lorebooks, in context.

    How much does API roleplay cost compared to a subscription?

    Usually less. A typical roleplay exchange runs a few thousand tokens, so on a model like DeepSeek V4 Flash (about $0.14 per million input tokens) even heavy daily use costs a fraction of a fixed chatbot subscription — and you only pay for what you use.

    • DevPass
    • Lounge
    • Models
    • Docs
    • Pricing
    • DevPass
    • Lounge
    • Pricing
    • Docs
    • Models
      • AI Gateway
      • DevPass
      • Lounge
      • Observability
      • Enterprise
      • Blog
      • Changelog
      • Integrations
      • Reliability
      • Guardrails
      • Providers
      • Partners
      • Rankings
      • Apps
      • Models
      • Model Timeline
      • Compare
      • Token Cost Calculator
      • Referral Program
      • MCP Server
      • CLI
      • Agents
      • AI SDK Provider
      • Agent Skills
      • Templates
      • Guides
    Log InGet Started
    1.7k