StoryClaw Docs

Model Guide

Understand what a model is, how to read the pricing format, and which model to choose for each task.

New here?

This guide explains what a model is, how to read the pricing format, and which model to choose for different tasks. If you're new, simply use storyclaw/auto to get started.

What Is a Model

A model is the core engine that powers an AI member's ability to think, understand, and generate content. Different models are trained by different AI companies and vary in speed, reasoning ability, language understanding, and cost. Choosing the right model helps you balance output quality and usage cost.

StoryClaw integrates large language models (LLMs) from major AI providers including Anthropic, DeepSeek, Google, OpenAI, xAI, Moonshot, and MiniMax. It also provides three intelligent routing modes (auto / eco / premium) that automatically route your request to a suitable underlying model based on task type — you are charged at that model's actual rate.

How to Switch Models

At the bottom-left of the chat input area in ClawBot, you can see the name of the currently active model (highlighted in red below). Click it to open the model selection panel.

Model switcher button at the bottom-left of the chat input

Clicking opens the model selector with a search field. The right side shows each model's input/output price per 1M tokens.

Model selection panel

Intelligent Routing Modes

StoryClaw offers three routing modes prefixed with storyclaw/. The platform automatically selects a suitable underlying model based on task type — no manual selection needed. The actual cost is determined by whichever underlying model the system routes to; there is no additional surcharge for the routing itself.

storyclaw/auto

  • Billing: Billed at the routed model's rate
  • Strategy: Platform automatically routes to the most suitable underlying model based on task type
  • Best For: General use; default when unsure what to pick

storyclaw/eco

  • Billing: Billed at the routed model's rate
  • Strategy: Economy mode — routes to faster, lower-cost underlying models
  • Best For: Simple Q&A, high-frequency or batch tasks, cost-sensitive scenarios

storyclaw/premium

  • Billing: Billed at the routed model's rate
  • Strategy: Premium mode — routes to more powerful flagship models
  • Best For: Complex creative tasks, in-depth analysis, quality-critical scenarios

Understanding the Price Format

The price displayed next to each model follows the format: Input price / Output price /1M — the cost per one million tokens processed. For example, $3 / $15 /1M means: $3 per 1M tokens you send to the model (input), and $15 per 1M tokens the model generates (output).

FieldMeaningExample
Input price (1st number)Cost per 1M tokens of content you send to the model (prompts, context, instructions)$3 /1M
Output price (2nd number)Cost per 1M tokens of content the model generates (usually higher than input)$15 /1M
/1M (per million tokens)Billing unit. 1M = 1,000,000 tokens. ~750 English words ≈ 1,000 tokens; ~1 Chinese character ≈ 1–2 tokens/1M
$0 / $0 /1MShown for routing modes (auto / eco / premium). No surcharge for the routing itself; actual cost is based on the underlying model the system selectsauto / eco / premium

Real-world cost

Real-world costs are usually very low. For example, with storyclaw/claude-sonnet-4-6 ($3/$15/1M): a 500-character Chinese prompt uses roughly 500–1,000 tokens, costing less than $0.01. Note that the AI member's system prompt and conversation history also count toward token usage, so actual consumption in multi-turn conversations will be higher than a single message alone.

Model Details

If you want precise control over which model is used, you can select one directly in the model selection panel. The main LLM models currently supported are listed below, grouped by provider. The actual available models are subject to what's shown in the console — the list is updated continuously as new models are added.

Anthropic — Claude Series

storyclaw/claude-opus-4-6

Flagship · $5 / $25 /1M · 1M ctx

Best For: Complex analytical reports, reviewing long contracts or legal documents, designing multi-step agent workflows, high-quality brand copywriting

Anthropic's flagship model with rigorous logical reasoning, precise language output, and excellent performance on complex instructions and long-form text.

storyclaw/claude-opus-4-7

Flagship · $5 / $25 /1M · 1M ctx

Best For: Step-by-step logic and math problems, complex strategic planning and decision analysis, multi-agent collaboration tasks, highest-quality creative writing

Latest Claude Opus with Extended Thinking — the model reasons step-by-step before responding, delivering notably higher accuracy on multi-step reasoning tasks.

storyclaw/claude-sonnet-4-6

Balanced · $3 / $15 /1M · 1M ctx

Best For: Everyday content creation and polishing, conversational assistants, meeting notes and document summaries, moderately complex analysis

The most balanced model in the Claude family — noticeably faster than Opus while maintaining first-tier output quality; best value for everyday use.

DeepSeek — DeepSeek Series

storyclaw/deepseek-v4-flash

Fast · $0.14 / $0.28 /1M · 1M ctx

Best For: Automated batch processing, simple Q&A chatbots, high-frequency workflow trigger nodes, highly cost-sensitive workloads

DeepSeek's lightweight fast model with low latency and among the lowest token costs available — ideal for high-concurrency scenarios with strict speed and cost requirements.

storyclaw/deepseek-v4-pro

Reasoning · $0.435 / $0.87 /1M · 1M ctx

Best For: Writing and debugging code (Python, JavaScript, etc.), solving math or algorithm problems, data processing and structured analysis, strong-reasoning tasks on a budget

DeepSeek's flagship reasoning model — excels at code generation, mathematical reasoning, and logical analysis; strong overall capability at a highly competitive price.

Google — Gemini Series

storyclaw/gemini-3-flash-preview

Balanced · $0.5 / $3 /1M · 1M ctx

Best For: Quickly summarizing large document sets, article generation and polishing, general-purpose Q&A, scenarios requiring fast processing of very long text

Google's latest fast language model with balanced speed and capability, supporting up to 1M token context for large-scale text processing in a single request.

storyclaw/gemini-3.1-flash-lite-preview

Lite · $0.25 / $1.5 /1M · 1M ctx

Best For: Text classification and key information extraction, simple translation and format conversion, high-concurrency low-cost pipelines, preliminary content screening

Lightweight variant of Gemini Flash that significantly reduces cost while retaining basic language capabilities — suited for tasks where top-tier quality isn't required.

OpenAI — GPT Series

storyclaw/gpt-5.4

Balanced · $2.5 / $15 /1M · 1M ctx

Best For: Agent tasks integrating external APIs, multi-tool collaborative workflows, reliable content generation and editing, business automation requiring tool-use capabilities

A balanced GPT-5 generation model with accurate instruction understanding, stable overall performance, and solid tool-use support — well-suited for agent scenarios with external system integrations.

storyclaw/gpt-5.5

Flagship · $5 / $30 /1M · 1M ctx

Best For: High-quality creative writing, complex business strategy analysis and proposal writing, demanding multi-step reasoning, quality-critical scenarios with sufficient budget

OpenAI's flagship model with top-tier performance across reasoning, creative writing, and instruction understanding — ideal for demanding tasks where output quality is paramount.

xAI — Grok Series

storyclaw/grok-4.20

Reasoning · $2.5 / $5 /1M · 1M ctx

Best For: In-depth research reports, synthesizing information from multiple sources, complex analysis requiring critical thinking, dialectical reasoning and comparative argument tasks

xAI's flagship reasoning model with deep thinking capability — excels at multi-perspective analysis and critical reasoning on complex problems, at a more competitive price than most flagship models.

Moonshot — Kimi Series

storyclaw/kimi-k2.5

Economy · $0.6 / $3 /1M · 256K ctx

Best For: Chinese article writing and editing, reading comprehension and summarization of uploaded documents, enterprise knowledge base Q&A, cost-sensitive Chinese content workflows

A high-value model from Moonshot with strong Chinese language understanding and generation, performing reliably on document reading and knowledge Q&A.

storyclaw/kimi-k2.6

Balanced · $0.95 / $4 /1M · 256K ctx

Best For: Long-form Chinese reports and proposals, intelligent enterprise knowledge base retrieval and Q&A, Chinese AI workflows requiring precise instruction-following, multi-step Chinese agent tasks

Enhanced Kimi with improved Chinese reasoning and instruction-following, plus better support for multi-step agent collaboration — suited for Chinese workflows requiring precise complex instruction understanding.

MiniMax — MiniMax Series

storyclaw/MiniMax-M2.7

Economy · $0.3 / $1.2 /1M · 200K ctx

Best For: Bulk social media content in Chinese, mass product description writing for e-commerce, budget-constrained Chinese chatbots, cost-first Chinese content automation

A domestically developed LLM with strong Chinese text generation and among the lowest token costs — designed for price-sensitive, high-volume Chinese content production.

Quick Selection Guide

Not sure which model to pick? Match your task type to the table below for a quick decision:

My TaskRecommended ModelWhy
Unsure, let the system decidestoryclaw/autoPlatform auto-matches; lowest friction
Casual chat, simple Q&Astoryclaw/ecoFast, cheap, and perfectly capable
WeChat / Feishu bot daily repliesstoryclaw/ecoHigh-frequency triggers; eco is fast and cost-controlled
Marketing copy, creative writingstoryclaw/premium or storyclaw/claude-opus-4-7Quality-first; flagship models produce more creative and polished output
Bulk automation, high-frequency processingstoryclaw/eco or storyclaw/deepseek-v4-flashLowest cost; ideal for large-scale usage
Code generation, math reasoning, structured tasksstoryclaw/deepseek-v4-proReasoning specialist at a very competitive price
Chinese content, enterprise knowledge base Q&Astoryclaw/kimi-k2.6Strong Chinese capability, high cost-efficiency
Chinese content + cost-firststoryclaw/MiniMax-M2.7Chinese-developed model; among the lowest-priced options
Maximum quality, budget not a concernstoryclaw/claude-opus-4-7 or storyclaw/gpt-5.5Top-tier model capabilities available

On this page