本頁介紹 Cursor 在改用基於用量的定價之前採用的舊版基於請求的定價模型。
概述¶
基於請求的定價模型按用戶發起的 AI 請求數量收費,而非按使用量收費。每個許可證均包含每月請求額度。超出包含的用量後,您可以按需購買額外用量。
請求¶
請求是發送給大多數模型的一條消息,包含您的消息、代碼庫中的相關上下文以及模型的響應。請查看模型表,瞭解各模型的請求次數。
- 按需用量按模型 API 費率加收 20%。
- **Max Mode**按模型 API 費率加收 20%。Max Mode 支持更大的上下文窗口、子智能體、圖像生成,以及在按請求計費的方案中按需使用最新的前沿模型。
模型¶
| Model | Provider | Default context | Max context | Capabilities | Requests | Notes |
|---|---|---|---|---|---|---|
| Claude 4 Sonnet | Anthropic | 200k | - | Agent, Thinking, Images | 1 | Hidden by default; Thinking variant counts as 2 requests in legacy pricing |
| Claude 4 Sonnet 1M | Anthropic | - | 1M | Agent, Thinking, Images | 1 | Hidden by default; Thinking variant counts as 2 requests in legacy pricing; This model can be very expensive due to the large context window; The cost is 2x when the input exceeds 200k tokens |
| Claude 4.5 Haiku | Anthropic | 200k | - | Thinking, Images | 1 | Hidden by default; Bedrock/Vertex: regional endpoints +10% surcharge; Cache: writes 1.25x, reads 0.1x |
| Claude 4.5 Opus | Anthropic | 200k | 200k | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans |
| Claude 4.5 Sonnet | Anthropic | 200k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge) |
| Claude 4.6 Opus | Anthropic | 200k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge) |
| Claude 4.6 Sonnet | Anthropic | 200k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge) |
| Claude 4.7 Opus | Anthropic | 300k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge) |
| Claude Fable 5 | Anthropic | 300k | 1M | Agent, Thinking, Images | - | Requires data retention approval for Enterprise customers, Teams and individual customers with Privacy Mode enabled; Anthropic stores agent input and output data for harm-prevention processes; this data is not used to train or improve Anthropic models or products; Requests that trip a security guardrail are automatically routed to Claude Opus; About 2x the cost of Claude Opus 5; Requires Max Mode on legacy request-based plans |
| Claude Opus 4.7 (fast mode) | Anthropic | 200k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Limited research preview; Up to 1M tokens with extended context at the same per-token rates as shorter context |
| Claude Opus 4.8 | Anthropic | 300k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Fast mode (`claude-opus-4-8-fast`) requires Max Mode on legacy request-based plans; Fast mode is 3x lower per-token pricing than Opus 4.7 fast mode; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge) |
| Claude Opus 5 | Anthropic | 300k | 1M | Agent, Thinking, Images | - | Requires Max Mode on legacy request-based plans; Fast mode (`claude-opus-5-fast`) requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge) |
| Claude Sonnet 5 | Anthropic | 200k | 1M | Agent, Thinking, Images | - | Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge); Uses an updated tokenizer, so the same input can map to more tokens |
| Composer 1 | Cursor | 200k | - | Agent, Images | 1 | Hidden by default |
| Composer 2.5 | Cursor | 200k | - | Agent, Thinking, Images | 2 | - |
| Gemini 2.5 Flash | 200k | 1M | Agent, Thinking, Images | 1 | Hidden by default | |
| Gemini 3 Flash | 200k | 1M | Agent, Thinking, Images | 1 | Hidden by default | |
| Gemini 3 Pro | 200k | 1M | Agent, Thinking, Images | 1 | Hidden by default | |
| Gemini 3 Pro Image Preview | 200k | 1M | Images | 1 | Hidden by default; Native image generation model optimized for speed, flexibility, and contextual understanding; Text input and output priced the same as Gemini 3 Pro; Image output: \(120/1M tokens (\~\)0.134 per 1K/2K image, \~$0.24 per 4K image); Preview models may change before becoming stable and have more restrictive rate limits | |
| Gemini 3.1 Pro | 200k | 1M | Agent, Thinking, Images | 1 | - | |
| Gemini 3.5 Flash | 200k | 1M | Agent, Thinking, Images | 1 | Hidden by default | |
| Gemini 3.6 Flash | 200k | 1M | Agent, Thinking, Images | 1 | Hidden by default | |
| Gemini 3.7 Flash | 200k | 1M | Agent, Thinking, Images | 1 | - | |
| GLM 5.2 | Z.ai | 200k | - | Agent, Thinking | 1 | Hidden by default |
| GPT-5 | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Agentic and reasoning capabilities; Available reasoning effort variant is gpt-5-high |
| GPT-5 Fast | OpenAI | 272k | - | Agent, Thinking, Images | 2 | Hidden by default; Faster speed but 2x price; Available reasoning effort variants are gpt-5-high-fast, gpt-5-low-fast |
| GPT-5 Mini | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default |
| GPT-5-Codex | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Agentic and reasoning capabilities |
| GPT-5.1 Codex | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Agentic and reasoning capabilities |
| GPT-5.1 Codex Max | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default |
| GPT-5.1 Codex Mini | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Agentic and reasoning capabilities; 4x rate limits compared to GPT-5.1 Codex |
| GPT-5.2 | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Agentic and reasoning capabilities; Available reasoning effort variant is gpt-5.2-high |
| GPT-5.2 Codex | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Agentic and reasoning capabilities |
| GPT-5.3 Codex | OpenAI | 272k | - | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; Available reasoning effort variant is gpt-5.3-codex-high |
| GPT-5.4 | OpenAI | 272k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; 90% discount on cached input tokens; Fast mode is 15% faster with 2x pricing; Long context supports up to 1M tokens with 2x input pricing |
| GPT-5.4 Mini | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Smaller, faster variant of GPT-5.4; 90% discount on cached input tokens |
| GPT-5.4 Nano | OpenAI | 272k | - | Agent, Thinking, Images | 1 | Hidden by default; Smallest GPT-5.4 variant, optimized for cost; 90% discount on cached input tokens |
| GPT-5.5 | OpenAI | 272k | 1M | Agent, Thinking, Images | - | Hidden by default; Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; More token-efficient than GPT-5.4 on comparable tasks; Improved persistence on long-running tasks; Fast mode is available at higher rates; Long context supports up to 1M tokens with 2x input pricing |
| GPT-5.6 Luna | OpenAI | 272k | - | Agent, Thinking, Images | - | Smallest GPT-5.6 variant, optimized for cost and speed; Agentic and reasoning capabilities; Fast mode is available at 2x pricing; Cache writes are billed at 1.25x the uncached input rate |
| GPT-5.6 Sol | OpenAI | 272k | 1M | Agent, Thinking, Images | - | Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; Fast mode is available at 2x pricing; Long context supports up to 1M tokens with 2x input pricing; Cache writes are billed at 1.25x the uncached input rate; Promotional pricing through November 21, 2026 |
| GPT-5.6 Terra | OpenAI | 272k | - | Agent, Thinking, Images | - | Mid-tier GPT-5.6 variant between Sol and Luna; Agentic and reasoning capabilities; Fast mode is available at 2x pricing; Cache writes are billed at 1.25x the uncached input rate |
| Grok 4.5 | Cursor | 256k | - | Agent, Thinking | - | Jointly trained by Cursor and SpaceXAI |
| Grok 4.6 | Cursor | 256k | - | Agent, Thinking | - | Jointly trained by Cursor and SpaceXAI |
| Kimi K2.7 Code | Moonshot | 262k | - | Agent, Thinking, Images | 1 | Hidden by default |
| Kimi K3 | Moonshot | 200k | 1M | Agent, Thinking, Images | 1 | Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge); No separate cache-write fee |
| Legacy Enterprise Auto | Cursor | - | - | Agent | - | Hidden by default |
舊版方案客戶¶
如果您使用的是舊版按請求計費方案,可以繼續使用至下次續訂。在續訂時,您將遷移至基於用量的定價模型。
遷移支持¶
我們的團隊可協助您從舊版定價遷移至現行定價模式。我們可以提供:
- 新舊定價的詳細成本對比分析
- 遷移時間表和規劃
- 面向企業客戶的定製解決方案
如需遷移協助,請聯繫 enterprise@cursor.com。