《Cursor文檔》-基於請求的定價 (舊版)

本頁介紹 Cursor 在改用基於用量的定價之前採用的舊版基於請求的定價模型。

概述

基於請求的定價模型按用戶發起的 AI 請求數量收費,而非按使用量收費。每個許可證均包含每月請求額度。超出包含的用量後,您可以按需購買額外用量。

請求

請求是發送給大多數模型的一條消息,包含您的消息、代碼庫中的相關上下文以及模型的響應。請查看模型表,瞭解各模型的請求次數。

  • 按需用量按模型 API 費率加收 20%。
  • **Max Mode**按模型 API 費率加收 20%。Max Mode 支持更大的上下文窗口、子智能體、圖像生成,以及在按請求計費的方案中按需使用最新的前沿模型。

模型

Model Provider Default context Max context Capabilities Requests Notes
Claude 4 Sonnet Anthropic 200k - Agent, Thinking, Images 1 Hidden by default; Thinking variant counts as 2 requests in legacy pricing
Claude 4 Sonnet 1M Anthropic - 1M Agent, Thinking, Images 1 Hidden by default; Thinking variant counts as 2 requests in legacy pricing; This model can be very expensive due to the large context window; The cost is 2x when the input exceeds 200k tokens
Claude 4.5 Haiku Anthropic 200k - Thinking, Images 1 Hidden by default; Bedrock/Vertex: regional endpoints +10% surcharge; Cache: writes 1.25x, reads 0.1x
Claude 4.5 Opus Anthropic 200k 200k Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans
Claude 4.5 Sonnet Anthropic 200k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge)
Claude 4.6 Opus Anthropic 200k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge)
Claude 4.6 Sonnet Anthropic 200k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge)
Claude 4.7 Opus Anthropic 300k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge)
Claude Fable 5 Anthropic 300k 1M Agent, Thinking, Images - Requires data retention approval for Enterprise customers, Teams and individual customers with Privacy Mode enabled; Anthropic stores agent input and output data for harm-prevention processes; this data is not used to train or improve Anthropic models or products; Requests that trip a security guardrail are automatically routed to Claude Opus; About 2x the cost of Claude Opus 5; Requires Max Mode on legacy request-based plans
Claude Opus 4.7 (fast mode) Anthropic 200k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Limited research preview; Up to 1M tokens with extended context at the same per-token rates as shorter context
Claude Opus 4.8 Anthropic 300k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Fast mode (`claude-opus-4-8-fast`) requires Max Mode on legacy request-based plans; Fast mode is 3x lower per-token pricing than Opus 4.7 fast mode; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge)
Claude Opus 5 Anthropic 300k 1M Agent, Thinking, Images - Requires Max Mode on legacy request-based plans; Fast mode (`claude-opus-5-fast`) requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge)
Claude Sonnet 5 Anthropic 200k 1M Agent, Thinking, Images - Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge); Uses an updated tokenizer, so the same input can map to more tokens
Composer 1 Cursor 200k - Agent, Images 1 Hidden by default
Composer 2.5 Cursor 200k - Agent, Thinking, Images 2 -
Gemini 2.5 Flash Google 200k 1M Agent, Thinking, Images 1 Hidden by default
Gemini 3 Flash Google 200k 1M Agent, Thinking, Images 1 Hidden by default
Gemini 3 Pro Google 200k 1M Agent, Thinking, Images 1 Hidden by default
Gemini 3 Pro Image Preview Google 200k 1M Images 1 Hidden by default; Native image generation model optimized for speed, flexibility, and contextual understanding; Text input and output priced the same as Gemini 3 Pro; Image output: \(120/1M tokens (\~\)0.134 per 1K/2K image, \~$0.24 per 4K image); Preview models may change before becoming stable and have more restrictive rate limits
Gemini 3.1 Pro Google 200k 1M Agent, Thinking, Images 1 -
Gemini 3.5 Flash Google 200k 1M Agent, Thinking, Images 1 Hidden by default
Gemini 3.6 Flash Google 200k 1M Agent, Thinking, Images 1 Hidden by default
Gemini 3.7 Flash Google 200k 1M Agent, Thinking, Images 1 -
GLM 5.2 Z.ai 200k - Agent, Thinking 1 Hidden by default
GPT-5 OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Agentic and reasoning capabilities; Available reasoning effort variant is gpt-5-high
GPT-5 Fast OpenAI 272k - Agent, Thinking, Images 2 Hidden by default; Faster speed but 2x price; Available reasoning effort variants are gpt-5-high-fast, gpt-5-low-fast
GPT-5 Mini OpenAI 272k - Agent, Thinking, Images 1 Hidden by default
GPT-5-Codex OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Agentic and reasoning capabilities
GPT-5.1 Codex OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Agentic and reasoning capabilities
GPT-5.1 Codex Max OpenAI 272k - Agent, Thinking, Images 1 Hidden by default
GPT-5.1 Codex Mini OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Agentic and reasoning capabilities; 4x rate limits compared to GPT-5.1 Codex
GPT-5.2 OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Agentic and reasoning capabilities; Available reasoning effort variant is gpt-5.2-high
GPT-5.2 Codex OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Agentic and reasoning capabilities
GPT-5.3 Codex OpenAI 272k - Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; Available reasoning effort variant is gpt-5.3-codex-high
GPT-5.4 OpenAI 272k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; 90% discount on cached input tokens; Fast mode is 15% faster with 2x pricing; Long context supports up to 1M tokens with 2x input pricing
GPT-5.4 Mini OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Smaller, faster variant of GPT-5.4; 90% discount on cached input tokens
GPT-5.4 Nano OpenAI 272k - Agent, Thinking, Images 1 Hidden by default; Smallest GPT-5.4 variant, optimized for cost; 90% discount on cached input tokens
GPT-5.5 OpenAI 272k 1M Agent, Thinking, Images - Hidden by default; Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; More token-efficient than GPT-5.4 on comparable tasks; Improved persistence on long-running tasks; Fast mode is available at higher rates; Long context supports up to 1M tokens with 2x input pricing
GPT-5.6 Luna OpenAI 272k - Agent, Thinking, Images - Smallest GPT-5.6 variant, optimized for cost and speed; Agentic and reasoning capabilities; Fast mode is available at 2x pricing; Cache writes are billed at 1.25x the uncached input rate
GPT-5.6 Sol OpenAI 272k 1M Agent, Thinking, Images - Requires Max Mode on legacy request-based plans; Agentic and reasoning capabilities; Fast mode is available at 2x pricing; Long context supports up to 1M tokens with 2x input pricing; Cache writes are billed at 1.25x the uncached input rate; Promotional pricing through November 21, 2026
GPT-5.6 Terra OpenAI 272k - Agent, Thinking, Images - Mid-tier GPT-5.6 variant between Sol and Luna; Agentic and reasoning capabilities; Fast mode is available at 2x pricing; Cache writes are billed at 1.25x the uncached input rate
Grok 4.5 Cursor 256k - Agent, Thinking - Jointly trained by Cursor and SpaceXAI
Grok 4.6 Cursor 256k - Agent, Thinking - Jointly trained by Cursor and SpaceXAI
Kimi K2.7 Code Moonshot 262k - Agent, Thinking, Images 1 Hidden by default
Kimi K3 Moonshot 200k 1M Agent, Thinking, Images 1 Hidden by default; Requires Max Mode on legacy request-based plans; Up to 1M tokens with extended context at the same per-token rates (no long-context surcharge); No separate cache-write fee
Legacy Enterprise Auto Cursor - - Agent - Hidden by default

舊版方案客戶

如果您使用的是舊版按請求計費方案,可以繼續使用至下次續訂。在續訂時,您將遷移至基於用量的定價模型。

遷移支持

我們的團隊可協助您從舊版定價遷移至現行定價模式。我們可以提供:

  • 新舊定價的詳細成本對比分析
  • 遷移時間表和規劃
  • 面向企業客戶的定製解決方案

如需遷移協助,請聯繫 enterprise@cursor.com

羽毛球分组比赛记分
小程序二维码

欢迎使用《羽毛球分组比赛记分》微信小程序

小夜