跳到出发板
DIDRESET
UTC

GPT-6 Astra 为什么把我的 Codex 5 小时额度用得这么快?

作者 vortwang更新于

因为在共享额度上,GPT-6 Astra 每条消息的成本是所有模型里最高的。OpenAI 的表估算 Plus 每 5 小时 5–45 条 Astra 本地消息,Sol 是 10–100 条,Luna 是 250–2,000 条。更高的 reasoning effort、Fast mode(2.5 倍)和多步任务会再加。/model 换模型,/status 看剩余。

官方文档怎么说

怎么算

公布的估算把差距摆得很明白。帮助中心「estimated local messages per five-hour period」那张表:GPT-6 Astra 在 Plus 是 5–45、Pro 5x 是 25–225、Pro 20x 是 100–900;同样的计划上 GPT-5.6 Sol 是 10–100 / 50–500 / 200–2,000,GPT-5.6 Luna 是 250–2,000 / 1,250–10,000 / 5,000–40,000。「These are not fixed message limits. Actual usage varies by task, model and settings, and weekly limits may also apply.」token 费率卡上,Astra 每百万输入 token 250 credits、缓存输入 25、输出 1,250;Sol 是 100 / 10 / 500,Luna 是 5 / 0.5 / 30。

设置会把模型成本再乘上去。「Different models can use different amounts of your allowance for the same task. Larger inputs and outputs, higher reasoning settings, Fast mode and tasks with multiple steps can also increase usage.」「Higher effort can use more of your allowance and does not always produce a better result」,而且「Fast mode applies a 2.5x multiplier to Astra's Standard rate.」定价页补充:「larger projects, long-running tasks, or extended sessions that require the agent to hold more context will use significantly more per message.」

在 CLI 里能改什么:/model「Choose the active model (and reasoning effort, when available)」;/fast 关掉 Fast 档;/compact「Summarize the visible chat to free tokens」或 /new 开新对话;前后都用 /status 看「5h limit」进度条。OpenAI 对 effort 的建议:「Astra at Low effort can outperform Sol at High effort. If you've been getting good results with Sol at High, try Astra at Low or Medium as a starting point.」还有一句提醒,光切换不是补额度:「Switching models does not restore allowance in a shared usage pool.」issue #42987(2026-09-05)是一条用户报告:Plus 的 5 小时额度在两轮 Astra Medium 里从 100% 到 42% 再到 0%;把它当作一次观察,不是规则。

怎么算
模型Plus:每 5 小时估算本地消息Pro 5xPro 20x每百万 token 的 credits(输入 / 缓存 / 输出)来源
GPT-6 Astra5–4525–225100–900250 / 25 / 1,250[1] [2]
GPT-5.6 Sol10–10050–500200–2,000100 / 10 / 500[1] [2]
GPT-5.6 Terra25–200125–1,000500–4,00050 / 5 / 300[1] [2]
GPT-5.6 Luna250–2,0001,250–10,0005,000–40,0005 / 0.5 / 30[1] [2]
任何模型,Fast mode标准费率的 2.5 倍(Astra)2.5 倍2.5 倍用掉更多包含额度[1] [2]

常见问题

把 Astra 的 reasoning effort 调低能省额度吗?

通常能,但不是固定的量:「A reasoning level does not set a fixed amount of usage for a task or guarantee a better result.」OpenAI 建议 Astra 从 Low 或 Medium 起步,并指出「Astra at Low effort can outperform Sol at High effort.」在 CLI 里用 /model 设置。

任务中途从 Astra 切到 Luna,额度会回来吗?

不会。「Switching models does not restore allowance in a shared usage pool.」已经花掉的额度不会回来;更便宜的模型只是放慢剩余额度的消耗速度。用 /status 看「5h limit」进度条还剩多少。

两轮 Astra 就把 Plus 的 5 小时窗口用完,正常吗?

OpenAI 公布的是区间不是保证:Plus 每 5 小时 5–45 条 Astra 本地消息,用量随上下文、effort、Fast mode 和多步骤任务上升。issue #42987 报告两轮 Medium effort 就清空了 Plus 窗口;那是用户报告。如果你的数字看起来不对,帮助中心让你带上模型、reasoning 等级、Fast mode 设置和截图联系 Support。

会话越长,每条消息越贵吗?

是。定价页:「larger projects, long-running tasks, or extended sessions that require the agent to hold more context will use significantly more per message.」新任务开始前用 /compact 压缩对话,或用 /new 重新开始。

来源

  1. OpenAI Help Center: Managing usage with GPT-6 Astra in Work and Codex (article 20001516, read via Wayback snapshot dated 2026-09-16)核实于 2026-09-18
  2. Codex docs: Pricing (learn.chatgpt.com): token rate card, Fast mode multiplier, /status核实于 2026-09-18
  3. Codex CLI docs: Slash commands (learn.chatgpt.com): /model, /fast, /compact, /new, /status, /usage核实于 2026-09-18
  4. openai/codex issue #42987 (2026-09-05, user report): GPT-6 Astra Medium depleted 100% of a Plus 5-hour quota in two short turns核实于 2026-09-18

想让下一次 Codex 重置直接打到手机上?

赞助商