Skip to content

#Costs/usage limits

0 items today
10/6Tue
  1. DEV Community · MCP78

    FP8 pitfall: GPU bill dropped 47%, but the model outputs “!!!!!!”

    The author ran Qwen2.5 7B/32B/72B on a single AMD MI300X with vLLM ROCm, priced at $2.99/GPU-hr. The BF16 baseline was 7B at $0.227/M, 32B at $0.77/M, and 72B at $1.67/M output tokens.

    Why it matters: The author benchmarked FP8 quantization on the MI300X and found that per-token billing can hide the model's output degrading into gibberish, then gave a reusable way to verify it.

  2. Reddit · ClaudeCode / Codex / VibeCoding12

    用户吐槽 Codex 付费后额度未重置:付了 $100 仍显示 0% 周用量

    一名用户续订 OpenAI Codex 的 5x 套餐并被扣款 $100 后,打开 Codex 发现周用量仍显示 0%,额度并未因付费而重置。他意识到此前的额度重置已把周期推移,只能等 Tibo 再次手动重置或再等 4-5 天自然恢复;若想立刻用上算力,本应先取消再重新订阅。他称退款机器人判定其不符合退款条件,表示打算再用一个月 Computer use 后就不再关注 GPT。

    Awaiting translation

  3. Reddit · ClaudeCode / Codex / VibeCoding20

    Claude Sonnet 与 Opus 5.5 被指 token 消耗过快,20x 用户 2 小时用掉 30% 额度

    一名付费 20x 的 Claude 用户反映,自 Sonnet 和 Opus 5.5 发布后,几乎每个任务都会消耗约 1% 用量,2 小时设计工作就用掉 30% 额度。该用户称已为 Claude 累计投入近 2400 美元,如今不得不每周四到周日改用 Codex 的 plus 计划,并认为 Claude 在用量上难以胜过 Codex,但在设计与输出质量上仍占优。

    Awaiting translation

10/5Mon
  1. Reddit · ClaudeCode / Codex / VibeCoding22

    Claude Code 用 Opus 5.5 跑两轮会话被计费 39 美元,/usage 却显示 Haiku 承担主要推理

    用户在 Claude Code 中以 API 认证方式使用 Opus 5.5(开启 1 小时 TTL 缓存和 ultracode 模式)跑了两轮会话,最终花费近 39 美元。但 /usage 显示 Haiku 承担了主要推理,包括 4.1m 输入和 204 次网页搜索,用户质疑为何付费使用 Opus 却由 Haiku 完成核心工作。

    Awaiting translation

  2. Reddit · ClaudeCode / Codex / VibeCoding22

    Codex Pro 20x quota shrinks, and now there's a vague "preview limit" too?

    A Codex Pro 20x subscriber reports that after the quota change, their 20x was effectively cut to 10x, and their always-on assistant, dot, even stopped working overnight because of a "preview limit." dot later couldn't say how big that quota was, how often it resets, or whether it shares a pool with the Codex quota—it didn't even give a retry time with a time zone. The user wants a clear limit, a visible usage meter, and a warning before the cutoff, and is asking whether anyone has found official docs for this restriction.

  3. Reddit · ClaudeCode / Codex / VibeCoding22

    每月付 200 美元用 Codex,工作日仍频繁遇到“模型容量已满”

    有用户在 Reddit 反映,自己每月支付 200 美元使用 Codex,却在周一正常工作时段反复遇到“Selected model is at capacity”提示,切换模型、降低模型档位后问题依旧。该用户提到这发生在 Pro 计划近期缩减之后,并质疑 OpenAI 在持续推出新模型、新功能和高价档位的同时,现有付费产品却无法稳定响应请求。

    Awaiting translation

9/29Tue
9/22Tue
9/18Fri
  1. V2EX · Codex28

    用户称 Codex 20x 订阅疑似被路由到 gpt-5.6-luna 降智

    有用户自建 AI 中转发现,Codex 20x 订阅请求的 gpt-6-astra 实际被路由到 gpt-5.6-luna,且返回的模型 ID 未作修改。该用户称 plus 用户似乎不受影响,问题主要集中在 Pro 的 20x 订阅。另有用户从官方模型调用统计中也看到不少 luna 记录,但自己并未使用过 luna。

    Awaiting translation

8/28Fri
8/17Mon