Skip to content

#Costs/usage limits

0 items today
10/6Tue
  1. DEV Community · Codex74

    Codex 与 Claude Code 同任务实测:质量打平,Codex 成本低 2.4 倍

    作者在 python-humanize 仓库上用四个任务(真实 issue 修 bug、按规格加功能、无行为变更重构、带 3 个植入 bug 的代码评审)各跑两遍,Claude Code 2.1.291(claude-opus-5-5)与 Codex CLI 0.160.0(gpt-6.1-sol)16 次运行全部通过检查,四个评审都找齐 3 个植入 bug 且无误报。

    Awaiting translation

  2. Reddit · ClaudeCode / Codex / VibeCoding30

    Claude Code 的 effort level 该怎么选?各档位适用场景与用户实践

    Claude Code 文档建议按任务选 effort level:low 用于自己复核的快速改动如重命名,medium 用于范围明确的功能(现为 Opus 5.5 和 Sonnet 5.5 默认),high 用于边界情况关键的 bug 修复,max 只留给需要 Claude 独立攻克的难题,因为它容易过度思考。

    Awaiting translation

  3. Habr · Claude Code64

    用英文给编码智能体写提示词更省 token 吗?40 次实测只差 12 个 token

    作者用 5 个 Python 任务、Haiku 4.5 和 Sonnet 5.5 各跑两遍共 40 次,对比俄语和英语提示词的 token 消耗与成本。俄语提示词平均每任务多 12 个 token,但单次运行平均读取约 22.2 万输入 token,其中 94% 是每轮从缓存重读的系统提示词、工具说明和项目文件,语言差异只占约 0.005%。

    Awaiting translation

  4. Reddit · ClaudeCode / Codex / VibeCoding12

    零后端经验用 Claude Code 搭建 Cloudflare 多人游戏后端,如何避免烧光 token?

    一位零后端和 Cloudflare 经验的开发者,正靠 Claude Code 设计并实现一个词汇学习 Web 应用的多人游戏基础设施,需求包括基于 WebSocket 的实时房间(6 人 + 10 名观众)、随机房间码、断线重连、服务端权威状态、货币与每日登录防刷,以及 Firebase Authentication 与 Cloudflare 后端集成。

    Awaiting translation

  5. DEV Community · MCP78

    FP8 pitfall: GPU bill dropped 47%, but the model outputs “!!!!!!”

    The author ran Qwen2.5 7B/32B/72B on a single AMD MI300X with vLLM ROCm, priced at $2.99/GPU-hr. The BF16 baseline was 7B at $0.227/M, 32B at $0.77/M, and 72B at $1.67/M output tokens.

    Why it matters: The author benchmarked FP8 quantization on the MI300X and found that per-token billing can hide the model's output degrading into gibberish, then gave a reusable way to verify it.

  6. Reddit · ClaudeCode / Codex / VibeCoding15

    同时用 Codex 和 Claude Code 的人,你们的工作流是怎么搭的?

    一位刚开始认真为 AI 工具付费的用户在 Reddit 提问:想保留 OpenAI 订阅以继续使用 Dot 和 Codex,同时考虑试用一个月 Claude Pro 跑 Claude Code,让两者在同一个本地 Git 仓库上交替工作——一个实现、另一个审查,并保持 Markdown 文档同步更新。他还在确认一个理解是否成立:API 每任务成本与订阅实际包含的用量衡量的是不同东西。

    Awaiting translation

  7. Reddit · ClaudeCode / Codex / VibeCoding12

    用户吐槽 Codex 付费后额度未重置:付了 $100 仍显示 0% 周用量

    一名用户续订 OpenAI Codex 的 5x 套餐并被扣款 $100 后,打开 Codex 发现周用量仍显示 0%,额度并未因付费而重置。他意识到此前的额度重置已把周期推移,只能等 Tibo 再次手动重置或再等 4-5 天自然恢复;若想立刻用上算力,本应先取消再重新订阅。他称退款机器人判定其不符合退款条件,表示打算再用一个月 Computer use 后就不再关注 GPT。

    Awaiting translation

  8. 宝玉71

    SemiAnalysis 实测 Anthropic、OpenAI 等九家 AI 订阅套餐后得出,同样 200 美元,Claude 订阅折算的 Token 用量约为 OpenAI 的 5 倍。

    Awaiting translation

    QuotedSemiAnalysis@SemiAnalysis_

    Anthropic Subscriptions Offer 5x+ More Value Than OpenAI Limit testing every AI subscription plan from Anthropic, OpenAI, Meta, SpaceXAI, MiniMax, Moonshot, Zdotai, Cursor, and Cognition https://newsletter.semianalysis.com/p/anthropic-subscriptions-offer-5x

  9. 宝玉71

    据 The Information 10 月 5 日报道,Meta 和微软都在减少员工内部使用 Anthropic 的 Claude,转向自家模型和工具。

    Awaiting translation

    QuotedNIK@ns123abc

    🚨BREAKING: Microsoft and META are aggressively cutting employee use of Claude ahead of Anthropic's IPO Microsoft has cut internal claude spend by more than 33%, nuked per-employee token budget from $100k/month to $10k/month, and forced Copilot to auto-route to cheaper models META used Claude code to build Muse, then cut active users from 60,000 to 30,000 (50% decline) after launch, and replaced it with Muse Code Palantir and Nvidia are also scaling back claude over soaring prices and data privacy fears it’s OVER…

  10. Reddit · ClaudeCode / Codex / VibeCoding76

    A local proxy spreads Claude Code requests across multiple Max accounts and switches before the quota runs out.

    The author open-sourced claudemanager, a local daemon that Claude Code points to via ANTHROPIC_BASE_URL. It only changes the request's Authorization header to route sessions to the Max account with the most remaining capacity in its 5-hour, weekly, and per-model windows, switching at custom thresholds before those windows fill up.

    Why it matters: The author also open-sourced a local proxy that automatically distributes Claude Code traffic across multiple Max accounts based on remaining quota, and logs requests along the way.

  11. Reddit · ClaudeCode / Codex / VibeCoding20

    Claude Sonnet 与 Opus 5.5 被指 token 消耗过快,20x 用户 2 小时用掉 30% 额度

    一名付费 20x 的 Claude 用户反映,自 Sonnet 和 Opus 5.5 发布后,几乎每个任务都会消耗约 1% 用量,2 小时设计工作就用掉 30% 额度。该用户称已为 Claude 累计投入近 2400 美元,如今不得不每周四到周日改用 Codex 的 plus 计划,并认为 Claude 在用量上难以胜过 Codex,但在设计与输出质量上仍占优。

    Awaiting translation

  12. Hacker News · MCP78

    Flash-Agents: an MCP plugin that hands Claude Code's coding tasks to a DeepSeek Flash worker

    Flash-Agents is a Claude Code plugin that delegates bounded coding work—implementing slices, porting tests, reviewing diffs, mapping out a codebase—to a DeepSeek V4.1 Flash worker, while Claude keeps architecture, acceptance criteria, and final review.

    Why it matters: The author outsources Claude Code's coding tasks to a DeepSeek Flash worker and shares the sandbox, patches, and measured data, so you can judge the cost and safety boundaries for yourself.

  13. Reddit · ClaudeCode / Codex / VibeCoding20

    用户反驳近期对 OpenAI 的批评:Sol 6.1 与 Astra 实际使用体验

    一名 OpenAI Pro 20x 订阅用户反驳近期对 OpenAI 的批评,称 Sol 6.1(多用 xhigh)智能可靠、不易跑偏,三天仅消耗 10% 用量,虽实测约 16 tok/sec 偏慢但产出稳定。他通过优化工作流将 token 消耗减半,并提到 Astra 消耗较大,以及 tibo 宣布 Sol 和 Astra 提速 50%。

    Awaiting translation

  14. 宝玉67

    OpenAI 在 28 天更新的第 1 天宣布,通过 ChatGPT 订阅使用 GPT-6 Astra 和 GPT-6.1 Sol 时默认速度提升约 50%,用户无需改设置,两小时内生效。

    Awaiting translation

    QuotedTibo@thsottiaux

    Day 1/ We have optimized the default speed to be ~50% faster across GPT-6 Astra and GPT-6.1 Sol through the subscription across all our products and partners using Sign in With ChatGPT (including OpenCode, Pi, Amp, Devin, ...). No changes needed on your end and this should be felt within the next two hours.

  15. DEV Community · Claude Code71

    quota-audit:按项目和 Skill 拆解 Claude Code 额度消耗的 Skill

    作者因 Claude Code 的 /usage 只给账号级百分比、无法看出哪个项目吃掉额度,写了名为 quota-audit 的 Skill,读取本机 ~/.claude/projects/*/*.jsonl 中的 token 用量,按项目、按 Skill 归因,并列出 5h、24h、7d 各窗口的估算花费、超过 150k context 的 session 和撞限记录。

    Awaiting translation

10/5Mon
  1. Reddit · ClaudeCode / Codex / VibeCoding22

    Claude Code 用 Opus 5.5 跑两轮会话被计费 39 美元,/usage 却显示 Haiku 承担主要推理

    用户在 Claude Code 中以 API 认证方式使用 Opus 5.5(开启 1 小时 TTL 缓存和 ultracode 模式)跑了两轮会话,最终花费近 39 美元。但 /usage 显示 Haiku 承担了主要推理,包括 4.1m 输入和 204 次网页搜索,用户质疑为何付费使用 Opus 却由 Haiku 完成核心工作。

    Awaiting translation

  2. Reddit · ClaudeCode / Codex / VibeCoding22

    Claude 用户该如何使用 subagent 控制 token 消耗?

    一位 Claude 用户为改善 token 用量,改用 Opus 做规划、按任务难度分配不同模型,结果收到提示:82% 的用量来自 subagent 密集的会话,因为每个 subagent 都会发起自己的请求。他感觉 subagent 比主对话更耗 token,但因使用更便宜的模型而认为影响不大,于是发帖询问自己的做法是否正确。

    Awaiting translation