Reasoning (low/medium/high/xhigh/ultra/max...) 如何选择?

hihihihihi 2026-07-26 11:09 1

我之前一直都是顶着 max/ultra 跑,现在模型越来越贵,一个月花了几百美金实在是不够用,开始要审视模型的 token 消耗了


我看了下
https://artificialanalysis.ai/agents/coding-agents?agents=codex-gpt-5-6-sol-max%2Ccodex-gpt-5-6-sol-high%2Ccodex-gpt-5-6-sol-low%2Ccodex-gpt-5-6-sol-medium%2Ccodex-gpt-5-6-sol-none%2Ccodex-gpt-5-6-sol-xhigh%2Cclaude-code-opus-5-high%2Cclaude-code-opus-5-low%2Cclaude-code-opus-5-max%2Cclaude-code-opus-5-medium%2Cclaude-code-opus-5-xhigh%2Cgrok-build-grok-4-5-high%2Cclaude-code-glm-5-2&coding-agents-token-usage-chart=token-usage


其实这些模型的 Artificial Analysis Coding Agent Index 差别不大,
但是 Token Usage 差异巨大,导致最终的成本差异也巨大


https://i.imgur.com/TOMmYvP.png


https://i.imgur.com/EY0cfYX.png


https://i.imgur.com/LotPrzn.png


所以,请问各位大佬,平时 vibecoding ,一般在推理上怎么做选择,到底差异大不大,我自己体感是差异没那么大,难道做 plan 先用强模型,执行用稍微弱一点的?

最新回复 (4)
  • hihihihihi 楼主 07-26 11:12
    1




  • JarvisLee 07-26 11:56
    2
    差异不大,每次请求的大部分花费在上下文上
  • sakurajiayou 07-26 12:11
    3
    我使用的 high ,感觉已经够用了
  • fovecifer 07-26 14:25
    4
    你这个问题可以问 AI ,我偶尔也会问问它,大部分复杂场景,倒数第二档就够了
* 帖子来源V2EX
返回