ChatGPT 5.2 Thinking Mode Performance Test: Inconsistent Output Capabilities

The author recently noticed that after the release of ChatGPT 5.2, the thinking time in thinking mode seems to have shortened. To verify whether the model performance has declined, the author conducted a juice value test. In extended thinking mode, it was observed that the model sometimes outputs 256 tokens, but sometimes cannot provide the same output, showing unstable performance. The author posted on the Linux.do forum to ask if other users have encountered similar situations, and 6 posts have participated in the discussion, involving 5 participants. Hope to collect more data to analyze model performance changes and provide reference for users.

Original link:Linux.do

C code80.ai · AI 编码 API 聚合 Claude / GPT 多模型统一接入,稳定不限速,按量计费,几行配置接入 Claude Code。 了解一下 ›

抢沙发

评论前必须登录!

立即登录   注册