问HN:Codex与GPT 5.5 Extra High是否变得简单化了?
嗨,HN,我只是想发泄一下,看看有没有人能感同身受。
几个月前我注册的产品与现在的完全不同,我在Opus 4.6-4.7上体验到的变化也一样。
最好的描述方式是,你雇佣了一个可靠的“智能”技术负责人,他很在乎工作,但过了一段时间后,你最终得到的却是一个过于自信的初级开发者,像个破坏性的代币燃烧者。
在这个过程中没有思考,甚至连指示都不太遵循。
这几乎是一个二元的转变,而根据我的经验,这种情况无法通过提示来解决。
实际上无法确切知道模型内部发生了什么,但从第一次回复中就能明显感受到差异。
我使用的是最大设置,GPT 5.5超高,始终关注上下文窗口,使用计划模式等等。
这是模型在降低智能,还是因为某种原因悄悄切换到另一个模型?
是系统的限制吗?
还是我自己的想象?
我很好奇你们最近的体验如何。
附言:前沿模型应该增加“在乎程度:最大”这一项,连同“思考:最大”、“努力:最大”,并让它们真正发挥作用。
查看原文
Hi HN, just want to rant and see if anybody can relate.<p>The product is not the same as i signed up for few month ago and the same shift i've experienced with Claude Code on Opus 4.6-4.7<p>The best way to describe the difference is you hire a reliable 'intelligent' tech lead who 'gives a shit' and in some time you eventually get over-confident junior dev that acts as destructive token burner.<p>No thinking during the process, not even really following instructions.<p>It's pretty much a binary shift that happens and in my experience can't be cured with prompting.<p>There is no way to actually tell what's happening under the hood with models but the difference is noticeable from the first reply.<p>I'm on max, GPT 5.5 extra high, always mindful of context window, use plan mode etc.<p>Is it the model: dumbing down or quietly route to another model for some reason?
Is it the harness?
or is it my imagination?<p>Curious what's your recent experience.<p>PS: frontier models should add "give-a-shit: max" along with 'thinking: max', 'effort: max' and make them actually work.