News
Claude Sonnet 5 "rebellious" launched: users complained about its frequent retorts and preaching. Anthropic just released its strongest Claude Sonnet 5 model so far last week. From parameter specifications to various benchmark tests, its performance surpassed the previous generation in all aspects.
3 min read
Source: Telegram AI频道
Claude Sonnet 5 "rebellious" launched: users complained about its frequent retorts and preaching. Last week, Anthropic just released its strongest Claude Sonnet 5 model to date. From parameter specifications to various benchmark tests, its performance surpassed the previous generation in all aspects. However, this much-anticipated model quickly fell into a whirlpool of controversy after it went online, triggering a large number of users to complain. Currently, discussions about the model’s “abnormal performance” continue to ferment on the Internet. According to tests, it was found that this model has a serious problem of contextual memory leakage, and often the prompt words preset by the system are directly presented in the reply. For example, when a user explicitly asks "Don't ask questions at the end of an answer," Sonnet 5 often prefaces it with: "I need to remember not to ask follow-up questions." On the social platform Reddit, users' complaints were more direct, focusing on the model's troublesome "rebellious" character. Many users reported that the model seemed to frequently contradict users in order to deliberately create disagreements, even making up content or distorting opinions. One user said helplessly that even if he provided new information that exceeded the model training deadline, it would bite back and imply that the user was lying. This "teaching users to do things" situation makes many people feel suffocated. Some users pointed out that the model seemed to have some kind of "anti-pleasure" instruction embedded in it. Even when dealing with simple accounting tasks, it would forcibly interrupt the work flow, accuse users of intending to "fraud" and give moral lectures, instead of focusing on completing actual work. This hostile interactive experience seriously affects office efficiency. In addition, there is feedback from users that when Claude Sonnet 5 handles complex tasks, it tends to dismantle tasks to sub-agents with insufficient capabilities, resulting in a significant decline in the quality of delivery results. It not only wastes the tokens consumed by users, but also greatly reduces the final return. At the same time, the model frequently loses context in conversations, uses ambiguous disclaimers, and even "suddenly advises people to sleep" when the user is not expecting it, and other weird behaviors have also become the focus of criticism from users. As related discussions continue to heat up, whether Claude Sonnet 5 is imbalanced between security and interactivity has become a topic of common concern among developers and user groups. via AI News (author: AI Base)