The study highlights the need for improved AI model evaluation metrics, emphasizing coherence and quality over mere correctness to enhance user experience.
The post Tencent paper reveals non-thinking mode increases response failures by up to 48% in multimodal AI models appeared first on Crypto Briefing.
