云聚 AI Token Plan 满 199 减 35 元
port:80 AI Junkie
AI 重度玩家的工程笔记本

Claude Wins Hallucination Test: Outperforms GPT and Gemini

云聚 AI Token Plan 满 199 减 35 元

On the Linux.do forum, a user conducted a web search capability test on mainstream AI models Claude, GPT, and Gemini, evaluating hallucination rates for questions with scarce information sources. The results showed that Claude Sonnet 4.5 performed best with a 0% hallucination rate, obtaining correct information in just three search rounds; GPT 5.2 had a 70% hallucination rate with low search efficiency; Gemini 3 Pro had a hallucination rate exceeding 90% with poor search results. The author emphasized that Claude is far ahead in tool usage capabilities, such as project management and file operations, and has switched from GPT to Claude as their primary tool. The article calls on AI companies to strengthen tool integration, enhance productivity, and break through model bottlenecks. This test provides practical reference for AI users, revealing performance differences and future development directions among models.

Original Link:Linux.do

阿里云 OPC 一人公司创业装备库
阿里云函数计算 一键部署 AI 大模型
赞(0)
未经允许不得转载:80aj » Claude Wins Hallucination Test: Outperforms GPT and Gemini
赞助推荐 FreeModel.dev Claude Code 中转
阿里云函数计算 一键部署 AI 大模型