I trust Claude for everything. This test made me rethink that.
xAI released Grok 4.5, a new version of their language model, which challenges the performance of Opus, a popular alternative. This raises questions about the reliability of current language models. Engineers should reassess their trust in these models and consider testing their own applications. Grok 4.5's performance may impact the choice of language model for future projects. Further testing is needed to confirm the results.