Moonshot AI's Kimi K3 successfully solved a Bongard problem by leveraging a tool, a task that both xAI's Grok 4.5 and Meta's Muse Spark 1.1 failed to complete. This performance places Kimi K3 in a competitive tier with Anthropic's Opus 4.8 and OpenAI's ChatGPT 5.6, which also demonstrated the ability to solve the same problem.
Moonshot AI's Kimi K3 demonstrates advanced tool-use capabilities, directly challenging the performance of xAI and Meta's models in complex reasoning tasks.