xAI’s Grok 4.6 has ranked third on the Artificial Analysis Healthcare and Medical Index, behind only Anthropic’s Claude Opus 5 (max) and Claude Fable 5 (with fallback). Grok 4.6 outperforms GPT-5.6 Sol and Moonshot AI’s Kimi K3 on this specialized medical ranking. The index evaluates models on medical knowledge at 35%, agentic knowledge work at 25%, non-hallucination rates at 15%, reasoning at 15%, and agentic customer interaction at 10%. Grok 4.6 scored 61 on the Intelligence Index, a 5-point improvement from Grok 4.5’s 56, priced at 2 dollars per million input tokens and 6 dollars per million output tokens.
Source: Read the original article

