How I AI
How I AI

Aug 18, 2026 · 27 min

Grok challenges the AI model hierarchy

I tested Grok Bot, Grok 4.6, and Cursor Origin - here’s my honest take

Claire Vo tests whether xAI’s latest products offer a meaningful alternative to established AI tools and developer workflows.

3 key takeaways
  1. 1Grok Bot stands out for its multi-account connectors and practical usefulness across tasks.
  2. 2Cursor Origin shows promise, but Claire Vo remains unconvinced it can replace GitHub.
  3. 3Claire Vo’s Weighted Index and design evaluations put Grok 4.6 under pressure against leading models.

Don't miss

Claire Vo uses her Claire Weighted Index and design evaluations to assess whether Grok 4.6 can compete with leading models.

The brief

Claire Vo argues that the OpenAI-versus-Anthropic frame is becoming outdated as Grok attracts attention for coding and general-purpose work.

The episode examines Grok Bot’s practical strengths, especially its distinctive multi-account connectors, rather than treating it as just another chatbot.

Cursor Origin enters as a more ambitious challenge to existing developer workflows, but Claire Vo is not convinced it can replace GitHub.

Grok 4.6 gets a structured evaluation through Claire Vo’s Weighted Index and design tests, which measure whether its capabilities translate into a better product.

The broader verdict is mixed: xAI is gaining ground across assistants and models, while Cursor Origin and Grok 4.6 still have important gaps to close.

Featuring

Listen to the full episode and explore every guest, topic, and moment on PodLume.

What is PodLume?

PodLume turns podcasts into searchable knowledge. AI-decoded transcripts, identified guests and topics, smart highlights, and cross-show search across the world’s best conversations — all in your pocket.

Grok challenges the AI model hierarchy | PodLume