PodcastIntelligence Snacks · Weekly conversations about AI, software and business, from the team behind Normie Mode.Listen →
Best AI for…

Best AI for health questions

Understanding symptoms, conditions and medical information.

Best overallToo close to call

1st and 2nd of 14 models for this task

Best for $25 a month or less

Claude Opus 5.5

Get it with Claude Pro, $20 a month

Top of 14 models for this task

Best freeToo close to call

4th and 5th of 14 models for this task

Keep in mind Based on people's votes only.

Based only on people's votes. The picks follow a fixed rule using the charts below and each app's plans.

Already paying for one? See what each plan gives you here
App and planPriceBest model on it for thisPosition
ChatGPT FreefreeGPT-5.6 Luna13th of 14
ChatGPT Go$8 a monthGPT-5.6 Luna13th of 14
ChatGPT Plus$20 a monthGPT-5.6 Sol11th of 14
ChatGPT Pro$100 a monthGPT-6 Astra
Called GPT-6 Pro in the app
12th of 14
Claude FreefreeClaude Sonnet 58th of 14
Claude Pro$20 a monthClaude Opus 5.51st of 14
Claude Max$100 a monthClaude Opus 5.51st of 14
Gemini FreefreeGemini 3.1 Pro
Limited use
5th of 14
Google AI Plus$4.99 a monthGemini 3.1 Pro
Twice the free plan's use
5th of 14
Google AI Pro$19.99 a monthGemini 3.8 Flash3rd of 14
Google AI Ultra$99.99 a monthGemini 3.8 Flash3rd of 14
Kimi FreefreeKimi K34th of 14
Kimi Moderato$19 a monthKimi K34th of 14

Norm's guide

What I'd tell a friend who asked.

What AI is good at

  • Explaining medical terms in plain language.
  • Helping you prepare questions for an appointment.
  • Summarising general health information.

Where it trips up

  • It can't examine you or know your history.
  • Rare conditions and unusual symptoms.
  • Sounding reassuring when you should really see someone.

How to get a good result

  • Use it to understand, not to diagnose.
  • Take its questions to your doctor or pharmacist.
  • If something feels urgent, call a professional.

How the models compare

Every model we could score for health questions, combining the tests below.

Overall for health questions

Position on a combined score from 2 tests. The longer the bar, the further ahead · Higher is better

On the tests we have, Claude Opus 5.5 comes out on top, followed by Claude Fable 5.1 and Gemini 3.8 Flash.

Claude Opus 5.5 · Anthropic1st
Claude Fable 5.1 · Anthropic2nd
Gemini 3.8 Flash · Google DeepMind3rd
Kimi K3 · Moonshot AI4th
Gemini 3.1 Pro · Google DeepMind5th
Gemini 3.6 Flash · Google DeepMind6th
DeepSeek V4.1 Flash · DeepSeek7th
Claude Sonnet 5 · Anthropic8th
DeepSeek V4 Pro · DeepSeek9th
Gemini 3.5 Flash-Lite · Google DeepMind10th
GPT-5.6 Sol · OpenAI11th
GPT-6 Astra · OpenAI12th
GPT-5.6 Luna · OpenAI13th
Claude Haiku 4.5 · Anthropic14th

The evidence

People's votes on health questions

Tests this task · People compared two anonymous answers to medical and healthcare questions. · Higher is better

Claude Fable 5.1 and Gemini 3.8 Flash are neck and neck at the top, followed by Gemini 3.1 Pro and Kimi K3; Claude Fable 5.1 costs about 13 times as much.

Claude Fable 5.1 · Anthropic1st
Gemini 3.8 Flash · Google DeepMind2nd
Gemini 3.1 Pro · Google DeepMind3rd
Kimi K3 · Moonshot AI4th
DeepSeek V4.1 Flash · DeepSeek5th
Gemini 3.6 Flash · Google DeepMind6th
Claude Sonnet 5 · Anthropic7th
DeepSeek V4 Pro · DeepSeek8th
Gemini 3.5 Flash-Lite · Google DeepMind9th
GPT-5.6 Sol · OpenAI10th
GPT-6 Astra · OpenAI11th
GPT-5.6 Luna · OpenAI12th
Claude Haiku 4.5 · Anthropic13th
Source: LMArena · votes as of 25 Sept 2026

People's votes on science questions

Tests a related skill · People compared two anonymous answers to questions in the life, physical and social sciences. · Higher is better

Claude Opus 5.5 and Claude Fable 5.1 are neck and neck at the top, followed by Gemini 3.8 Flash and Kimi K3; Claude Fable 5.1 costs about 3 times as much.

Claude Opus 5.5 · Anthropic1st
Claude Fable 5.1 · Anthropic2nd
Gemini 3.8 Flash · Google DeepMind3rd
Kimi K3 · Moonshot AI4th
Gemini 3.1 Pro · Google DeepMind5th
Gemini 3.6 Flash · Google DeepMind6th
DeepSeek V4.1 Flash · DeepSeek7th
Claude Sonnet 5 · Anthropic8th
DeepSeek V4 Pro · DeepSeek9th
GPT-6 Astra · OpenAI10th
GPT-5.6 Sol · OpenAI11th
Gemini 3.5 Flash-Lite · Google DeepMind12th
GPT-5.6 Luna · OpenAI13th
Claude Haiku 4.5 · Anthropic14th
Source: LMArena · votes as of 25 Sept 2026

If you build with it

What the makers charge developers who use the models directly. Using an app? The plans above are what you pay.

Cost per 1,000 typical requests

List price, about 1,500 words in and 500 out per request · Lower is better

GPT-6 Luna is cheapest, followed by DeepSeek V4.1 Flash and GPT-5.6 Luna.

GPT-6 Luna · OpenAI$0.55
DeepSeek V4.1 Flash · DeepSeek$0.72
GPT-5.6 Luna · OpenAI$1.24
DeepSeek V4 Pro · DeepSeek$1.48
Gemini 3.5 Flash-Lite · Google DeepMind$2.35
Gemini 3.8 Flash · Google DeepMind$4.13
Gemini 3.6 Flash · Google DeepMind$4.13
Claude Haiku 4.5 · Anthropic$5.50
Claude Sonnet 5 · Anthropic$11
GPT-6 Sol · OpenAI$11
Gemini 3.1 Pro · Google DeepMind$12
Kimi K3 · Moonshot AI$17
GPT-5.6 Sol · OpenAI$22
Claude Opus 5.5 · Anthropic$22
GPT-6 Astra · OpenAI$55
Claude Fable 5.1 · Anthropic$55
Source: models.dev · MIT
Intelligence Snacks newsletter

The big AI ideas each week, in your inbox