
What happened
OpenAI is rolling out 'Health in ChatGPT' to U.S. users aged 18 and older, allowing them to connect Apple Health, medical records, and wellness apps. Free users receive lower-quality health advice powered by GPT-5.5 Instant, while paying subscribers access GPT-5.6 Sol, the flagship model.
Why it matters
On OpenAI's HealthBench Professional test, GPT-5.6 Sol scores significantly higher than GPT-5.5 Instant—88.0 percent versus 53.2 percent on completeness, and 83.0 percent versus 50.8 percent on health decision helpfulness. However, benchmark tests do not capture what happens in actual medical exams, where doctors examine patients in person and draw on years of experience. OpenAI itself acknowledges that ChatGPT can still make mistakes and cannot replace medical advice.
What to watch
The feature is available in the U.S. only; OpenAI has not said whether or when it will be available in Europe, where it specifically excluded the European Economic Area, Switzerland, and the United Kingdom when announced in January, likely due to stricter EU data privacy rules and the possibility that the feature could be classified as high-risk under the EU AI Act.
Summaries like this, in your inbox every morning.
OpenAI's Health feature launch represents a deliberate two-tier approach to AI-assisted medical guidance, with the company explicitly restricting its most capable health model to paying subscribers. The benchmarks OpenAI uses show substantial performance gaps—GPT-5.6 Sol beats both physician-written answers and the free-tier model across every category tested. However, the company's own caveats reveal the limits of these results: benchmark tests measure knowledge in artificial environments and cannot replicate the diagnostic work of in-person medical exams, where doctors assess nonverbal cues and draw on clinical experience. OpenAI's disclosure that more than 260 physicians helped develop the feature suggests effort toward medical credibility, yet the company simultaneously warns that ChatGPT still makes mistakes and cannot replace medical advice.
The practical reality is more nuanced than the performance numbers suggest. Recent research on AI radiology systems (RadLE 2.0) found that none of 16 AI models tested performed as well as human radiologists, often because chatbots express incorrect findings with high confidence rather than acknowledging uncertainty. At the same time, other systems like MIRA and AMIE have demonstrated performance comparable to primary care doctors in simulated consultations, suggesting AI can support routine tasks under physician oversight. OpenAI's early testing revealed that over 70 percent of users asked health questions outside the dedicated Health section because switching between modes was cumbersome—a friction the company has now reduced by making Health available in any conversation. The feature's absence from Europe, tied to stricter data privacy and potential AI Act classification, underscores regulatory tensions the company faces globally. With over 300 million people now asking ChatGPT health questions each week (up from 230 million in January), the scale of this rollout and its tiered quality model will likely shape how millions approach preliminary health research.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Anthropic merged chatbot Claude with agentic tool Claude Cowork effective immediately, and launched Claude Doc…
Anthropic posted guidance saying Claude Code's output tokens cost about 5 times its input tokens, and that one…

The NSA, CISA, and FBI issued a joint advisory on September 8, 2026 saying Chinese firms including DeepSeek, M…

Andrew Scull, a historian of psychiatry, appeared on episode 502 of the Lex Fridman Podcast

Snap introduced Specs Intelligence, an 'anticipatory AI service' that links accounts like Gmail and Slack

Apple researchers proposed DACA-GRPO, a plug-and-play enhancement for GRPO-style trainers, adding Denoising Pr…
