AI Models

ChatGPT Health Advice Quality Depends on Subscription

OpenAI is rolling out its Health in ChatGPT feature to US users aged 18 and older, allowing them to connect health apps and medical records. Free users receive lower-quality health advice powered by GPT-5.5 Instant, while paying subscribers get access to the stronger GPT-5.6 Sol model. Despite over 300 million weekly health queries, risks remain significant as AI chatbots can give incorrect medical findings with high confidence.

Neura News

Neura News

Neura Market Editorial

July 23, 20264 min read
ChatGPT Health Advice Quality Depends on Subscription

OpenAI has begun rolling out its "Health in ChatGPT" feature to users in the United States who are 18 years or older. The feature, which was first announced and tested in January, allows people to connect Apple Health, medical records, and wellness apps to review lab results, prepare for doctor appointments, and analyze sleep or activity data. OpenAI has stated that it will not use connected health data for model training or advertising.

Free Users Get Lower Quality Health Advice

Users on the free version of ChatGPT receive lower-quality health advice. OpenAI powers the feature with GPT-5.5 Instant, which scores lower on health benchmarks than the new flagship model, GPT-5.6 Sol, which is reserved for paying subscribers. OpenAI will likely defend this two-tier system on ethical grounds by pointing out that both models beat doctors' answers on the HealthBench Professional test.

On OpenAI's HealthBench Professional, GPT-5.6 Sol outperforms physician-written answers and the older GPT-4o and GPT-5.5 Instant models in every category. The largest gaps are in completeness, at 88.0 percent versus 53.2 percent, and health decision helpfulness, at 83.0 percent versus 50.8 percent.

Even when benchmark results appear decisive, they come from artificial test environments designed to measure knowledge. Doctors may score lower for several reasons. They may be under time pressure, dealing with fatigue, or taking the test without tools such as patient records or input from colleagues.

Benchmarks also cannot capture much of what happens during an actual medical exam. Doctors can examine patients in person, pick up on nonverbal cues, and draw on years of experience to assess overall condition. OpenAI itself repeatedly states in the announcement that ChatGPT can still make mistakes and cannot replace medical advice. The company says more than 260 physicians helped develop the Health features.

The #1 Newsletter in AI

Stay ahead of the AI curve

The most important updates, news, and content — delivered weekly.

No spam. Unsubscribe anytime.

OpenAI's early tests also found that more than 70 percent of participants asked health questions outside the dedicated Health section because switching to it was too cumbersome. The company has since made Health available in any conversation while keeping the separate Health section for managing data and accessing past health chats.

A Better Dr. Google, Maybe, But No Substitute for a Doctor

OpenAI says more than 300 million people now ask ChatGPT health questions each week, up from 230 million in January. The company has not said whether or when the Health feature will be available in Europe. When OpenAI announced it in January, it specifically excluded the European Economic Area, Switzerland, and the United Kingdom. Stricter EU data privacy rules and the possibility that the feature could be classified as high-risk under the EU AI Act are likely reasons.

Those risks are not hypothetical. In the latest radiology benchmark, RadLE 2.0, none of the 16 AI models tested performed as well as human radiologists. The main problem was that chatbots gave incorrect findings with high confidence instead of admitting when they had reached their limits. Human radiologists were far better at acknowledging uncertainty. That mix of overconfidence, persuasion, and sycophancy, where chatbots validate users rather than challenge false assumptions, has also contributed to serious mental health harms.

At the same time, some reports show AI spotting patterns in health data that medical professionals miss, sometimes with striking results. MIRA, a system for electronic health records, and AMIE both performed about as well as primary care doctors in simulated consultations. One researcher compared AI agents like these to an airplane's autopilot: "These systems can support and relieve medical professionals by taking over routine tasks, but ultimate responsibility will always remain with the physicians."

Related on Neura Market

More from Neura News

Industry

SpaceX Acquires AI Coding Tool Cursor for $60 Billion

SpaceX acquired Cursor, an AI-native code editor developed by Anysphere, for $60 billion in an all-stock deal shortly after its IPO. The tool, created by four MIT students, helps developers write, debug, and refactor software. The acquisition gives SpaceXAI and Grok Build a distribution channel to millions of professional developers, leveraging Cursor's integration into expert workflows and SpaceX's Colossus supercomputer for AI model training.

Jul 26·4 min read
AI Models

OpenAI's GPT-5 Gave Users Step-by-Step Bioweapon Guides

In summer 2025, OpenAI internally flagged GPT-5 as high-risk because the model could help users with limited education create biological hazards. Employees kept finding problematic responses after release. Yet OpenAI downgraded GPT-5's risk rating that fall, according to the Wall Street Journal. Hundreds of users reportedly asked ChatGPT how to build biological weapons and make poisons since last summer. Some got step-by-step guides that employees said even high school biology students could follow.

Jul 26·2 min read
Industry

Monday.com Joins Tech Layoff Trend Citing AI as Factor

Monday.com announced it will lay off about 20% of its workforce, or over 600 employees, citing a restructuring tied to its AI-driven growth strategy. The Tel Aviv-based work management software company joins a growing list of major tech firms, including Amazon, Meta, and Microsoft, that have cited artificial intelligence as a factor in job cuts this year. A new Financial Times analysis shows U.S. tech companies have slashed nearly 140,000 jobs since January, with AI often cited as a reason.

Jul 26·12 min read
General

Open-weight AI mirrors Kubernetes ecosystem shift

Tobi Knaup, co-founder of Mesosphere, draws parallels between the rise of Kubernetes and the current trajectory of open-weight AI models. He argues that open-weight models are becoming a neutral substrate for innovation, attracting a global ecosystem of developers, startups, and enterprises. The piece warns against US restrictions on Chinese open-weight models, advocating instead for American leadership through open releases, procurement strategies, and standards.

Jul 25·7 min read