
I've been watching OpenAI's latest move very closely. They have rolled out a new feature called ChatGPT Health directly to adult users in the US. Whether you are on their free plan or a paid tier, you can now ask medical questions right in your main chat window. Under the hood, this new tool runs on their GPT-5.6 Sol model.
To me, what they have built here is technically very impressive, but it raises some massive red flags. Let's look at how it works, where it fails, and why we need to be incredibly careful.
The way OpenAI set up the plumbing for this system is quite clever. They built a private pipeline to handle your personal health records and tracking data.
Here is how your information moves through their system:
While this setup sounds highly secure, the actual medical advice it gives is a different story. I always tell people that impressive-sounding accuracy numbers can easily hide where the real danger lives.
Take a look at other companies in this space. For example, a company called Hippocratic AI has been bragging about a 99.38% accuracy rate for their Polaris 3.0 model. But what are they actually doing? They are just handling administrative phone calls and paperwork, not making real medical decisions.
When we look at actual clinical testing for sorting sick patients—what doctors call triaging—the numbers tell a much scarier story. A big study published in Nature Medicine tested OpenAI's tool on 960 different medical situations across 21 specialties. Here is what they found:
In my view, this proves that when it comes to health, the average score does not matter. What matters is how the system works at the extremes. If a system is wrong nearly half the time during an emergency, it is simply too dangerous to rely on.
These are not just theoretical worries. I've been following a lawsuit filed in Florida by a man named Scott Winters. He used an older OpenAI model, GPT-4o, to ask about high blood pressure and groin pain. Instead of telling him to see a doctor immediately, the AI told him to rest and stay still. He ended up in the hospital with a life-threatening blood clot in his lungs. This is exactly why we cannot treat AI like a real doctor.
If you do decide to use these AI tools for minor health questions, I highly recommend following a few strict rules to keep yourself safe:
At the end of the day, AI can be a helpful assistant, but it is not a doctor. When your life is on the line, always trust a human expert.
Anthropic's Claude 5 models (Sonnet, Opus, Fable, Mythos) offer varied speeds, costs, and capabilities. Safety routing directs sensitive queries.
Small businesses need AI governance to avoid vendor lock-in, liability, and regulatory risks. Control data & processes; don't just use free tools.
AI pipelines transform scattered data into knowledge, accelerating rare disease research by connecting disparate information and forming new insights.