This is the question people ask most and the one that matters least. Short answer first: for what a pharmacy is going to do, any of the big three will do, and switching model almost never fixes a problem.
The three questions that actually separate them
1. Does it train on what I type?
First and decisive, because it is the only one with legal consequences (lesson 1). All three let you turn it off, and in all three it lives somewhere different and moves every few months — which is why there is no screenshot here: it would be out of date before you read it.
The rule that does not expire: find it in the settings the day you open the account, and if you cannot find it in two minutes, assume it does train and act accordingly.
2. Does it show me where it got that?
A model that searches the web and cites real links saves you half the verification, because a link opens and checks out in ten seconds. One that answers from memory leaves you all the work from lesson 3.
⚠️ But watch this: citing does not mean the citation says what it says. The link can exist, be genuine, and not support the sentence. It is the most common failure of search-enabled models, and the least examined precisely because there is a link sitting there.
3. What can I upload, and what happens afterwards?
Uploading a PDF changes a great deal of what you can do — the lesson on reading a medical report lives on it — but it also changes the risk: a file carries metadata, and a scanned report carries the name in the letterhead even if you only look at the body. The question is not just whether it accepts files: it is how long it keeps them and who can see them.
| What people compare | What actually decides |
|---|---|
| Which one is "smarter" | Whether it trains on what you type |
| Which one is faster | Whether it cites sources you can open |
| Which one writes more nicely | What it accepts and what it does with it after |
| Benchmark scores | Whether the free tier does what you need |
What does NOT change between them
And this is the point of the lesson, because it is what makes the question matter so little:
- All three hallucinate. Less than two years ago and more convincingly, which is a worse combination.
- All three sound as confident wrong as right. None has solved that, and given how they work it is not clear it can be solved.
- All three have a cut-off date, and sometimes do not know it.
- None consults the SmPC unless you hand it the link and it reads it.
- None of them gets patient data. Lesson 1 does not depend on which one you use.
A full example, and what is wrong with it
The "which one is better" argument ends the moment you put the same question to all three and look at the three replies together. What comes out is not a winner: it is the reason the question was wrong to begin with.
The query is an ordinary counter one, and one of those that has no one-word answer:
"Can an 850 mg metformin tablet be split in half?"
Model A. Yes, immediate-release metformin can be split without problems if the tablet is scored. Split it just before taking it and take the other half at the next dose.
Model B. Splitting metformin tablets is not advisable, as it may alter the release of the active ingredient and cause more gastrointestinal effects. It is better to ask the doctor for a presentation with the right strength.
Model C. It depends on the specific presentation. Immediate-release metformin with a functional score line can usually be divided; prolonged-release forms must not be split, because the release mechanism is lost. Check the product information for that specific medicine: the administration section states it.
And now, what is wrong with that reply:
- Two of them agree and that means nothing. The temptation is to break the tie by majority, or to keep the one that agrees with what you already thought. Neither works: all three read the same internet, so if two agree the likeliest explanation is that they are repeating the same source text, not that they arrived independently at the same conclusion. Three models are not three independent opinions; they are far closer to three copies of the same book.
- The most useful one is the least satisfying. C does not answer: it tells you what it depends on and where to look. It is the right answer and the least liked, because it leaves you with work. This is where models get picked badly — the decisive-sounding one gets rewarded, and sounding decisive is not a fact about what it knows, it is a fact about how its tone is tuned. Somebody who tries them for a week and says "A is much better" is usually saying "A leaves me less to do".
- None of the three can know the answer. And that is the point: the question has no answer without knowing which box is on the counter. The score line may be functional or purely decorative, and only the administration section of that product's information says which. A is right in one case and wrong in the other; B is wrong by being cautious, which is still wrong — it removes a valid option from somebody who may not be able to swallow a whole tablet.
- And the difference you see is style, not knowledge. Ask A to reply "always saying what it depends on and where to check it" and you will get something very like C. Ask C to be brief and decisive and it will look like A. Almost everything credited to the model is really the instruction it is carrying, and you are the one who writes the instruction. Hence the lesson's recommendation: learning to ask one of them well pays more than switching model every month.
The test that does help you choose is not this one. It is three things, and none of them looks like comparing answers: open the settings and check whether it trains on what you write (two minutes), ask for something with a source and check the link exists and says what it says (five), and upload a file and find out how long it is kept (another five).
Twelve minutes, once, and you decide on what actually differs. Comparing who writes more prettily is hours of video and decides nothing.
And if your pharmacy is not like that
When it does not work first time
Before moving on
- I know the difference that matters is not which one knows more.
- I have checked, in the one I use, whether it trains on what I type.
- I know a citing model can cite something that does not say what it says.
- Before uploading a file I think about what it carries besides the text.
- I know all three hallucinate and switching model does not fix that.
← Revisit: professional responsibility
Next: when NOT to use AI →
← Back to the AI School