Exactly. So the organisations creating and serving these models need to be clearer about the fact that they're not general purpose intelligence, and are in fact contextual language generators.
I've seen demos of the models used as actual diagnostic aids, and they're not LLMs (plus require a doctor to verify the result).
Precisely. Many of the narrowly scoped solutions work really well, too (for what they're advertised for).
As of today though, they're nowhere near reliable enough to replace doctors, and any breakthrough on that front is very unlikely to be a language model IMO.