So, on one hand you can get an expert take that might be wrong, but is linked to an actual person, with a reputation and some level of ownership. On the other hand you have an over confident LLM that might be wrong and has no reputation, no ownership. How does that actually improve things compared to the older status quo?
It is entirely clear that is it not the case when I compare agent generated apps with what contracted software teams have produced.
For medicine it is likely worse.
Doctors do care, but they have to give an advice based on their 15 year old knowledge - they simply can't read through 83 papers in a quick session.
I think it is a matter of time before we see the first insurance companies assign greater risk to human advice (legal, tech, medicine, etc.) than to agentic advice.
You likely have to change your idea about this. Heck, this view was wrong 6 months ago. Sticking with is becomes a hazard to patients.
That's not my argument... I didn't draw the conclusion that because of those 2 factors humans generally provide better answers. I personally have no idea if that's the case, and for sure wouldn't rely on my personal feelings to evaluate that
> So, on one hand you can get an expert take that might be wrong, but is linked to an actual person, with a reputation and some level of ownership.
Regardless, I agree.
The newest studies still works on llms that are two years old. It doesn't appear that proper medical harnesses with frontier models has been evaluated.
My slight intuition is that we will already now see results that are much better than average human doctors.
That's the whole point, yes. LLMs are inherently unreliable. Both human experts and LLMs are unreliable in their own ways. The human expert has a reputation and some level of responsibility, the LLM doesn't