LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
1–10 of 17 posts
Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#2Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#3Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#4The fact that even the abstract is very clearly AI-generated does not fill me with confidence in the rigour of the research.
EDIT: And Pangram agrees that the abstract is 100% AI generated.
Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#5The fact that even the abstract is very clearly AI-generated does not fill me with confidence in the rigour of the research.
Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#6The fact that even the abstract is very clearly AI-generated does not fill me with confidence in the rigour of the research.
Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#7The fact that even the abstract is very clearly AI-generated does not fill me with confidence in the rigour of the research.
Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#8The fact that even the abstract is very clearly AI-generated does not fill me with confidence in the rigour of the research.
I know arXiv has taken some measures to combat spam like this, but it seems like they’ll need to do more. There’s just very little barrier now to creating giant slop papers like this and then dumping them anywhere that won’t reject them. It is an insult to everyone’s time, and I can’t imagine they expect people to actually read this. If the expectation is that everyone will use an LLM to interpret it, then maybe they should have at least had a few more rounds of tightening and polishing the paper, even via LLM, to save the redundant token use.
Re: LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes
#9The fact that even the abstract is very clearly AI-generated does not fill me with confidence in the rigour of the research.
I understand the answer to the question I'm about to ask, but how does any human read that abstract and not think, "This is entirely too many colons." EDIT: And Pangram agrees that the abstract is 100% AI generated.