> not-entirely-useless but still quite crappy15 months ago, general-purpose LLMs that have not been specifically trained on legal reasoning could score better than 90% of humans on the multistate bar exam, and these are humans who actually completed law school.
General-purpose LLMs get similar results in medicine, and when the models are fine-tuned for medical diagnosis they're even better.
And that was more than a year ago. Those who have seen current models not yet released tell us they'll make the current state of the art look like toys.
Progress is still tracking the steep part of the S-curve, and there's no indication that they're near the top yet.