Earlier quoted context omitted.
LLMs have been improving exponentially for a few years. let's at least wait until exponential improvements slow down to make a judgement about their potential
In some domains (math and code), progress is still very fast. In others it has slowed or arguably stopped. We see little progress in "soft" skills like creative writing. EQBench is a benchmark that tests LLM ability to write stories, narratives, and poems. The winning models are mostly tiny Gemma finetunes with single-digit parameter counts. Huge foundation models with hundreds of billions of parameters (Claude 3 Opu…
Isn't that kind of obvious? Even human speakers and writers have problems changing people's minds, let alone reliably.