Earlier quoted context omitted.
Not really. They have a way of squaring this circle, by changing their inference code. Speculative sampling [1] would still make their first claim a lie – sure, there'd still be the original GPT-4 model, plus a smaller draft worker. But early exit decoding [2] allows you to get almost as good results for much cheaper from exactly the same checkpoint. We know that this line of research for large-scale inference is goi…
I don't get why you're jumping to cloak and daggers style operations: OpenAI would not kneecap their commercial offering by randomly changing how it works. At the end of the day 99% of the confusion comes from people using the web interface, which undoubtedly does change much more often than the API versions they share. The web app they host isn't a simple API wrapper, it does summarization, has some sort of system p…
1. https://chat.openai.com/share/44a0c5b6-c629-470a-992f-8cdbbe...