how is GPT4-Turbo higher than GPT-4?
No clue, but it wouldn't surprised me if they identified inputs and training that weren't actually helping. Just as human beings aren't necessarily more helpful by having more varied input, the same seems to apply to LLMs. The interesting thing is it still has a 128k context length. This is awesome because GPT became way more useful to me once it reached this level of context.
Google's Bard shows big leap on LLM performance leaderboard
41–50 of 96 posts
Re: Google's Bard shows big leap on LLM performance leaderboard
#42this leaderboard seems easily cheated/gamed. once enough eyes are on it it will be worthless
Re: Google's Bard shows big leap on LLM performance leaderboard
#43It's not a valid comparison. Bard uses Google (ie the Internet) to answer things, while GPT4 doesn't. So bard can answer things like, "what's the weather in SF today" or "who won the basketball game last night".
Re: Google's Bard shows big leap on LLM performance leaderboard
#44It's not a valid comparison. Bard uses Google (ie the Internet) to answer things, while GPT4 doesn't. So bard can answer things like, "what's the weather in SF today" or "who won the basketball game last night".
Re: Google's Bard shows big leap on LLM performance leaderboard
#45> Bard, powered by the Gemini Pro-scale model, debuts at the #2 position on the independent lmsys leaderboard.
According to Jeff Dean's tweet, it looks like they have a new "Gemini Pro-scale model" being rolled out, not sure what it means by "Pro-scale" though. Also not sure if everyone already got it...
Re: Google's Bard shows big leap on LLM performance leaderboard
#46It's not a valid comparison. Bard uses Google (ie the Internet) to answer things, while GPT4 doesn't. So bard can answer things like, "what's the weather in SF today" or "who won the basketball game last night".
Chat GPT4 does do internet searches now.
Re: Google's Bard shows big leap on LLM performance leaderboard
#47It's not a valid comparison. Bard uses Google (ie the Internet) to answer things, while GPT4 doesn't. So bard can answer things like, "what's the weather in SF today" or "who won the basketball game last night".
I'm not sure what you mean. ChatGPT with GPT4 uses Bing to search the web.
Re: Google's Bard shows big leap on LLM performance leaderboard
#48Re: Google's Bard shows big leap on LLM performance leaderboard
#49Wow. I've suspected for a while that Bard's performance has been limited mostly by cost. Google isn't charging for Bard and they didn't want to run a gigantic model for everyone for free forever. Maybe they made a breakthrough in inference cost for their better models? Or maybe they got tired of everyone clowning on them for being behind and decided to eat the cost for a while. I still think they ought to launch a su…
The trick is to access the "bard-jan-24-gemini-pro" model, available in direct chat mode here: https://chat.lmsys.org/ . Significantly better than the prior model.
Re: Google's Bard shows big leap on LLM performance leaderboard
#50From all free LLMs I find bard to be most useful. Chatgpt 3.5 is not even close and it lazy.