Anyone know the difference between Bard Gemini Pro and Gemini Pro Dev API on the leaderboard?
Google's Bard shows big leap on LLM performance leaderboard
31–40 of 96 posts
Re: Google's Bard shows big leap on LLM performance leaderboard
#32Re: Google's Bard shows big leap on LLM performance leaderboard
#33I'm curious about how the benchmark is done. I suspect it can be improved in order to represent user's / usability expectations. I gave Bard a go, after seeing Jeff Dean's tweet. It's just as frustrating as it was, compared to GPT-4. It's simply off the question and unable to realize it's off. I asked it to generate a chart and 3 times it came back with "here's a chart" with no chart, finally saying it doesn't have t…
Bard is free, while GPT-4 is not, so it doesn’t seem like a totally fair comparison. Also what a wild comparison, because afaik chat gpt can’t make charts either.
Wait so if Google suddenly started charging for Bard, it would be instantly better?
Re: Google's Bard shows big leap on LLM performance leaderboard
#34Wow. I've suspected for a while that Bard's performance has been limited mostly by cost. Google isn't charging for Bard and they didn't want to run a gigantic model for everyone for free forever. Maybe they made a breakthrough in inference cost for their better models? Or maybe they got tired of everyone clowning on them for being behind and decided to eat the cost for a while. I still think they ought to launch a su…
Re: Google's Bard shows big leap on LLM performance leaderboard
#35Re: Google's Bard shows big leap on LLM performance leaderboard
#36Re: Google's Bard shows big leap on LLM performance leaderboard
#37Earlier quoted context omitted.
Bard is free, while GPT-4 is not, so it doesn’t seem like a totally fair comparison. Also what a wild comparison, because afaik chat gpt can’t make charts either.
> Bard is free, while GPT-4 is not, so it doesn’t seem like a totally fair comparison. Wait so if Google suddenly started charging for Bard, it would be instantly better?
Re: Google's Bard shows big leap on LLM performance leaderboard
#38I'm curious about how the benchmark is done. I suspect it can be improved in order to represent user's / usability expectations. I gave Bard a go, after seeing Jeff Dean's tweet. It's just as frustrating as it was, compared to GPT-4. It's simply off the question and unable to realize it's off. I asked it to generate a chart and 3 times it came back with "here's a chart" with no chart, finally saying it doesn't have t…
Re: Google's Bard shows big leap on LLM performance leaderboard
#39how is GPT4-Turbo higher than GPT-4?
The interesting thing is it still has a 128k context length. This is awesome because GPT became way more useful to me once it reached this level of context.