Earlier quoted context omitted.
> I can only hope this data is being incorporated in some way that makes hallucinations less likely. The key word is "real-time". LLMs can't be trained in realtime, so it's obviously going to call an API that pulls up and reads from AP news, just like their search engine.
i don't think you can assume that - "real time" in this context could just mean they feed every article into their training system as soon as it's published.
Working with The Associated Press to provide fresh results for the Gemini app
21–30 of 68 posts
Re: Working with The Associated Press to provide fresh results for the Gemini app
#22Earlier quoted context omitted.
> I can only hope this data is being incorporated in some way that makes hallucinations less likely. The key word is "real-time". LLMs can't be trained in realtime, so it's obviously going to call an API that pulls up and reads from AP news, just like their search engine.
i don't think you can assume that - "real time" in this context could just mean they feed every article into their training system as soon as it's published.
Fine-tuning, which is cheaper and faster, has been proven to not be a good solution to "teach" models new facts.
I think what's most likely here is that Gemini will have access to a form of RAG based on a database of AP articles that gets updated in real-time as new articles are published.
Re: Working with The Associated Press to provide fresh results for the Gemini app
#23Earlier quoted context omitted.
Gemini is the leading model with the lowest hallucination rate: https://www.visualcapitalist.com/ranked-ai-models-with-the-l... I would expect that number to go down from 1.3% to below 1% over the course of the year. There's always a chance what you're reading is wrong - due to purposeful deception, negligence, or accident. Realistically, hardly anything is 100% accurate besides math.
1.3% isn't great. I'd rather just go, and pay, directly to trusted news sources. Everyone has different tolerance for falsehoods and priorities I guess.
Re: Working with The Associated Press to provide fresh results for the Gemini app
#24Earlier quoted context omitted.
Gemini is the leading model with the lowest hallucination rate: https://www.visualcapitalist.com/ranked-ai-models-with-the-l... I would expect that number to go down from 1.3% to below 1% over the course of the year. There's always a chance what you're reading is wrong - due to purposeful deception, negligence, or accident. Realistically, hardly anything is 100% accurate besides math.
The Gemini models themselves may score well on this, but Google's feature implementations are a whole other thing. AI Overviews frequently take untrustworthy search results (like a fan fiction plot outline for Encanto 2) and turn those into confidently incorrect answers. https://simonwillison.net/2024/Dec/29/encanto-2/
Re: Working with The Associated Press to provide fresh results for the Gemini app
#25Earlier quoted context omitted.
> I can only hope this data is being incorporated in some way that makes hallucinations less likely. The key word is "real-time". LLMs can't be trained in realtime, so it's obviously going to call an API that pulls up and reads from AP news, just like their search engine.
i don't think you can assume that - "real time" in this context could just mean they feed every article into their training system as soon as it's published.
Re: Working with The Associated Press to provide fresh results for the Gemini app
#26Earlier quoted context omitted.
Gemini is the leading model with the lowest hallucination rate: https://www.visualcapitalist.com/ranked-ai-models-with-the-l... I would expect that number to go down from 1.3% to below 1% over the course of the year. There's always a chance what you're reading is wrong - due to purposeful deception, negligence, or accident. Realistically, hardly anything is 100% accurate besides math.
1.3% isn't great. I'd rather just go, and pay, directly to trusted news sources. Everyone has different tolerance for falsehoods and priorities I guess.
Re: Working with The Associated Press to provide fresh results for the Gemini app
#27Earlier quoted context omitted.
i don't think you can assume that - "real time" in this context could just mean they feed every article into their training system as soon as it's published.
Have you ever asked an LLM what time it is? It takes months to train them...
Re: Working with The Associated Press to provide fresh results for the Gemini app
#28Earlier quoted context omitted.
> I can only hope this data is being incorporated in some way that makes hallucinations less likely. The key word is "real-time". LLMs can't be trained in realtime, so it's obviously going to call an API that pulls up and reads from AP news, just like their search engine.
i don't think you can assume that - "real time" in this context could just mean they feed every article into their training system as soon as it's published.
Re: Working with The Associated Press to provide fresh results for the Gemini app
#29As someone who works in the news industry I find it pretty sad that we've just capitulated to big tech on this one. There are countless examples of AI summaries getting things catastrophically wrong, but I guess Google has long since decided that pushing AI was more important than accurate or relevant results, as can also be seen with their search results that simply omit parts of your query. I can only hope this dat…
The on device model that it uses is also literally 1% the size of the large models like Gemini