Live data from Hacker News

Claude 3 beats Google Translate

arxiv.org

81–90 of 127 posts

Re: Claude 3 beats Google Translate

#81

Im surprised they didn’t try out Gemini. It’s a lot better than Google translate and in my experience this is one use case where Gemini outshines even other frontier models like ChatGPT4

I've also found gemini to be better than chatGPT (3.5, I'm using both of them for free).

The results have felt much more like natural language to me.

Re: Claude 3 beats Google Translate

#82
Sorry in advance for a rant against the proliferation of LLM benchmarks:

I wrote a book on using LLMs in applications 14 months ago, I am sort of a fan, within reason. I don’t like all the mind-space taken up on comparisons between LLMs on standard test suites because I personally think most instruction tuned models are specifically tuned for the standard tests.

I like to see which models people are choosing to use for their projects, and of course, tools like Ollama make it easy to try many models and get a least a subjective feel for what they can do.

For model comparison I have my own standard little tests that are my own and private, and I can be sure models are not tuned specifically for.

I do the same for commercial LLM APIs and products surfacing models. For example, I have a few short tests I run in Bard to test integration with Google Workspace data, something I am interested in, and it is interesting to track at least subjectively the slow improvement.

Re: Claude 3 beats Google Translate

#83

Is Google Translate considered "the benchmark to beat"? At least between German and English (both not exactly "fringe languages"), Google Translate quality is still very hit and miss.

Deepl has the best rep for European languages - though it now supports many others. I can say it's German translations are pretty good (as a non-native German speaker).

But ChatGPT blows it out of the water if you have a specific context and don't need something exactly translated (e.g. formal letter/e-mail writing).

Re: Claude 3 beats Google Translate

#84
post #27

Earlier quoted context omitted.

Orders of magnitude more energy necessary though.

The GPT 3.5 API is cheaper than Google Translate. And in our testing (over a year ago) was better for translation. So I assume in that case the energy use is less?

I doubt energy usage dictate the final price in such a way. You can't compare two widely different products from two company at completely different development stage to infer the energy usage of their product

Re: Claude 3 beats Google Translate

#85
post #35
post #27

Earlier quoted context omitted.

The GPT 3.5 API is cheaper than Google Translate. And in our testing (over a year ago) was better for translation. So I assume in that case the energy use is less?

Isn't OpenAI still operating at a loss? You can't infer much from an API bring cheaper if it's being subsidised.

I've not heard any claim that they're making a loss, only that they're structured as a kinda-but-its-weird not-for-profit.

Given they tripped and fell over a money printing machine and then chose to lower their API prices, it would be pretty surprising (but not impossible) if their API prices are currently subsidised.

Re: Claude 3 beats Google Translate

#87
post #45

DeepL already beat Google Translate years ago.

DeepL has its moments but it also has many many failure cases. It particularly is very bad with long text blocks, it commonly completely misses sections of text or repeats the same block of text multiple times.

Yeah this is my experience with DeepL as well. It might (sometimes) do better than Google Translate with lone sentences, but give it a few paragraphs and it'll entirely ignore some, and other times it starts repetitively rambling about stuff not even in the original text, quite frequently cars/houses/money.

Re: Claude 3 beats Google Translate

#88

Yeah this is old news. When i was in China last year google translate seemed to literally never work (looks of confusion trying to do basic interactions in stores). GPT-4 worked perfectly every time, I think google translate might be another soft abandoned project from google

It's kind of perplexing how Google Translate has not improved at all (apparently), given that the "transformer" paper ("Attention is all you need") that kickstarted this whole LLM thing was published by Google(rs) intended for translation (between English and French)...

As a sibling mentioned it's probably something about cost, but given the narrow domain and the performance of smaller models on translation tasks, I'm surprised they're still doing the same old thing in 2024....

Re: Claude 3 beats Google Translate

#89

Earlier quoted context omitted.

Where is the hallucination? It seems in line with the others.

You think this: "Swarming like a swarm of bees. He was carried among the people, hanging from the handle. No matter how good you think about the situation you're in, it's disgusting. Where are you now?" is a comparable translation to this: "No matter how you couch it, riding the subway feels disgusting: you dangle like ripe fruit from a hanging vine, squeezed in among humans swarming like bees." Or this ?: "Being cra…

I do think those are comparable.

Could be better, but it communicates the key concepts and emotional tone?

Re: Claude 3 beats Google Translate

#90

I still find it amazing how LLM translation capabilities are an almost accidental feature. Yet they still managed to leapfrog decades of research and billions of investment dollars in traditional machine translation.

And still, there are people saying that LLMs are useless. I find that even more amazing.

They probably mean useless to themselves, not useless to everyone.
Post reply on HN