I think this drama was blown up a bit out of proportion. The entire discourse I am seeing online seems to revolve around this: > While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models I mean... yeah? What do you expect? What else can they say? How could you prove a negative in this case? I do not want to comment on specific OAI employee chat messa…
A very simple "these two pipelines don't connect up in our architecture, here's our internal high level network diagram combined with our data ingestion opt-out feature flag that we will stand by in court" as opposed to "yeah, we don't even entirely know how our own customer facing systems are connected to our training pipeline, but it probably didn't happen".
The Navier–Stokes Millennium Prize Problem
81–90 of 235 posts
Re: The Navier–Stokes Millennium Prize Problem
#82LLMs can't contribute good code to some of the good OSS math libraries, How is it even solving these problems?
Re: The Navier–Stokes Millennium Prize Problem
#83My take home from this entire drama is that one should not use LLM services for confidential or proprietary information as they all seem to be run by assholes. And you’re sending them everything you are doing. Would you send your lab notebook to an asshole? Hell no. I say that as a mathematician (on paper) who perhaps surprisingly doesn’t give a crap about the problem itself.
Their privacy policy for normie subscribers says in plain English they use your Personal Data for research. I think it’s pretty unreasonable to use the service and expect otherwise.
Re: The Navier–Stokes Millennium Prize Problem
#84Occam's Razor says: "They heard this problem is solved or about to be solved amongst the rest of the other problems. They prioritized this and put substantial compute with their newest model and solved it." I know everyone loves juicy rumors, theories etc. but honestly that is the simplest and most plausible explanation given the state of AI improvement now. Obviously spending 15 million on a problem is not a slam du…
Why is that simple or plausible? Why is simpler or more plausible than lifting an almost-finished solution from a researcher's account?
Re: The Navier–Stokes Millennium Prize Problem
#85I don't get the sales pitch, spend 15 million dollars to win a 1 million dollar price? Showing of the model's capabilities - okay, but it's not like it solved the problem on its own, and apparently not particularly efficient. Are there practical applications that justify the investment?
Re: The Navier–Stokes Millennium Prize Problem
#86> ... we heard rumors that two Millennium Prize problems had been resolved. Inspired by these rumors ... I've observed this exact effect last week. I made a discovery regarding a stepwise performance improvement in a codebase. I shared the benchmark results with a peer and within 12 hours they replicated the same. We had both been looking for this for years. I think giving someone hope that an answer exists might as…
It's something that happened before LLMs - multiple discovery. Calculus is a classic example.
Re: The Navier–Stokes Millennium Prize Problem
#87>While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models. What do you mean, as OpenAI employee, you cannot tell that his work has entered the training data ? But also correct me if I'm wrong, if the two mathematician were really close to finish this problem, and their conversation were used by OpenAI, shouldn't the Agent have succeeded way faster/e…
The text for all the Goosebumps books are certainly in the training data and to some small amount influenced the solve. But their contribution was so vanishingly small it would seem absurd to say R L Stein should have recourse for contibuting to the solve.
Re: The Navier–Stokes Millennium Prize Problem
#88It could be much worse. OpenAI can easily identify these outstanding human behind their accounts. Human in OpenAI constantly check their logs for breakthrough. When they find something interesting, they brute force the result using their massive computing power. No LLM is even needed.
Re: The Navier–Stokes Millennium Prize Problem
#89People talk about the "compute" but what about the "storage"? Is storage exponentially greater, or soon to be, than the compute? Is the storage going to slow down growing to some constant rate, i.e. all people on earth using chatgpt, or no, on the contrary, it will keep growing?
If there were any shady business, I do not condone it, but technologically we are not there yet for said shady business to happen.
Re: The Navier–Stokes Millennium Prize Problem
#90My take home from this entire drama is that one should not use LLM services for confidential or proprietary information as they all seem to be run by assholes. And you’re sending them everything you are doing. Would you send your lab notebook to an asshole? Hell no. I say that as a mathematician (on paper) who perhaps surprisingly doesn’t give a crap about the problem itself.
Wait until you learn how many other SaaS and web 2.0 and cloud based things are also run by assholes.
You know this metaphor? https://en.wikipedia.org/wiki/Turtles_all_the_way_down
But instead of turtles, it's assholes.
But more seriously, no, none of what I wrote above is an attempt to excuse or play down the specific role of assholes in large AI companies.