Live data from Hacker News

Notes on OpenAI o3-mini

simonwillison.net

21–30 of 81 posts

Re: Notes on OpenAI o3-mini

#23
post #9
post #6

Hasn't Gemini pricing been lower than this (or even free) for awhile? https://ai.google.dev/pricing

Are you insinuating Gemini is similar in performance to o3-mini?

Definitely varies by application, but the blind "taste test" vibes are very good for Gemini: https://lmarena.ai/?leaderboard

Re: Notes on OpenAI o3-mini

#24

Open AI really needs to work on their naming conventions for these things.

It's all based on omni which to me has weird religious connotations. It just occurred to me to put it together with sama's other project, scanning everyone's eyes. That's one aspect of omniscience - keeping track of every soul.

Another thing it seems similar to is how Jeff Bezos registered relentless.com. There seems to be a gap between the ideal branding from the perspective of the creators and branding that makes sense to consumers.

Re: Notes on OpenAI o3-mini

#25
post #9
post #6

Hasn't Gemini pricing been lower than this (or even free) for awhile? https://ai.google.dev/pricing

Are you insinuating Gemini is similar in performance to o3-mini?

I've only had o3-mini for a day, but Gemini 2.0 Flash Thinking is still clearly better for my use cases.

And it's currently free in aistudio.google.com and in the API.

And it handles a million tokens.

Re: Notes on OpenAI o3-mini

#26

Earlier quoted context omitted.

In what way would you say doctors have largely failed Western society lately?

You obviously don't have a chronic illness or you wouldn't be asking that. Either that or you're rich.

Are you conflating doctors with insurance like the larger health care industrial complex?

Re: Notes on OpenAI o3-mini

#27

Earlier quoted context omitted.

You obviously don't have a chronic illness or you wouldn't be asking that. Either that or you're rich.

Are you conflating doctors with insurance like the larger health care industrial complex?

I think doctors generally do their best, but it's still disappointing. Doctors can't make food less enticing. https://www.instagram.com/lukesmithrd/p/DBRoWvRSMKx/

Re: Notes on OpenAI o3-mini

#28
post #23
post #9

Earlier quoted context omitted.

Are you insinuating Gemini is similar in performance to o3-mini?

Definitely varies by application, but the blind "taste test" vibes are very good for Gemini: https://lmarena.ai/?leaderboard

that reminds me that a week ago there was a (now deleted but has a copy of the content available in the comments) post on Reddit where the author claimed they have attempted manipulating/manipulated voting on lmarena in favor of Gemini to tip the scale on Polymarket where on a question like "which AI model will be the best one by $date" (with the outcome decided based on the scoring on lmarena) they have supposedly made O(USD10k).

Original deleted post: https://old.reddit.com/r/MachineLearning/comments/1i83mhj/lm...

A copy of the content: https://old.reddit.com/r/MachineLearning/comments/1i83mhj/lm...

Re: Notes on OpenAI o3-mini

#29
post #16
post #9

Earlier quoted context omitted.

Are you insinuating Gemini is similar in performance to o3-mini?

Are you implying it isn't? (evidence please, everyone)

Simple example: o3-mini-high gets this [1] right, whereas Gemini 2.0 Flash 01-21 gets it wrong.

[1] https://chatgpt.com/share/679d9579-5bb8-8008-ac4a-38cef65b45...

Re: Notes on OpenAI o3-mini

#30
post #28
post #23

Earlier quoted context omitted.

Definitely varies by application, but the blind "taste test" vibes are very good for Gemini: https://lmarena.ai/?leaderboard

that reminds me that a week ago there was a (now deleted but has a copy of the content available in the comments) post on Reddit where the author claimed they have attempted manipulating/manipulated voting on lmarena in favor of Gemini to tip the scale on Polymarket where on a question like "which AI model will be the best one by $date" (with the outcome decided based on the scoring on lmarena) they have supposedly m…

[deleted]
Post reply on HN