"Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
41–50 of 110 posts
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#42GPT 5.6 Sol had the best two drawings (rose and starry nights) but even more impressive was how efficient it was RE cost/time/tokens vs Fable (3.4M vs 14.6M / $7.74 vs $161!). OpenAI has quietly innovated around inference - this is will be a growing differentiator even against open models.
At work, even if Fable is technically better I much prefer Sol because it is so much faster and concise.
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#43As I looked through the images I was unimpressed entirely, at first. But, then I started thinking, these look a little... "childish" to me. Childish as in... A newish artist who is drawing a concept rather than light / forms (Which is something artists typically do as they understand drawing more and more). The rose in the vase specifically - some models understood that there was supposed to be shading, reflections,…
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#44Earlier quoted context omitted.
Parrots can also sound extremely human, but it’s only mimicry. How do we know the difference?
Yes, but Parrots don't mimic the stages of learning speech development like children do. They just memorize a phrase. These SOTA LLMs aren't trying to mimic existing children's drawings, but interestingly they're following somewhat similar progression that human children do as they develop.
AI labs are improving the ML techniques used to build better models; it’s a big difference.
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#45Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#46Earlier quoted context omitted.
What's interesting is that the way in which they're childish is actually extremely human. In fact, one of the ways that you're often taught to draw more realistically is to stop thinking of the concepts as icons you're drawing the outlines of, and instead sort of blur your eyes and see things as they are: hues and values. In other words, become a camera or a printer that has no idea what it's capturing or printing ot…
Parrots can also sound extremely human, but it’s only mimicry. How do we know the difference?
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#47GPT 5.6 Sol had the best two drawings (rose and starry nights) but even more impressive was how efficient it was RE cost/time/tokens vs Fable (3.4M vs 14.6M / $7.74 vs $161!). OpenAI has quietly innovated around inference - this is will be a growing differentiator even against open models.
I do think they are going to stretch their lead in value if Anthropic doesn't wake up and stop YOLOing tokens. Kimi is an amazing achievement, but it has the same (or worse) kitchen sink approach as Fable. At work, even if Fable is technically better I much prefer Sol because it is so much faster and concise.
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#48Earlier quoted context omitted.
Not useless. LLMs are the most general purpose computer algorithms ever created. They are getting smarter and cheaper at a geometric rate. What is a bad idea today could have useful applications tomorrow.
> cheaper at a geometric rate Citation needed - my company is paying more than ever for code generation. I have no reason to believe (given anecdotes) that anyone finds themselves in the opposite situation.
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#49Earlier quoted context omitted.
What's interesting is that the way in which they're childish is actually extremely human. In fact, one of the ways that you're often taught to draw more realistically is to stop thinking of the concepts as icons you're drawing the outlines of, and instead sort of blur your eyes and see things as they are: hues and values. In other words, become a camera or a printer that has no idea what it's capturing or printing ot…
Parrots can also sound extremely human, but it’s only mimicry. How do we know the difference?
Re: "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
#50I wonder how much better a harness could get for drawing
- the best image related stuff I've seen is where the harness is constantly cropping and looking closer at things (likely helps a lot for computer vision in general) - also it would be interesting because the harness could almost have its own "palette" as if it could play with blank squares and different strokes over lapping or blending before applying to the main canvas