DreamFusion: Text-to-3D using 2D Diffusion
81–90 of 208 posts
Re: DreamFusion: Text-to-3D using 2D Diffusion
#82The samples are lacking definition, but they're otherwise spatially stable across perspectives.
That's something that's been struggled with for years.
Re: DreamFusion: Text-to-3D using 2D Diffusion
#83The most incredible thing here is that this demonstrates a level of 3D understanding that I didn't believe existed in 2D image models yet. All of the 3D information in the output was inferred from the training set, which is exclusively uncurated and unsorted 2D still images. No 3D models, no camera parameters, no depth maps. No information about picture content other than a text label (scraped from the web and often…
So I wonder if unusual angles that normally do not get photographed will be distorted? For example, underneath a table looking up.
The example with a squirrel wearing a hoodie demonstrates an interesting edge case, the "front" of the squirrel (with hoodie over the head) show a normal hooded face as expected, but when you rotate to the "back" you get another face where the hoodie is low over the eyes. Each looks fine in isolation, but in aggregate it seems like we have a two-faced squirrel.
Re: DreamFusion: Text-to-3D using 2D Diffusion
#84The most incredible thing here is that this demonstrates a level of 3D understanding that I didn't believe existed in 2D image models yet. All of the 3D information in the output was inferred from the training set, which is exclusively uncurated and unsorted 2D still images. No 3D models, no camera parameters, no depth maps. No information about picture content other than a text label (scraped from the web and often…
So I wonder if unusual angles that normally do not get photographed will be distorted? For example, underneath a table looking up.
It'll just make up some colours and geometries that don't contradict anything it already knows from the defined perspectives.
Or leave it empty.
Re: DreamFusion: Text-to-3D using 2D Diffusion
#85Earlier quoted context omitted.
> This seems like basically plugging a couple of techniques together that already existed [...] In his Lex Fridman interview, John Carmack makes similar assertions about this prospect for AGI: That it will likely be the clever combination of existing primitives (plus maybe a couple novel new ones) that make the first AGI feasible in just a couple thousand lines of code.
That's a great example that reminds me of another one: there was nothing new about Bitcoin conceptually, it was all concepts we already had just in a new combination. IRC, Hashing, Proof of Work, Distributed Consensus, Difficulty algorithms, you name it. Aside from Base58 there wasn't much original other than the combination of those elements.
Re: DreamFusion: Text-to-3D using 2D Diffusion
#86Re: DreamFusion: Text-to-3D using 2D Diffusion
#87Earlier quoted context omitted.
That's a great example that reminds me of another one: there was nothing new about Bitcoin conceptually, it was all concepts we already had just in a new combination. IRC, Hashing, Proof of Work, Distributed Consensus, Difficulty algorithms, you name it. Aside from Base58 there wasn't much original other than the combination of those elements.
Base58 really should have been base57.
Re: DreamFusion: Text-to-3D using 2D Diffusion
#88Did we hit some sort of technical inflection point in the last couple of weeks or is this just coincidence that all of these ML papers around high quality procedural generation are just dropping every other day?
The pace at which methods scale up is currently a lot faster than hardware improvements, so unless these scaled up methods become incredibly lucrative (not impossible), I think it's quite likely we'll soon-ish (a couple years from now) see a slowdown.
Re: DreamFusion: Text-to-3D using 2D Diffusion
#89Did we hit some sort of technical inflection point in the last couple of weeks or is this just coincidence that all of these ML papers around high quality procedural generation are just dropping every other day?
Re: DreamFusion: Text-to-3D using 2D Diffusion
#90Earlier quoted context omitted.
Base58 really should have been base57.
Hello Stavros, I agree. When I look at the goals that base58 sought to achieve, (eliminating visually similar characters) I couldn't help but wonder why more characters were not eliminated. There is quite a bit of typeface androgyny when you consider case and face.