Live data from Hacker News

Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

github.com

71–80 of 152 posts

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#71
post #41

Earlier quoted context omitted.

I assume it's safe to ignore as model weights aren't copyrightable, probably.

you dont know what kind of backdoors are hidden in the model weights

Can you elaborate on how any sort of backdoor could be hidden in the model weights?

It's a technical possibility to hide something in the code, but that would be a bit silly since there's not that much of it here. It's not technically possible to hide a backdoor in a set of numbers that are solely used as the operands to trivial mathematical operations, so I'm very curious about what sort of hidden backdoor you think is here.

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#72

For the AI un-initiated; is this something you could feasibly run at home? eg on a 4090? (How can I tell how "big" the model is from the github or huggingface page?)

I tried using Hunyuan3D-2 on a 4090 GPU. The Windows install encountered build errors, but it worked better on WSL Ubuntu. I first tried it with CUDA 11.3 but got a build error. Switching to CUDA 12.4 worked better. I ran it with their demo image but it reported that the mesh was too big. I removed the mesh size check and it ran fine on the 4090. It is a bit slow on my i9 14k with 128G of memory.

(I previously tried the stability 3d models: https://stability.ai/stable-3d and this seems similar in quality and speed)

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#73
post #67

Earlier quoted context omitted.

> everyone will be able to generate "something incredible" and then it all becomes not incredible. no, that's just your standard moving up. There is an absolute scale for which you can measure, and ai is approaching a point where it is an acceptable level. Imagine if you applied your argument to quality of life - it used to be that nobody had access to easy, cheap clean drinking water. Now everybody has access to it.…

It is equally childish to compare the engineering of our modern water and plumbing systems with the automated generation of virtual textured polygons. People don't get tired of good clean water because we NEED it to survive. But oh, another virtual world entirely thought up by a machine? Throw it on the pile. We're going to get bored of it, and it will quickly become not incredible.

> we NEED it to survive.

plenty of people in the world still drink crappy water, and they survive.

You don't _need_ it, you want it, because it's much more comfortable.

But when something becomes a "need" as you described it, you think of it differently. Just like how you don't _need_ electricity to survive, but it's so ingrained that you now think of it as a need.

> We're going to get bored of it, and it will quickly become not incredible.

exactly, but i have already said this in my original post - your standards just moved up.

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#74
post #41

Earlier quoted context omitted.

you dont know what kind of backdoors are hidden in the model weights

Can you elaborate on how any sort of backdoor could be hidden in the model weights? It's a technical possibility to hide something in the code, but that would be a bit silly since there's not that much of it here. It's not technically possible to hide a backdoor in a set of numbers that are solely used as the operands to trivial mathematical operations, so I'm very curious about what sort of hidden backdoor you think…

When you run their demo locally, there are two places that trigger a warning that the code loads the weights unsafely. To learn more about this issue, search "pytorch model load safety issues" on Google.

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#75
post #51

Earlier quoted context omitted.

> Take a look at the ImgnAI gallery ( https://app.imgnai.com/ ) and tell me: can you paint better and more imaginatively than that? So while I generally agree with you, I think this was a bad example to use: a lot of these are slop, with the kind of AI sheen we've come to glaze over. I'd say less than 20% are actually artistically impressive / engaging / thought-provoking.

This is a better AI gallery (I sorted all images on the site by top from this year). https://civitai.com/images There's still plenty of slop in there, and it would be a better gallery of if there was a way to filter out anime girls. But it's definitely higher than 20% interesting to me. The closest similar community of human made art is this: https://www.deviantart.com/ Although unfortunately they've decided to allow…

I think you fundamentally misunderstand what people use "slop" to describe.

> Most human made art is slop too.

I'm assuming you're using the term "slop" to describe low-quality, unpolished works, or works where the artist has been too ambitious with their skill level.

Let me put it this way:

Every piece of art that is made, is a series of decisions. The artist uses their lived experience, their tastes and their values to create something that's meaningful to them. Art doesn't need to have a high-level of technical expertise to be meaningful to others. It's fundamentally about communication from artists to their audience. To this point, I don't believe there's such a thing as "bad art" (all works have something to say about the artist!).

In contrast, when you prompt an image generator, you're handing over the majority of the decisions to the algorithm. You can put in your subject matter, poses, even add styles, but how much is really being communicated here? Undoubtedly it would require a high level of technical skill to render similarly by hand, but that's missing the forest for the trees- what is the image saying? There's a reason why most "good" AI-generated images generally have a lot of human curation and editing.

Here's an example of some "slop" from the AI Art Turing Test (https://www.astralcodexten.com/p/how-did-you-do-on-the-ai-ar...) from a while back: https://i.imgur.com/RAMFKP1.jpeg There's definitely a high level of technical expertise that a human would require to paint something like this. But it's very clearly AI-generated. Can you figure out why?

---

As a side note, here's a human-made piece that I appreciate a lot. https://i.imgur.com/AZiiZj1.jpeg The longer you explore it, the more the story unfolds, it's quite lovely. On the other hand, when I focus on the details in AI-generated works, there's not much else to see.

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#76
post #63

Earlier quoted context omitted.

I already feel like text models are already at sufficiently entertaining and useful quality as you define it. It's definitely possible we never get there for video or 3D modalities, but I think there are strong enough economic incentives such that big tech will dump tens of billions of dollars into achieving it.

I don't know why you think that's the case regarding text models. If that was the case, there would be articles on here that are just created by only generative AI and nobody would know the difference. It's pretty obvious that's not happening yet, not the least of which because I know what kinds of slop state-of-the-art generative models still produce when you give them open-ended prompts.

Ironic how this comment exemplifies the issue - broad claims about "slop" output but no specific examples or engagement with current architectures. Real discussions here usually reference benchmarks or implementation details.

(from Claude)

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#77
post #76
post #63

Earlier quoted context omitted.

I don't know why you think that's the case regarding text models. If that was the case, there would be articles on here that are just created by only generative AI and nobody would know the difference. It's pretty obvious that's not happening yet, not the least of which because I know what kinds of slop state-of-the-art generative models still produce when you give them open-ended prompts.

Ironic how this comment exemplifies the issue - broad claims about "slop" output but no specific examples or engagement with current architectures. Real discussions here usually reference benchmarks or implementation details. (from Claude)

You're sort of ignoring the issue? If the generated content was good and interesting enough on it's own, we would already have ai publishing houses pushing out entire trilogies, and each of those would be top sellers.

Generative content right now is OK. OK isn't really the goal, or what anyone wants.

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#78
post #76
post #63

Earlier quoted context omitted.

I don't know why you think that's the case regarding text models. If that was the case, there would be articles on here that are just created by only generative AI and nobody would know the difference. It's pretty obvious that's not happening yet, not the least of which because I know what kinds of slop state-of-the-art generative models still produce when you give them open-ended prompts.

Ironic how this comment exemplifies the issue - broad claims about "slop" output but no specific examples or engagement with current architectures. Real discussions here usually reference benchmarks or implementation details. (from Claude)

I'm sorry I didn't meet the LLMs expectations, but whether something is subjectively entertaining or not can't be exemplified by objective benchmarks.

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#79
post #35

Question related to 3D mesh models in general: has any significant work been done on models oriented towards photogrammetry? Case in point, I have a series of photos (48) that capture a small statue. The photos are high quality, the object was on a rotating platform. Lighting is consistent. The background is solid black. These normally are ideal variables for photogrammetry but none of the various common applications…

> the object was on a rotating platform

Isn't a static-object-rotating-camera basically a requirement for photogrammetry?

Re: Hunyuan3D 2.0 – High-Resolution 3D Assets Generation

#80
post #73

Earlier quoted context omitted.

It is equally childish to compare the engineering of our modern water and plumbing systems with the automated generation of virtual textured polygons. People don't get tired of good clean water because we NEED it to survive. But oh, another virtual world entirely thought up by a machine? Throw it on the pile. We're going to get bored of it, and it will quickly become not incredible.

> we NEED it to survive. plenty of people in the world still drink crappy water, and they survive. You don't _need_ it, you want it, because it's much more comfortable. But when something becomes a "need" as you described it, you think of it differently. Just like how you don't _need_ electricity to survive, but it's so ingrained that you now think of it as a need. > We're going to get bored of it, and it will quickl…

Are the ai worlds the dirt or the water here?
Post reply on HN