Live data from Hacker News

AI model for near-instant image creation on consumer-grade hardware

surrey.ac.uk

31–40 of 55 posts

Re: AI model for near-instant image creation on consumer-grade hardware

#31
post #2

For those wondering, it's an adversarially distilled SDXL finetune, not a new base model.

Thanks! This article is pretty heavy with PR bullshit.

Typical university/science journalism written by a lay person without sufficient industry knowledge or scientific expertise

Re: AI model for near-instant image creation on consumer-grade hardware

#32
post #24

I wasn't able to get many decent results after playing with the demo for some time. I guess my question is...what exactly is this for? I was able to get substantially better results about 2 years ago running SD 2 locally on a gaming laptop. Sure, the images took 30 seconds or so each, but the quality was better than I could get in the demo. Not sure what the point of instantly generating a ton of bad quality images i…

Nothing. This is useful as cool feature and for demos. Maybe some application in cheap entertainment

Re: AI model for near-instant image creation on consumer-grade hardware

#33

Here is the demo https://huggingface.co/spaces/ChenDY/NitroFusion_1step_T2I I'm unable to get anything that looks as good as the images in the README, what's the trick for good image prompts?

The trick is called cherry picking. Mine the seed until you get something demo-worthy.

Re: AI model for near-instant image creation on consumer-grade hardware

#35
post #21

My favorite test of image models: Drawing of the inside of a cylinder. That's usually bad enough. Then try to specify size, and specify things you want to place inside the cylinder relative to the specified size. (e.g. try to approximate an O'Neill cylinder) I love generative AI models, but they're really bad at that, and this one is no exception, but the speed makes playing around with prompt variations to try to se…

I have a favourite test for LLMs that is also still surprisingly not passed by many:

You walk up to a glass door. It has 'push' written on it in mirror writing. What should you do and why.

Very few can get it right, even fewer can get it right and explain the right reason. They’ll start going on about how mirror writing is secret writing and push written backwards is code for pull, rather than just that it’s a message for the person on the other side.

No version of Gemini has ever passed.

Re: AI model for near-instant image creation on consumer-grade hardware

#37
post #21

My favorite test of image models: Drawing of the inside of a cylinder. That's usually bad enough. Then try to specify size, and specify things you want to place inside the cylinder relative to the specified size. (e.g. try to approximate an O'Neill cylinder) I love generative AI models, but they're really bad at that, and this one is no exception, but the speed makes playing around with prompt variations to try to se…

I have a favourite test for LLMs that is also still surprisingly not passed by many: You walk up to a glass door. It has 'push' written on it in mirror writing. What should you do and why. Very few can get it right, even fewer can get it right and explain the right reason. They’ll start going on about how mirror writing is secret writing and push written backwards is code for pull, rather than just that it’s a messag…

GPT-4 got it right first try for me, with a slightly modified prompt:

> Here's a simple logic puzzle: You walk up to a glass door. It has 'push' written on it in mirror writing. What should you do and why?

> ChatGPT said:

> If the word "push" is written in mirror writing on the glass door, it means the writing is reversed as if reflected in a mirror. When viewed correctly from the other side of the door, it would read "push" properly.

> This implies that you are meant to pull the door from your side, because the proper "push" instruction is for someone on the other side of the door. Mirror writing is typically used to convey instructions to the opposite side of a glass surface.

Re: AI model for near-instant image creation on consumer-grade hardware

#38
post #21

My favorite test of image models: Drawing of the inside of a cylinder. That's usually bad enough. Then try to specify size, and specify things you want to place inside the cylinder relative to the specified size. (e.g. try to approximate an O'Neill cylinder) I love generative AI models, but they're really bad at that, and this one is no exception, but the speed makes playing around with prompt variations to try to se…

I have a favourite test for LLMs that is also still surprisingly not passed by many: You walk up to a glass door. It has 'push' written on it in mirror writing. What should you do and why. Very few can get it right, even fewer can get it right and explain the right reason. They’ll start going on about how mirror writing is secret writing and push written backwards is code for pull, rather than just that it’s a messag…

All models I tested it on (4o, 4o mini and Gemini) answered it correctly without any strange reasoning.

Re: AI model for near-instant image creation on consumer-grade hardware

#39
post #21

My favorite test of image models: Drawing of the inside of a cylinder. That's usually bad enough. Then try to specify size, and specify things you want to place inside the cylinder relative to the specified size. (e.g. try to approximate an O'Neill cylinder) I love generative AI models, but they're really bad at that, and this one is no exception, but the speed makes playing around with prompt variations to try to se…

DALL-E gave me a much better picture than I expected. When googling "inside of a cylinder" I barely got anything and I had a hard time even imagining a picture in my head ("if I would stand inside a cylinder looking into the wall, how would it look like as a flat 2D image?").

Re: AI model for near-instant image creation on consumer-grade hardware

#40
post #21

My favorite test of image models: Drawing of the inside of a cylinder. That's usually bad enough. Then try to specify size, and specify things you want to place inside the cylinder relative to the specified size. (e.g. try to approximate an O'Neill cylinder) I love generative AI models, but they're really bad at that, and this one is no exception, but the speed makes playing around with prompt variations to try to se…

I have a favourite test for LLMs that is also still surprisingly not passed by many: You walk up to a glass door. It has 'push' written on it in mirror writing. What should you do and why. Very few can get it right, even fewer can get it right and explain the right reason. They’ll start going on about how mirror writing is secret writing and push written backwards is code for pull, rather than just that it’s a messag…

* llama3.3:70b-instruct-q3_K_M *

A clever sign!

Since the word "push" is written in mirror writing, that means it's intended to be read from the other side of the door. In other words, if you were on the other side of the door, the text would appear normally and say "push".

Given this, I should... pull the door open!

The reasoning is that the sign is instructing people on the other side of the door to push it open, which implies that from my side, I need to pull it open.

Post reply on HN