Earlier quoted context omitted.
As a (web app) developer I never quite sure what to put in alt. Figured you might have some advice here?
The question to ask is, what a sighted person learns after looking at the image? The answer is the alt text. E.g if the image is a floppy, maybe you communicate that this is the save button. If it shows a cat sleeping on the windowsill, the alt text is yep: "my cat looking cute while sleeping on the windowsill".
Apple Releases Open Weights Video Model
41–50 of 178 posts
Re: Apple Releases Open Weights Video Model
#42Looking at text to video examples ( https://starflow-v.github.io/#text-to-video ) I'm not impressed. Those gave me the feeling of the early Will Smith noodles videos. Did I miss anything?
Re: Apple Releases Open Weights Video Model
#43Re: Apple Releases Open Weights Video Model
#44Looking at text to video examples ( https://starflow-v.github.io/#text-to-video ) I'm not impressed. Those gave me the feeling of the early Will Smith noodles videos. Did I miss anything?
As far as I know, this might be the most advanced text-to-video model that has been released? I'm not sure whether the license will qualify as open enough in everyone's eyes, though.
Re: Apple Releases Open Weights Video Model
#45Earlier quoted context omitted.
I guess that auto-generated audio descriptions for (almost?) any video you want is a very, very nice feature for a blind person.
My two cents, this seems like a case where it’s better to wait for the person’s response instead of guessing.
Re: Apple Releases Open Weights Video Model
#46Earlier quoted context omitted.
The question to ask is, what a sighted person learns after looking at the image? The answer is the alt text. E.g if the image is a floppy, maybe you communicate that this is the save button. If it shows a cat sleeping on the windowsill, the alt text is yep: "my cat looking cute while sleeping on the windowsill".
I really like how you framed this as the takeaway or learning that needs to happen as what should be in the alt and not a recitation of the image. Where I've often had issues is more for things like business charts and illustrations and less cute cat photos.
Re: Apple Releases Open Weights Video Model
#47Earlier quoted context omitted.
The question to ask is, what a sighted person learns after looking at the image? The answer is the alt text. E.g if the image is a floppy, maybe you communicate that this is the save button. If it shows a cat sleeping on the windowsill, the alt text is yep: "my cat looking cute while sleeping on the windowsill".
I really like how you framed this as the takeaway or learning that needs to happen as what should be in the alt and not a recitation of the image. Where I've often had issues is more for things like business charts and illustrations and less cute cat photos.
Re: Apple Releases Open Weights Video Model
#48Re: Apple Releases Open Weights Video Model
#49Earlier quoted context omitted.
The question to ask is, what a sighted person learns after looking at the image? The answer is the alt text. E.g if the image is a floppy, maybe you communicate that this is the save button. If it shows a cat sleeping on the windowsill, the alt text is yep: "my cat looking cute while sleeping on the windowsill".
I really like how you framed this as the takeaway or learning that needs to happen as what should be in the alt and not a recitation of the image. Where I've often had issues is more for things like business charts and illustrations and less cute cat photos.
Re: Apple Releases Open Weights Video Model
#50Earlier quoted context omitted.
My two cents, this seems like a case where it’s better to wait for the person’s response instead of guessing.
My two cents, this seems like a comment it should be up to the OP to make instead of virtue signaling.
A question directed to GP, directly asking about their life and pointing this out is somehow virtue signalling, OK.