Live data from Hacker News

Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

github.com

241–250 of 266 posts

Re: Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

#241
post #126

Earlier quoted context omitted.

Are there any LLMs being widely used for audio classification? I know VLMs are being used a lot in image stuff. It always seems kind of silly to me to throw everything at an LLM. I know they’re huge and can automatically handle a huge number of tasks but something in me finds it wasteful when we could be creating easily trainable, cheap to run bespoke models for a lot of stuff

The harness that connects to a chatbot, API or voice interaction is the place to route requests to different systems. If you remember the early days of ChatGPT it explicitly said it was routing image generation to Dall-E after embellishing your request itself first. Determining which tool to use should be a lightweight operation but I’m not expert enough to understand exactly how much lighter than a full LLM call jus…

For sure, I just know it’s tempting given the power of large transformers to throw things at an existing model.

For instance, OCR is something that can be done locally with no access to a GPU but people (including me) still often use cloud hosted multi-modal large language models for it.

Re: Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

#243

The underlying classifier, BirdNET, is a traditional neural network and not an LLM: https://doi.org/10.1016/j.ecoinf.2021.101236

Why would anyone assume this uses an LLM? It classifies bird sounds, not human language. I don't mean this as an attack, I'm genuinely curious! This seemed obvious to me, and I want to know what line of thinking might lead one to believe that an LLM is the better (or more likely) tool for this job over a purpose-built classifier.

Re: Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

#244
post #214

Earlier quoted context omitted.

They are used in book readers. Amazon Kindle Colorsoft, for example. Or I have android phone with such screen (though, not latest generation). Rather opposite: I never seen consumer products with two-layer displays.

Nope, that's my point, on your kindle, the screen is a Kaleido 3, which is a black&white+lcd layer: https://www.reddit.com/r/eink/comments/1i8e26x/seriously_how... Vs Spectra (the expensive screen): https://www.eink.com/upload/2023_11_14/3_20231114082849vy2as...

I don't see LCD layer here. Passive color filters, B/W capsules (which reflects or doesn't reflect incoming light, depends on orientation), TFT (thin layer transistors, not LCD — liquid crystal display!) as control to orient capsules.

Yes, it is not 4 color capsules as in Spectra, but where is LCD?

LCD what is caught my attention (and surprise) in your commentary.

But now I see difference, thank you.

Re: Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

#247

I think this is a cool idea but I would never want ai generated art on my wall.

Author here: it's not AI-generated. All art is public domain and made by real people. This project doesn't run generative AI models.

See https://github.com/arnegiacomo/fugleramme/blob/main/assets/a... for all the sources

Re: Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

#249

Earlier quoted context omitted.

It's been around in slang for a few years with a few variants, e.g. "I haven't seen her in a hot minute." I think it's just filtered into wider cultural vernacular.

people have been saying "it's been a minute" to mean a long time since at least the 90s in NY

Yeah, I’ve been saying hot minutes for decades myself!
Post reply on HN