Live data from Hacker News

Pink Trombone

dood.al

51–60 of 80 posts

Re: Pink Trombone

#52

Earlier quoted context omitted.

I'm guessing the thing from '96 was a Java applet, so indeed an app.

or a desktop/windows app. Back then we'd be more likely to call them programs.

And also sometimes referred to as applications. And while the abbreviation “app” only became mainstream around the time of the first iPhone, the warez scene had a habit of referring to software applications as “appz” since way before then, which admittedly is not the exact same word as “app”, appz being an intentional misspelling of a pluralization of the abbreviation, but it’s close to it. They sure loved the letter z. Appz, crackz, mp3z, moviez, gamez, ebookz.

This way of writing gave people some unique words to query search engines for, making it easier to find warez sites, torrent trackers and ftp servers hosting pirated files but I think it orginated long before the web was born — even before any sort of networked computing existed.

Consider the word “phreaking” which was invented at the time of mechanical telephone switching systems. This word came from combining the words “phone” and “freaking”. I think that word could have inspired hackers to use use ph in place of f in other words, and then once you start making that substitution, other substitutions follow, like using z instead of s.

I dunno, I grew up in the 90s so there is a lot of hacker culture that precedes my time. What I do know is that a lot of early hacker culture spawned other subcultures, and that several influences of the origins remain central in these. For example, the demoscene.

Re: Pink Trombone

#53

I've actually been looking for the opposite of this (i.e. sound in, mouth representation out) for a while. Does anyone know of such a thing?

Yes! Oculus makes an SDK for this. You can use it in Unity 3D, Unreal, or directly in a native app. https://developer.oculus.com/documentation/audiosdk/latest/c...

Re: Pink Trombone

#54
post #53

I've actually been looking for the opposite of this (i.e. sound in, mouth representation out) for a while. Does anyone know of such a thing?

Yes! Oculus makes an SDK for this. You can use it in Unity 3D, Unreal, or directly in a native app. https://developer.oculus.com/documentation/audiosdk/latest/c...

Thanks! That's like 80% of the way there. It looks to be missing a lot of state internal to the mouth (understandable given that it's targeting avatar lipsyncing), and appears to discretize the values somewhat, making it less useful for linguistics practice. But I bet the underlying technology could be adapted easily.

Re: Pink Trombone

#55
post #5

This reminds me of the Sprechmaschine [1] ("speaking machine") built at the end of the 18th century by Wolfgang von Kempelen (the guy who build the original mechanical turk [2]). Here is a YouTube video showing it in action (for example, the machine says "Mama" around 1:14): https://www.youtube.com/watch?v=k_YUB_S6Gpo [1] https://de.wikipedia.org/wiki/Wolfgang_von_Kempelen#Die_Spre... [2] https://en.wikipedia.org/wik…

also, the 1939 voder comes to mind :

https://en.wikipedia.org/wiki/Voder

https://www.youtube.com/watch?v=0rAyrmm7vv0

Re: Pink Trombone

#56
This would be useful to demonstrate the difference between p/f and l/r for those brought up without those distinctions.

I'd also (as an English speaker) like to see/hear Dutch g and Xhosan clicks.

Re: Pink Trombone

#57
Reminds me of Xiph's Speex/CELP model of speech as a mix of noise and frequency to achieve high compression, requiring as little as 2.15 kilobits (275 bytes) per second. It sounds perceptibly similar to the original recording, even though the difference between the input and output sampled data may be high:

https://www.speex.org/docs/manual/speex-manual/node9.html

Bitrate comparison:

https://www.speex.org/comparison/

Samples:

https://www.speex.org/samples/

Maybe higher compression can be achieved with better prediction, aka machine learning.

Re: Pink Trombone

#58
post #8

In a similar vein: http://www.adultswim.com/etcetera/choir/ The most interesting thing about this one is the chord progressions it generates.

In a historical vein, the absolutely magnificent Voder, from the late 1930s: https://youtu.be/TsdOej_nC1M?t=16 https://en.wikipedia.org/wiki/Voder

I found it difficult to tell how well it's actually speaking because the announcer is priming the audience for every utterance. Very cool though!
Post reply on HN