Live data from Hacker News

Speech and Language Processing, 3rd ed. draft (2017)

web.stanford.edu

11–20 of 41 posts

Re: Speech and Language Processing, 3rd ed. draft (2017)

#12

Often this stuff is used for surveillance. If you choose to study this area, please be careful how you apply your knowledge. There are plenty of positive ways to use it: contributing content classification systems to sci-hub or libgen, building tools for the disabled, automating multilingual visual design aesthetics with computational linguistics and machine learning...

Inventing ships also invented shipwreks

Re: Speech and Language Processing, 3rd ed. draft (2017)

#14
post #8
post #6

Earlier quoted context omitted.

Great, these who want to build these things will simply hire people who don't subscribe to this oath ... And now the organization will be populated by a whole tribe of people who have no qualms about doing this -- does that sound like something that will make things better? Aside: the book looked interesting. Skimmed it, beautiful typesetting, not gratuitously mathematical, looked readable -- I may be wrong, but it's…

> Great, these who want to build these things will simply hire people who don't subscribe to this oath ... How many doctors are willing to give up their title to do unethical stuff? An ethical code will change people's perception, and that's a big thing. It will morph the perception of Google and Facebook from cool tech companies into dark places where no decent people want to work; unless of course these companies a…

Um, how are they supposed to make money if they don't track you to a certain extent? Do they have flaws? Yes -- it's not funny what kind of things you can target for based AdWords, their internal staff probably aren't creative enough to think up ways in which their tools can be used to inflict misery. Software engineers aren't soldiers, doctors, or civil servants what kind of ethical code do you expect? How will people be held accountable to it? What will happen to these who make transgressions?

Re: Speech and Language Processing, 3rd ed. draft (2017)

#16
Why link to the pdf?

The webpage https://web.stanford.edu/~jurafsky/slp3/ links directly to the PDF, and gives context and other download options. It's not so easy to go back from the PDF to the webpage.

Mods: I'd change the link, and the title to "...3rd Edition draft".

Everyone else: please stop linking to PDFs when there is an obvious html page to link to instead.

Re: Speech and Language Processing, 3rd ed. draft (2017)

#17
I need a voice activity detection module (VAD) for my wearable computer. Should I roll my own, or use someone else's (open-source). My immediate need is speaker-dependent (just me), but it would be nice if I could offer up a speaker-independent version eventually.

Re: Speech and Language Processing, 3rd ed. draft (2017)

#18
post #16

Why link to the pdf? The webpage https://web.stanford.edu/~jurafsky/slp3/ links directly to the PDF, and gives context and other download options. It's not so easy to go back from the PDF to the webpage. Mods: I'd change the link, and the title to "...3rd Edition draft". Everyone else: please stop linking to PDFs when there is an obvious html page to link to instead.

Good point, moreover it would be easy to read the simple html than wait for a minute to load the PDF

Re: Speech and Language Processing, 3rd ed. draft (2017)

#19
I found speech and language processing to be one of the most interesting courses of my degree. Recently I decided to take a look at speech synthesis again and discovered a book by Paul Taylor on this subject (http://svr-www.eng.cam.ac.uk/~pat40/ and draft PDF at http://svr-www.eng.cam.ac.uk/~pat40/ttsbook_draft_2.pdf). It is more engineering focused than other books in this area.

Re: Speech and Language Processing, 3rd ed. draft (2017)

#20
post #19

I found speech and language processing to be one of the most interesting courses of my degree. Recently I decided to take a look at speech synthesis again and discovered a book by Paul Taylor on this subject ( http://svr-www.eng.cam.ac.uk/~pat40/ and draft PDF at http://svr-www.eng.cam.ac.uk/~pat40/ttsbook_draft_2.pdf ). It is more engineering focused than other books in this area.

Note that the book was released 2009 and the draft is from 2007. While most of the content is still very relevant for understanding the basics and challenges in building TTS systems, recent progress in DNN-based synthesis (including WaveNet, GANs, and end-to-end approaches like Tacotron) is obviously not covered.
Post reply on HN