While this technology is fascinating, I can't help but feel a little unsettled reading that.
Google Cloud Vision API enters Beta
51–60 of 110 posts
Re: Google Cloud Vision API enters Beta
#52Earlier quoted context omitted.
There actually are other APIs, though with a smaller scope. For text extraction for example there is the OnDemand Api, https://dev.havenondemand.com/apis/ocrdocument#overview , backed by HP. They also have logo detection. I'd be surprised if no replacement for the category detection exists. Though I admit I also hesitate to replace that API with the google offering for the one app where I actually use it. The results…
There's also this API by Microsoft Research. https://www.projectoxford.ai/
If you need to understand emotional reaction on video sources, our API can fill in the gaps not currently filled by Google's Cloud Vision API: https://www.kairos.com/emotion-analysis-api
Disclosure: I'm CTO of Kairos.com
Re: Google Cloud Vision API enters Beta
#53I don't know for certain, but I suspect that Google utilized images from the web in training this system. Even if they didn't, suppose they had. I think this can raise an interesting question around copyright. In training an AI system with hundreds/thousands of bits of data, no single piece of training data makes much of a difference. If one of my images on the web that I had captioned with the keyword 'dog' was used…
By using this API, we're effectively training Google's system to be more and more accurate. Shouldn't Google pay us for using it? Just saying :-)
Re: Google Cloud Vision API enters Beta
#54I feel like this is some really compelling tech. It would be so amazing to build stuff with this in mind. I wouldn't be comfortable doing it, though. This sort of API is available only until Google decide that they don't want it to be available. There's not really anything close to equivalent that you could drop in to replace it if it were being shut down, the price were being hiked, or you had some sort of other iss…
Re: Google Cloud Vision API enters Beta
#55I don't know for certain, but I suspect that Google utilized images from the web in training this system. Even if they didn't, suppose they had. I think this can raise an interesting question around copyright. In training an AI system with hundreds/thousands of bits of data, no single piece of training data makes much of a difference. If one of my images on the web that I had captioned with the keyword 'dog' was used…
By using this API, we're effectively training Google's system to be more and more accurate. Shouldn't Google pay us for using it? Just saying :-)
Besides, it's a call without any feedback, so it's not that valuable as far as training goes.
Re: Google Cloud Vision API enters Beta
#56If the OCR is good then they're totally burying the lede, it's pricing is 100x cheaper than commercial OCR APIs. It's potentially a game changer, plenty of industries have piles of scanned documents. Cheap OCR means this data suddenly becomes accessible even if the value per individual document is low (i.e. for input into machine learning).
http://www.educatingsilicon.com/wp-content/uploads/2013/10/p...
A lot better for text in photographs. Comparison might be different on dense document text though.
Re: Google Cloud Vision API enters Beta
#57I feel like this is some really compelling tech. It would be so amazing to build stuff with this in mind. I wouldn't be comfortable doing it, though. This sort of API is available only until Google decide that they don't want it to be available. There's not really anything close to equivalent that you could drop in to replace it if it were being shut down, the price were being hiked, or you had some sort of other iss…
There is also http://www.deepdetect.com
Re: Google Cloud Vision API enters Beta
#58Re: Google Cloud Vision API enters Beta
#59Does something like this exist for sound? Any open source projects worth looking at?
Re: Google Cloud Vision API enters Beta
#60Does something like this exist for sound? Any open source projects worth looking at?