Live data from Hacker News

Show HN: I turned my face rec system into a video codec

vertigo.ai

81–90 of 150 posts

Re: Show HN: I turned my face rec system into a video codec

#82
post #73

Tried it out on my phone, app crashes every time I press "go to room". I'm using iPhone SE from 2016 so maybe my phone is too old for the GPU features? Alas! Was looking forward to trying it. As a side note the UI looks like a toy or joke app. I'm not sure what market you're going for (it seems like a general purpose video chat app?) but you might want to reconsider the aesthetic.

Could you post the exact specs? I've tested on iPhone 6s until recently and it used to work, but it could be I am doing something silly.

Re: Show HN: I turned my face rec system into a video codec

#84
post #74
post #56

Earlier quoted context omitted.

> incorrect detail generation Absolutely. Reminds me of Xerox number mangling: https://www.dkriesel.com/en/blog/2013/0802_xerox-workcentres...

Or when AI resolution enhancement inserts Ryan Goslings face. https://petapixel.com/2020/08/17/gigapixel-ai-accidentally-a...

Also white Obama! https://www.theverge.com/21298762/face-depixelizer-ai-machin...

Re: Show HN: I turned my face rec system into a video codec

#85

Earlier quoted context omitted.

Hi, we don't collect any data from the app, and have filled in the privacy etc. statements on the App Store accordingly. Ideally I would like to collect faces to train the compression on, in which case we would have to consult with lawyers to come up with an EULA allowing us to do this. The advantage compared to using broadly available datasets to train on would be more realistic shot noise, low light images, and so…

> We've been sitting on the face recognition tech since 2018, so if we'd wanted to become Clearview.ai we probably would have a long time ago. It says right at the beginning of the post you were doing quite well until the pandemic shut down businesses. I try not to be overly critical (I really do) but this is one of those cases I just can’t help myself, I see no reason individual businesses should be running facial r…

There is a difference between doing well and becoming China in terms of surveillance. Most of our revenue were from just plain head counting and from tracking cars in a smart city project.

The one face rec system we actually sold was used to measure waiting times in a retail setting, and replaced an existing system that was using AWS Rekognition in the same manner, except with all the video footage going to the cloud. That license has long expired and the system is no longer running.

In any case, what is stated in the app's privacy statement is what we are doing. At the moment I don't even collect the IP addresses of users connecting.

Re: Show HN: I turned my face rec system into a video codec

#86

> Because it uses face rec, you can ONLY show your face, and if you disappear from view your audio will stop after a while, to avoid situations like when you need to go the the restroom but forget to mute. This also solves dick-pics etc. Who would have thought anti dick pic technology would become a product feature but here we are in 2022.

Stock ticker: NOTHOTDOG

Re: Show HN: I turned my face rec system into a video codec

#87
post #83

You lost me at "patent pending". This idea has been obvious and in progress for a while now with lots of prior work. The issue is more the standards. Please don't sell this to a patent troll who will harass the industry for 20 years.

Nvidia already showed this last year I think.

So whatever he is trying to patent, big companies already patent something.

Re: Show HN: I turned my face rec system into a video codec

#88

WOW. This is amazing. I really believe your project can be game changing for the video-call industry. Have you considered entering into the YC Program? I think it could be an awesome match. There are many startups I know they may want to take use of your service, and even fly.io is part of YC family! Also, have you thought about open-sourcing it? (perhaps using a dual license could work great for an enterprise offeri…

Thanks! I tried entering YC in the fall 2021 batch, and got to the top 10%. I believe my main problem wrt YC is that I currently lack a co-founder, so I did not apply in the Spring as this was still the case. I am seriously thinking about open source, I believe for instance WebRTC found a good model with dual-licensing, where you have something like AGPL with the option of buying exceptions. I have had multiple advis…

I entered YC as a solo-founder with Wasmer, so I think it might just be a circumstantial thing (they receive a lot of applications so it's always hard to judge who should enter with the limited time and data they have to make a decision). I would really encourage you to apply again!

In any case I'd love to help you on both aspects (YC application and OSS), I believe your idea has really great potential. Please ping me to syrus@wasmer.io and we can schedule some time to chat!

Re: Show HN: I turned my face rec system into a video codec

#89

> is a new from-the-ground-up patent pending What is the invention? The models are just complex mathematical formulas and these cannot be patented.

The patent that I filed is not around the models, but in how it boosts the parts of the face most relevant to face-to-face conversations. I am not a super fan of patents, but for background please consider that Asger Jensen and I could have patented VM live migration in 2002 and chose not to for idealistic reasons, just to see VMware do it.

Would be interesting if a foundation with a charter could own a a patent, to prevent later trolling.

Re: Show HN: I turned my face rec system into a video codec

#90
post #48

> Because it uses face rec, you can ONLY show your face, and if you disappear from view your audio will stop after a while, to avoid situations like when you need to go the the restroom but forget to mute. Of course, the real killer app for Zoom calls is the opposite of this: some kind of deep fakery that makes it seem we're there when we're not. Yet as it is, this is a fantastic idea. It's surprising video codecs de…

Well if you just want to trick people into thinking you’re present when you aren’t, you can just use a video file as your camera feed. There are a bunch of tools to do this. Now, the main issue would be if you want to speak using this setup, but I guess one solution to that would be to use a model to lip sync your non-speaking video on-the-fly, which seems to be being discussed here[0].

[0] https://github.com/Rudrabha/Wav2Lip/issues/358

Post reply on HN