Serverless development experience for embedded computer vision
1–9 of 9 posts
Re: Serverless development experience for embedded computer vision
#2Re: Serverless development experience for embedded computer vision
#3It is so cool you shared this repo. I am the developer behind it, hope you enjoy it and can provide some valuable feedback!
Re: Serverless development experience for embedded computer vision
#4There are a few serverless GPU companies like Banana.dev and Modal; I really want to give them a shot. Anyone have experience using them in prod?
Re: Serverless development experience for embedded computer vision
#5I found this OSS tool interesting since it abstracts pretty much all the plumbing required to set up a computer vision pipeline.
Re: Serverless development experience for embedded computer vision
#6I found this OSS tool interesting since it abstracts pretty much all the plumbing required to set up a computer vision pipeline.
This looks like a really cool project; would you be open to us PR'ing support for the 50k fine-tuned models on Roboflow Universe[1] via an `inference`[2] integration? [1] https://roboflow.com/universe [2] https://github.com/roboflow/inference
Re: Serverless development experience for embedded computer vision
#7Getting started in AI + vision with just 3 commands is amazing! I will definitely try it for some personal projects with IP cameras.
Good stuff :)
Re: Serverless development experience for embedded computer vision
#8Pretty neat! We've been using Lambda for ML serving low-volume CV models (and my understanding is AWS' Sagemaker Serverless is a lambda wrapper) for a couple of years at Roboflow and it is really good for low-volume and bursty use-cases. The latency is surprisingly not bad. It gets really expensive relative to GPUs for high load (and especially predictable high-load like monitoring security cameras 24/7) though so we…
Re: Serverless development experience for embedded computer vision
#9Pretty neat! We've been using Lambda for ML serving low-volume CV models (and my understanding is AWS' Sagemaker Serverless is a lambda wrapper) for a couple of years at Roboflow and it is really good for low-volume and bursty use-cases. The latency is surprisingly not bad. It gets really expensive relative to GPUs for high load (and especially predictable high-load like monitoring security cameras 24/7) though so we…
We've been building with Modal over the past few months (though no prod-scale tests yet) and were slightly disappointed by very large (10-20 second) cold start times. In the long term we're more interested in inference servers that use compiled/optimized models instead of running plain old PyTorch (which adds another few seconds to cold start on its own).