Moondream 3 Preview: Frontier-level reasoning at a blazing speed
1–10 of 46 posts
Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#2Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#3Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#4That’s actually kinda impressive for an 8b model. Normally my experience with them is that they’re not really useful.
Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#5I tried it out on their website and it seems pretty legit, it gets stuff wrong but so do all the vision models in my experience
Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#6Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#7One oddity is that I haven't seen the claimed improvements beyond the 2025-01-09 tag - subsequent releases improve recall but degrade precision pretty significantly. It'd be amazing if object detection VLMs like this reported class confidences to better address this issue. That said, having a dedicated object detection API is very nice and absent from other models/wrappers AFAIK.
Looking forward to Moondream 3 post-inference optimizations. Congrats to the team. The founder Vik is a great follow on X if that's your thing.
Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#8Earlier quoted context omitted.
Only 2b active also - very fast
Can run it on a phone then? Seems like it could be somewhat useful for people with poor eyesight or blindness
couple people got it running on a raspberry pi though
Re: Moondream 3 Preview: Frontier-level reasoning at a blazing speed
#9Moondream 2 has been very useful for me: I've been using it to automatically label object detection datasets for novel classes and distill an orders of magnitude smaller but similarly accurate CNN. One oddity is that I haven't seen the claimed improvements beyond the 2025-01-09 tag - subsequent releases improve recall but degrade precision pretty significantly. It'd be amazing if object detection VLMs like this repor…