Viewing profile — visioninmyblood
visioninmyblood
HN member- Joined
- Tue, Nov 18, 2025, 11:09 PM UTC
- HN karma
- 27
- Public activity
- 79 items
- HN profile
- View on Hacker News ↗
About visioninmyblood
No profile information was provided.
Recent public activity
-
story
Show HN: Visual Agents with Code Mode
Rather than calling tools one by one, Orion 2 generates a program and runs it end to end, meaning fewer round-trips and lower latency. You can try it out at https://chat.vlm.run Wh…
-
comment
Comment #48115254
LLM-based agents handle text incredibly well, but images, videos, or PDFs with visual content are hard to interpret. mm-ctx gives your CLI agent multi-modal skills. Try it interact…
- story
- story
- story
-
comment
Comment #47692689
https://meta.ai/ this is where you can try it seems like the API is not publicly accessable yet. I feel they are very late to the game and do not show value to customers over other…
- story
- comment
- story
- story
-
comment
Comment #46861507
The roles seems more oriented towards visual agents. Is this is similar to google vision agent but with more capabilities?
- story
-
comment
Comment #46518060
Blog: https://vlm.run/blog/introducing-orion-artifacts Cookbook: https://github.com/vlm-run/vlmrun-cookbook Docs: https://docs.vlm.run/agents/artifacts
- story
-
comment
Comment #46517773
Blog: https://vlm.run/blog/introducing-orion-artifacts Cookbook: https://github.com/vlm-run/vlmrun-cookbook Docs: https://docs.vlm.run/agents/artifacts
- story
-
comment
Comment #46427160
Meta was lacking behind on the agents space. This is a good capture but they are making crazy good offers but not turning them into killer products so far. The AI agents space is p…
-
comment
Comment #46333338
was fun generating the as well
-
comment
Comment #46318859
you can chat with vlm.run to generate these assets from image generation without needing a gpu. Modi: https://chat.vlm.run/c/bdbaf1dc-b3c2-4b8a-ad17-6e26d87475fd Musk: https://chat…
-
comment
Comment #46177383
I tried this by using an gemini visual agent build with orion from vlm.run. it was able to produce two different images with five leg dog. you need to make it play with itself to i…
- story
- story
- story
-
comment
Comment #46130290
you want get the exact coordinated by running a key point network to pinpoint which coordinates does the next click point is you can. here I show a example simple prompt which retu…
-
comment
Comment #46128909
I agree claude and chatgpt and even gemini does a poor job in detecting and cropping into a region. Some of the simplest tasks, Qwen also is great at summerization but not into sol…