Live data from Hacker News

Viewing profile — visioninmyblood

visioninmyblood

HN member
Joined
Tue, Nov 18, 2025, 11:09 PM UTC
HN karma
27
Public activity
79 items

About visioninmyblood

No profile information was provided.

Recent public activity

  1. story
    Show HN: Visual Agents with Code Mode

    Rather than calling tools one by one, Orion 2 generates a program and runs it end to end, meaning fewer round-trips and lower latency. You can try it out at https://chat.vlm.run Wh…

  2. comment
    Comment #48115254

    LLM-based agents handle text incredibly well, but images, videos, or PDFs with visual content are hard to interpret. mm-ctx gives your CLI agent multi-modal skills. Try it interact…

  3. story
  4. story
  5. story
  6. comment
    Comment #47692689

    https://meta.ai/ this is where you can try it seems like the API is not publicly accessable yet. I feel they are very late to the game and do not show value to customers over other…

  7. story
  8. comment
  9. story
  10. story
  11. comment
    Comment #46861507

    The roles seems more oriented towards visual agents. Is this is similar to google vision agent but with more capabilities?

  12. story
  13. comment
    Comment #46518060

    Blog: https://vlm.run/blog/introducing-orion-artifacts Cookbook: https://github.com/vlm-run/vlmrun-cookbook Docs: https://docs.vlm.run/agents/artifacts

  14. story
  15. comment
    Comment #46517773

    Blog: https://vlm.run/blog/introducing-orion-artifacts Cookbook: https://github.com/vlm-run/vlmrun-cookbook Docs: https://docs.vlm.run/agents/artifacts

  16. story
  17. comment
    Comment #46427160

    Meta was lacking behind on the agents space. This is a good capture but they are making crazy good offers but not turning them into killer products so far. The AI agents space is p…

  18. comment
    Comment #46333338

    was fun generating the as well

  19. comment
    Comment #46318859

    you can chat with vlm.run to generate these assets from image generation without needing a gpu. Modi: https://chat.vlm.run/c/bdbaf1dc-b3c2-4b8a-ad17-6e26d87475fd Musk: https://chat…

  20. comment
    Comment #46177383

    I tried this by using an gemini visual agent build with orion from vlm.run. it was able to produce two different images with five leg dog. you need to make it play with itself to i…

  21. story
  22. story
  23. story
  24. comment
    Comment #46130290

    you want get the exact coordinated by running a key point network to pinpoint which coordinates does the next click point is you can. here I show a example simple prompt which retu…

  25. comment
    Comment #46128909

    I agree claude and chatgpt and even gemini does a poor job in detecting and cropping into a region. Some of the simplest tasks, Qwen also is great at summerization but not into sol…