Live data from Hacker News

Viewing profile — taesiri

taesiri

HN member
Joined
Fri, Apr 24, 2015, 3:04 PM UTC
HN karma
238
Public activity
26 items

About taesiri

No profile information was provided.

Recent public activity

  1. story
  2. comment
    Comment #47944338

    would be sick with meta glasses; just look at broken things, draws what you mean, and get help fixing it. not just fixing but anything

  3. story
  4. story
  5. story
  6. comment
    Comment #44172326

    for overly represented concepts, like popular brands, it seems that the model “ignores” the details once it detects that the overall shapes or patterns are similar. Opening up the …

  7. comment
    Comment #44169414

    State-of-the-art Vision Language Models achieve 100% accuracy counting on images of popular subjects (e.g. knowing that the Adidas logo has 3 stripes and a dog has 4 legs) but are …

  8. story
  9. story
  10. comment
    Comment #44073894

    tldr; We find that GenAI can satisfy 1/3 of everyday image editing requests, while 2/3 of the requests are better handled by human image editors.

  11. story
  12. story
  13. comment
    Comment #43290866

    Abstract: An Achilles heel of Large Language Models (LLMs) is their tendency to hallucinate non-factual statements. A response mixed of factual and non-factual statements poses a c…

  14. story
  15. comment
    Comment #43075634

    Abstract: Large Multimodal Models (LMMs) exhibit major shortfalls when interpreting images and, by some measures, have poorer spatial cognition than small children or animals. Desp…

  16. comment
    Comment #43075572

    All frontier models, (o1, o1-pro, QVQ, gemini-flash-thinking) score exactly 0% on main questions of this benchmark.

  17. story
  18. comment
    Comment #40926735

    This paper examines the limitations of current vision-based language models, such as GPT-4 and Sonnet 3.5, in performing low-level vision tasks. Despite their high scores on numero…

  19. story
  20. comment
    Comment #37106966

    Not one place, but there are some people tweeting about new papers daily (@arankomatsuzaki, @_akhaliq, @omarsar0) other people summarizing papers (@davisblalock, @rasbt). Latent Sp…

  21. comment
    Comment #37106054

    Coool! Would be nice to have an option to send commands to an LLM and show the results to the "user"! :D

  22. story
  23. story
  24. story
  25. story