Viewing profile — taesiri
taesiri
HN member- Joined
- Fri, Apr 24, 2015, 3:04 PM UTC
- HN karma
- 238
- Public activity
- 26 items
- HN profile
- View on Hacker News ↗
About taesiri
No profile information was provided.
Recent public activity
- story
-
comment
Comment #47944338
would be sick with meta glasses; just look at broken things, draws what you mean, and get help fixing it. not just fixing but anything
- story
- story
- story
-
comment
Comment #44172326
for overly represented concepts, like popular brands, it seems that the model “ignores” the details once it detects that the overall shapes or patterns are similar. Opening up the …
-
comment
Comment #44169414
State-of-the-art Vision Language Models achieve 100% accuracy counting on images of popular subjects (e.g. knowing that the Adidas logo has 3 stripes and a dog has 4 legs) but are …
- story
- story
-
comment
Comment #44073894
tldr; We find that GenAI can satisfy 1/3 of everyday image editing requests, while 2/3 of the requests are better handled by human image editors.
- story
- story
-
comment
Comment #43290866
Abstract: An Achilles heel of Large Language Models (LLMs) is their tendency to hallucinate non-factual statements. A response mixed of factual and non-factual statements poses a c…
- story
-
comment
Comment #43075634
Abstract: Large Multimodal Models (LMMs) exhibit major shortfalls when interpreting images and, by some measures, have poorer spatial cognition than small children or animals. Desp…
-
comment
Comment #43075572
All frontier models, (o1, o1-pro, QVQ, gemini-flash-thinking) score exactly 0% on main questions of this benchmark.
- story
-
comment
Comment #40926735
This paper examines the limitations of current vision-based language models, such as GPT-4 and Sonnet 3.5, in performing low-level vision tasks. Despite their high scores on numero…
- story
-
comment
Comment #37106966
Not one place, but there are some people tweeting about new papers daily (@arankomatsuzaki, @_akhaliq, @omarsar0) other people summarizing papers (@davisblalock, @rasbt). Latent Sp…
-
comment
Comment #37106054
Coool! Would be nice to have an option to send commands to an LLM and show the results to the "user"! :D
- story
- story
- story
- story