Viewing profile — spindump8930
spindump8930
HN member- Joined
- Thu, Mar 07, 2024, 7:39 PM UTC
- HN karma
- 134
- Public activity
- 50 items
- HN profile
- View on Hacker News ↗
About spindump8930
No profile information was provided.
Recent public activity
-
comment
Comment #48755413
I agree with your recomendation, but converting a pdf to an image is by no means smaller. PDFs are much closer to SVGs then to jpegs.
-
comment
Comment #48667623
> Claude and ChatGPT are both blocked in China So it's presumably cheaper than attempting to spin up your own method of circumventing the blocks.
-
comment
Comment #48416103
[dead]
-
comment
Comment #48415848
Exactly. Good peer reviewers understand that you can also move down on the scaling curve, not just up. Also laughable to try a "yolo" run without validating a scaling ladder/curve.…
-
comment
Comment #48415830
Can you share the specific part of this work that demonstrates better scaling than original transformers? Also note that many of the changes to that architecture, that have been pr…
-
comment
Comment #48415795
That's why you do several small and medium scale tests, fit a curve, and ideally show that the trend persists at several scales. Not a single large or medium run - see the other co…
-
comment
Comment #48356527
I think folks looking for more on this incident are better off reading the original threads linked elsewhere in the comments. This blog doesn't seem to add any information and is i…
-
comment
Comment #48136880
Likely in this case the time vault was the collapse of Mt Gox, which has now recently been paying back holders.
-
comment
Comment #48068454
Some combination of reporting bias given concerns about LLM security capabilities and actual new vulnerabilities found with LLM assistance. Even if exploits and outages are unrelat…
-
comment
Comment #48067495
It's very common if you improperly seed, as others in the thread brought up! Or in your framing, as rare as earth getting hit if it were surrounded by a sci-fi density asteroid fie…
-
comment
Comment #47978421
Sure, this is cute and interesting, but there's no validation or baselines and those examples are not particularly compelling. The o3 example just lists some terms!
-
comment
Comment #47948233
Between the neo and the chances for privacy respecting local model inference, all the new apple hardware has me excited.
-
comment
Comment #47948184
That artificial analysis page has some great references for this, thanks for sharing.
-
comment
Comment #47940085
Remember that models on different inference platforms might not necessarily give exactly the same results, adding another axis of non-determinism to development. Things like quanti…
-
comment
Comment #47939938
Any more context on the copilot training note? More pointers would be very interesting, but we'd need to keep in mind how many different underlying models were (are?) branded as co…
-
comment
Comment #47894542
> The researchers tested five LLMs: OpenAI’s GPT-4o (before the highly sycophantic and since-sunset GPT-5) Interesting, I always thought the sycophancy peaked with 4o and the assoc…
-
comment
Comment #47894489
Hopefully this money means more compute infrastructure to help Anthropic counter the efficiency changes that have created this perceived downtrend in claude quality.
-
comment
Comment #47894459
Having known some folks who did recurse, I think places like this want to select for those who consider coding a type of craft or art or self-expression. You can use LLMs, but stan…
-
comment
Comment #47783580
Not clear that they even have any GPUs yet: > Allbirds, which will be renamed “NewBird AI,” said it executed a $50 million deal with an unnamed institutional investor to acquire “h…
-
comment
Comment #47783543
Yes, the paper itself tells a different story than the bullet points in this article.
-
comment
Comment #47783530
The article seems quite editorialized, shifting between describing "large-scale AI models" and "neural network-based approaches". The underlying paper itself is more precise, compa…
-
comment
Comment #47694397
Yes, it's far more certain that meta released this, which is less convincing on evals, as a result of the mythos previews.
-
comment
Comment #47694383
Re: changes, there's been enormous turnover in AI organizations, and in theory this one was developed by a "new" org. Whether that means less or more benchmaxxing is anyone's guess…
-
comment
Comment #47694362
Spending tons of money on Claude and the recent token benchmarks came WELL after Meta's huge investments in compute infrastructure for AI as well as the long history of language mo…
-
comment
Comment #47660603
Only for poor quality systems. Unfortunately there are many systems that tried to make easy hype, but are the equivalent of an ML 101 classifier class project. If one measures for …